plugin-evaluation

SKILLFlujo de trabajocomunidad
v0.0.0viktorbezdekMITActualizado hace 2 mFuente →

Measures whether a Claude Code plugin actually works by running triggering evals (does the model pick the skill?) and output evals (does it produce correct results?). Use when you need to evaluate a plugin, run skill activation testing, set up an eval harness, measure plugin quality, write trigger r

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
10Estrellas del repo
1Clientes
1Formatos
hace 2 mÚltima actualización
Skill
Autorviktorbezdek
Versión0.0.0
LicenciaMIT
CategoríaFlujo de trabajo
Formatosskill.md
PromptNo publicado
Compatibilidad
Claude✓ Compatible
Cursor
Copilot
ChatGPT
Gemini
Acerca de

Measures whether a Claude Code plugin actually works by running triggering evals (does the model pick the skill?) and output evals (does it produce correct results?). Use when you need to evaluate a plugin, run skill activation testing, set up an eval harness, measure plugin quality, write trigger rate tests, check output quality, compare plugin iterations, or iterate on a SKILL description based

Palabras clave
skillclaude