plugin-evaluation

SKILLFlusso di lavorocommunity
v0.0.0viktorbezdekMITAggiornato 2 mesi faFonte →

Measures whether a Claude Code plugin actually works by running triggering evals (does the model pick the skill?) and output evals (does it produce correct results?). Use when you need to evaluate a plugin, run skill activation testing, set up an eval harness, measure plugin quality, write trigger r

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
10Stelle del repo
1Client
1Formati
2 mesi faUltimo aggiornamento
Skill
Autoreviktorbezdek
Versione0.0.0
LicenzaMIT
CategoriaFlusso di lavoro
Formatiskill.md
PromptNon pubblicato
Compatibilità
Claude✓ Supportato
Cursor
Copilot
ChatGPT
Gemini
Descrizione

Measures whether a Claude Code plugin actually works by running triggering evals (does the model pick the skill?) and output evals (does it produce correct results?). Use when you need to evaluate a plugin, run skill activation testing, set up an eval harness, measure plugin quality, write trigger rate tests, check output quality, compare plugin iterations, or iterate on a SKILL description based

Parole chiave
skillclaude