agentic-evaluation-framework

SKILLFlusso di lavorocommunity
v0.0.0borgheiNOASSERTIONAggiornato 8 g faFonte →

This skill should be used when the user asks to "evaluate LLM output quality", "set up LLM-as-judge", "build an eval rubric", "compare model outputs pairwise", or "measure agent quality".

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
481Stelle del repo
1Client
1Formati
8 g faUltimo aggiornamento
Skill
Autoreborghei
Versione0.0.0
LicenzaNOASSERTION
CategoriaFlusso di lavoro
Formatiskill.md
PromptNon pubblicato
Compatibilità
Claude✓ Supportato
Cursor
Copilot
ChatGPT
Gemini
Descrizione

This skill should be used when the user asks to "evaluate LLM output quality", "set up LLM-as-judge", "build an eval rubric", "compare model outputs pairwise", or "measure agent quality".

Parole chiave
skillclaude