metric-design

SKILLFlusso di lavorocommunity
v0.0.0agentscope-aiApache-2.0Aggiornato 17 g faFonte →

Use when the user has evaluation principles or a dataset but needs help choosing the right graders, designing evaluation metrics, creating LLM-as-judge prompts, combining multiple metrics into a composite score, or building an automated evaluation pipeline. Also use when the user mentions grader sel

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
789Stelle del repo
1Client
1Formati
17 g faUltimo aggiornamento
Skill
Autoreagentscope-ai
Versione0.0.0
LicenzaApache-2.0
CategoriaFlusso di lavoro
Formatiskill.md
PromptApri (vedi la scheda Prompt)
Compatibilità
Claude✓ Supportato
Cursor
Copilot
ChatGPT
Gemini
Descrizione

Use when the user has evaluation principles or a dataset but needs help choosing the right graders, designing evaluation metrics, creating LLM-as-judge prompts, combining multiple metrics into a composite score, or building an automated evaluation pipeline. Also use when the user mentions grader selection, metric design, judge prompt engineering, rubric design, evaluation pipeline code, or "how to

Parole chiave
skillclaude