metric-design

SKILLFlujo de trabajocomunidad
v0.0.0agentscope-aiApache-2.0Actualizado hace 17 dFuente →

Use when the user has evaluation principles or a dataset but needs help choosing the right graders, designing evaluation metrics, creating LLM-as-judge prompts, combining multiple metrics into a composite score, or building an automated evaluation pipeline. Also use when the user mentions grader sel

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
789Estrellas del repo
1Clientes
1Formatos
hace 17 dÚltima actualización
Skill
Autoragentscope-ai
Versión0.0.0
LicenciaApache-2.0
CategoríaFlujo de trabajo
Formatosskill.md
PromptAbrir (ver la pestaña Prompt)
Compatibilidad
Claude✓ Compatible
Cursor
Copilot
ChatGPT
Gemini
Acerca de

Use when the user has evaluation principles or a dataset but needs help choosing the right graders, designing evaluation metrics, creating LLM-as-judge prompts, combining multiple metrics into a composite score, or building an automated evaluation pipeline. Also use when the user mentions grader selection, metric design, judge prompt engineering, rubric design, evaluation pipeline code, or "how to

Palabras clave
skillclaude