metric-design

SKILLWorkflowCommunity
v0.0.0agentscope-aiApache-2.0Aktualisiert vor 18 TQuelle →

Use when the user has evaluation principles or a dataset but needs help choosing the right graders, designing evaluation metrics, creating LLM-as-judge prompts, combining multiple metrics into a composite score, or building an automated evaluation pipeline. Also use when the user mentions grader sel

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
789Repo-Sterne
1Clients
1Formate
vor 18 TLetzte Aktualisierung
Skill
Autoragentscope-ai
Version0.0.0
LizenzApache-2.0
KategorieWorkflow
Formateskill.md
PromptÖffnen (siehe Tab „Prompt“)
Kompatibilität
Claude✓ Unterstützt
Cursor
Copilot
ChatGPT
Gemini
Über

Use when the user has evaluation principles or a dataset but needs help choosing the right graders, designing evaluation metrics, creating LLM-as-judge prompts, combining multiple metrics into a composite score, or building an automated evaluation pipeline. Also use when the user mentions grader selection, metric design, judge prompt engineering, rubric design, evaluation pipeline code, or "how to

Schlagwörter
skillclaude