diegosouzapw-advanced-evaluation

SKILLFlusso di lavorocommunity
v0.0.0diegosouzapwMITAggiornato 1 mesi faFonte →

Advanced Evaluation workflow skill. Use this skill when the user needs This skill should be used when the user asks to \"implement LLM-as-judge\", \"compare model outputs\", \"create evaluation rubrics\", \"mitigate evaluation bias\", or mentions direct scoring, pairwise comparison, position bias, e

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
116Stelle del repo
1Client
1Formati
1 mesi faUltimo aggiornamento
Skill
Autorediegosouzapw
Versione0.0.0
LicenzaMIT
CategoriaFlusso di lavoro
Formatiskill.md
PromptNon pubblicato
Compatibilità
Claude✓ Supportato
Cursor
Copilot
ChatGPT
Gemini
Descrizione

Advanced Evaluation workflow skill. Use this skill when the user needs This skill should be used when the user asks to \"implement LLM-as-judge\", \"compare model outputs\", \"create evaluation rubrics\", \"mitigate evaluation bias\", or mentions direct scoring, pairwise comparison, position bias, evaluation pipelines, or automated quality assessment and the operator should preserve the upstream w

Parole chiave
skillclaude