diegosouzapw-advanced-evaluation

SKILLFlusso di lavorocommunity
v0.0.0diegosouzapwMITAggiornato 2 mesi faFonte →

Advanced Evaluation workflow skill. Use this skill when the user needs This skill should be used when the user asks to \"implement LLM-as-judge\", \"compare model outputs\", \"create evaluation rubrics\", \"mitigate evaluation bias\", or mentions direct scoring, pairwise comparison, position bias, e

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
116Stelle del repo
1Client
1Formati
2 mesi faUltimo aggiornamento
Skill
Autorediegosouzapw
Versione0.0.0
LicenzaMIT
CategoriaFlusso di lavoro
Formatiskill.md
PromptNon pubblicato
Compatibilità
Claude✓ Supportato
Cursor—
Copilot—
ChatGPT—
Gemini—
Descrizione

Advanced Evaluation workflow skill. Use this skill when the user needs This skill should be used when the user asks to \"implement LLM-as-judge\", \"compare model outputs\", \"create evaluation rubrics\", \"mitigate evaluation bias\", or mentions direct scoring, pairwise comparison, position bias, evaluation pipelines, or automated quality assessment and the operator should preserve the upstream w

Parole chiave
skillclaude

Nessuna copertura delle dipendenze

Questa voce non pubblica alcun pacchetto npm, quindi Forge non ha un albero delle dipendenze per essa. È una lacuna di copertura, non l'affermazione che non abbia dipendenze.