guanyang-advanced-evaluation

SKILLWorkflowcommunity
v0.0.0guanyangMITUpdated 1d agoSource →

This skill should be used for advanced LLM evaluation: LLM-as-judge systems, direct scoring, pairwise comparison, rubric calibration, evaluator bias mitigation, confidence scoring, and automated quality assessment.

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
948Repo stars
1Clients
1Formats
1d agoLast update
Skill
Authorguanyang
Version0.0.0
LicenseMIT
CategoryWorkflow
Formatsskill.md
PromptOpen (see Prompt tab)
Compatibility
Claude✓ Supported
Cursor
Copilot
ChatGPT
Gemini
About

This skill should be used for advanced LLM evaluation: LLM-as-judge systems, direct scoring, pairwise comparison, rubric calibration, evaluator bias mitigation, confidence scoring, and automated quality assessment.

Keywords
skillclaude