ml-model-eval-benchmark

SKILLWorkflowcommunity
v0.0.00x-ProfessorApache-2.0Updated 5mo agoSource →

Compare model candidates using weighted metrics and deterministic ranking outputs. Use for benchmark leaderboards and model promotion decisions.

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
10Repo stars
1Clients
1Formats
5mo agoLast update
Skill
Author0x-Professor
Version0.0.0
LicenseApache-2.0
CategoryWorkflow
Formatsskill.md
PromptNot published
Compatibility
Claude✓ Supported
Cursor
Copilot
ChatGPT
Gemini
About

Compare model candidates using weighted metrics and deterministic ranking outputs. Use for benchmark leaderboards and model promotion decisions.

Keywords
skillclaude