openjudge

SKILLWorkflowcommunauté
v0.0.0agentscope-aiApache-2.0Mis à jour il y a 18 jSource →

Build custom LLM evaluation pipelines using the OpenJudge framework. Covers selecting and configuring graders (LLM-based, function-based, agentic), running batch evaluations with GradingRunner, combining scores with aggregators, applying evaluation strategies (voting, average), auto-generating grade

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
789Étoiles du dépôt
1Clients
1Formats
il y a 18 jDernière mise à jour
Skill
Auteuragentscope-ai
Version0.0.0
LicenceApache-2.0
CatégorieWorkflow
Formatsskill.md
PromptOuvrir (voir l’onglet Prompt)
Compatibilité
Claude✓ Pris en charge
Cursor
Copilot
ChatGPT
Gemini
À propos

Build custom LLM evaluation pipelines using the OpenJudge framework. Covers selecting and configuring graders (LLM-based, function-based, agentic), running batch evaluations with GradingRunner, combining scores with aggregators, applying evaluation strategies (voting, average), auto-generating graders from data, and analyzing results (pairwise win rates, statistics, validation metrics). Use when t

Mots-clés
skillclaude