iblai-api-agent-eval

SKILLWorkflowcommunauté
v0.0.0iblaiMITMis à jour il y a 10 jSource →

Measure and improve agent quality via the platform API — evaluation datasets, dataset items (JSON, CSV upload, or from chat traces), experiment runs, LLM-as-Judge and human-annotation scoring, score configs, and CSV export. Use to test an agent against a dataset and grade the results.

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
15Étoiles du dépôt
1Clients
1Formats
il y a 10 jDernière mise à jour
Skill
Auteuriblai
Version0.0.0
LicenceMIT
CatégorieWorkflow
Formatsskill.md
PromptNon publié
Compatibilité
Claude✓ Pris en charge
Cursor
Copilot
ChatGPT
Gemini
À propos

Measure and improve agent quality via the platform API — evaluation datasets, dataset items (JSON, CSV upload, or from chat traces), experiment runs, LLM-as-Judge and human-annotation scoring, score configs, and CSV export. Use to test an agent against a dataset and grade the results.

Mots-clés
skillclaude