build-eval-dataset

SKILLWorkflowcommunauté
v0.0.0ContextJet-aiNOASSERTIONMis à jour il y a 11 jSource →

Use this to build a good evaluation dataset for an LLM app, the part everyone underestimates. Trigger on "make an eval set", "what should I test my LLM on", "I don't have test data for my prompt", "build a golden dataset", or before setting up evals. A great eval set beats a great metric; garbage-in

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
29Étoiles du dépôt
1Clients
1Formats
il y a 11 jDernière mise à jour
Skill
AuteurContextJet-ai
Version0.0.0
LicenceNOASSERTION
CatégorieWorkflow
Formatsskill.md
PromptNon publié
Compatibilité
Claude✓ Pris en charge
Cursor
Copilot
ChatGPT
Gemini
À propos

Use this to build a good evaluation dataset for an LLM app, the part everyone underestimates. Trigger on "make an eval set", "what should I test my LLM on", "I don't have test data for my prompt", "build a golden dataset", or before setting up evals. A great eval set beats a great metric; garbage-in means your evals lie to you.

Mots-clés
skillclaude