add-llm-evals

SKILLWorkflowcommunauté
v0.0.0ContextJet-aiNOASSERTIONMis à jour il y a 11 jSource →

Use this when adding evaluation to an LLM/agent app - measuring output quality (correctness, faithfulness, relevance, safety) rather than just watching traces. Trigger on "add evals", "test my prompt", "is my RAG accurate", "catch regressions", "score outputs", or setting up an eval suite in CI. Cov

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
29Étoiles du dépôt
1Clients
1Formats
il y a 11 jDernière mise à jour
Skill
AuteurContextJet-ai
Version0.0.0
LicenceNOASSERTION
CatégorieWorkflow
Formatsskill.md
PromptNon publié
Compatibilité
Claude✓ Pris en charge
Cursor
Copilot
ChatGPT
Gemini
À propos

Use this when adding evaluation to an LLM/agent app - measuring output quality (correctness, faithfulness, relevance, safety) rather than just watching traces. Trigger on "add evals", "test my prompt", "is my RAG accurate", "catch regressions", "score outputs", or setting up an eval suite in CI. Covers offline (CI) and online (production LLM-as-a-judge) evaluation.

Mots-clés
skillclaude