add-llm-evals

SKILLWorkflowCommunity
v0.0.0ContextJet-aiNOASSERTIONAktualisiert vor 11 TQuelle →

Use this when adding evaluation to an LLM/agent app - measuring output quality (correctness, faithfulness, relevance, safety) rather than just watching traces. Trigger on "add evals", "test my prompt", "is my RAG accurate", "catch regressions", "score outputs", or setting up an eval suite in CI. Cov

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
29Repo-Sterne
1Clients
1Formate
vor 11 TLetzte Aktualisierung
Skill
AutorContextJet-ai
Version0.0.0
LizenzNOASSERTION
KategorieWorkflow
Formateskill.md
PromptNicht veröffentlicht
Kompatibilität
Claude✓ Unterstützt
Cursor
Copilot
ChatGPT
Gemini
Über

Use this when adding evaluation to an LLM/agent app - measuring output quality (correctness, faithfulness, relevance, safety) rather than just watching traces. Trigger on "add evals", "test my prompt", "is my RAG accurate", "catch regressions", "score outputs", or setting up an eval suite in CI. Covers offline (CI) and online (production LLM-as-a-judge) evaluation.

Schlagwörter
skillclaude