add-llm-evals

SKILLWorkflowcommunity
v0.0.0ContextJet-aiNOASSERTIONUpdated 6d agoSource →

Use this when adding evaluation to an LLM/agent app - measuring output quality (correctness, faithfulness, relevance, safety) rather than just watching traces. Trigger on "add evals", "test my prompt", "is my RAG accurate", "catch regressions", "score outputs", or setting up an eval suite in CI. Cov

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
29Repo stars
1Clients
1Formats
6d agoLast update
Skill
AuthorContextJet-ai
Version0.0.0
LicenseNOASSERTION
CategoryWorkflow
Formatsskill.md
PromptNot published
Compatibility
Claude✓ Supported
Cursor
Copilot
ChatGPT
Gemini
About

Use this when adding evaluation to an LLM/agent app - measuring output quality (correctness, faithfulness, relevance, safety) rather than just watching traces. Trigger on "add evals", "test my prompt", "is my RAG accurate", "catch regressions", "score outputs", or setting up an eval suite in CI. Covers offline (CI) and online (production LLM-as-a-judge) evaluation.

Keywords
skillclaude