add-llm-evals

SKILLFlujo de trabajocomunidad
v0.0.0ContextJet-aiNOASSERTIONActualizado hace 11 dFuente →

Use this when adding evaluation to an LLM/agent app - measuring output quality (correctness, faithfulness, relevance, safety) rather than just watching traces. Trigger on "add evals", "test my prompt", "is my RAG accurate", "catch regressions", "score outputs", or setting up an eval suite in CI. Cov

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
29Estrellas del repo
1Clientes
1Formatos
hace 11 dÚltima actualización
Skill
AutorContextJet-ai
Versión0.0.0
LicenciaNOASSERTION
CategoríaFlujo de trabajo
Formatosskill.md
PromptNo publicado
Compatibilidad
Claude✓ Compatible
Cursor
Copilot
ChatGPT
Gemini
Acerca de

Use this when adding evaluation to an LLM/agent app - measuring output quality (correctness, faithfulness, relevance, safety) rather than just watching traces. Trigger on "add evals", "test my prompt", "is my RAG accurate", "catch regressions", "score outputs", or setting up an eval suite in CI. Covers offline (CI) and online (production LLM-as-a-judge) evaluation.

Palabras clave
skillclaude