add-llm-evals

SKILLFlusso di lavorocommunity
v0.0.0ContextJet-aiNOASSERTIONAggiornato 11 g faFonte →

Use this when adding evaluation to an LLM/agent app - measuring output quality (correctness, faithfulness, relevance, safety) rather than just watching traces. Trigger on "add evals", "test my prompt", "is my RAG accurate", "catch regressions", "score outputs", or setting up an eval suite in CI. Cov

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
29Stelle del repo
1Client
1Formati
11 g faUltimo aggiornamento
Skill
AutoreContextJet-ai
Versione0.0.0
LicenzaNOASSERTION
CategoriaFlusso di lavoro
Formatiskill.md
PromptNon pubblicato
Compatibilità
Claude✓ Supportato
Cursor
Copilot
ChatGPT
Gemini
Descrizione

Use this when adding evaluation to an LLM/agent app - measuring output quality (correctness, faithfulness, relevance, safety) rather than just watching traces. Trigger on "add evals", "test my prompt", "is my RAG accurate", "catch regressions", "score outputs", or setting up an eval suite in CI. Covers offline (CI) and online (production LLM-as-a-judge) evaluation.

Parole chiave
skillclaude