ai-eval-ci

SKILLFlusso di lavorocommunity
v0.0.0TerminalSkillsApache-2.0Aggiornato 26 g faFonte →

Run AI agent and LLM evaluations in CI/CD pipelines — automated quality gates that fail the build when AI output quality drops. Use when someone asks to "test my AI agent", "add evals to CI", "catch prompt regressions", "compare models", "evaluate LLM output quality", "set up AI quality gates", or "

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
132Stelle del repo
1Client
1Formati
26 g faUltimo aggiornamento
Skill
AutoreTerminalSkills
Versione0.0.0
LicenzaApache-2.0
CategoriaFlusso di lavoro
Formatiskill.md
PromptNon pubblicato
Compatibilità
Claude✓ Supportato
Cursor
Copilot
ChatGPT
Gemini
Descrizione

Run AI agent and LLM evaluations in CI/CD pipelines — automated quality gates that fail the build when AI output quality drops. Use when someone asks to "test my AI agent", "add evals to CI", "catch prompt regressions", "compare models", "evaluate LLM output quality", "set up AI quality gates", or "benchmark my agent before deploying". Covers eval frameworks (Cobalt, Promptfoo, Braintrust), LLM-as

Parole chiave
skillclaude