ai-eval-ci

SKILLWorkflowCommunity
v0.0.0TerminalSkillsApache-2.0Aktualisiert vor 26 TQuelle →

Run AI agent and LLM evaluations in CI/CD pipelines — automated quality gates that fail the build when AI output quality drops. Use when someone asks to "test my AI agent", "add evals to CI", "catch prompt regressions", "compare models", "evaluate LLM output quality", "set up AI quality gates", or "

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
132Repo-Sterne
1Clients
1Formate
vor 26 TLetzte Aktualisierung
Skill
AutorTerminalSkills
Version0.0.0
LizenzApache-2.0
KategorieWorkflow
Formateskill.md
PromptNicht veröffentlicht
Kompatibilität
Claude✓ Unterstützt
Cursor
Copilot
ChatGPT
Gemini
Über

Run AI agent and LLM evaluations in CI/CD pipelines — automated quality gates that fail the build when AI output quality drops. Use when someone asks to "test my AI agent", "add evals to CI", "catch prompt regressions", "compare models", "evaluate LLM output quality", "set up AI quality gates", or "benchmark my agent before deploying". Covers eval frameworks (Cobalt, Promptfoo, Braintrust), LLM-as

Schlagwörter
skillclaude