agent-evals-and-observability

SKILLFlusso di lavorocommunity
v0.0.0magnus919MITAggiornato 6 g faFonte →

Design, run, review, or release framework- and vendor-neutral evaluations and observability for AI agents. Use when defining agent evals, datasets, graders, trajectory review, regression analysis, release gates, production traces, or privacy-aware telemetry. Covers task and trajectory contracts, sta

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
44Stelle del repo
1Client
1Formati
6 g faUltimo aggiornamento
Skill
Autoremagnus919
Versione0.0.0
LicenzaMIT
CategoriaFlusso di lavoro
Formatiskill.md
PromptNon pubblicato
Compatibilità
Claude✓ Supportato
Cursor
Copilot
ChatGPT
Gemini
Descrizione

Design, run, review, or release framework- and vendor-neutral evaluations and observability for AI agents. Use when defining agent evals, datasets, graders, trajectory review, regression analysis, release gates, production traces, or privacy-aware telemetry. Covers task and trajectory contracts, statistical comparisons, and incident-to-case learning; route framework implementation to pydanticai or

Parole chiave
skillclaude