agent-evals-and-observability

SKILLWorkflowCommunity
v0.0.0magnus919MITAktualisiert vor 6 TQuelle →

Design, run, review, or release framework- and vendor-neutral evaluations and observability for AI agents. Use when defining agent evals, datasets, graders, trajectory review, regression analysis, release gates, production traces, or privacy-aware telemetry. Covers task and trajectory contracts, sta

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
44Repo-Sterne
1Clients
1Formate
vor 6 TLetzte Aktualisierung
Skill
Autormagnus919
Version0.0.0
LizenzMIT
KategorieWorkflow
Formateskill.md
PromptNicht veröffentlicht
Kompatibilität
Claude✓ Unterstützt
Cursor
Copilot
ChatGPT
Gemini
Über

Design, run, review, or release framework- and vendor-neutral evaluations and observability for AI agents. Use when defining agent evals, datasets, graders, trajectory review, regression analysis, release gates, production traces, or privacy-aware telemetry. Covers task and trajectory contracts, statistical comparisons, and incident-to-case learning; route framework implementation to pydanticai or

Schlagwörter
skillclaude