write-ai-evals

SKILLWorkflowcommunity
v0.0.0dineshrevunuruMITUpdated 2mo agoSource →

Designs and runs evals for any AI feature the way Dinesh does — golden sets, grading rubrics, LLM-as-judge, hard safety gates, failure-mode taxonomies fed back into structural fixes, and AI-quality product metrics (acceptance, regeneration, edit-distance). ALSO owns the pre-design model capability a

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
1Repo stars
1Clients
1Formats
2mo agoLast update
Skill
Authordineshrevunuru
Version0.0.0
LicenseMIT
CategoryWorkflow
Formatsskill.md
PromptNot published
Compatibility
Claude✓ Supported
Cursor—
Copilot—
ChatGPT—
Gemini—
About

Designs and runs evals for any AI feature the way Dinesh does — golden sets, grading rubrics, LLM-as-judge, hard safety gates, failure-mode taxonomies fed back into structural fixes, and AI-quality product metrics (acceptance, regeneration, edit-distance). ALSO owns the pre-design model capability assessment: what can this model actually do for this use case, cost-latency-quality tradeoffs, model-

Keywords
skillclaude

No dependency coverage

This entry publishes no npm package, so Forge has no dependency tree for it. That is a gap in coverage — not a statement that it has no dependencies.