Skill
write-ai-evals
Designs and runs evals for any AI feature the way Dinesh does — golden sets, grading rubrics, LLM-as-judge, hard safety gates, failure-mode taxonomies fed back into structural fixes, and AI-quality product metrics (acceptance, regeneration, edit-distance). ALSO owns the pre-design model capability a
Claim this listing
Connect your GitHub to prove you own or maintain this listing. We verify repo access automatically — most publishers are confirmed in seconds.
1Connect GitHub
2Submit your claim
3Auto-verified, or reviewed within 48h