pmstack-transcript-review

SKILLWorkflowcommunity
v0.0.0RyanAlbertsMITUpdated 1mo agoSource →

Walks a PM through Anthropic's Step 6 ritual — reading transcripts from many trials to diagnose every failed eval task as one of three things — model mistake, grader mistake, or task-spec error. Implements the practice Anthropic describes as "critical" — without it, badly-calibrated graders mask rea

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
8Repo stars
1Clients
1Formats
1mo agoLast update
Skill
AuthorRyanAlberts
Version0.0.0
LicenseMIT
CategoryWorkflow
Formatsskill.md
PromptNot published
Compatibility
Claude✓ Supported
Cursor
Copilot
ChatGPT
Gemini
About

Walks a PM through Anthropic's Step 6 ritual — reading transcripts from many trials to diagnose every failed eval task as one of three things — model mistake, grader mistake, or task-spec error. Implements the practice Anthropic describes as "critical" — without it, badly-calibrated graders mask real model improvements. Use when the user has a /run-eval result with failures, asks "why did this fai

Keywords
skillclaude