evaluator

SKILLWorkflowCommunity
v0.0.0moses607Apache-2.0Aktualisiert vor 1 Mon.Quelle →

An auto-evaluation and hallucination guard that scores an agent's output against a task-derived rubric, checks every claim for support, and returns an accept / revise / reject verdict with specific fixes. It derives the rubric from the task, tests factual grounding against provided context, probes e

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
1Clients
1Formate
vor 1 Mon.Letzte Aktualisierung
Skill
Autormoses607
Version0.0.0
LizenzApache-2.0
KategorieWorkflow
Formateskill.md
PromptNicht veröffentlicht
Kompatibilität
Claude✓ Unterstützt
Cursor
Copilot
ChatGPT
Gemini
Über

An auto-evaluation and hallucination guard that scores an agent's output against a task-derived rubric, checks every claim for support, and returns an accept / revise / reject verdict with specific fixes. It derives the rubric from the task, tests factual grounding against provided context, probes edge cases, and never rubber-stamps. Use when the user says "evaluate this", "verify", "grade this",

Schlagwörter
skillclaude