ai-evals

SKILLWorkflowcommunity
v0.0.0RefoundAIMITUpdated 2mo agoSource →

Help users build robust infrastructure for measuring, monitoring, and iterating on AI product performance using human, code-based, and LLM-as-a-judge methodologies.

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
1kRepo stars
1Clients
1Formats
2mo agoLast update
Skill
AuthorRefoundAI
Version0.0.0
LicenseMIT
CategoryWorkflow
Formatsskill.md
PromptOpen (see Prompt tab)
Compatibility
Claude✓ Supported
Cursor—
Copilot—
ChatGPT—
Gemini—
About

Help users build robust infrastructure for measuring, monitoring, and iterating on AI product performance using human, code-based, and LLM-as-a-judge methodologies.

Keywords
skillclaude

No dependency coverage

This entry publishes no npm package, so Forge has no dependency tree for it. That is a gap in coverage — not a statement that it has no dependencies.