model-evaluation

SKILLWorkflowcommunity
v0.0.0awslabsApache-2.0Updated 1mo agoSource →

Generates python code that evaluates SageMaker models. Supports two evaluation types: LLM-as-Judge and Custom Scorer. Use when the user says "evaluate my model", "run a benchmark", "test model performance", "how did my model perform", "compare models", or other similar requests.

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
862Repo stars
1Clients
1Formats
1mo agoLast update
Skill
Authorawslabs
Version0.0.0
LicenseApache-2.0
CategoryWorkflow
Formatsskill.md
PromptOpen (see Prompt tab)
Compatibility
Claude✓ Supported
Cursor—
Copilot—
ChatGPT—
Gemini—
About

Generates python code that evaluates SageMaker models. Supports two evaluation types: LLM-as-Judge and Custom Scorer. Use when the user says "evaluate my model", "run a benchmark", "test model performance", "how did my model perform", "compare models", or other similar requests.

Keywords
skillclaude

No dependency coverage

This entry publishes no npm package, so Forge has no dependency tree for it. That is a gap in coverage — not a statement that it has no dependencies.