model-evaluation

SKILLWorkflowcommunity
v0.0.0awslabsApache-2.0Updated 2d agoSource →

Generates python code that evaluates SageMaker models. Supports two evaluation types: LLM-as-Judge and Custom Scorer. Use when the user says "evaluate my model", "run a benchmark", "test model performance", "how did my model perform", "compare models", or other similar requests.

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
862Repo stars
1Clients
1Formats
2d agoLast update
Skill
Authorawslabs
Version0.0.0
LicenseApache-2.0
CategoryWorkflow
Formatsskill.md
PromptOpen (see Prompt tab)
Compatibility
Claude✓ Supported
Cursor
Copilot
ChatGPT
Gemini
About

Generates python code that evaluates SageMaker models. Supports two evaluation types: LLM-as-Judge and Custom Scorer. Use when the user says "evaluate my model", "run a benchmark", "test model performance", "how did my model perform", "compare models", or other similar requests.

Keywords
skillclaude