model-evaluation

SKILLWorkflowCommunity
v0.0.0awslabsApache-2.0Aktualisiert vor 7 TQuelle →

Generates python code that evaluates SageMaker models. Supports two evaluation types: LLM-as-Judge and Custom Scorer. Use when the user says "evaluate my model", "run a benchmark", "test model performance", "how did my model perform", "compare models", or other similar requests.

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
862Repo-Sterne
1Clients
1Formate
vor 7 TLetzte Aktualisierung
Skill
Autorawslabs
Version0.0.0
LizenzApache-2.0
KategorieWorkflow
Formateskill.md
PromptÖffnen (siehe Tab „Prompt“)
Kompatibilität
Claude✓ Unterstützt
Cursor
Copilot
ChatGPT
Gemini
Über

Generates python code that evaluates SageMaker models. Supports two evaluation types: LLM-as-Judge and Custom Scorer. Use when the user says "evaluate my model", "run a benchmark", "test model performance", "how did my model perform", "compare models", or other similar requests.

Schlagwörter
skillclaude