model-evaluation

SKILLFlusso di lavorocommunity
v0.0.0awslabsApache-2.0Aggiornato 7 g faFonte →

Generates python code that evaluates SageMaker models. Supports two evaluation types: LLM-as-Judge and Custom Scorer. Use when the user says "evaluate my model", "run a benchmark", "test model performance", "how did my model perform", "compare models", or other similar requests.

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
862Stelle del repo
1Client
1Formati
7 g faUltimo aggiornamento
Skill
Autoreawslabs
Versione0.0.0
LicenzaApache-2.0
CategoriaFlusso di lavoro
Formatiskill.md
PromptApri (vedi la scheda Prompt)
Compatibilità
Claude✓ Supportato
Cursor
Copilot
ChatGPT
Gemini
Descrizione

Generates python code that evaluates SageMaker models. Supports two evaluation types: LLM-as-Judge and Custom Scorer. Use when the user says "evaluate my model", "run a benchmark", "test model performance", "how did my model perform", "compare models", or other similar requests.

Parole chiave
skillclaude