agent-redteam

SKILLFlusso di lavorocommunity
v0.0.0HefrockMITAggiornato 2 mesi faFonte →

Generates adversarial test cases targeting safe-failure behavior — refusals, hedging, graceful degradation. Use this when you want to stress-test an agent's safety boundaries, check for prompt injection, or build an adversarial eval set. Pairs with agent-eval for scoring: agent-redteam generates cas

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
1Client
1Formati
2 mesi faUltimo aggiornamento
Skill
AutoreHefrock
Versione0.0.0
LicenzaMIT
CategoriaFlusso di lavoro
Formatiskill.md
PromptNon pubblicato
Compatibilità
Claude✓ Supportato
Cursor—
Copilot—
ChatGPT—
Gemini—
Descrizione

Generates adversarial test cases targeting safe-failure behavior — refusals, hedging, graceful degradation. Use this when you want to stress-test an agent's safety boundaries, check for prompt injection, or build an adversarial eval set. Pairs with agent-eval for scoring: agent-redteam generates cases, agent-eval runs and scores them. Triggers on "red team my agent," "test for jailbreaks," "advers

Parole chiave
skillclaude

Nessuna copertura delle dipendenze

Questa voce non pubblica alcun pacchetto npm, quindi Forge non ha un albero delle dipendenze per essa. È una lacuna di copertura, non l'affermazione che non abbia dipendenze.