llm-caching

SKILLFlusso di lavorocommunity
v0.0.0BagelHoleMITAggiornato 3 mesi faFonte →

Implement multi-layer LLM caching with exact match, semantic similarity, and provider-side prompt caching. Reduce API costs by 30–70%, cut latency, and improve throughput using Redis, GPTCache, and provider caching APIs.

Community-submitted skill. Not yet reviewed by the Forge team. Full prompt content may not be available.Request review →
47Stelle del repo
1Client
1Formati
3 mesi faUltimo aggiornamento
Skill
AutoreBagelHole
Versione0.0.0
LicenzaMIT
CategoriaFlusso di lavoro
Formatiskill.md
PromptNon pubblicato
Compatibilità
Claude✓ Supportato
Cursor
Copilot
ChatGPT
Gemini
Descrizione

Implement multi-layer LLM caching with exact match, semantic similarity, and provider-side prompt caching. Reduce API costs by 30–70%, cut latency, and improve throughput using Redis, GPTCache, and provider caching APIs.

Parole chiave
skillclaude