Skip to main content

llm-cache-designer

Étoiles4
Forks0
Mis à jour18 juillet 2026 à 15:25

Design caching layers in front of LLM inference — exact-match caches, semantic caches, provider prefix caching, and negative caching — so repeated questions never re-burn inference. Use this skill whenever the user mentions repeated/similar LLM queries, wants to cut inference cost or carbon for FAQ/support/RAG workloads, asks about semantic caching, or has any high-volume LLM endpoint. Part of Lean Agentic AI Skills; emits lean-findings.json plus a cache design.

Installation

Installer avec Codex ou Claude Copiez ce prompt, collez-le dans Codex, Claude ou un autre assistant, puis laissez-le vérifier la page du skill et l'installer pour vous.

Explorateur de fichiers
3 fichiers
SKILL.md
readonly