Skip to main content

llm-cache-designer

النجوم٤
التفرعات٠
آخر تحديث١٨ يوليو ٢٠٢٦ في ١٥:٢٥

Design caching layers in front of LLM inference — exact-match caches, semantic caches, provider prefix caching, and negative caching — so repeated questions never re-burn inference. Use this skill whenever the user mentions repeated/similar LLM queries, wants to cut inference cost or carbon for FAQ/support/RAG workloads, asks about semantic caching, or has any high-volume LLM endpoint. Part of Lean Agentic AI Skills; emits lean-findings.json plus a cache design.

التثبيت

التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.

مستكشف الملفات
3 ملفات
SKILL.md
readonly