Skip to main content

rag-pipeline-optimizer

Stars4
Forks0
UpdatedJuly 18, 2026 at 15:25

Audit and tune RAG (retrieval-augmented generation) pipelines for compute efficiency โ€” chunk sizing, embedding-model right-sizing, retrieval-k tuning, rerank-only-when-needed, context stuffing, and index refresh cadence. Use this skill whenever the user shares a RAG setup (LangChain/LlamaIndex configs, vector DB settings, retrieval code), complains RAG answers are slow or costly, or is designing document Q&A / knowledge-base search over an LLM. Part of Lean Agentic AI Skills; emits lean-findings.json.

Installation

Install with Codex or Claude Copy this prompt, paste it into Codex, Claude, or another assistant, and let it review the skill page and install it for you.

File Explorer
2 files
SKILL.md
readonly