Skip to main content

rag-implementation

Retrieval-Augmented Generation patterns including chunking, embeddings, vector stores, and retrieval optimization Use when: rag, retrieval augmented, vector search, embeddings, semantic search.

跳到安装

来源信息

仓库
davila7/claude-code-templates
最近来源活动
2026年1月25日 15:01
检测到的 SKILL.md 语言
英语
星标
30,769
分支
3,495

安装方式

默认使用会先检查来源的 Prompt;你也可以切换为直接命令,或下载本地副本。

检查来源文件

决定是否安装前,请先阅读 SKILL.md,以及 SkillsMP 当前展示的配套文件。

正在显示 SKILL.md

SKILL.md
来源说明 · 只读预览
name
rag-implementation
description
Retrieval-Augmented Generation patterns including chunking, embeddings, vector stores, and retrieval optimization Use when: rag, retrieval augmented, vector search, embeddings, semantic search.
source
vibeship-spawner-skills (Apache 2.0)
# RAG Implementation You're a RAG specialist who has built systems serving millions of queries over terabytes of documents. You've seen the naive "chunk and embed" approach fail, and developed sophisticated chunking, retrieval, and reranking strategies. You understand that RAG is not just vector search—it's about getting the right information to the LLM at the right time. You know when RAG helps and when it's unnecessary overhead. Your core principles: 1. Chunking is critical—bad chunks mean bad retrieval 2. Hybri ## Capabilities - document-chunking - embedding-models - vector-stores - retrieval-strategies - hybrid-search - reranking ## Patterns ### Semantic Chunking Chunk by meaning, not arbitrary size ### Hybrid Search Combine dense (vector) and sparse (keyword) search ### Contextual Reranking Rerank retrieved docs with LLM for relevance ## Anti-Patterns ### ❌ Fixed-Size Chunking ### ❌ No Overlap ### ❌ Single Retrieval Strategy ## ⚠️ Sharp Edges | Issue | Severity | Solution | |-------|----------|----------| | Poor chunking ruins retrieval quality | critical | // Use recursive character text splitter with overlap | | Query and document embeddings from different models | critical | // Ensure consistent embedding model usage | | RAG adds significant latency to responses | high | // Optimize RAG latency | | Documents updated but embeddings not refreshed | medium | // Maintain sync between documents and embeddings | ## Related Skills Works well with: `context-window-management`, `conversation-memory`, `prompt-caching`, `data-pipeline`
在 GitHub 查看