Topic · AI & LLM Engineering
Best retrieval-augmented generation skills for Claude Code, Codex and other agents.
- skills
- 360
- official
- 46
Retrieval-augmented generation skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Reference of Claude API examples and guides covering tool use, vision, RAG, classification, summarization, text-to-SQL, prompt caching and agent patterns. | 2025Emma/ | 23k | 1 repo | ~2.2k | Automated safety check: Pass | MIT | 9 mo ago |
| 2 | Converts PDFs, Office files, HTML, images and other documents into a unified DoclingDocument with Markdown or JSON output, through the docling CLI, Python SDK or a remote service. | docling-project/ | 68k | — | ~1.1k | Automated safety check: Pass | MIT | yesterday |
| 3 | Shows how to store documents and embeddings in Chroma, query them by similarity with metadata filters, and persist them to disk for RAG and semantic search projects. | Orchestra-Research/ | 13k | 8 repos | ~2.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 4 | Answers questions from a local knowledge base folder by walking hierarchical index files and searching with grep, pdfplumber and pandas instead of loading whole files. | ConardLi/ | 714 | 2 repos | ~1.7k | Automated safety check: Pass | No licence | 5 mo ago |
| 5 | 5.Qmd Local hybrid search for markdown notes and docs. An agent skill from alsk1992/CloddsBot. | alsk1992/ | 2.9k | 3 repos | ~1.2k | Automated safety check: Pass | MIT | 4 days ago |
| 6 | Runs and debugs evaluations of how Chatbox models answer questions about large attached files, using synthetic and real long-document fixtures. | chatboxai/ | 42k | — | ~758 | Automated safety check: Pass | GPL-3.0 | 13 days ago |
| 7 | Generates vector embeddings through the 9Router /v1/embeddings endpoint, using models from providers such as OpenAI, Gemini, Mistral and Voyage for RAG and semantic search. | decolua/ | 30k | — | ~604 | Automated safety check: Pass | MIT | 6 days ago |
| 8 | TRIGGER when working with ai-sdk which is Laravel official first-party AI SDK. | trypostit/ | 676 | 2 repos | ~3.5k | Automated safety check: Pass | MIT | today |
| 9 | A skill your agent uses for anything involving the Synalinks neuro-symbolic LM framework (Keras-inspired): DataModel/Field/Input, JSON operators (+ & | ^ ~), synalinks.ops… | SynaLinks/ | 907 | — | ~4.8k | Automated safety check: Pass | Apache-2.0 | 9 days ago |
| 10 | Searches, summarizes, compares and answers questions from an already configured AutoRAG librarian agent over local documents and authorized datasources. | Marker-Inc-Korea/ | 5.1k | — | ~2k | Automated safety check: Pass | MIT | yesterday |
| 11 | Sets up FAISS for fast nearest-neighbor search over large collections of dense vectors, choosing between Flat, IVF, HNSW and product quantization indexes. | Orchestra-Research/ | 13k | 7 repos | ~1.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 12 | Python guidance for the Azure AI Search SDK covering vector, hybrid and semantic search, index management and indexers, with Entra ID authentication preferred over keys. | microsoft/ | 3.1k | 6 repos | ~4.4k | Automated safety check: Pass | MIT | yesterday |
| 13 | Guides building Tavily integrations for web search, URL extraction, site crawling and AI-assisted research in Python or JavaScript agent and RAG projects. | andrewyng/ | 14k | — | ~1.1k | Automated safety check: Pass | MIT | 4 mo ago |
| 14 | Building applications with Large Language Models - prompt engineering, RAG patterns, and LLM integration. | MoizIbnYousaf/ | 1.1k | 2 repos | ~1.3k | Automated safety check: Pass | MIT | 16 days ago |
| 15 | Diagnoses and repairs a broken AutoRAG install so every configured datasource is both indexed and returns real search hits. | Marker-Inc-Korea/ | 5.1k | — | ~3.3k | Automated safety check: Pass | MIT | yesterday |
| 16 | Guide to building Node.js and TypeScript apps on the agent-squad package: orchestrator, agent types, classifier routing, storage, retrievers and MCP tools. | 2FastLabs/ | 7.8k | — | ~4.3k | Automated safety check: Pass | Apache-2.0 | today |
| 17 | Provides reference guides and Python scripts for prompt optimization, RAG evaluation, and agent orchestration when building or tuning LLM systems. | maslennikov-ig/ | 259 | 4 repos | ~1.4k | Automated safety check: Pass | Unknown | 7 mo ago |
| 18 | A Chinese-language guide and reference for building agents with LangGraph 1.0, from a first ReAct agent through middleware, memory, MCP, RAG and web search. | luochang212/ | 457 | — | ~837 | Automated safety check: Notes | Unknown | 26 days ago |
| 19 | 19.Tw Legal RAG Retrieve real Taiwan court judgments with verifiable citations before answering any question about Taiwan law or case law. | aa0101181514/ | 327 | — | ~580 | Automated safety check: Pass | Unknown | 3 days ago |
| 20 | Searches, saves, and maintains a local document index through a local RAG MCP server. | shinpr/ | 407 | — | ~4.4k | Automated safety check: Pass | MIT | today |
| 21 | Comprehensive guide for building Agentic RAG systems using Microsoft Agent Framework in C. | shuyu-labs/ | 278 | — | ~1.1k | Automated safety check: Pass | Unknown | 3 mo ago |
| 22 | Quick reference for wdoc, a command-line and Python tool that summarizes, searches and answers questions over documents of many file types. | thiswillbeyourgithub/ | 545 | — | ~1.1k | Automated safety check: Pass | AGPL-3.0 | 1 mo ago |
| 23 | Efficiently perform web searches using the mcp-local-rag server with semantic similarity ranking. | nkapila6/ | 134 | 1 repo | ~1.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 24 | A skill your agent uses when polishing, diagnosing, tailoring, or exporting resumes for LLM, RAG, Agent, Agentic RL, post-training, pretraining, AIGC, search/ranking, multimodal, AI backend, or LLM… | wanyichen06/ | 324 | — | ~1.2k | Automated safety check: Pass | MIT | 2 mo ago |
| 25 | Process documents with Blockify API to create optimized IdeaBlocks for RAG. | iternal-technologies-partners/ | 316 | — | ~6.2k | Automated safety check: Notes | Unknown | 5 mo ago |
| 26 | Builds a retail product search agent on Google Cloud, from catalog ingestion into BigQuery and Vector Search to ADK scaffolding, evaluation and Cloud Run deployment. | google/ | 10k | — | ~3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 27 | A skill your agent uses for setting up vector similarity search with pgvector for AI/ML embeddings, RAG applications, or semantic search. | timescale/ | 1.9k | 1 repo | ~3.8k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 28 | 28.Gno Search local documents, files, notes, and knowledge bases. An agent skill from gmickel/gno. | gmickel/ | 115 | 1 repo | ~1.6k | Automated safety check: Pass | MIT | today |
| 29 | 29.LLM Wiki Build and maintain an LLM-curated knowledge base from papers, articles, transcripts, notes and project findings. | praneybehl/ | 117 | — | ~5.7k | Automated safety check: Pass | MIT | 24 days ago |
| 30 | Installs, configures, and repairs AutoRAG's search model, approved folders, indexes, and datasources, and registers its Lite MCP server. | Marker-Inc-Korea/ | 5.1k | — | ~5.2k | Automated safety check: Pass | MIT | yesterday |
| 31 | Give a Python agent (such as Hermes Agent by Nous Research) durable, local-first memory plus a queryable SPARQL knowledge graph, backed by CortexDB through its gRPC sidecar and the cortexdb-client… | liliang-cn/ | 273 | — | ~1.7k | Automated safety check: Pass | MIT | today |
| 32 | 32.Clawmem ClawMem operational reference for agents at query time — the 3-rule escalation gate, MCP tool routing, the 4 query-optimization levers, pipeline behavior (query vs intentsearch), composite scoring… | yoloshii/ | 210 | — | ~7.5k | Automated safety check: Pass | MIT | yesterday |
| 33 | 33.Memex Search Discover prior agent work across sessions, projects, providers, and machines. | nicosuave/ | 246 | — | ~3.3k | Automated safety check: Pass | MIT | yesterday |
| 34 | 34.Edgeparse Extract structured content from any PDF for AI agents, RAG pipelines, and Copilot Skills. | raphaelmansuy/ | 143 | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 35 | Review RAG designs or accessible implementations for extraction failures, deterministic facts, evidence routing, source versioning, code retrieval, and memory boundaries. | FinanceFlash/ | 141 | — | ~766 | Automated safety check: Pass | Apache-2.0 | today |
| 36 | 使用本地 Bilibili RAG 服务进行检索与问答。用户询问 B 站收藏夹内容、视频要点总结、来源追溯、入库状态时使用。内容问答时优先通过 sessionid 与 folderids 限定范围,避免空范围导致 fallback。 | via007/ | 1.3k | — | ~326 | Automated safety check: Pass | Apache-2.0 | 10 days ago |
| 37 | Builds AI agents on the Convex agent component: threads, messages, tools that call queries and mutations, streaming, RAG with vector search, and workflows for multi step jobs. | waynesutton/ | 404 | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | 9 days ago |
| 38 | Search, query, and manage Weaviate vector database collections. | weaviate/ | 105 | — | ~1.8k | Automated safety check: Pass | BSD-3-Clause | yesterday |
| 39 | 39.Graphmemory Build and query embedded GraphRAG knowledge graphs with DuckDB-backed vector, full-text, and hybrid search. | bradAGI/ | 160 | — | ~3.1k | Automated safety check: Pass | MIT | 5 mo ago |
| 40 | 40.Evaluate RAG Guides evaluation of a RAG system by diagnosing failures in traces, building a retrieval test set and scoring retrieval and generation separately. | ai-evals-course/ | 1.5k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | 13 days ago |
| 41 | Give a Node.js agent (such as OpenClaw) durable, local-first memory plus a queryable SPARQL knowledge graph, backed by CortexDB through its gRPC sidecar and the cortexdb-client npm package. | liliang-cn/ | 273 | — | ~1.6k | Automated safety check: Pass | MIT | today |
| 42 | 用于升级 AI 系统、agent workflow、Codex skill、prompt、memory、RAG、tool routing、schema、eval set 或 feedback loop;也用于把 AI 工作单从指令单升级为意图单,并对研究、检索、测试和 AI 对话做 VOI 决策门审计。需要 Intent Work Order、WOOP… | DY-2026/ | 410 | — | ~2.1k | Automated safety check: Pass | MIT | 1 mo ago |
| 43 | Teaches an agent to build LM pipelines, RAG systems and agents in DSPy using signatures, modules and optimizers instead of hand-tuned prompts. | Orchestra-Research/ | 13k | 10 repos | ~3.8k | Automated safety check: Pass | MIT | 3 mo ago |
| 44 | A skill your agent uses when the user has changed a prompt (system prompt, RAG template, agent instruction, etc.) and wants to know whether the candidate is better or worse than the baseline. | agentscope-ai/ | 867 | — | ~2.8k | Automated safety check: Pass | Apache-2.0 | 26 days ago |
| 45 | 45.RAG Skills RAG-specific best practices for LlamaIndex, ChromaDB, and Celery workers. | llama-farm/ | 837 | 1 repo | ~1.3k | Automated safety check: Pass | Apache-2.0 | 3 mo ago |
| 46 | Search 2500+ curated ChatGPT and LLM open-source repositories. | taishi-i/ | 3.3k | — | ~3.8k | Automated safety check: Pass | CC0-1.0 | 3 days ago |
| 47 | Shows how to use Pinecone, a managed vector database, for production RAG, semantic search and recommendations: indexes, upserts, queries, filters and namespaces. | Orchestra-Research/ | 13k | 6 repos | ~2k | Automated safety check: Pass | MIT | 3 mo ago |
| 48 | Guides and best practices for working with Lakebase Postgres, the database behind Neon. | usenotra/ | 255 | — | ~4.1k | Automated safety check: Notes | AGPL-3.0 | today |
Questions, answered from the data.
What is the best retrieval-augmented generation skill?
Claude Cookbooks Reference from 2025Emma/vibe-coding-cn ranks first of the 360 retrieval-augmented generation skills listed here, with the highest score: its repository has 23k GitHub stars, 1 other GitHub owner carry a copy, its SKILL.md loads about 2.2k tokens and it passes the automated safety check with no findings. Next come Docling Document Conversion and Chroma Vector Database.
Which retrieval-augmented generation skills are official?
46 of the 360 retrieval-augmented generation skills are official, published by the vendor's own GitHub organization: Azure AI Search Python SDK, Retail Product Search Agent, Weaviate, Release, Qdrant Advisor and 41 more.
How are these skills ranked?
By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.
Explore related skills
Category
More topics in AI & LLM Engineering
- Building AI agents525
- Deep learning408
- Embeddings381
- LLM inference and serving364
- Prompt engineering350
- Fine-tuning313
- LLM evaluation303
- Speech recognition and synthesis272
- Structured output and tool calling271
- LLM cost and token optimization256
- LLM API integration218
- LLM observability217
- Model routing and gateways217
- LLM guardrails208
- Computer vision206
- Model hubs and datasets180
- GPU and accelerator computing171
- Diffusion and image models167
- Natural language processing141
- Reinforcement learning67
- AI interpretability23