RAG & retrieval
retrieval and knowledge layers for agents
infiniflow/ragflow infiniflow/ragflow is the category leader; if you need a production-grade RAG engine that combines agent orchestration with template-based document understanding, start there. Otherwise pick a specialist project—some focus on deep document indexing, truly local search, semantic governance, domain workflows, or audit-safe CLIs.
Which one matches your setup?
Answer any of the questions — the shortlist updates as you go. Recommendations come from the capability passports below, nothing else.
Comparison matrix
Axes are extracted from each project's docs by our review pipeline; the maturity score is computed from stars, growth and commit activity — not an opinion. Click a column to sort.
| Runs in | Models | Context | Cost to run | |||||
|---|---|---|---|---|---|---|---|---|
⭐ 90.0k +558/7d | Web app | BYOKOpenAIGemini | Whole-repo analysis | Re-ranking & citations | ●●●●● | Self-hostable | ●●●●● | Your API key (paid LLMs); self-host or use cloud |
⭐ 52.0k +111/7d | CLIWeb appCI | BYOKOpenAIOllama / local | Whole-data ingestion | Rerankers + filters | ●●●●● | Local model support | ●●●●● | Your API key (or optional LlamaParse cloud) |
⭐ 35.5k +151/7d | CLIWeb app | BYOKOpenAI | Whole-repo analysis | Basic filters/config | ●●●●● | Self-hostable | ●●●●● | Uses your LLM API key (e.g. OpenAI); cloud service may be paid. |
⭐ 32.1k +702/7d | CLIIDECoding-agent pluginWeb app | AnthropicOpenAIGemini | Indexed frontmatter + selected skills | Matching + verification | ●●●●● | Self-hostable | ●●●●● | Free to clone; your model/API costs |
⭐ 29.2k +49/7d | CLIWeb app | BYOK | No repo context | Basic filters | ●●●●● | Runs locally | ●●●●● | Free to self-host; optional Chroma Cloud hosted service |
⭐ 2.4k | CLICoding-agent plugin | BYOK | Whole-repo index | Ranked + filters | ●●●●● | Fully local | ●●●●● | Free local use; external models may require your API key. |
⭐ 1.6k +7/7d | CLICoding-agent plugin | BYOKAnthropicOpenAIFixed provider | Whole-repo analysis | Dedup & human review | ●●●●● | Self-hostable | ●●●●● | Your LLM API key; ktx adds no extra usage billing |
⭐ 1.2k +1/7d | CLICoding-agent plugin | BYOK | Local-only | Dry-runs & confirmations | ●●●●● | Cloud APIs with key | ●●●●● | Volcengine account + optional LLM API key |
Capability profiles
Six axes, 0–5 each. The shape tells you the strategy: a wide hexagon is a generalist, a spike is a specialist.
Combines a production-grade RAG engine with agent orchestration and template-based document understanding to provide an agentic context layer for LLMs.
A data-first RAG framework focused on document agents and agentic OCR with extensive indexing, retrieval, and 300+ integrations.
Vectorless, reasoning-based hierarchical retrieval that builds a human-like table-of-contents tree for traceable, context-aware RAG without vector DBs or chunking.
Maps 817 structured cybersecurity skills to six major industry frameworks (e.g., MITRE ATT&CK and MITRE F3), providing practitioner-grade, cross-framework workflows for AI agents.
Provides a minimal, 4-function client API for fast in-memory prototyping plus an optional hosted Chroma Cloud for production-ready vector search.
Unifies ripgrep, BM25, and vector search into a local-first search layer that serves both humans and AI agents.
Automatically ingests wiki, semantic layers, and raw table metadata to build a reconciled context layer so agents reuse approved metrics instead of inventing SQL.
Provides an automation-safe, reviewable CLI with installable "Viking skills" and explicit dry-run/confirmation/read-after-write verification for production AI search and retrieval workflows.
All repositories (8)
RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with…
LlamaIndex is the leading document agent and OCR platform
📑 PageIndex: Document Index for Vectorless, Reasoning-based RAG
817 structured cybersecurity skills for AI agents · Mapped to 6 frameworks: MITRE ATT&CK, NIST CSF 2.0, MITRE…
Local-first search across your workspace, built for humans and AI agents.
ktx is an executable context layer for data and analytics agents 🐙 Allow Claude Code, Codex, or other AI agen…
Open CLI for integrating AI search, recommendation, and conversational retrieval into agent systems and busine…
Hosted alternatives
If running your own reviewer is more ops than you want, these managed services cover the same job.
A hosted RAG pipeline: ingest, chunk, embed and retrieve as a managed service, so you skip running the vector store and the sync jobs around it.
Try Vectorize →Managed vector database that underpins many retrieval stacks, run and scaled for you rather than self-hosted next to your app.
Try Pinecone →A fully managed RAG-as-a-service API with connectors to common sources, an alternative to wiring ingestion and retrieval yourself.
Try Ragie →