Toolbox/Applied agents

Research agents

deep research and web research

13 repositories, most starred first
State of the category

assafelovic/gpt-researcher assafelovic/gpt-researcher is the clear leader by adoption and breadth of features for general-purpose deep-research agents; before choosing, pick based on a single constraint you care about most (self-hosting/privacy, extreme long‑context/verifiability, or tight domain integrations), because many projects specialize to solve those gaps.

Local & privacy
Teams that must keep data and models on-prem or offline need projects explicitly engineered for fully local runs or BYO keys rather than cloud-first stacks.
Long‑horizon workflows
Sustained investigations require very large context windows, high tool-call budgets, and persistent memory so teams can verify multi-step reasoning over long documents and timeframes.
Specialized integrations
Research teams often need first-class integrations with existing tooling (Zotero, LaTeX/arXiv pipelines, patent workflows) to fit agents into established authoring and review processes.
Auditability & reproducibility
Regulated outputs or publishable papers require strict citation rules, ledgers, deterministic pipelines and edit-safety measures so results are defensible and reproducible.
Skill & workflow packaging
Groups that want repeatable human-in-the-loop processes prefer projects that package research skills, reviewer loops, and decision workflows as reusable skillsets or agent plugins.
Find your fit

Which one matches your setup?

Answer any of the questions — the shortlist updates as you go. Recommendations come from the capability passports below, nothing else.

Where should reviews happen?
Can code leave your infrastructure?
What matters most?
Model access?
Pick at least one answer to get a shortlist.
Side by side

Comparison matrix

Axes are extracted from each project's docs by our review pipeline; the maturity score is computed from stars, growth and commit activity — not an opinion. Click a column to sort.

Runs inModelsContextCost to run
29.3k +102/7d
Web appCoding-agent pluginCLIBYOKOpenAIAnthropicGeminiMulti-source contextAggregation & filteringSelf-hostableYour API key; cloud model costs (≈$0.4 per deep research with o3-mini as documented)
2.9k +89/7d
Coding-agent pluginBYOKOpenAIAnthropicGeminiOllama / localFixed providerWhole-library accessCitation gating & confirmationsLocal model supportYour API key or local models
2.1k +52/7d
Coding-agent pluginCLIAnthropicFixed providerWeb-onlyHuman-in-the-loopCloud APIsDepends on model provider
1.1k +73/7d
CLIWeb appBYOKOllama / localWhole-project accessClarifying prompts & reviewFully local optionalYour API key (pay-per-use) or free local models
1.1k +73/7d
Coding-agent pluginCLIAnthropicWhole-repo analysisSeverity gating & dedupCloud LLM via sessionRuns in your Claude session — model usage billed by the provider
239 +3/7d
CLIWeb appCIBYOKOpenAIFixed providerWhole-library analysisRerank + quotasSelf-hostableYour API key for LLM/embeddings (optional local-only use)
8.4k +2/7d
Web appFixed providerDiff onlyTrace collectionFully localFree; self-host models or use HuggingFace
3.1k
CLIWeb appBYOKOpenAIAnthropicGeminiDiff + related filesNone mentionedFully local possibleOpenRouter API key to start, or self-hosted local GPU
892 +7/7d
CLIWeb appFixed providerSession-limitedProtocol enforcedFully localFree, self-host
423 +5/7d
Coding-agent pluginCLIOpenAIAnthropicGeminiWhole-repo analysisGated workflow & verificationCloud API requiredRequires access to Codex or Claude (your provider access); local LaTeX required
76 +1/7d
CLIWeb appCIGitHub ActionBYOKFile-level onlyGovernance gatingCloud via API keySelf-hosted; your API key for AI services
58 +1/7d
CLIWeb appCICoding-agent pluginBYOKOpenAIFixed providerMulti-module pipelineStrict evidence gatingSelf-hostableRequires your LLM + search API keys
Ranked by maturity — 1 more in the full catalog below.
At a glance

Capability profiles

Six axes, 0–5 each. The shape tells you the strategy: a wide hexagon is a generalist, a spike is a specialist. Showing the 8 most established — the rest are in the full catalog.

ContextNoiseCustomPrivacyModelsMaturity

An open-source multi-agent deep-research pipeline that combines web scraping, local document analysis, and MCP integrations to produce long, cited research reports.

ContextNoiseCustomPrivacyModelsMaturity

Deep Zotero-native integration that provides grounded, citation-linked paper chat and library-wide agent actions (reads, writes, tagging, notes) not offered by generic agents.

ContextNoiseCustomPrivacyModelsMaturity

Provides a structured two-phase (outline then deep investigation) human-in-the-loop research workflow packaged specifically as skills for Claude Code / OpenCode / Codex, including parallel web research modules and report generation.

ContextNoiseCustomPrivacyModelsMaturity

A local-first, bring-your-own-keys AI research assistant that runs on your computer and bundles domain-specific scientific skills, workflows, and a living lab notebook.

ContextNoiseCustomPrivacyModelsMaturity

Bundles a bounded, audit-ready review→decide→patch→recheck loop with ledgered issues and edit-safety safeguards, exposed as a Claude Code skill.

ContextNoiseCustomPrivacyModelsMaturity

A local-first research agent that reads original source windows and runs an investigation loop to produce cited, source-grounded reports instead of one-shot RAG.

ContextNoiseCustomPrivacyModelsMaturity

Provides open-source deep research agents with extremely long (256K) context windows and very high tool-call budgets, enabling long‑horizon, verifiable multi-step research workflows.

ContextNoiseCustomPrivacyModelsMaturity

An open-source research-agent framework that reproduces state-of-the-art results on multiple agentic benchmarks while remaining fully runnable locally.

All repositories (13)

An autonomous agent that conducts deep research on any data using any LLM providers

29.3k
+1027d
Python
3 yrs

MiroThinker is a deep research agent optimized for complex research and prediction tasks. Our latest models, M…

8.4k
+27d
Python
1 yr

🏆 Top-1 on 5+ benchmarks | Web UI | Supports MiroThinker, Claude, Kimi, OpenAI

3.1k
Python
1 yr

An open-sourced research agent system deeply rooted in your Zotero library.

2.9k
+897d
TypeScript
7 mo

Structured deep research skill for Claude Code/Open Code/Codex with human-in-the-loop control

2.1k
+527d
Python
8 mo

An AI co-scientist running on your desktop. Claude Science but better.

1.1k
+737d
TypeScript
5 mo

Pre-submission AI review stress-test for research papers. A Claude Code skill: review, verdict, revise, verify…

1.1k
+737d
JavaScript
3 mo

🔬🦞 A self-evolving AI research colleague for scientists. 285 skills, zero hallucination, persistent memory.

892
+77d
TypeScript
5 mo

A highly customizable agentic harness for arXiv-ready ML/AI review papers (and beyond). It drives agentic AI l…

423
+57d
TeX
1 yr

A library-science-inspired personal knowledge management system with LLM agents

239
+37d
Python
3 mo

Literature-grounded research idea exploration for CLI agents. 文献驱动的研究选题与方向探索工具。

135
+17d
Python
5 mo

AI-native macro investment research infrastructure with native MCP, terminal CLI, agent runtime, and disciplin…

76
+17d
Python
8 mo

专利侵权分析系统 —— 输入专利公开号,产出竞品侵权分析报告;同时打包成 skill,可被任意 agent(dsh, codex, claudecode 等) 调用。

58
+17d
Python
4 mo
Don't want to self-host?

Hosted alternatives

If running your own reviewer is more ops than you want, these managed services cover the same job.

OpenAI Deep Research

A hosted agent that browses, reads and writes a cited report from one prompt, so you get long-horizon research without running a retrieval-and-agent stack.

Try OpenAI Deep Research
Perplexity

Managed answer engine with a research mode that gathers and cites sources for you, the hosted alternative to standing up your own web-research agent.

Try Perplexity
Elicit

A hosted research assistant focused on the academic literature: it finds papers, extracts findings and summarizes across them without a pipeline you maintain.

Try Elicit

More in Applied agents