Toolbox/Coding agents

Autonomous SWE agents

self-directed agents that go from issue to PR

46 repositories, most starred first
State of the category

code-yeongyu/oh-my-openagent code-yeongyu/oh-my-openagent is the clear leader by adoption and breadth — it’s the default if you want a well‑adopted, multi-agent orchestration harness for end‑to‑end repo workflows. Pick a different project only if you need a narrow capability (security, self‑hosting, sandboxed PRs, CI/test loops, or experimental autonomy) that the leader doesn’t prioritize.

Vertical specialization
Teams need out‑of‑the‑box, domain‑specific automation (SAST, website cloning, issue‑fixers) because general orchestration frameworks still require heavy customization for mission‑critical tasks.
Safety & governance
Real engineering teams demand auditability, mandatory approval gates, sandboxed execution, and thread‑native receipts so agents can’t silently introduce risky changes into repos or CI pipelines.
CI, tests & team workflows
Integrating autonomous agents into existing CI/PR/test loops and providing reproducible, test‑driven runs matters to teams who need reliable, reviewable outputs rather than exploratory toy runs.
Self‑hosted & enterprise deployment
Enterprises and security‑sensitive teams need self‑hostable platforms, edge/enterprise deployability, or managed offerings rather than cloud‑only demos so they can meet compliance and scale requirements.
Experimentation & autonomy
Researchers and advanced teams are iterating on self‑modifying agents, evolutionary improvement, reusable skill marketplaces, and persistent project memory because true long‑running autonomy is still an open research and engineering problem.
Find your fit

Which one matches your setup?

Answer any of the questions — the shortlist updates as you go. Recommendations come from the capability passports below, nothing else.

Where should reviews happen?
Can code leave your infrastructure?
What matters most?
Model access?
Pick at least one answer to get a shortlist.
Side by side

Comparison matrix

Axes are extracted from each project's docs by our review pipeline; the maturity score is computed from stars, growth and commit activity — not an opinion. Click a column to sort.

Runs inModelsContextCost to run
68.7k +214/7d
CLICoding-agent pluginIDEBYOKAnthropicGeminiOpenAIFixed providerWhole-repo analysisBasic filtersCloud APIs (your key)Requires your provider subscriptions / API keys
33.8k +524/7d
CLIIDECoding-agent pluginAnthropicOpenAIGeminiFixed providerWhole-repo analysisBasic QA (visual diff)Cloud agent APIsRequires an AI coding agent — provider billing applies
27.0k +123/7d
PR botCICLIWeb appFixed providerDiff + related filesHuman acceptance gatingSelf-hostableSelf-hosted — runs on your infrastructure
20.2k +57/7d
CLIIDEBYOKDiff + related filesBasic config filtersYour API keyYour API key (cloud LMs)
16.5k +43/7d
CLIWeb appCIBYOKWhole-repo analysisPermission & validationBYO API keyRequires your model API key (OpenRouter or other provider)
10.7k +34/7d
PR botWeb appCLIBYOKWhole-repo accessBasic filters & middlewareCloud APIs (your key)Your API key + cloud sandbox costs
4.6k +62/7d
Web appPR botCIFixed providerDiff + related filesNone mentionedSelf-hostableFree to start; online service or self-hosted
11.4k +68/7d
CLIFixed providerDiff onlyNo controlsThird-party cloudUnclear; depends on deployment
5.3k +12/7d
Web appCLICIGeminiDiff + related filesLinting & checksCloud APIs (your key)Cloudflare Workers paid plan + Google Gemini API key
4.8k +28/7d
CLICoding-agent pluginBYOKOllama / localRelated files + contextApproval gates + validationFully local availableYour model provider costs (use your API key)
3.4k +50/7d
CLIIDECoding-agent pluginOpenAIAnthropicWhole-repo analysisVerified completion & reviewCloud APIs (BYO key)Your API key, model provider charges
1.3k +29/7d
CLIBYOKOllama / localWhole-repo analysisReviewed workflowsFully local capableYour API key or a local GGUF model
Ranked by maturity — 6 more in the full catalog below.
At a glance

Capability profiles

Six axes, 0–5 each. The shape tells you the strategy: a wide hexagon is a generalist, a spike is a specialist. Showing the 8 most established — the rest are in the full catalog.

ContextNoiseCustomPrivacyModelsMaturity

Orchestrates multiple agent harnesses and parallel "discipline" agents (Team Mode) to run whole-repo autonomous workflows rather than a single-model single-agent approach.

ContextNoiseCustomPrivacyModelsMaturity

Reconstructs a live website into a production-ready Next.js codebase by extracting design tokens, assets, and exact computed styles, then dispatching parallel builder agents to assemble the site.

ContextNoiseCustomPrivacyModelsMaturity

Orchestrates isolated, autonomous implementation runs so teams manage work instead of supervising individual coding agents.

ContextNoiseCustomPrivacyModelsMaturity

State-of-the-art open-source agent focused on autonomously fixing GitHub issues and offering a specialized EnIGMA mode for offensive cybersecurity, all governed by a single YAML.

ContextNoiseCustomPrivacyModelsMaturity

DeepCode pairs multi-agent delegation and reusable SKILL playbooks with autonomous, test-driven loops and parallel workers so it can decompose large features, run them in isolation, and iterate until tests pass.

ContextNoiseCustomPrivacyModelsMaturity

Provides pluggable isolated cloud sandboxes and subagent orchestration out of the box for safe, parallelized repo edits and automatic PR creation.

ContextNoiseCustomPrivacyModelsMaturity

Provides an enterprise-grade, end-to-end AI development platform that combines cloud development environments, AI task and model management, and requirement tracking for team workflows.

ContextNoiseCustomPrivacyModelsMaturity

This repository is an archived/deprecated reference implementation and public issue history that points to an actively rebuilt service at humanlayer.com.

All repositories (46)

OmO: Drop your tokens. Ultrawork. Done.

68.7k
+2147d
TypeScript
9 mo

Clone any website with one command using AI coding agents

33.8k
+5247d
JavaScript
5 mo

Symphony turns project work into isolated, autonomous implementation runs, allowing teams to manage work inste…

27.0k
+1237d
Elixir
6 mo

SWE-agent takes a GitHub issue and tries to automatically fix it, using your LM of choice. It can also be empl…

20.2k
+577d
Python
2 yrs

"DeepCode: Open Agentic Coding (Agent Harness & Loop Engineering & Multi-Agent Orchestration)"

16.5k
+437d
Python
1 yr

The best way to get AI coding agents to solve hard problems in complex codebases.

11.4k
+687d
TypeScript
2 yrs

An Open-Source Asynchronous Coding Agent

10.7k
+347d
Python
1 yr

An open-source vibe coding platform that helps you build your own vibe-coding platform, built entirely on Clou…

5.3k
+127d
TypeScript
1 yr

AI agent framework for plan-first development workflows with approval-based execution. Multi-language support…

4.8k
+287d
TypeScript
1 yr

AI coding platform for teams

4.6k
+627d
TypeScript
1 yr

The one and only agent harness for complex codebases. Project memory, planning, execution, and verified comple…

3.4k
+507d
TypeScript
3 mo

Visa Vulnerability Agentic Harness

2.7k
+837d
Python
3 mo

Ship your code, on autopilot. An open source agent that lives on your machines 24/7 and keeps your apps runnin…

1.8k
+107d
Rust
1 yr

Mention any ACP coding agent from Slack, GitHub, GitLab, Linear, or Lark. OpenTag runs Claude Code, Codex, Cur…

1.4k
+87d
TypeScript
2 mo

Ouroboros — self-creating AI agent. Born Feb 16, 2026.

1.3k
+297d
Python
6 mo

Codebase harness + loop engineer

1.2k
+207d
Python
2 mo

Your AI forgets. This remembers. Spec-driven coding harness for vibecoders, product owners, CEOs and real buil…

1.1k
+147d
JavaScript
3 mo

OpenAlpha_Evolve is an open-source Python framework inspired by the groundbreaking research on autonomous codi…

1.0k
-17d
Python
1 yr

Autonomous software engineering fleet of AI agents for production-grade PRs on AgentField: plan, code, test, a…

990
+47d
Go
7 mo

Autoprompt is a coding-agent skill that cuts failures by 45% on agentic coding tasks.

984
+707d
JavaScript
2 wk

Spec-driven development workflow for AI coding agents: architecture-first planning, task decomposition, GitHub…

976
+17d
Shell
5 mo

Official AHE code — Agentic Harness Engineering: observability-driven automatic evolution of coding-agent harn…

868
+177d
Python
4 mo

pi coding agent in a technicolor web trenchcoat

847
+177d
TypeScript
6 mo

Repeatable agents-plus-code workflows, packaged as one skill, stamped into any repo. Deterministic Python owns…

797
+427d
Python
1 mo

Claude Code plugin for Elixir/Phoenix/LiveView — 26 specialist agents, Iron Laws enforcement, and Tidewave MCP…

538
+47d
Python
6 mo

Bring your own agent and build a self-improving agentic system. Automatically mine failures, optimize the agen…

536
+17d
Python
5 mo

A Claude Code plugin that interviews you, designs the whole architecture, and writes a self-contained blueprin…

485
+87d
5 mo

Codex/Claude Code/Cursor的开源增强版,专注一句话实现复杂长程任务(教育、编程、办公、生活、娱乐、游戏)。部署在你自己的服务器上,团队用浏览器打开就能编程——包括手机。CLI & Web UI 双入…

480
+77d
Java
5 mo

The agent skills I actually use to build software with coding agents. The PIV loop, planning, worktrees, and t…

474
Python
1 mo

A Boris-style agentic orchestrator TUI that supervises headless Claude Code agents through a gated software-de…

448
+17d
Python
2 mo

DeepAgent Code: AI coding agent with persistent memory and control plane

434
TypeScript
2 mo

ATLAS: a senior-engineer layer for Claude Code. Explore with wireframes & prototypes, clarify the essentials,…

393
-17d
Python
1 yr

Local-LLM-first agentic coding assistant, with everything you need out of the box.

285
+247d
TypeScript
5 mo

Broccoli turns Linear tickets into shipped PRs — powered by Claude and Codex, running on your own Google Cloud…

285
Python
4 mo

From ticket to reviewed pull request. Free and open-source, on your machine.

254
Python
1 mo

Bivor — a macOS desktop workbench for the pi coding agent

208
+17d
TypeScript
3 wk

A local-first, open-source coding agent for your desktop. Bring your own LLM key; your code stays on your mach…

189
-57d
Rust
2 mo

Universal, model-agnostic operating harness for AI agents (Claude, Codex, Gemini, …) — a lean core + work-type…

185
-87d
Python
1 mo

A local-first, cross-platform Electron desktop workspace for Pi Coding Agent, with sessions, project files, br…

183
+227d
TypeScript
1 mo

Open-source OpenAI Codex alternative | Multi-agent desktop GUI for OpenCode | manage coding sessions, visualiz…

178
+27d
TypeScript
6 mo

A plugin harness for Claude Code and Codex, built on bounded-autonomy architecture

150
+87d
Shell
3 mo

Terminal-first, knowledge-grounded multi-agent software delivery pipeline: scope requirements, implement chang…

143
+17d
Python
1 mo

Terminal-first, knowledge-grounded multi-agent software delivery pipeline: scope requirements, implement chang…

143
+17d
Python
1 mo

The token-efficient agentic coding workbench. Built for a future where every token counts — it optimizes token…

135
Go
4 mo

Finish-First Autonomous Agent Loop for Claude Opus 4.7 – 2026 Edition

123
HTML
2 mo

The agent harness that makes small local models finish hard tasks — and prove it. Done is enforced by the engi…

83
-17d
TypeScript
2 mo
Don't want to self-host?

Hosted alternatives

If running your own reviewer is more ops than you want, these managed services cover the same job.

Devin

Cognition's hosted autonomous engineer: assign it a ticket and it plans, edits and opens a PR in its own cloud sandbox. Closest managed analogue to running SWE-agent yourself.

Try Devin
Google Jules

Async agent that clones your repo into a VM, works the issue and returns a diff; free tier for GitHub, no self-hosting.

Try Google Jules
OpenHands Cloud

Managed hosting from the OpenHands team — the same open-source agent, run for you with GitHub integration instead of your own infra.

Try OpenHands Cloud

More in Coding agents