WA

weijt606/ai-agent-map

Developer tools
66 stars 품질 45 트렌드 45

A practical, visual-first guide to comparing AI agents, platforms, runtimes, and orchestration tools so you can shortlist the right options faster.

개요

AI Agent Map is a practical, visual-first guide for comparing mainstream AI agents, agent platforms, runtimes, and orchestration tools. The goal is simple: help readers get to a sensible shortlist faster. - The agent landscape is crowded. - Many resources explain ideas, but not fit, anti-fit, or operating cost. - People usually need a comparison layer, not another pile of links. This repo stays focused on selection: what a system is good at, where it breaks down, and what kind of operator cost comes with it. This table tracks projects that showed up as especially hot in the latest weekly GitHub snapshot. The rank follows the 7-day gain. The total star counts below were checked when this repo was updated. 2026-09-17 · 2026-09-09 → 2026-09-17 (gain since last update, against last window's 3, so raw gains are not comparable; every claim below is stated on the ) · checked at update time Project names link to the upstream GitHub repo.

README

AI Agent Map

AI Agent Map is a practical, visual-first guide for comparing mainstream AI agents, agent platforms, runtimes, and orchestration tools.

The goal is simple: help readers get to a sensible shortlist faster.

What This Repo Is For

  • The agent landscape is crowded.
  • Many resources explain ideas, but not fit, anti-fit, or operating cost.
  • People usually need a comparison layer, not another pile of links.

This repo stays focused on selection: what a system is good at, where it breaks down, and what kind of operator cost comes with it.

Where To Start

If your question is… Start here
I need a shortlist first
I need help choosing for coding automation
I already have candidates and want a side-by-side view
I care about dimensions like approval, memory, scheduling, and deployment
I want every project scored on those dimensions, side by side
I want to know what it actually costs to run, and which model tier is worth it
My agents already run — I need to know whether they still work
I want the stock rankings and the weekly trend chart
I want problem-first guides or the full comparison list

Recent Heat Ranking

Popularity is not fit.

This table tracks projects that showed up as especially hot in the latest weekly GitHub snapshot. The rank follows the 7-day gain. The total star counts below were checked when this repo was updated.

Last updated: 2026-09-17 · Snapshot window: 2026-09-09 → 2026-09-17 (gain since last update, 8 days against last window’s 3, so raw gains are not comparable; every claim below is stated on the weekly rate) · Star counts: checked at update time

Project names link to the upstream GitHub repo. When this map has a written profile, it is linked separately in the “Map status” column.

Rank Project Current stars Snapshot gain Map status How to read it
#1 (new) Open Code Review 33.1k +10,935 In scope · profile From off the table to #1 on a 26x jump in weekly rate (~370 → ~9,570), the largest this board has computed. The cause is documented rather than inferred: it hit #1 on GitHub Trending on Sept 16 (+3,215 that day, +2,751 the day before). Grew 49% of its own size in eight days
#2 (=) DeepSeek Harness 227.2k +10,049 In scope · profile Accelerated 17% in its second measured window and crossed 225k. By this board’s own rule that is a confirmed trend, not a spike
#3 (↓) mattpocock/skills 263.9k +6,317 Watchlist (Skills Wave) Lost #1, down 27% on the rate. It held the seat by 11 stars last window; this one it fell two places. Past 260k
#4 (↓) Superpowers 287.8k +4,018 In scope · profile Up 7% and cleared 285k, but down a seat as two larger movers came through above it
#5 (↑) TradingAgents 107.2k +3,541 Out of scope (finance-research vertical) Second consecutive acceleration (+44% after +189%), which is the confirmation this board said it was waiting for. Crossed 105k; still not an in-scope agent surface
#6 (=) Pi 106.5k +3,075 In scope · profile Up 12%, reversing two windows of decline, and crossed 105k
#7 (↓) Hermes Agent 246.3k +2,616 In scope · profile Down 20% and three seats; crossed 245k. Present in all 22 recorded windows
#8 (new) addyosmani/agent-skills 95.7k +2,529 Watchlist (Skills Wave) Rate up 61%, back on the table after one window off; past 95k
#9 (↑) Codex CLI 124.9k +2,097 In scope · profile Essentially flat (−3%) and up a seat, which is what a steady line looks like when the board reshuffles around it
#10 (↓) OpenCode 208.0k +1,920 In scope · profile Down 21% in its second window, past 208k. Its first measured gain now reads as the higher of the two
  • Heat is useful for discovery, not for selection by itself.
  • This window is 8 days against last window’s 3. Raw gains between the two columns are not comparable, so every claim here is on the weekly rate (this window: gain ÷ 8 × 7; last window: gain ÷ 3 × 7).
  • Open Code Review went from off the table to #1 on a 26x jump in weekly rate, the largest this board has computed on the basis it uses for these (the previous high was 14x, K-Dense on 2026-09-01). The cause is documented, not guessed: the project hit #1 on GitHub Trending on September 16, taking +3,215 stars that day and +2,751 the day before. It grew 49% of its own size in eight days. By this board’s own rule that is a spike until a second window confirms it, and a Trending placement is the most transient cause there is.
  • DeepSeek Harness accelerated 17% in its second measured window. That is the rule paying out in the other direction: two consecutive windows of growth is the confirmation this board asks for, and it now reads as a trend rather than an arrival spike. Crossed 225k.
  • mattpocock/skills lost #1, down 27%. It held the seat by 11 stars last window and fell two places this one, which is what a narrow lead usually means.
  • Of the two spikes this board flagged last window, one confirmed and one did not. TradingAgents added a second consecutive acceleration (+44% after +189%) and is now a confirmed trend. OpenHands gave back 32% and stayed off the table. Reporting both is the point of flagging them.
  • A correction to last window’s read. This board wrote that Ruflo had confirmed rather than reversed and called it the first breakout in five windows to survive its follow-up. It then gave back 59% and left the table. The claim was made on a three-day window, which was too short a base to carry it; the pattern held for one window and not two.
  • Browser Use fell 70% after entering at #5 on its first measured gain. A first measured window is not a trend either, and this board should read those entries the same way it reads spikes.
  • The skills wave recovered to 3 of 10 from last window’s 2, its lowest since the wave began. addyosmani/agent-skills is back at +61%, while K-Dense (−41%) and academic-research-skills (−32%) both kept cooling.
  • Milestones: Open Code Review crossed 30k and 33k, DeepSeek Harness 225k, mattpocock/skills 260k, Superpowers 285k, Hermes Agent 245k, TradingAgents and Pi both 105k, OpenCode 208k, addyosmani 95k.
  • OpenClaw remains the absolute leader at 389.9k stars (+651); it is profiled but stays out of the gain-ranked table because reliable week-over-week deltas for a project this large are noisy.

Ranking Trend

How the weekly top 10 has shifted since tracking began — each line is one project, breaks mean it fell off the board that week:

And the same windows read as seats per layer — the quantitative version of the skills-wave story the bullets tell in prose:

Full stock rankings by category — agents, agent infra, skills, and their verticals, sorted by total stars — live in rankings/.

Beyond The Rank

Popularity tells you what to look at. These four pages tell you what to pick:

  • Capability matrix — every project scored side by side (●/◐/○/—) across the nine shared capability dimensions, grouped by route. The answer to “for this capability, who treats it as a core strength.”
  • Cost & benchmarks — frontier-model capability vs per-token price, plus how each coding agent actually bills. Since the model layer went tiered and metered, “which tier for this task” is the selection decision.
  • Memory approaches — six different things projects mean by “has memory,” from self-editing stores to passive semantic recall, and which to pick for what you need to persist.
  • Observability & evaluation — the layer under everything above: once an agent runs unattended, failure stops looking like a crash and starts looking like silent quality drift. Compares Langfuse, Opik, Phoenix, Helicone, LangSmith and others — and untangles the four different things “open source” means in that field.

Market Pulse

The three structural stories shaping selection right now — full records with dates and sources live in market-events.md:

  • The .claude/skills wave keeps compounding — and is now concentrating into one directory (May 2026 → ongoing): curated skill collections and skills frameworks have held roughly half of the weekly heat top 10 for three months, and through August the count stopped moving entirely — four of ten, three windows running, with no rotation in the last one. What is still moving is the split inside the wave: mattpocock/skills now out-gains the other three combined, where a month ago it was level with them. For many tasks the skill layer now matters as much as the underlying agent; read the concentration as key-person risk, not a broadening ecosystem. This map profiles the framework end through Superpowers and tracks collections on the skill boards.
  • The model layer became a budget decision — and in the first week of September both ceilings moved to the same price: Claude Fable 5.1 (Sept 1) and GPT-6 Astra (Sept 3) both list at $10 / $50, so the frontier comparison is no longer about the sticker. It is about cache reads ($0.25 vs $1) and long-context shape — Anthropic bills its 1M window at flat rates, OpenAI doubles input past 272k tokens. Underneath, the tier ladders are what most work should run on: Opus 5 ($5/$25, the Claude Code default since July 24) and Sonnet 5 ($2/$10) on one side, GPT-5.6 Sol/Terra/Luna on the other, with Sol’s promotional $4/$20 running to at least Nov 21. Full table in cost & benchmarks; lineage in GPT-5.5.
  • Product boundaries are collapsing upward: OpenAI merged Codex into the ChatGPT app (July 9) — on the OpenAI side, “which coding agent” is turning into “how you use ChatGPT.” Since Codex CLI rust-v0.153.4 (Sept 4) the bundled default model there is GPT-6 Astra, which means the product’s default now inherits Astra’s restricted cybersecurity behaviour. See Codex.

The First Cut Of The Map

Route Representative projects Typical user
Direct execution Claude Code, Aider, Codex, Kimi Code, MiMoCode, CodeWhale, ZCode, OpenCode, Gemini CLI, Qwen Code, Grok Build, Devin, Jules Someone who wants to hand a concrete coding task to an agent (see the terminal coding CLI comparison)
Agent harness framework DeepSeek Harness, Pi, jcode, OpenHands, SWE-agent, mini-swe-agent, OpenHarness, QM, Omnigent, TrueForge Someone who wants to own the agent loop, tool surface, and permissions instead of inheriting a vendor’s product — QM and Omnigent extend this to running several harnesses under one layer (see the harness comparison)
Frontier agentic model Claude Fable 5.1, Claude Opus 5, GPT-6 Astra, GPT-5.5 Someone choosing which model to wire into their own agent system or evaluating the capability ceiling of Anthropic / OpenAI surfaces — the ceilings (Fable 5.1, Astra) and the default you actually run (Opus 5) are separate decisions
Open-weights agentic model Kimi K3, GLM-5.3, DeepSeek V4, Qwen3-Coder Someone who wants frontier-class capability on weights they host and licence themselves — a different decision from picking between the closed ceilings
Agentic skills framework Superpowers Someone who wants a methodology + composable skills layer that plugs into Claude Code, Codex, Cursor, and similar agents
Workflow / orchestration layer oh-my-claudecode, oh-my-codex, Ruflo Someone who already likes Claude Code or Codex and wants stronger orchestration on top (Ruflo extends this to multi-machine federation and 100+ specialized agents)
Editor-centric AI workflow Cursor, Windsurf, Continue Someone who wants the editor itself to stay central
Review-first automation Cline, GitHub Copilot, Froge Code, CoStrict, Open Code Review Someone who wants review and human control to stay central (CoStrict adds enterprise strict-workflow + private deployment; Open Code Review is review only, tuned for precision in CI)
Managed background path Claude Managed Agents Someone who needs scheduled, cloud, or detached Anthropic workflows
General-purpose autonomous agent AutoGPT, Agent Zero, BabyAGI, Julep, GenericAgent, ml-intern, WorkBuddy, Kimi Work Someone who wants autonomous, general-purpose task execution — ml-intern is the ML-specialized variant, while WorkBuddy and Kimi Work aim the same loop at desktop knowledge work rather than repositories
Build-your-own system LangChain, LangGraph, CrewAI, LlamaIndex, Haystack, Semantic Kernel, DSPy, Pydantic AI, Microsoft Agent Framework Teams building their own agent platform instead of buying one
Runtime and tools n8n, MemGPT, Open Interpreter, LiteLLM, Flowise, CodeGraph, CLI-Anything Teams that need workflow automation, code execution, LLM gateways, agent context infrastructure, agent-driven CLIs, or visual builders
Observability and evals Langfuse Someone whose agents already run in production and needs to know what they did, what they cost, and whether quality is drifting (see observability & evaluation)
Browser agent Browser Use Someone whose task lives on a website with no API — a different question from how an agent edits files, because it decides what an agent may do as you on the open web
Self-hosted / local runtime AI Edge Gallery, Goose, Hermes Agent, OpenClaw, Mercury Agent, OpenHuman Users who need on-device privacy, long-running agents, local control, channels, devices, or personal-data life integration

Current Mainstream Coverage

77 profiled projects, grouped by what they are. Expand a group, or browse the full route/coverage tables in agents/.

Example Reading Paths

If you are still deciding where to begin, use one of these quick routes and then branch out.

If you sound like this… Follow this path What it helps you answer
I want a day-to-day coding agent and need to choose terminal vs editor AiderClaude Codeterminal coding CLI comparisonCursorClinecoding automation guide Which vendor CLI fits your model, terminal-first local loop vs editor-led flow vs approval-first control
I already like Claude Code or Codex but want stronger orchestration Claude Codeoh-my-claudecodeCodexoh-my-codexmainstream matrix When the base agent is enough and when a workflow layer actually adds value
I want to understand how the 2026 model race changes agent choice Claude Fable 5.1Claude Opus 5GPT-6 AstraCodexClaude Codemarket events How the frontier tiers and the default tier under them shift the capability ceiling, and what it means for product choice
I want a dedicated AI IDE instead of stitching tools together CursorWindsurfGitHub Copilotmainstream matrix Dedicated AI editor vs ecosystem platform
I want to hand off tickets and check back later CodexJulesDevinClaude Managed Agentsmainstream matrix Async cloud delegation vs managed background automation
I need something open-source or self-hosted AiderOpenHandsGooseHermes Agentcapabilities Terminal control, open-source execution, and local runtime ownership
I am building an internal agent stack, not buying a product LangChainLangGraphcapabilitiesmainstream matrix Framework vs runtime vs product boundaries

Disclaimer

Star counts and 7-day gains are point-in-time GitHub snapshots taken when the repo is updated; numbers shift quickly between weekly refreshes and small rounding differences are expected. Project descriptions, vendors, and capability summaries reflect public information at the time of writing and may change as projects evolve, get acquired, or pivot. This map is selection guidance — not endorsement, financial advice, or a production-readiness guarantee. Verify against each project’s own docs before committing to a choice.

View this README on GitHub

추천 도구

다른 키워드를 입력하거나 필터를 제거해 보세요.

설치

npx skillfish add weijt606/ai-agent-map