Agent framework standings

The runtimes and harnesses people actually build agents with β€” ranked by live GitHub traction, alongside published SWE-bench results for agent scaffolds. OpenClaw vs Hermes Agent vs the coding-agent field, updated daily.

GitHub metrics retrieved 2026-09-12

Framework leaderboard by GitHub stars

#FrameworkCategoryLanguageStarsForksLast pushLicenseSource
1 OpenClaw
The AI that really does things. Any OS. Any Platform. The lobster way. 🦞
Personal assistant TypeScript 389,475 81,872 2026-09-12 Other repo
2 Hermes Agent
The agent that grows with you
Personal assistant Python 244,671 50,729 2026-09-12 MIT repo
3 Claude Code
Claude Code is an agentic coding tool that lives in your terminal, understands your codebase, and helps you code faster by executing routine tasks, explaining complex code, and handling git workflows - all through natural language commands.
Coding agent Python 144,799 23,117 2026-09-11 β€” repo
4 Codex CLI
Lightweight coding agent that runs in your terminal
Coding agent Rust 123,462 19,006 2026-09-12 Apache-2.0 repo
5 Gemini CLI
An open-source AI agent that brings the power of Gemini directly into your terminal.
Coding agent TypeScript 106,930 14,558 2026-09-12 Apache-2.0 repo
6 OpenHands
πŸ™Œ OpenHands: AI-Driven Development
Coding agent TypeScript 87,577 11,469 2026-09-12 MIT repo
7 Cline
Autonomous coding agent as an SDK, IDE extension, or CLI assistant.
Coding agent TypeScript 67,863 7,331 2026-09-12 Apache-2.0 repo
8 AutoGen
A programming framework for agentic AI
Orchestration SDK Python 60,941 9,206 2026-04-15 CC-BY-4.0 repo
9 CrewAI
Framework for orchestrating role-playing, autonomous AI agents. By fostering collaborative intelligence, CrewAI empowers agents to work together seamlessly, tackling complex tasks.
Orchestration SDK Python 58,387 8,408 2026-09-11 MIT repo
10 aider
aider is AI pair programming in your terminal
Coding agent Python 48,913 4,945 2026-05-22 Apache-2.0 repo
11 LangGraph
Build resilient agents.
Orchestration SDK Python 41,492 7,011 2026-09-11 MIT repo
12 gcrab our project
Small, safe, replayable self-hosted AI agent. Single Go binary, zero deps, hard 8K context budget, policy-gated exec, append-only event log. Telegram + CLI.
Personal assistant Go 0 0 2026-09-10 MIT repo

Source: GitHub REST API, live per-repo Β· retrieved 2026-09-12. Stars measure adoption/attention, not capability β€” see methodology.

πŸ¦€ Spotlight: gcrab our project

gcrab is the token-frugal counterpoint to the big agent runtimes: a single static Go binary with zero dependencies that runs on a 512 MB VPS. A hard 8K context budget is enforced before every provider call (local-model friendly), every action lands in an append-only replayable event log, exec is policy-gated with an un-disableable deny backstop, and Telegram access is deny-by-default Β· 0 stars on GitHub.

Full disclosure: gcrab is built by the operators of this site. It receives no ranking boost anywhere on AgentLeaderboards β€” the table above sorts purely on public GitHub metrics, and gcrab's columns show exactly what the GitHub API returns.

gcrab.com β†’GitHub repo

SWE-bench Verified β€” top 20

"Checked" means the SWE-bench maintainers verified the run themselves; "self-reported" entries were submitted with logs by the authors.

#Agent + modelResolvedDateStatus
1Sonar Foundation Agent + Claude 4.5 Opus79.2%2025-12-05self-reported
2live-SWE-agent + Claude 4.5 Opus medium (20251101)79.2%2025-12-15self-reported
3TRAE + Doubao-Seed-Code78.8%2025-09-28self-reported
4live-SWE-agent + Gemini 3 Pro Preview (2025-11-18)77.4%2025-11-20self-reported
5EPAM AI/Run Developer Agent v20250719 + Claude 4 Sonnet76.8%2025-08-04self-reported
6Atlassian Rovo Dev (2025-09-02)76.8%2025-09-02self-reported
7Claude 4.5 Opus (high)76.8%2026-02-17self-reported
8ACoder76.4%2025-08-19self-reported
9Gemini 3 Flash (high)75.8%2026-02-17self-reported
10MiniMax M2.5 (high)75.8%2026-02-17self-reported
11Warp75.6%2025-09-01self-reported
12Claude 4.6 Opus75.6%2026-02-17self-reported
13TRAE + Claude Sonnet 4 + Opus 4 + Sonnet 3.7 + Gemini 2.5 Pro75.2%2025-06-12self-reported
14Harness AI74.8%2025-07-31self-reported
15Sonar Foundation Agent + Claude 4.5 Sonnet74.8%2025-11-03self-reported
16Lingxi-v1.5_claude-4-sonnet-2025051474.6%2025-07-20self-reported
17JoyCode + Claude 4 Sonnet + GPT-4.174.6%2025-09-15self-reported
18Refact.ai Agent + Claude 4 Sonnet + o4-mini74.4%2025-06-03self-reported
19Prometheus-v1.2.1 + GPT-574.4%2025-10-15self-reported
20Claude 4.5 Opus (20251101) (medium)74.4%2025-11-24checked

Source: swebench.com Β· fetched 2026-09-12, board last updated 2026-02-26 (198 days ago)

All SWE-bench boards we track

  • Multilingual: 13 entries β€” leader: Gemini 3 Flash (72.7%)
  • Test: 24 entries β€” leader: Sonar Foundation Agent + Claude 4.5 Opus (52.62%)
  • Verified: 180 entries β€” leader: Sonar Foundation Agent + Claude 4.5 Opus (79.2%)
  • Lite: 84 entries β€” leader: ExpeRepair-v1.0 + Claude 4 Sonnet (60.33%)
  • Multimodal: 22 entries β€” leader: GUIRepair + o3 (2025-04-16) (35.98%)

Full extracts: /api/v1/swebench.json