Skip to content

Arena

AI Code Review arenaAI Code Review

AI code review agents — the bots that review every pull request: summarizing changes, catching real bugs with full-codebase context, enforcing team standards, and increasingly holding the line on the flood of agent-authored code no human team could review alone — judged on review accuracy and noise discipline, codebase understanding and team memory, native GitHub/GitLab integration, in-repo configuration and custom rules, PR chat, agentic autofixes and pre-merge checks, merge gating and analytics, and IDE/CLI review surfaces. The 2026 market moved fast and the arena models it honestly: CodeRabbit raised a $143M Series C at a $1.5B valuation and rebranded around agentic change management; Graphite retired its Diamond branding into Graphite AI reviews; Qodo handed the open-source PR-Agent to a community org (The-PR-Agent/pr-agent now states it is "not the Qodo free tier") and sells credit-metered Qodo Review; Cursor’s Bugbot switched to usage-based billing (~$1.00–$1.50 per review run); cubic (formerly the mrge stacking tool, YC X25) ships custom review agents and Ultrareview. Ellipsis (pivoted to a managed cloud for coding agents) and Bito (pivoted to the Governor model router) were excluded as pivots, not omissions.

53 user stories · 318 judged cells · updated 2026-09-16 · Evidence as of 2026-09-16

Buyer checklist →Procurement report →

Leaderboard — every product ranked by evidenceLeaderboard

Rank by
1CodeRabbit logoCodeRabbit
free-tier
vs Qodo
35/100
2Qodo logoQodo
usage-based
vs CodeRabbit
37/100
3cubic logocubic
free-tier
vs CodeRabbit
36/100
4Greptile logoGreptile
free-tier
vs CodeRabbit
39/100
5Graphite logoGraphite
free-tier
vs CodeRabbit
30/100
6Cursor Bugbot logoCursor Bugbot
usage-based
vs CodeRabbit
26/100

Best by user type — persona-weighted winnersBest by user type

Per persona, the product with the highest persona-weighted coverage over just that persona's stories — not the same ranking as the overall PA Score leaderboard above.

Best for developer

CodeRabbit logo

CodeRabbit

68/100

Runner-up: cubic logo cubic (64/100)

12 developer stories scored

Best for engineering-lead

CodeRabbit logo

CodeRabbit

57/100

Runner-up: cubic logo cubic (42/100)

7 engineering-lead stories scored

Best for security-engineer

Cursor Bugbot logo

Cursor Bugbot

80/100

Runner-up: Greptile logo Greptile (60/100)

1 security-engineer story scored

Best for ai-native

CodeRabbit logo

CodeRabbit

40/100

Runner-up: cubic logo cubic (37/100)

33 ai-native stories scored

Story matrix — every product × every judged storyStory matrix

53/53 stories shown · legend

Agenticness — how well agents can access and operate the productAgenticness

Agent access

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Agenticness — how well agents can access and operate the productPoint an agent at llms.txt or agent-oriented docsai-native
fullT
9/10
fullT
8/10
partialT
5/10
fullT
8/10
fullT
8/10
fullT
9/10
Agenticness — how well agents can access and operate the productRun the product headlessly / in CI for automationai-native
partialT
6/10
fullT
8/10
partialT
4/10
fullT
7/10
partialC
6/10
partialT
6/10
Agenticness — how well agents can access and operate the productPlug MCP servers into this product so it can use their toolsai-native
fullC
7/10
none
0/10
none
0/10
none
0/10
n/a
none
0/10
Agenticness — how well agents can access and operate the productConnect an agent via an official MCP serverai-native
none
0/10
fullT
8/10
fullT
6/10
fullT
7/10
n/a
fullT
8/10
Agenticness — how well agents can access and operate the productUse an official CLIai-native
fullT
8/10
fullT
8/10
fullT
9/10
fullT
8/10
n/a
fullT
8/10
Agenticness — how well agents can access and operate the productDrive the product through a documented public APIai-native
partialT
6/10
partialT
6/10
partialT
5/10
partialT
6/10
none
0/10
partialT
6/10
Agenticness — how well agents can access and operate the productIssue scoped/least-privilege API credentials for an agentai-native
none
0/10
none
0/10
none
0/10
none
0/10
n/a
none
0/10
Agenticness — how well agents can access and operate the productBuild against official SDKsai-native
none
0/10
none
0/10
partialT
6/10
none
0/10
n/a
none
0/10
Agenticness — how well agents can access and operate the productSubscribe to events via webhooksai-native
none
0/10
none
0/10
none
0/10
none
0/10
none
0/10
none
0/10

Agentic features

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Agenticness — how well agents can access and operate the productGet AI-generated insights and suggestions from my data inside the productai-native
fullX
9/10
disputedD
6/10
fullX
8/10
fullX
8/10
fullX
8/10
fullX
9/10
Agenticness — how well agents can access and operate the productSet up automations that run autonomously in the backgroundai-native
fullC
7/10
fullC
7/10
partialC
6/10
partialC
6/10
fullC
7/10
fullC
8/10
Agenticness — how well agents can access and operate the productDelegate tasks to a built-in AI assistant inside the productai-native
fullX
8/10
disputedD
5/10
fullX
8/10
fullX
7/10
partialC
5/10
fullX
7/10
Agenticness — how well agents can access and operate the productOperate the product with natural-language commandsai-native
fullC
7/10
partialX
5/10
fullC
7/10
partialC
6/10
partialC
6/10
fullC
7/10

Api quality

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Agenticness — how well agents can access and operate the productExplore an interactive API reference with runnable examplesai-native
none
0/10
none
0/10
none
0/10
none
0/10
n/a
none
0/10
Agenticness — how well agents can access and operate the productDownload a machine-readable API spec (OpenAPI or equivalent)ai-native
fullT
9/10
none
0/10
none
0/10
none
0/10
n/a
none
0/10
Agenticness — how well agents can access and operate the productTest against a sandbox environment without touching production dataai-native
none
0/10
fullC
6/10
none
0/10
n/a
n/a
n/a
Agenticness — how well agents can access and operate the productRely on versioned APIs with a documented deprecation policyai-native
none
0/10
none
0/10
none
0/10
none
0/10
n/a
none
0/10

Autofix agents — stories about autofix agents in this arenaAutofix agents

Ai authored

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Autofix agents — stories about autofix agents in this arenaThe reviewer holds the line on AI-generated PRs — it verifies agent-authored code at a volume no human team could reviewai-native
disputedD
6/10
fullX
8/10
partialX
6/10
partialX
7/10
disputedD
6/10
fullX
8/10

Checks

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Autofix agents — stories about autofix agents in this arenaI define custom agentic pre-merge checks in plain language — 'docs updated', 'tests cover new paths' — that run on every PRai-native
partialC
6/10
partialC
6/10
partialC
5/10
partialC
5/10
partialC
6/10
fullX
7/10

Fixes

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Autofix agents — stories about autofix agents in this arenaI turn a review finding into an applied fix — a committed patch or an agent-generated follow-up — without leaving the PRdeveloper
fullC
8/10
fullC
8/10
fullC
7/10
fullC
7/10
partialX
6/10
fullC
8/10

Handoff

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Autofix agents — stories about autofix agents in this arenaReview findings hand off cleanly to my coding agent — copyable fix prompts or direct integration with Claude Code, Cursor, or Codexai-native
partialC
7/10
fullC
9/10
partialT
5/10
fullT
7/10
partialC
6/10
fullT
8/10

Automation depth — how much of the product can run unattendedAutomation depth

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Automation depth — how much of the product can run unattendedPerform bulk operations across many items at onceai-native
partialC
4/10
partialC
6/10
partialX
6/10
partialC
5/10
partialC
4/10
partialC
5/10
Automation depth — how much of the product can run unattendedDefine rules that trigger actions automatically on eventsai-native
partialC
7/10
partialC
5/10
partialC
5/10
partialC
6/10
partialC
6/10
partialX
5/10
Automation depth — how much of the product can run unattendedSchedule recurring jobs or workflowsai-native
partialC
3/10
none
0/10
none
0/10
none
0/10
none
0/10
n/a
Automation depth — how much of the product can run unattendedVersion, review, and roll back my automationsai-native
partialC
3/10
partialC
4/10
partialC
5/10
partialC
4/10
n/a
partialC
3/10

Codebase understanding — how deeply the tool maps your repo — cross-file context, architecture awareness, historyCodebase understanding

Context

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Codebase understanding — how deeply the tool maps your repo — cross-file context, architecture awareness, historyThe reviewer understands changes that span multiple repositories or a large monorepo and reviews them coherentlyengineering-lead
partialC
7/10
none
0/10
partialC
4/10
partialC
7/10
none
0/10
partialC
7/10
Codebase understanding — how deeply the tool maps your repo — cross-file context, architecture awareness, historyReview comments reflect the whole repository — call sites, related modules, existing conventions — not just the changed hunksdeveloper
fullC
7/10
partialX
6/10
fullC
7/10
fullC
8/10
partialC
4/10
partialC
7/10

Memory

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Codebase understanding — how deeply the tool maps your repo — cross-file context, architecture awareness, historyThe reviewer builds a persistent memory of my team's conventions and past review decisions and applies it to future PRsai-native
fullC
8/10
fullC
7/10
partialC
4/10
fullC
7/10
partialC
4/10
fullX
8/10

Interaction — how you steer it — commands, replies, review conversations, configurability in the loopInteraction

Chat

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Interaction — how you steer it — commands, replies, review conversations, configurability in the loopI reply to the reviewer in the PR thread to ask questions, get explanations, or issue commands — and it answers in contextdeveloper
fullC
8/10
partialX
5/10
fullC
7/10
fullC
7/10
partialC
4/10
fullX
8/10

Control

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Interaction — how you steer it — commands, replies, review conversations, configurability in the loopI control when reviews run — skip drafts, trigger on demand, filter by branch or label — so the bot shows up only when wanteddeveloper
fullC
8/10
partialC
5/10
partialC
4/10
partialC
6/10
partialC
5/10
partialC
4/10

Openness — open source, data portability, and self-hosting storiesOpenness

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Openness — open source, data portability, and self-hosting storiesDo everything through the API that I can do in the UIai-native
partialT
3/10
partialT
5/10
partialT
4/10
partialT
4/10
none
0/10
partialT
5/10
Openness — open source, data portability, and self-hosting storiesExport all of my data in open formats and leaveai-native
partialC
2/10
none
0/10
partialX
4/10
none
0/10
none
0/10
none
0/10
Openness — open source, data portability, and self-hosting storiesRead the product's source under an open licenseai-native
none
0/10
none
0/10
none
0/10
none
0/10
none
0/10
none
0/10
Openness — open source, data portability, and self-hosting storiesSelf-host the core productai-native
fullC
6/10
fullC
7/10
none
0/10
fullC
7/10
n/a
none
0/10

Pr integration — stories about pr integration in this arenaPr integration

Platforms

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Pr integration — stories about pr integration in this arenaThe reviewer installs as a GitHub/GitLab app and posts reviews as native inline comments on my pull requests within minutesdeveloper
fullX
8/10
partialX
7/10
partialX
6/10
fullC
7/10
fullX
8/10
fullX
8/10

Suggestions

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Pr integration — stories about pr integration in this arenaReview comments include committable suggested diffs I can apply with one clickdeveloper
fullC
9/10
partialC
5/10
fullC
7/10
fullC
7/10
partialC
5/10
fullC
7/10

Summaries

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Pr integration — stories about pr integration in this arenaEvery PR gets an auto-generated summary and change walkthrough so human reviewers orient fastdeveloper
fullX
9/10
fullX
7/10
partialC
6/10
fullC
9/10
none
0/10
fullX
9/10

Updates

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Pr integration — stories about pr integration in this arenaPushing new commits triggers an incremental re-review that tracks what was fixed instead of repeating old commentsdeveloper
fullC
8/10
partialC
7/10
none
0/10
partialC
4/10
fullT
7/10
partialC
6/10

Privacy posture — data-handling and privacy storiesPrivacy posture

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Privacy posture — data-handling and privacy storiesChoose where my data is stored (region/residency)ai-native
partialC
5/10
partialC
6/10
none
0/10
partialC
5/10
none
0/10
none
0/10
Privacy posture — data-handling and privacy storiesPrevent my data from being used to train AI modelsai-native
fullC
9/10
partialC
4/10
fullC
8/10
fullC
8/10
partialC
6/10
fullC
8/10
Privacy posture — data-handling and privacy storiesControl data retention and deletionai-native
partialC
6/10
partialC
5/10
none
0/10
partialC
4/10
partialC
4/10
partialC
3/10
Privacy posture — data-handling and privacy storiesOpt out of telemetry and usage trackingai-native
partialC
5/10
partialC
4/10
none
0/10
none
0/10
partialC
4/10
none
0/10

Quality gates — stories about quality gates in this arenaQuality gates

Analytics

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Quality gates — stories about quality gates in this arenaI see dashboards of findings, acceptance rates, and review coverage across my orgengineering-lead
fullC
8/10
none
0/10
partialC
3/10
partialC
6/10
partialC
4/10
fullC
7/10

Gates

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Quality gates — stories about quality gates in this arenaThe reviewer can gate merges — a required status check or blocking review that enforces resolution of critical findingsengineering-lead
fullC
7/10
none
0/10
partialC
5/10
none
0/10
none
0/10
none
0/10

Review accuracy — stories about review accuracy in this arenaReview accuracy

Detection

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Review accuracy — stories about review accuracy in this arenaThe reviewer catches real bugs in my PR — logic errors, race conditions, broken edge cases — not just style nitsdeveloper
disputedD
6/10
partialX
6/10
partialX
4/10
partialX
5/10
disputedD
6/10
fullX
7/10

Learning

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Review accuracy — stories about review accuracy in this arenaPush back on a bad review comment and the reviewer learns — it stops repeating the same rejected feedbackdeveloper
partialX
7/10
partialC
7/10
partialC
4/10
none
0/10
partialX
4/10
fullX
7/10

Noise

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Review accuracy — stories about review accuracy in this arenaThe reviewer keeps noise low — few false positives, deduplicated comments, severity labels — so my team doesn't tune it outengineering-lead
disputedD
4/10
partialX
6/10
partialX
5/10
partialX
6/10
partialX
5/10
partialX
6/10

Security

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Review accuracy — stories about review accuracy in this arenaReviews flag security problems in the diff — injection risks, leaked secrets, insecure patterns — alongside functional bugssecurity-engineer
partialX
6/10
fullX
6/10
none
0/10
partialC
5/10
fullT
8/10
partialX
6/10

Surfaces — where it meets your workflow — IDE, CLI, web, PR comments, CI checksSurfaces

Cli

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Surfaces — where it meets your workflow — IDE, CLI, web, PR comments, CI checksI run reviews from a CLI against local diffs or in CI scripts, with machine-readable output my tooling can consumedeveloper
partialT
5/10
partialT
5/10
none
0/10
partialT
5/10
none
0/10
partialT
5/10

Ide

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Surfaces — where it meets your workflow — IDE, CLI, web, PR comments, CI checksI get the same review inside my IDE before I push, catching issues while the code is still in my editordeveloper
fullT
8/10
partialT
6/10
none
0/10
partialT
6/10
none
0/10
fullT
7/10

Workflow config — stories about workflow config in this arenaWorkflow config

Config

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Workflow config — stories about workflow config in this arenaI configure the reviewer with a versioned config file in my repo — path filters, per-path instructions, review profilesengineering-lead
fullC
8/10
partialC
6/10
partialC
5/10
partialC
6/10
partialC
5/10
fullX
7/10

Governance

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Workflow config — stories about workflow config in this arenaI roll out org-level review defaults across hundreds of repos and manage exceptions centrallyengineering-lead
partialC
7/10
partialC
7/10
partialC
5/10
fullC
8/10
partialC
5/10
partialC
5/10

Rules

StoryPersona
CodeRabbit logoCodeRabbit
Greptile logoGreptile
Graphite logoGraphite
Qodo logoQodo
Cursor Bugbot logoCursor Bugbot
cubic logocubic
Workflow config — stories about workflow config in this arenaI encode my team's own review guidelines — natural-language rules, AST patterns, or linked style guides — and the reviewer enforces themengineering-lead
fullC
9/10
fullC
8/10
partialC
6/10
partialC
7/10
fullC
7/10
partialX
7/10
Verdict✓ fullclear evidence~ partialwith caveats! disputedevidence conflicts— noneno evidence foundn/aquestion doesn't apply to this kind of product
ProofT probedtested by usX communityusers back itC claimedvendor claim onlyD contradictedevidence disagrees⚿ auth-gatedprobe hit a live sign-in wall — verified reachable, untestable keylessly
quality 0–10 · PA Score /100 · A–D = evidence confidence · full guide

Adjacent arenas — categories often shopped togetherAdjacent arenas

Shopping this category often means shopping these too.