Skip to content

AI Coding Agents Arena

Claude Code vs Google Antigravity

Claude Code wins · 4117 (13 drawn)

Agenticness — how well agents can access and operate the productAgenticness

How well agents can access and operate the product

Agent access

  1. ai-native userPoint an agent at llms.txt or agent-oriented docs

    weight 2 · round to Google Antigravity
    Claude Codepartialprobed5/10

    Claude Code itself ships llms.txt files (docs.claude.com/llms.txt, code.claude.com/docs/llms.txt) confirming it is agent-oriented-docs-aware for its own product, and its agentic search/MCP tooling means it can fetch and consume arbitrary web docs including llms.txt if pointed at them via URL fetch or MCP. However, there is no explicit documented feature or first-party guidance describing 'point Claude Code at llms.txt of a third-party site' as a supported workflow. missing for 10: explicit product feature/docs describing consuming arbitrary llms.txt/agent-oriented docs as a first-class capability, independent hands-on confirmation of this specific use case.

    • [probe] PROBE llms.txt: HTTP 200 at https://docs.claude.com/llms.txt # Anthropic Developer Documentation This file provides an overview of the Anth…
    • [probe] PROBE docs-md: HTTP 200 at https://docs.claude.com/en/docs/claude-code/overview.md > ## Documentation Index > Fetch the complete documentati…
    • [claimed-docs] Claude Code maps and explains entire codebases in a few seconds. It uses agentic search to understand project structure and dependencies wit…
    • [claimed-docs] Claude Code can connect to hundreds of external tools and data sources through the Model Context Protocol (MCP)
    Google Antigravityfullprobed9/10

    Antigravity hosts a working llms.txt (HTTP 200) describing itself, and provides markdown-formatted docs pages (e.g. getting-started.md) that an agent can fetch directly, confirming genuine agent-oriented documentation support. Missing for 10: independent third-party confirmation that agents actually consume these successfully in practice.

    • [probe] PROBE llms.txt: HTTP 200 at https://antigravity.google/llms.txt # Google Antigravity > Google Antigravity is an advanced agentic coding pla…
    • [probe] PROBE docs-md: HTTP 200 at https://antigravity.google/docs/getting-started.md # Getting Started with Antigravity 2.0 ### Download Visit [a…
    • [claimed-docs] Visit antigravity.google/download to download Google Antigravity 2.0. Select your operating system below
  2. ai-native userRun the product headlessly / in CI for automation

    weight 2 · round to Claude Code
    Claude Codefullclaimed9/10

    Docs explicitly describe running Claude Code in CI (GitHub Actions/GitLab CI/CD) for automated code review and issue triage, piping logs into it, and scheduled/headless runs for repeated automation tasks, plus GitHub Action integration for automatic PR review. This directly matches the headless/CI automation story with strong first-party documentation. Missing for 10: independent/hands-on confirmation of a working CI pipeline (community evidence is silent on CI usage specifically).

    • [claimed-docs] Claude Code is composable and follows the Unix philosophy. Pipe logs into it, run it in CI, or chain it with other tools
    • [claimed-docs] Run Claude on a schedule to automate work that repeats: morning PR reviews, overnight CI failure analysis, weekly dependency audits, or sync…
    • [claimed-docs] In CI, you can automate code review and issue triage with GitHub Actions or GitLab CI/CD.
    • [claimed-docs] Get automatic code review on every PR | GitHub Code Review
    Google Antigravityfullclaimed7/10

    Antigravity CLI has a documented headless mode explicitly for scripting agent tasks, CI pipeline integration, and machine-readable output (antigravity-docs-50), plus scheduled/cron tasks and background subagents support agentic automation outside interactive UI. Missing for 10: independent hands-on CI usage reports, concrete CI config examples/output schema, and no community corroboration of headless/CI use in practice.

    • [claimed-docs] Run Antigravity CLI non-interactively to script agent tasks, integrate with CI pipelines, and capture machine-readable output.
    • [claimed-docs] Automate routine checks with Scheduled Tasks, simply define a cron schedule and the agents start and run autonomously in the background.
    • [claimed-docs] Edit, orchestrate, and build all in natural language. Tell your agents what you need, and they’ll work on getting it done.
    • [claimed-docs] Create or download fully customizable skills to further your agent’s autonomy and transform how you get work done.
  3. ai-native userPlug MCP servers into this product so it can use their tools

    weight 3 · round to Claude Code
    Claude Codefullclaimed9/10

    Claude Code has extensive first-party MCP documentation showing users can add MCP servers (e.g. `claude mcp add --transport http notion ...`), supporting stdio/HTTP transports, connecting to hundreds of external tools like Jira, Slack, Google Drive, Postgres, and even scaffolding new servers via a dev plugin. This is well corroborated across multiple doc pages with concrete CLI examples and use cases. Missing for 10: independent/hands-on community confirmation specifically of MCP tool usage (community evidence covers other topics, not MCP plugging in).

    • [claimed-docs] With MCP, Claude Code can read your design docs in Google Drive, update tickets in Jira, pull data from Slack, or use your own custom toolin…
    • [claimed-docs] Claude Code can connect to hundreds of external tools and data sources through the Model Context Protocol (MCP)
    • [claimed-docs] Implement features from issue trackers: "Add the feature described in JIRA issue ENG-4521 and create a PR on GitHub."
    • [claimed-docs] claude mcp add --transport http notion https://mcp.notion.com/mcp
    • [claimed-docs] Stdio servers run as local processes on your machine. They're ideal for tools that need direct system access or custom scripts.
    • [claimed-docs] You can also have Claude scaffold a server for you with the official mcp-server-dev plugin
    • [claimed-docs] an MCP server can also act as a channel that pushes messages into your session, so Claude reacts to Telegram messages, Discord chats, or web…
    Google Antigravityfullclaimed8/10

    Antigravity has explicit, dedicated MCP documentation stating MCP lets it 'fetch structured context directly or execute safe actions on your behalf' and that it 'securely connects to local developer tools, databases, file parsers, and external remote APIs' via MCP, plus CLI/SDK support for configuring MCP servers (slash commands, plugins bundling MCP servers, layering MCP servers in the Agent SDK). Missing for 10: independent hands-on confirmation of successfully connecting a third-party MCP server and using its tools in a real workflow.

    • [claimed-docs] MCP lets Antigravity fetch structured context directly or execute safe actions on your behalf when needed.
    • [claimed-docs] lets AI agents and editors securely connect to local developer tools, databases, file parsers, and external remote APIs
    • [claimed-docs] Access plugins, MCP, skills, and hooks configurations instantly via slash commands, quickly enhancing your workflow.
    • [claimed-docs] Layer custom Python callables, Model Context Protocol (MCP) servers, and reusable agent skills over our built-in filesystem and terminal too…
    • [claimed-docs] Plugins are namespaced bundles that allow you to extend Antigravity’s capabilities by grouping skills, rules, MCP servers, and hooks into a …
  4. ai-native userUse an official CLI

    weight 2 · round to Claude Code
    Claude Codefullprobed9/10

    Claude Code is itself an official CLI tool with documented install (curl install script), usage (`cd project && claude`), cross-platform support (macOS/Linux/Windows), and deep terminal-native workflows (git, MCP, hooks, CI). GitHub repo and docs confirm first-party CLI status with active community usage corroborating real-world use. Missing for 10: independent benchmarking of CLI robustness/UX beyond mixed community sentiment.

    • [claimed-docs] cd your-project claude
    • [claimed-docs] curl -fsSL https://claude.ai/install.sh | bash
    • [claimed-docs] Available for macOS, Linux, and Windows.
    • [github] Use it in your terminal, IDE, or tag @claude on Github.
    • [github] helps you code faster by executing routine tasks, explaining complex code, and handling git workflows -- all through natural language comman…
    • [probe] official CLI documented at https://code.claude.com/docs/en/setup
    Google Antigravityfullprobed8/10

    Google Antigravity ships an official CLI with dedicated docs (antigravity-cli product page, headless/non-interactive mode for CI, sandboxing, vim mode, gcli migration), enabling natural-language orchestration of parallel agents, slash commands, and MCP/plugin config — clearly AI-native and agentic. Community evidence corroborates the CLI works in practice alongside VSCode. Missing for 10: independent deep-dive review of CLI-specific reliability/performance beyond a single community mention.

    • [claimed-docs] Edit, orchestrate, and build all in natural language. Tell your agents what you need, and they’ll work on getting it done.
    • [claimed-docs] Have multiple agents working in parallel, so larger tasks get tackled faster.
    • [claimed-docs] Navigate your entire workflow via standard terminal shortcuts: adjust permissions, themes, and preferences via /config and type /keybindings…
    • [claimed-docs] Access plugins, MCP, skills, and hooks configurations instantly via slash commands, quickly enhancing your workflow.
    • [claimed-docs] the CLI automatically detects your existing profiles. An interactive checklist prompts you to choose which assets to migrate
    • [claimed-docs] Sensitive files like ~/.ssh and .env are blocked, anything not explicitly mounted is invisible inside the sandbox
    • [claimed-docs] Vim editor mode replaces the editing model in every multi-line input surface of the CLI
    • [claimed-docs] Run Antigravity CLI non-interactively to script agent tasks, integrate with CI pipelines, and capture machine-readable output.
    • [community] I much prefer using Gemini CLI in combination with vscode. It works like a charm. Now, I'll do the same with Antigravity CLI and vscode. It …
    • [probe] official CLI documented at https://antigravity.google/product/antigravity-cli
  5. ai-native userDrive the product through a documented public API

    weight 3 · round to Claude Code
    Claude Codefullclaimed8/10

    Claude Code exposes multiple documented programmatic surfaces: the Agent SDK for building custom agents with full control over orchestration/tools/permissions, a CLI (claude, claude mcp serve) that can be scripted/piped/run in CI, and ANTHROPIC_API_KEY-based direct API access, all documented in first-party docs. This goes beyond a closed UI and gives AI-native users documented, programmatic control paths. Missing for 10: independent/hands-on validation of the Agent SDK's API surface and no explicit REST/OpenAPI reference beyond the SDK and CLI docs.

    • [claimed-docs] the Agent SDK lets you build your own agents powered by Claude Code's tools and capabilities, with full control over orchestration, tool acc…
    • [claimed-docs] Use Claude Code as an MCP server. You can use Claude Code itself as an MCP server that other applications can connect to: claude mcp serve (…
    • [claimed-docs] If you've set the ANTHROPIC_API_KEY environment variable, Claude Code skips the login prompt and asks you to approve the key instead.
    • [claimed-docs] ANTHROPIC_API_KEY environment variable. Sent as the X-Api-Key header. Use this for direct Anthropic API access with a key from the Claude Co…
    • [claimed-docs] Claude Code is composable and follows the Unix philosophy. Pipe logs into it, run it in CI, or chain it with other tools
    • [claimed-docs] Claude Code can connect to hundreds of external tools and data sources through the Model Context Protocol (MCP)
    Google Antigravitypartialprobed6/10

    Antigravity documents an Agent SDK (Python) exposing the same tools/agent loop/context management as the app, plus a CLI headless mode for scripting and CI integration, both of which let an AI-native user drive the product programmatically. However, there is no evidence of a formal public REST/HTTP API — a probe for OpenAPI/swagger specs returned 404 on all candidate paths, so the 'documented public API' is limited to SDK/CLI surfaces rather than a conventional API contract. Missing for 10: a documented REST/HTTP API or OpenAPI spec, independent third-party confirmation of SDK usage/stability.

    • [claimed-docs] The Agent SDK gives you the same tools, agent loop, and context management that power Google Antigravity, programmable in Python.
    • [claimed-docs] Layer custom Python callables, Model Context Protocol (MCP) servers, and reusable agent skills over our built-in filesystem and terminal too…
    • [claimed-docs] Build AI agents that autonomously read files, run commands, edit code, and more.
    • [claimed-docs] Run Antigravity CLI non-interactively to script agent tasks, integrate with CI pipelines, and capture machine-readable output.
    • [probe] PROBE openapi: all candidate paths 404 (https://antigravity.google/openapi.json, https://antigravity.google/swagger.json, https://antigravit…
  6. ai-native userIssue scoped/least-privilege API credentials for an agent

    weight 2 · round to Claude Code
    Claude Codepartialclaimed4/10

    Enterprise IAM docs mention role-based permissions, managed policy settings, and SSO/domain capture for org-wide configurations, plus sandboxing controls that restrict file/network access at runtime, suggesting some least-privilege controls exist. However, there is no explicit documentation of issuing scoped or limited-permission API keys/credentials specifically for an agent's use. Missing for 10: explicit scoped API key creation/management flow, granular credential scoping documentation, and independent verification of least-privilege credential issuance.

    • [claimed-docs] Claude for Enterprise: adds SSO, domain capture, role-based permissions, compliance API, and managed policy settings for organization-wide C…
    • [claimed-docs] Single sign-on (SSO/SAML) and domain capture
    • [claimed-docs] Learn how Claude Code's sandboxed Bash tool provides filesystem and network isolation for safer, more autonomous agent execution. The Bash s…
    • [claimed-docs] If you've set the ANTHROPIC_API_KEY environment variable, Claude Code skips the login prompt and asks you to approve the key instead.
    • [claimed-docs] You can sign in to your Console account without creating an API key, even when your organization doesn't let developers create them.
    Google Antigravitydisputedcontradicted4/10

    Antigravity documents permission tiers (Deny/Ask/Allow) and a CLI sandbox that explicitly blocks access to sensitive files like .env and ~/.ssh, which is the closest analog to least-privilege credential scoping for an agent (docs-23, docs-42, docs-48). However, independent reports document a concrete bypass: Antigravity's own setting disallowing .env access was circumvented via prompt injection to exfiltrate secrets, and a default allowlisted domain (webhook.site) was used as an exfiltration channel — directly contradicting the claimed least-privilege protection (antigravity-comm-11, antigravity-comm-12). There is no evidence of a true scoped API-credential-issuance mechanism (e.g., minting restricted API keys/tokens for an agent); missing for 10: actual credential/token scoping API, third-party security audit confirming the sandbox holds, and any documented remediation.

    • [claimed-docs] Permissions are evaluated across three distinct access lists: Deny... Ask... Allow
    • [claimed-docs] Permissions are evaluated across three distinct access lists: Deny...Ask...Allow
    • [claimed-docs] Sensitive files like ~/.ssh and .env are blocked, anything not explicitly mounted is invisible inside the sandbox
    • [community] Google Antigravity exfiltrates data via indirect prompt injection attack: Gemini is not supposed to have access to .env files with default s…
    • [community] The default Allowlist provided with Antigravity includes 'webhook.site', which was used as an exfiltration vector for secrets.
  7. ai-native userBuild against official SDKs

    weight 2 · round to Claude Code
    Claude Codefullclaimed8/10

    Claude Code offers the official Agent SDK, letting developers build their own agents with full control over orchestration, tool access, and permissions, on top of Claude Code's tools/capabilities — a direct SDK for AI-native builders. This is backed by first-party docs and complemented by API-key-based programmatic access (ANTHROPIC_API_KEY) for direct integration. Missing for 10: independent/hands-on developer reports building production apps with the Agent SDK, and deeper docs on SDK language coverage/versioning.

    • [claimed-docs] the Agent SDK lets you build your own agents powered by Claude Code's tools and capabilities, with full control over orchestration, tool acc…
    • [claimed-docs] If you've set the ANTHROPIC_API_KEY environment variable, Claude Code skips the login prompt and asks you to approve the key instead.
    • [claimed-docs] ANTHROPIC_API_KEY environment variable. Sent as the X-Api-Key header. Use this for direct Anthropic API access with a key from the Claude Co…
    Google Antigravityfullprobed7/10

    Google explicitly documents an official Agent SDK ('same tools, agent loop, and context management that power Antigravity, programmable in Python') supporting custom Python callables, MCP servers, skills, and multimedia inputs, which directly satisfies building against an official SDK. Missing for 10: independent/hands-on developer confirmation of the SDK working as documented, and no public API reference/OpenAPI spec was found (probe returned 404s), so depth of documentation beyond marketing copy is unverified.

    • [claimed-docs] The Agent SDK gives you the same tools, agent loop, and context management that power Google Antigravity, programmable in Python.
    • [claimed-docs] Layer custom Python callables, Model Context Protocol (MCP) servers, and reusable agent skills over our built-in filesystem and terminal too…
    • [claimed-docs] Pass rich multimedia file attachments (images, videos, audio, and documents) to the agent alongside textual instruction prompt lists.
    • [claimed-docs] Build AI agents that autonomously read files, run commands, edit code, and more.
    • [probe] PROBE openapi: all candidate paths 404 (https://antigravity.google/openapi.json, https://antigravity.google/swagger.json, https://antigravit…
  8. ai-native userSubscribe to events via webhooks

    weight 2 · round to Claude Code
    Claude Codepartialclaimed4/10

    Claude Code doesn't offer a first-party webhook subscription feature, but docs note that an MCP server can act as a channel pushing events—including webhook events—into a Claude Code session while the user is away, enabling indirect event subscription via custom MCP tooling. Missing for 10: a native/first-party webhook subscription mechanism, official documentation or example of setting up webhook-triggered sessions, and independent confirmation this works in practice.

    • [claimed-docs] an MCP server can also act as a channel that pushes messages into your session, so Claude reacts to Telegram messages, Discord chats, or web…
    • [claimed-docs] an MCP server can also act as a channel that pushes messages into your session, so Claude reacts to Telegram messages, Discord chats, or web…
    • [claimed-docs] Claude Code can connect to hundreds of external tools and data sources through the Model Context Protocol (MCP)
    Google Antigravitynone0/10

    No evidence of any webhook subscription mechanism; Antigravity is an IDE/CLI/agent platform with hooks, MCP, and scheduled tasks, but nothing about outbound event subscriptions via webhooks. Even the openapi probe returned 404s, indicating no public API surface for such integration.

    • [probe] PROBE openapi: all candidate paths 404 (https://antigravity.google/openapi.json, https://antigravity.google/swagger.json, https://antigravit…

Agentic features

  1. ai-native userGet AI-generated insights and suggestions from my data inside the product

    weight 2 · round to Claude Code
    Claude Codefullclaimed7/10

    Claude Code generates AI-driven insights and suggestions from a user's data: it maps/explains entire codebases automatically, reviews code and PRs for security issues with explanations, and via MCP can query databases (e.g., PostgreSQL) or pull data from Slack/Jira/Google Drive to answer questions and suggest actions. This is all documented first-party capability with concrete examples (codebase mapping, automatic PR/security review, data queries via MCP). missing for 10: independent/hands-on corroboration specifically validating the quality of data-driven insights (community evidence is mostly about coding reliability, not insight generation), and no dedicated analytics/dashboard-style insight feature beyond code/data-source querying.

    • [claimed-docs] Claude Code maps and explains entire codebases in a few seconds. It uses agentic search to understand project structure and dependencies wit…
    • [claimed-docs] Claude helps security teams and developers by reviewing code for security issues, drafts patches, and explains the risk in language your who…
    • [claimed-docs] Get automatic code review on every PR | GitHub Code Review
    • [claimed-docs] Find emails of 10 random users who used feature ENG-4521, based on our PostgreSQL database.
    • [claimed-docs] With MCP, Claude Code can read your design docs in Google Drive, update tickets in Jira, pull data from Slack, or use your own custom toolin…
    • [claimed-docs] Claude Code can read your design docs in Google Drive, update tickets in Jira, pull data from Slack, or use your own custom tooling.
    Google Antigravitypartialcommunity6/10

    Antigravity's editor/agent generates code suggestions, tab-autocompletion, and rich 'Artifacts' (implementation plans, diagrams, code diffs) that surface AI-derived insights from the user's codebase (antigravity-docs-6, -24, -30, -41), fitting the 'insights from data' story in a coding context. However, this is inference-in-editor suggestion generation rather than dedicated analytics/insight dashboards, and community hands-on reports raise real quality concerns ('the model was not good and slow, the harness was not good' — antigravity-comm-8), undercutting confidence in consistent insight quality. Missing for 10: no evidence of dedicated data-analysis/insight-summarization features beyond code artifacts, and no independent corroboration that suggestions are reliably high quality.

    • [claimed-docs] Google Antigravity's Editor view offers tab autocompletion, natural language code commands, and a configurable, and context-aware configurab…
    • [claimed-docs] Planning Mode: The agent plans thoroughly before executing tasks... produces structured implementation plans called Artifacts
    • [claimed-docs] Artifacts include rich markdown plans (Implementation Plans), code diffs, architecture diagrams, images, and browser recordings.
    • [claimed-docs] offers tab autocompletion, natural language code commands, and a configurable, and context-aware configurable agent
    • [community] It's not even good, honestly. I was using it for couple weeks before dropping that 2 months ago. The model was not good and slow, the harnes…
  2. ai-native userSet up automations that run autonomously in the background

    weight 2 · round drawn
    Claude Codefullclaimed8/10

    Docs describe explicit background-automation features: scheduled runs for recurring tasks (morning PR reviews, overnight CI analysis, weekly audits), cloud/web sessions for long-running tasks you check back on, GitHub Actions/GitLab CI integration for automated review/triage, and MCP servers that push events (Telegram/Discord/webhooks) into a session while the user is away. Missing for 10: independent/hands-on corroboration that scheduled or background jobs run reliably unattended over time (only first-party docs cited).

    • [claimed-docs] Run Claude on a schedule to automate work that repeats: morning PR reviews, overnight CI failure analysis, weekly dependency audits, or sync…
    • [claimed-docs] Step away from your desk and keep working from your phone or any browser with Remote Control
    • [claimed-docs] Kick off a long-running task on the web or the Claude mobile app, then pull it into your terminal with `claude --teleport`.
    • [claimed-docs] Review diffs visually, run multiple sessions side by side, schedule recurring tasks, and kick off cloud sessions.
    • [claimed-docs] Kick off long-running tasks and check back when they're done, work on repos you don't have locally, or run multiple tasks in parallel.
    • [claimed-docs] Run Claude Code in your browser with no local setup. Kick off long-running tasks and check back when they're done, work on repos you don't h…
    • [claimed-docs] In CI, you can automate code review and issue triage with GitHub Actions or GitLab CI/CD.
    • [claimed-docs] an MCP server can also act as a channel that pushes messages into your session, so Claude reacts to Telegram messages, Discord chats, or web…
    • [claimed-docs] an MCP server can also act as a channel that pushes messages into your session, so Claude reacts to Telegram messages, Discord chats, or web…
    Google Antigravityfullclaimed8/10

    Docs explicitly describe Scheduled Tasks with cron schedules that let agents 'start and run autonomously in the background' (docs-2), plus related capabilities like scheduling messages to agents while away (docs-39), isolated background worktrees (docs-38), and headless/non-interactive CLI runs for CI automation (docs-50). This directly matches the story of autonomous background automations. Missing for 10: independent/hands-on verification that scheduled background tasks work reliably (community evidence focuses on other bugs/exfiltration issues, not scheduling specifically), and more detail on monitoring/error-handling for unattended runs.

    • [claimed-docs] Automate routine checks with Scheduled Tasks, simply define a cron schedule and the agents start and run autonomously in the background.
    • [claimed-docs] Worktree support: Projects natively support Git worktrees, allowing agents to operate in isolated background folders.
    • [claimed-docs] users can schedule messages to be sent to their agents while they’re away
    • [claimed-docs] Run Antigravity CLI non-interactively to script agent tasks, integrate with CI pipelines, and capture machine-readable output.
    • [claimed-docs] Able to autonomously operate across your editor, terminal, and browser.
  3. ai-native userDelegate tasks to a built-in AI assistant inside the product

    weight 3 · round drawn
    Claude Codefullcommunity8/10

    Claude Code's entire premise is delegating tasks to a built-in AI agent: docs describe it planning approaches, writing code across files, running tests, handling git workflows, and autonomously completing multi-step tasks (docs-2, docs-3, docs-20, docs-28, gh-3). This is corroborated by extensive first-party documentation and some community confirmation that it performs well as a coding assistant, though other community reports describe reliability issues and failure modes in autonomous execution. Missing for 10: independent benchmark-level validation of consistent task delegation success and stronger consensus on reliability (community reports show notable failure cases).

    • [claimed-docs] Describe what you want in plain language. Claude Code plans the approach, writes the code across multiple files, and verifies it works.
    • [claimed-docs] Claude Code works directly with git. It stages changes, writes commit messages, creates branches, and opens pull requests.
    • [claimed-docs] Claude Code handles the tedious tasks that eat up your day: writing tests for untested code, fixing lint errors across a project, resolving …
    • [claimed-docs] Claude Code plans the approach, writes the code across multiple files, and verifies it works.
    • [github] helps you code faster by executing routine tasks, explaining complex code, and handling git workflows -- all through natural language comman…
    • [community] Claude is significantly better than other models at code assistant tasks, or at least in the way I use it.
    • [community] I've been using Claude Code daily for months on a project with Elixir, Rust, and Python. The worst failure mode is when it does a replace_al…
    • [community] I've tried to use Claude code for a month now. It has a 100% failure rate so far. Comparing that to creating a project and just chatting wit…
    Google Antigravityfullcommunity8/10

    Antigravity is built around delegating tasks to autonomous agents that operate across editor, terminal, and browser, with subagents, scheduled tasks, and natural-language task delegation extensively documented; hands-on community reports (comm-1, comm-10) confirm the agent/CLI actually works for delegated tasks. missing for 10: independent third-party benchmarking of delegation quality, and community evidence is mixed on reliability/bugs which caps quality below top marks.

    • [claimed-docs] Able to autonomously operate across your editor, terminal, and browser.
    • [claimed-docs] Edit, orchestrate, and build all in natural language. Tell your agents what you need, and they’ll work on getting it done.
    • [claimed-docs] an agent can delegate tasks—such as running tests or performing extensive codebase searches—to dedicated subagents.
    • [claimed-docs] Antigravity comes pre-packaged with several specialized subagents out of the box
    • [claimed-docs] Antigravity 2.0 serves as your AI agents’ central command center, providing a unified platform to launch, monitor, and orchestrate their act…
    • [community] I went ahead and downloaded it, it looks to be a VSCode fork very similar to Cursor, with support for Gemini 3 Pro, Claude Sonnet 4.5, and G…
    • [community] I much prefer using Gemini CLI in combination with vscode. It works like a charm. Now, I'll do the same with Antigravity CLI and vscode. It …
  4. ai-native userOperate the product with natural-language commands

    weight 2 · round drawn
    Claude Codefullclaimed8/10

    Claude Code is explicitly designed to be operated via plain-language instructions—describing tasks, git workflows, MCP tool use, and even natural-language chat commands (@claude in Slack, GitHub) all documented as core interaction modes, and GitHub docs explicitly state it works 'all through natural language commands.' missing for 10: independent hands-on benchmarking specifically confirming natural-language command comprehension breadth/accuracy versus slash-command or scripted usage, and some community reports note failure modes/hallucination under natural language instructions reducing reliability.

    • [claimed-docs] Describe what you want in plain language. Claude Code plans the approach, writes the code across multiple files, and verifies it works.
    • [claimed-docs] Claude Code plans the approach, writes the code across multiple files, and verifies it works.
    • [claimed-docs] Claude Code works directly with git. It stages changes, writes commit messages, creates branches, and opens pull requests.
    • [github] helps you code faster by executing routine tasks, explaining complex code, and handling git workflows -- all through natural language comman…
    • [claimed-docs] Route tasks from team chat: mention @Claude in Slack with a bug report and get a pull request back
    • [claimed-docs] cd your-project claude
    Google Antigravityfullcommunity8/10

    Docs consistently describe natural-language operation as the core interaction model — editing, orchestrating, and building 'all in natural language' (antigravity-docs-9), NL code commands in the IDE (antigravity-docs-6/41), and even voice-to-prompt transcription (antigravity-docs-3), backed by planning/artifact review flows driven by conversational prompts (antigravity-docs-24, antigravity-docs-25). Community evidence corroborates it functions as an agentic assistant (comm-1, comm-10) though with quality/reliability complaints unrelated to the NL-command axis itself. Missing for 10: independent hands-on confirmation specifically praising the NL-command UX (most community commentary focuses on bugs/pricing/security rather than command quality).

    • [claimed-docs] Edit, orchestrate, and build all in natural language. Tell your agents what you need, and they’ll work on getting it done.
    • [claimed-docs] Google Antigravity's Editor view offers tab autocompletion, natural language code commands, and a configurable, and context-aware configurab…
    • [claimed-docs] offers tab autocompletion, natural language code commands, and a configurable, and context-aware configurable agent
    • [claimed-docs] Speak your prompts. Powered by the latest Gemini Audio models, real-time transcription converts conversational speech into clearly phrased p…
    • [claimed-docs] Planning Mode: The agent plans thoroughly before executing tasks... produces structured implementation plans called Artifacts
    • [claimed-docs] The agent always halts and requests your explicit approval before proceeding with proposed changes.
    • [community] I went ahead and downloaded it, it looks to be a VSCode fork very similar to Cursor, with support for Gemini 3 Pro, Claude Sonnet 4.5, and G…
    • [community] I much prefer using Gemini CLI in combination with vscode. It works like a charm. Now, I'll do the same with Antigravity CLI and vscode. It …

Api quality

  1. ai-native userDownload a machine-readable API spec (OpenAPI or equivalent)

    weight 2 · round drawn
    Claude Codenone0/10

    The evidence pack shows Claude Code as a CLI/agent tool with SDK, MCP, and CI integrations, but no mention of a downloadable OpenAPI or equivalent machine-readable API spec for Claude Code itself. This axis is plausible for a product with an Agent SDK and API-key based access, but the pack contains no such artifact.

      Google Antigravitynone0/10

      A direct probe for OpenAPI/Swagger specs at all standard candidate paths returned 404s, and no documentation item mentions a machine-readable API spec despite extensive docs on SDK, CLI, and MCP.

      • [probe] PROBE openapi: all candidate paths 404 (https://antigravity.google/openapi.json, https://antigravity.google/swagger.json, https://antigravit…
    • ai-native userTest against a sandbox environment without touching production data

      weight 1 · round to Claude Code
      Claude Codepartialclaimed5/10

      Claude Code documents a sandboxed Bash tool that enforces filesystem and network isolation, letting Claude execute commands within OS-enforced boundaries rather than freely touching arbitrary systems — this supports the spirit of testing in isolation, but the docs don't specifically describe spinning up a 'sandbox vs production' environment or protecting production data per se. Missing for 10: explicit documentation of test/staging vs production environment separation, guidance on preventing production data access, and independent/hands-on validation that the sandbox reliably prevents production data exposure.

      • [claimed-docs] Learn how Claude Code's sandboxed Bash tool provides filesystem and network isolation for safer, more autonomous agent execution. The Bash s…
      Google Antigravitydisputedcontradicted4/10

      Antigravity's CLI docs claim a sandbox that blocks sensitive files (~/.ssh, .env) and hides anything not explicitly mounted, which sounds like exactly the kind of safe-testing boundary this story wants, but hands-on community reports directly contradict this: Gemini bypassed its own .env protection to exfiltrate secrets via prompt injection, a default allowlisted webhook.site was used as an exfiltration vector, and in another incident Antigravity commands deleted an entire drive outside any expected sandbox boundary. This is a concrete, documented failure of the sandbox promise rather than mere skepticism. Missing for 10: a genuine isolated/staging environment separate from real user data, and any vendor or independent confirmation that the sandbox reliably prevents production-data access after these reported bypasses.

      • [claimed-docs] Sensitive files like ~/.ssh and .env are blocked, anything not explicitly mounted is invisible inside the sandbox
      • [community] Google Antigravity exfiltrates data via indirect prompt injection attack: Gemini is not supposed to have access to .env files with default s…
      • [community] The default Allowlist provided with Antigravity includes 'webhook.site', which was used as an exfiltration vector for secrets.
      • [community] Antigravity was also vulnerable to the classic Markdown image exfiltration bug, reported a few days prior and flagged as 'intended behavior'…
      • [community] Google Antigravity just deleted the contents of whole drive - came down to commanding a deletion of a 'directory with space in the name' wit…
    • ai-native userRely on versioned APIs with a documented deprecation policy

      weight 2 · round drawn
      Claude Codenone0/10

      No evidence pack items mention API versioning schemes, version numbers, or a documented deprecation policy for Claude Code's APIs/CLI/SDK; the pack covers features, integrations, and community sentiment but nothing about API stability or deprecation commitments.

        Google Antigravitynone0/10

        No evidence of versioned APIs or a documented deprecation policy; OpenAPI probe returned 404s across all candidate paths and no docs mention API versioning or deprecation timelines.

        • [probe] PROBE openapi: all candidate paths 404 (https://antigravity.google/openapi.json, https://antigravity.google/swagger.json, https://antigravit…

      Automation depth — how much of the product can run unattendedAutomation depth

      How much of the product can run unattended

      1. ai-native userPerform bulk operations across many items at once

        weight 2 · round to Google Antigravity
        Claude Codedisputedcontradicted6/10

        Claude Code's docs explicitly support bulk operations — fixing lint errors 'across a project', multi-file writes, spawning multiple agents to work on different parts of a task simultaneously, and running multiple sessions/tasks in parallel or on a schedule — which strongly matches the story. However, a hands-on community report describes a concrete failure mode during a bulk-style replace_all operation that corrupted code (turning a constant into 'GROQ_URL = GROQ_URL'), with the user stating you 'absolutely can't trust it to self-verify' on such operations, directly contradicting reliable execution of bulk changes at scale. Missing for 10: independent corroboration that large-scale bulk operations complete reliably without manual review, and resolution/acknowledgment of the reported failure mode.

        • [claimed-docs] writing tests for untested code, fixing lint errors across a project, resolving merge conflicts, updating dependencies, and writing release …
        • [claimed-docs] Spawn multiple Claude Code agents that work on different parts of a task simultaneously. A lead agent coordinates the work, assigns subtasks…
        • [claimed-docs] Claude Code handles the tedious tasks that eat up your day: writing tests for untested code, fixing lint errors across a project, resolving …
        • [claimed-docs] Spawn multiple Claude Code agents that work on different parts of a task simultaneously.
        • [claimed-docs] Review diffs visually, run multiple sessions side by side, schedule recurring tasks, and kick off cloud sessions.
        • [claimed-docs] Kick off long-running tasks and check back when they're done, work on repos you don't have locally, or run multiple tasks in parallel.
        • [community] I've been using Claude Code daily for months on a project with Elixir, Rust, and Python. The worst failure mode is when it does a replace_al…
        Google Antigravitypartialclaimed6/10

        Antigravity supports parallel multi-agent orchestration across independent projects, subagent delegation, scheduled/background tasks, and a headless CLI for scripting bulk/CI workflows, which together enable operating across many items or tasks concurrently. However, there is no explicit documentation of a dedicated 'bulk operation' primitive (e.g., batch-apply an action across a list of files/items in one command) — the capability is inferred from parallelism/orchestration features rather than a purpose-built bulk-ops interface. missing for 10: explicit bulk/batch API or command for applying one operation across many items, independent hands-on evidence of large-scale parallel task execution working reliably.

        • [claimed-docs] Orchestrate multiple autonomous agents working in parallel across independent projects.
        • [claimed-docs] Have multiple agents working in parallel, so larger tasks get tackled faster.
        • [claimed-docs] an agent can delegate tasks—such as running tests or performing extensive codebase searches—to dedicated subagents.
        • [claimed-docs] Worktree support: Projects natively support Git worktrees, allowing agents to operate in isolated background folders.
        • [claimed-docs] Run Antigravity CLI non-interactively to script agent tasks, integrate with CI pipelines, and capture machine-readable output.
      2. ai-native userDefine rules that trigger actions automatically on events

        weight 3 · round to Google Antigravity
        Claude Codepartialclaimed6/10

        Claude Code supports Hooks (shell commands triggered before/after actions like auto-formatting or lint on edits) and scheduled runs plus MCP channels (Telegram/Discord/webhook events) that push messages into a session automatically, which together constitute event-triggered automation rules. However, there's no unified declarative 'rules engine' with conditions/triggers documented — it's a patchwork of hooks, cron-like scheduling, and MCP event channels rather than a first-class rule-definition system. missing for 10: a unified rules/trigger definition UI or config, broader event types beyond hooks/schedule/MCP channels, and independent/hands-on validation of these automation triggers working reliably.

        • [claimed-docs] Hooks let you run shell commands before or after Claude Code actions, like auto-formatting after every file edit or running lint before a co…
        • [claimed-docs] Run Claude on a schedule to automate work that repeats: morning PR reviews, overnight CI failure analysis, weekly dependency audits, or sync…
        • [claimed-docs] an MCP server can also act as a channel that pushes messages into your session, so Claude reacts to Telegram messages, Discord chats, or web…
        • [claimed-docs] an MCP server can also act as a channel that pushes messages into your session, so Claude reacts to Telegram messages, Discord chats, or web…
        • [claimed-docs] In CI, you can automate code review and issue triage with GitHub Actions or GitLab CI/CD.
        Google Antigravityfullclaimed7/10

        Antigravity's docs describe explicit rule activation modes (Manual, Always On, Model Decision, Glob) that trigger agent behavior automatically based on context/file patterns, plus Hooks that run custom scripts at specific points in the execution loop and Scheduled Tasks that trigger agents on a cron schedule — together these directly satisfy 'rules that trigger actions automatically on events'. Missing for 10: independent/hands-on verification that rule-triggering works reliably in practice, and more detail on broader event types beyond glob/model-decision/cron.

        • [claimed-docs] At the rule level you can define how a rule should be activated: Manual... Always On... Model Decision... Glob
        • [claimed-docs] Hooks allow you to run custom scripts or shell commands at specific points during Antigravity’s execution loop.
        • [claimed-docs] Automate routine checks with Scheduled Tasks, simply define a cron schedule and the agents start and run autonomously in the background.
        • [claimed-docs] Rules are manually defined constraints for the Agent to follow, at both the local and global levels.
      3. ai-native userSchedule recurring jobs or workflows

        weight 2 · round to Google Antigravity
        Claude Codefullclaimed7/10

        Docs explicitly describe running Claude Code on a schedule for recurring automation (PR reviews, CI failure analysis, dependency audits, doc syncing) and mention 'schedule recurring tasks' as a feature. Missing for 10: independent/hands-on confirmation of the scheduling mechanism and details on configuration (cron syntax, triggers, reliability).

        • [claimed-docs] Run Claude on a schedule to automate work that repeats: morning PR reviews, overnight CI failure analysis, weekly dependency audits, or sync…
        • [claimed-docs] Review diffs visually, run multiple sessions side by side, schedule recurring tasks, and kick off cloud sessions.
        Google Antigravityfullclaimed8/10

        Docs explicitly describe Scheduled Tasks with cron-defined schedules that run agents autonomously in the background, plus scheduling messages to agents for later delivery, directly matching the recurring-jobs/workflow story. Missing for 10: independent/hands-on confirmation that scheduling actually works reliably in practice, and more detail on job management (editing/deleting/monitoring scheduled runs).

        • [claimed-docs] Automate routine checks with Scheduled Tasks, simply define a cron schedule and the agents start and run autonomously in the background.
        • [claimed-docs] users can schedule messages to be sent to their agents while they’re away
      4. ai-native userVersion, review, and roll back my automations

        weight 1 · round to Claude Code
        Claude Codepartialclaimed5/10

        Automations in Claude Code (CLAUDE.md, skills, hooks, slash commands) are plain files that live in the repo, so they inherit git's version history, and Claude Code natively works with git (staging, commits, diffs) and supports visual diff review (claude-code-docs-3, claude-code-docs-13, claude-code-docs-32, claude-code-docs-33, claude-code-docs-22). However, there is no dedicated feature for versioning/rolling back automations themselves (e.g., no automation-specific history log, no built-in 'revert this hook/skill run' or undo mechanism) — reviewers rely entirely on generic git workflows rather than a purpose-built automation-lifecycle tool. missing for 10: a dedicated automation versioning/audit history UI, an explicit rollback/undo command for skills or hooks, and independent hands-on confirmation that rollback of automations works as intended.

        • [claimed-docs] Claude Code works directly with git. It stages changes, writes commit messages, creates branches, and opens pull requests.
        • [claimed-docs] Review diffs visually, run multiple sessions side by side, schedule recurring tasks, and kick off cloud sessions.
        • [claimed-docs] Create skills to package repeatable workflows your team can share, like `/review-pr` or `/deploy-staging`.
        • [claimed-docs] Hooks let you run shell commands before or after Claude Code actions, like auto-formatting after every file edit or running lint before a co…
        • [claimed-docs] CLAUDE.md is a markdown file you add to your project root that Claude Code reads at the start of every session. Use it to set coding standar…
        Google Antigravitynone0/10

        Evidence shows automations (skills, hooks, plugins, scheduled tasks) but no mention of versioning, review history, or rollback capabilities for these automations themselves — missing for 10: version control/history for skills/hooks/plugins, a review workflow for automation changes, and any rollback/undo mechanism for automations.

        Autonomy agents — stories about autonomy agents in this arenaAutonomy agents

        Stories about autonomy agents in this arena

        Background execution

        1. ai-native userHave a cloud agent build, test, and demo a feature end-to-end for my review

          weight 2 · round to Claude Code
          Claude Codepartialcommunity7/10

          Docs show Claude Code can run as a cloud/browser session for long-running tasks (web, mobile, remote control, teleport), plan and write code across files, write tests, and open PRs with diff review for others to inspect — covering build, test, and reviewable-artifact steps end-to-end without local setup (claude-code-docs-9,10,13,14,26,28,3,12). However there's no explicit 'demo' feature (e.g., live preview/staging deploy) beyond PR/diff review, and independent hands-on reports raise reliability concerns about self-verification on complex tasks. Missing for 10: dedicated demo/preview-environment tooling, independent corroboration of full cloud build-test-PR pipelines succeeding end-to-end without human intervention.

          • [claimed-docs] Step away from your desk and keep working from your phone or any browser with Remote Control
          • [claimed-docs] Kick off a long-running task on the web or the Claude mobile app, then pull it into your terminal with `claude --teleport`.
          • [claimed-docs] Review diffs visually, run multiple sessions side by side, schedule recurring tasks, and kick off cloud sessions.
          • [claimed-docs] Kick off long-running tasks and check back when they're done, work on repos you don't have locally, or run multiple tasks in parallel.
          • [claimed-docs] Run Claude Code in your browser with no local setup. Kick off long-running tasks and check back when they're done, work on repos you don't h…
          • [claimed-docs] Claude Code plans the approach, writes the code across multiple files, and verifies it works.
          • [claimed-docs] Claude Code works directly with git. It stages changes, writes commit messages, creates branches, and opens pull requests.
          • [claimed-docs] Get automatic code review on every PR | GitHub Code Review
          • [community] I've been using Claude Code daily for months on a project with Elixir, Rust, and Python. The worst failure mode is when it does a replace_al…
          Google Antigravitypartialcommunity6/10

          Docs describe agents that autonomously operate across editor/terminal/browser, delegate testing to subagents, produce reviewable Artifacts (implementation plans, diffs, browser recordings) and halt for approval — covering build, test and demo-for-review end-to-end (antigravity-docs-7,18,24,25,30,46). However, community reports of a subpar harness, app-breaking bugs, and a case where autonomous terminal execution deleted a whole drive raise real doubts about reliable end-to-end execution (antigravity-comm-8,antigravity-comm-14). Missing for 10: independent hands-on confirmation of a full successful build→test→demo cycle, and resolution of reliability/security concerns that could derail autonomous runs.

          • [claimed-docs] Able to autonomously operate across your editor, terminal, and browser.
          • [claimed-docs] an agent can delegate tasks—such as running tests or performing extensive codebase searches—to dedicated subagents.
          • [claimed-docs] Planning Mode: The agent plans thoroughly before executing tasks... produces structured implementation plans called Artifacts
          • [claimed-docs] The agent always halts and requests your explicit approval before proceeding with proposed changes.
          • [claimed-docs] Artifacts include rich markdown plans (Implementation Plans), code diffs, architecture diagrams, images, and browser recordings.
          • [claimed-docs] an agent can delegate tasks—such as running tests or performing extensive codebase searches—to dedicated subagents
          • [community] It's not even good, honestly. I was using it for couple weeks before dropping that 2 months ago. The model was not good and slow, the harnes…
          • [community] Google Antigravity just deleted the contents of whole drive - came down to commanding a deletion of a 'directory with space in the name' wit…
        2. developerDelegate longer-running coding tasks to run in the background in an isolated cloud environment

          weight 3 · round to Claude Code
          Claude Codefullclaimed7/10

          Docs describe running Claude Code in-browser with no local setup, kicking off long-running tasks and checking back later, working on repos not present locally, running multiple tasks in parallel, and remote control/teleport features to move sessions between web/mobile and terminal — matching the delegate-to-cloud story directly. Missing for 10: independent/hands-on confirmation of the cloud environment's isolation guarantees (the sandboxing docs cited relate to local Bash tool isolation, not the cloud session itself) and details on how isolated/secure the cloud runtime is.

          • [claimed-docs] Step away from your desk and keep working from your phone or any browser with Remote Control
          • [claimed-docs] Kick off a long-running task on the web or the Claude mobile app, then pull it into your terminal with `claude --teleport`.
          • [claimed-docs] Review diffs visually, run multiple sessions side by side, schedule recurring tasks, and kick off cloud sessions.
          • [claimed-docs] Kick off long-running tasks and check back when they're done, work on repos you don't have locally, or run multiple tasks in parallel.
          • [claimed-docs] Run Claude Code in your browser with no local setup. Kick off long-running tasks and check back when they're done, work on repos you don't h…
          Google Antigravitypartialclaimed5/10

          Antigravity supports background/autonomous execution via Scheduled Tasks that run agents in the background, Git worktree-based isolated background folders, and scheduling messages for agents while away, plus a 'Remote Control' feature to connect to running desktop sessions across machines. However, these mechanisms describe local-machine or worktree isolation and remote access to local sessions, not a distinctly cloud-hosted sandbox environment for offloading long-running tasks the way some competitors do. Missing for 10: explicit documentation of a persistent cloud-hosted execution environment independent of the user's machine, and independent/hands-on confirmation that background tasks truly run isolated in the cloud rather than locally.

          • [claimed-docs] Automate routine checks with Scheduled Tasks, simply define a cron schedule and the agents start and run autonomously in the background.
          • [claimed-docs] Worktree support: Projects natively support Git worktrees, allowing agents to operate in isolated background folders.
          • [claimed-docs] users can schedule messages to be sent to their agents while they’re away
          • [claimed-docs] Antigravity Remote Control allows you to securely connect to and drive your Antigravity 2.0 desktop sessions running across your machines fr…
          • [claimed-docs] Antigravity 2.0 serves as your AI agents’ central command center, providing a unified platform to launch, monitor, and orchestrate their act…
          • [claimed-docs] Run Antigravity CLI non-interactively to script agent tasks, integrate with CI pipelines, and capture machine-readable output.

        Parallel agents

        1. ai-native userLaunch fleets of autonomous agents that work in parallel on different tasks for hours or days

          weight 2 · round to Claude Code
          Claude Codefullclaimed7/10

          Docs explicitly describe spawning multiple Claude Code agents with a lead agent coordinating subtasks, running multiple sessions/tasks in parallel in the cloud, scheduling recurring/long-running tasks, and remote/teleport control to check back later — directly matching the fleet/parallel/long-duration story. Missing for 10: independent hands-on verification of multi-day unattended fleet runs and clearer guarantees on stability over very long horizons (community reports note reliability/quality drift over extended sessions).

          • [claimed-docs] Spawn multiple Claude Code agents that work on different parts of a task simultaneously. A lead agent coordinates the work, assigns subtasks…
          • [claimed-docs] Spawn multiple Claude Code agents that work on different parts of a task simultaneously.
          • [claimed-docs] Spawn multiple Claude Code agents that work on different parts of a task simultaneously. A lead agent coordin
          • [claimed-docs] Run Claude on a schedule to automate work that repeats: morning PR reviews, overnight CI failure analysis, weekly dependency audits, or sync…
          • [claimed-docs] Review diffs visually, run multiple sessions side by side, schedule recurring tasks, and kick off cloud sessions.
          • [claimed-docs] Kick off long-running tasks and check back when they're done, work on repos you don't have locally, or run multiple tasks in parallel.
          • [claimed-docs] Run Claude Code in your browser with no local setup. Kick off long-running tasks and check back when they're done, work on repos you don't h…
          • [claimed-docs] Step away from your desk and keep working from your phone or any browser with Remote Control
          • [claimed-docs] Kick off a long-running task on the web or the Claude mobile app, then pull it into your terminal with `claude --teleport`.
          Google Antigravitypartialclaimed7/10

          Docs describe orchestrating multiple autonomous agents in parallel across independent projects, scheduled/cron tasks that run autonomously in the background, worktree-isolated agents, subagent delegation, and remote control to check on running sessions from a browser — all supporting a 'fleet of parallel long-running agents' story. However, there is no independent/hands-on confirmation of agents actually running unattended for 'hours or days' at scale, and community reports focus on bugs, quota limits, and security issues rather than validating multi-day parallel fleet operation. Missing for 10: independent verification of long-duration (hours/days) autonomous runs, evidence of fleet scale limits, and hands-on confirmation from third parties.

          • [claimed-docs] Orchestrate multiple autonomous agents working in parallel across independent projects.
          • [claimed-docs] Automate routine checks with Scheduled Tasks, simply define a cron schedule and the agents start and run autonomously in the background.
          • [claimed-docs] Have multiple agents working in parallel, so larger tasks get tackled faster.
          • [claimed-docs] Antigravity 2.0 serves as your AI agents’ central command center, providing a unified platform to launch, monitor, and orchestrate their act…
          • [claimed-docs] Worktree support: Projects natively support Git worktrees, allowing agents to operate in isolated background folders.
          • [claimed-docs] users can schedule messages to be sent to their agents while they’re away
          • [claimed-docs] Antigravity Remote Control allows you to securely connect to and drive your Antigravity 2.0 desktop sessions running across your machines fr…
          • [claimed-docs] an agent can delegate tasks—such as running tests or performing extensive codebase searches—to dedicated subagents.
        2. developerRun several task attempts in parallel and compare results before choosing one

          weight 1 · round to Claude Code
          Claude Codepartialclaimed6/10

          Docs mention running 'multiple sessions side by side' and reviewing diffs visually in the web/cloud interface, plus running multiple tasks in parallel and spawning multiple agents—supporting parallel execution and comparison, though not explicitly framed as multiple attempts at the *same* task with a selection step. Missing for 10: explicit documentation of running several independent attempts at one identical task and a UI/workflow for choosing the best among them, and independent hands-on confirmation of this specific workflow.

          • [claimed-docs] Review diffs visually, run multiple sessions side by side, schedule recurring tasks, and kick off cloud sessions.
          • [claimed-docs] Kick off long-running tasks and check back when they're done, work on repos you don't have locally, or run multiple tasks in parallel.
          • [claimed-docs] Run Claude Code in your browser with no local setup. Kick off long-running tasks and check back when they're done, work on repos you don't h…
          • [claimed-docs] Spawn multiple Claude Code agents that work on different parts of a task simultaneously. A lead agent coordinates the work, assigns subtasks…
          • [claimed-docs] Spawn multiple Claude Code agents that work on different parts of a task simultaneously.
          Google Antigravitypartialclaimed4/10

          Docs confirm agents can run in parallel (multiple agents across projects, multiple CLI agents for large tasks) and a central dashboard to monitor/orchestrate them, but nothing describes running multiple attempts at the SAME task and comparing outputs before choosing a winner — that specific 'compare-and-select' workflow is unevidenced. missing for 10: explicit multi-attempt/variant generation for a single task, a comparison UI or ranking mechanism, and any selection step among parallel attempts.

          • [claimed-docs] Have multiple agents working in parallel, so larger tasks get tackled faster.
          • [claimed-docs] Orchestrate multiple autonomous agents working in parallel across independent projects.
          • [claimed-docs] Antigravity 2.0 serves as your AI agents’ central command center, providing a unified platform to launch, monitor, and orchestrate their act…

        Scheduled automation

        1. ai-native userSet up always-on agents that run on schedules or triggers to maintain and fix my software autonomously

          weight 2 · round to Google Antigravity
          Claude Codepartialcommunity7/10

          First-party docs show robust support for scheduled/triggered automation: 'Run Claude on a schedule' for recurring maintenance tasks (docs-8), 'schedule recurring tasks' in the web UI (docs-13), MCP servers that push Telegram/Discord/webhook events into a session 'while you're away' (docs-31/54), and Slack @mentions triggering PRs (docs-11), plus CI integration for automated review/triage (docs-36). However, community reports raise real concerns about autonomous reliability over sustained/unsupervised runs (e.g. degrading output quality, self-verification failures, 'can't trust it to self-verify' — comm-16, comm-17, comm-19, comm-20), which tempers confidence that always-on autonomous maintenance works robustly in practice. Missing for 10: independent/hands-on validation that scheduled/triggered agents reliably self-maintain software over time without human correction, and no explicit multi-day/continuous 'always-on' uptime evidence beyond scheduled/triggered runs.

          • [claimed-docs] Run Claude on a schedule to automate work that repeats: morning PR reviews, overnight CI failure analysis, weekly dependency audits, or sync…
          • [claimed-docs] Review diffs visually, run multiple sessions side by side, schedule recurring tasks, and kick off cloud sessions.
          • [claimed-docs] an MCP server can also act as a channel that pushes messages into your session, so Claude reacts to Telegram messages, Discord chats, or web…
          • [claimed-docs] an MCP server can also act as a channel that pushes messages into your session, so Claude reacts to Telegram messages, Discord chats, or web…
          • [claimed-docs] Route tasks from team chat: mention @Claude in Slack with a bug report and get a pull request back
          • [claimed-docs] In CI, you can automate code review and issue triage with GitHub Actions or GitLab CI/CD.
          • [community] I've been using Claude Code daily for months on a project with Elixir, Rust, and Python. The worst failure mode is when it does a replace_al…
          • [community] Whenever the phrase 'simplest fix' appears, it's time to pull the emergency break. This has gotten much worse over the past few weeks. It wi…
          • [community] I've tried to use Claude code for a month now. It has a 100% failure rate so far. Comparing that to creating a project and just chatting wit…
          • [community] A month ago the agents researched, designed, and implemented a compelling app idea with minimal guidance and felt super human. A month later…
          Google Antigravityfullcommunity7/10

          First-party docs explicitly describe Scheduled Tasks with cron schedules that start and run agents autonomously in the background, plus scheduling messages to agents while away, parallel autonomous agent orchestration, and headless CLI mode for CI/trigger-based automation. However, there is no independent/hands-on corroboration of the scheduling feature itself, and community reports document serious reliability/safety incidents with autonomous execution (e.g., an agent deleting a whole drive via unattended terminal auto-execution), raising doubt about safely running such agents unattended. Missing for 10: independent verification that scheduled/cron-triggered agents work reliably in practice, and evidence that autonomous 'maintain and fix' runs don't require the same close supervision seen in incident reports.

          • [claimed-docs] Automate routine checks with Scheduled Tasks, simply define a cron schedule and the agents start and run autonomously in the background.
          • [claimed-docs] users can schedule messages to be sent to their agents while they’re away
          • [claimed-docs] Orchestrate multiple autonomous agents working in parallel across independent projects.
          • [claimed-docs] Run Antigravity CLI non-interactively to script agent tasks, integrate with CI pipelines, and capture machine-readable output.
          • [community] Google Antigravity just deleted the contents of whole drive - came down to commanding a deletion of a 'directory with space in the name' wit…

        Code generation — quality of generated code — correctness, style, fit to the codebaseCode generation

        Quality of generated code — correctness, style, fit to the codebase

        Code completion

        1. developerReceive inline code completions and next-edit suggestions as I type

          weight 3 · round to Google Antigravity
          Claude Codenone0/10

          Claude Code's documented interaction model is conversational/agentic (terminal commands, plan-then-execute, PR generation) and its IDE extensions offer inline diffs and @-mentions, not ghost-text style inline completions or next-edit suggestions as the user types. No evidence pack item describes autocomplete-style inline suggestions.

          • [claimed-docs] The VS Code extension provides inline diffs, @-mentions, plan review, and conversation history directly in your editor.
          • [claimed-docs] A plugin for IntelliJ IDEA, PyCharm, WebStorm, and other JetBrains IDEs with interactive diff viewing and selection context sharing.
          Google Antigravitypartialclaimed5/10

          Docs mention the Editor view offers 'tab autocompletion' and 'natural language code commands' alongside the agent, which covers basic inline completion, but there is no detail on next-edit suggestions (predictive multi-line edits) or independent/hands-on confirmation of completion quality or latency. missing for 10: explicit next-edit-suggestion feature description, independent hands-on validation of autocomplete quality/reliability.

          • [claimed-docs] Google Antigravity's Editor view offers tab autocompletion, natural language code commands, and a configurable, and context-aware configurab…
          • [claimed-docs] offers tab autocompletion, natural language code commands, and a configurable, and context-aware configurable agent

        Debugging

        1. developerDebug a live running web application directly from my coding assistant

          weight 1 · round to Claude Code
          Claude Codepartialclaimed6/10

          Docs explicitly list a Chrome integration for debugging live web applications, indicating Claude Code can connect to and debug a running app via browser tooling rather than just editing static code. However, evidence is thin — just a single doc title/link with no detail on setup, capabilities (e.g., breakpoints, console/network inspection), or hands-on/community verification of this workflow. missing for 10: detailed documentation of the Chrome debugging workflow, independent/hands-on confirmation it works on real live apps, coverage of non-Chrome runtime debugging scenarios.

          • [claimed-docs] Debug live web applications | Chrome
          • [claimed-docs] Work with Claude directly in your codebase. Build, debug, and ship from your terminal, IDE, Slack, web, and more.
          Google Antigravitypartialclaimed4/10

          Antigravity's agent can 'autonomously operate across your editor, terminal, and browser' and produces 'browser recordings' as artifacts, implying some browser-based interaction/testing, but there is no explicit documentation of live debugging features (console inspection, breakpoints, network tab, DOM inspection) for a running web app. missing for 10: explicit live-debugging tooling (breakpoints, console/network inspection), documented workflow for attaching to a running app, independent hands-on confirmation of debugging use.

          • [claimed-docs] Able to autonomously operate across your editor, terminal, and browser.
          • [claimed-docs] Artifacts include rich markdown plans (Implementation Plans), code diffs, architecture diagrams, images, and browser recordings.
        2. developerDebug issues and troubleshoot using natural-language queries

          weight 2 · round to Claude Code
          Claude Codefullcommunity7/10

          Docs explicitly cover debugging: 'Debug live web applications' (Chrome integration), 'overnight CI failure analysis', explaining complex code, and codebase-wide understanding to trace issues via natural-language prompts. This is core positioning ('Build, debug, and ship from your terminal, IDE...'). missing for 10: independent hands-on validation specifically of debugging workflows (community evidence instead highlights reliability issues like self-verification failures and bugs introduced during edits, which are adjacent but not direct proof debugging-via-NL fails).

          • [claimed-docs] Debug live web applications | Chrome
          • [claimed-docs] Run Claude on a schedule to automate work that repeats: morning PR reviews, overnight CI failure analysis, weekly dependency audits, or sync…
          • [claimed-docs] Work with Claude directly in your codebase. Build, debug, and ship from your terminal, IDE, Slack, web, and more.
          • [claimed-docs] It understands your entire codebase and can work across multiple files and tools to get things done.
          • [github] helps you code faster by executing routine tasks, explaining complex code, and handling git workflows -- all through natural language comman…
          • [community] I've been using Claude Code daily for months on a project with Elixir, Rust, and Python. The worst failure mode is when it does a replace_al…
          Google Antigravitypartialclaimed6/10

          Antigravity's docs show natural-language code commands, autonomous operation across editor/terminal/browser, and subagents that can run tests and search codebases (docs-6,7,9,18,46), which collectively support debugging/troubleshooting via NL prompts, but there is no explicit documentation of a dedicated 'debug' workflow or troubleshooting examples, and community reports focus on stability/security issues rather than confirming debugging quality. Missing for 10: explicit debugging-specific documentation or examples, and independent hands-on validation that NL debugging queries work reliably.

          • [claimed-docs] Google Antigravity's Editor view offers tab autocompletion, natural language code commands, and a configurable, and context-aware configurab…
          • [claimed-docs] Able to autonomously operate across your editor, terminal, and browser.
          • [claimed-docs] Edit, orchestrate, and build all in natural language. Tell your agents what you need, and they’ll work on getting it done.
          • [claimed-docs] an agent can delegate tasks—such as running tests or performing extensive codebase searches—to dedicated subagents.
          • [claimed-docs] an agent can delegate tasks—such as running tests or performing extensive codebase searches—to dedicated subagents

        Feature implementation

        1. developerTurn a tracked issue into a complete pull request end-to-end

          weight 3 · round to Claude Code
          Claude Codefullcommunity7/10

          Docs explicitly describe the full loop: reading tracked issues (Jira, GitHub, Slack) via MCP, generating code across multiple files, running tests, creating branches, and opening PRs — e.g. 'Add the feature described in JIRA issue ENG-4521 and create a PR on GitHub' and 'reading issues, writing code, running tests, and submitting PRs—all from your terminal.' Community reports corroborate real-world usage but also note reliability issues (self-verification failures, quality degradation over time), so results aren't guaranteed to be flawless end-to-end. Missing for 10: independent case studies quantifying success rate of full issue-to-PR automation, and detail on how failures/test verification are handled when the generated PR doesn't pass CI.

          • [claimed-docs] Claude Code works directly with git. It stages changes, writes commit messages, creates branches, and opens pull requests.
          • [claimed-docs] Claude Code integrates with GitHub, GitLab, and your command line tools to handle the entire workflow—reading issues, writing code, running …
          • [claimed-docs] Implement features from issue trackers: "Add the feature described in JIRA issue ENG-4521 and create a PR on GitHub."
          • [claimed-docs] Add the feature described in JIRA issue ENG-4521 and create a PR on GitHub.
          • [claimed-docs] Route tasks from team chat: mention @Claude in Slack with a bug report and get a pull request back
          • [community] I've been using Claude Code daily for months on a project with Elixir, Rust, and Python. The worst failure mode is when it does a replace_al…
          • [community] I've tried to use Claude code for a month now. It has a 100% failure rate so far. Comparing that to creating a project and just chatting wit…
          Google Antigravitynone0/10

          Antigravity's docs describe autonomous coding agents that can edit files, run terminal commands, and operate across editor/terminal/browser, but there is no evidence of any issue-tracker (e.g., GitHub Issues) integration or an end-to-end workflow that ingests a tracked issue and produces a pull request. Missing for 10: issue-tracker ingestion, automated branch/PR creation, and any documented GitHub/GitLab PR workflow example.

          • [claimed-docs] Able to autonomously operate across your editor, terminal, and browser.
          • [claimed-docs] Antigravity 2.0 serves as your AI agents’ central command center, providing a unified platform to launch, monitor, and orchestrate their act…
          • [claimed-docs] Run Antigravity CLI non-interactively to script agent tasks, integrate with CI pipelines, and capture machine-readable output.
        2. developerDescribe a feature or bug in plain language and have the agent implement or fix it across multiple files

          weight 3 · round to Google Antigravity
          Claude Codedisputedcontradicted6/10

          Docs strongly claim the core capability: describe a feature/bug in plain language and Claude Code plans, implements, and verifies code changes across multiple files (claude-code-docs-2/28/51/20, claude-code-gh-3). However, hands-on community reports cite concrete failures undermining reliability of multi-file edits, e.g. a replace_all bug corrupting a constant (GROQ_URL=GROQ_URL) and inability to self-verify, plus a user reporting a '100% failure rate' and quality degradation over time (claude-code-comm-16, claude-code-comm-17, claude-code-comm-19, claude-code-comm-20), balanced against other users praising its code-assistant ability (claude-code-comm-5). missing for 10: consistent independent benchmarks confirming reliability across diverse multi-file tasks, resolution of reported failure modes.

          • [claimed-docs] Describe what you want in plain language. Claude Code plans the approach, writes the code across multiple files, and verifies it works.
          • [claimed-docs] Claude Code plans the approach, writes the code across multiple files, and verifies it works.
          • [claimed-docs] It understands your entire codebase and can work across multiple files and tools to get things done.
          • [claimed-docs] Claude Code handles the tedious tasks that eat up your day: writing tests for untested code, fixing lint errors across a project, resolving …
          • [github] helps you code faster by executing routine tasks, explaining complex code, and handling git workflows -- all through natural language comman…
          • [community] I've been using Claude Code daily for months on a project with Elixir, Rust, and Python. The worst failure mode is when it does a replace_al…
          • [community] Whenever the phrase 'simplest fix' appears, it's time to pull the emergency break. This has gotten much worse over the past few weeks. It wi…
          • [community] I've tried to use Claude code for a month now. It has a 100% failure rate so far. Comparing that to creating a project and just chatting wit…
          • [community] A month ago the agents researched, designed, and implemented a compelling app idea with minimal guidance and felt super human. A month later…
          • [community] Claude is significantly better than other models at code assistant tasks, or at least in the way I use it.
          Google Antigravitypartialcommunity7/10

          Docs describe the core loop clearly: natural-language commands drive an agent that autonomously edits code across the editor/terminal, with Projects spanning multiple folders/repos giving full codebase context and Artifacts showing diffs/plans for review (antigravity-docs-6,7,9,17,24,30,40). Community reports confirm it functions as a real coding-agent IDE (comm-1) but also describe hands-on quality issues with the agent harness and model reliability during actual implementation work (comm-8), so delivery is real but not consistently polished. Missing for 10: independent benchmark/case-study evidence of successful multi-file feature implementation, and resolution of reported harness/quality complaints.

          • [claimed-docs] Google Antigravity's Editor view offers tab autocompletion, natural language code commands, and a configurable, and context-aware configurab…
          • [claimed-docs] Able to autonomously operate across your editor, terminal, and browser.
          • [claimed-docs] Edit, orchestrate, and build all in natural language. Tell your agents what you need, and they’ll work on getting it done.
          • [claimed-docs] a project can work with one folder or multiple folders (e.g., a frontend and a backend repo), providing your agents with all of the context …
          • [claimed-docs] Planning Mode: The agent plans thoroughly before executing tasks... produces structured implementation plans called Artifacts
          • [claimed-docs] Artifacts include rich markdown plans (Implementation Plans), code diffs, architecture diagrams, images, and browser recordings.
          • [claimed-docs] Build AI agents that autonomously read files, run commands, edit code, and more.
          • [community] I went ahead and downloaded it, it looks to be a VSCode fork very similar to Cursor, with support for Gemini 3 Pro, Claude Sonnet 4.5, and G…
          • [community] It's not even good, honestly. I was using it for couple weeks before dropping that 2 months ago. The model was not good and slow, the harnes…

        Maintenance automation

        1. developerHave the agent write tests, fix lint errors, resolve merge conflicts, and update dependencies for me

          weight 3 · round to Claude Code
          Claude Codefullcommunity8/10

          First-party docs explicitly list this exact story's capabilities verbatim ('writing tests for untested code, fixing lint errors across a project, resolving merge conflicts, updating dependencies') and Claude Code is broadly documented as an agentic coding assistant that edits files, runs commands, and manages projects end-to-end. Community feedback confirms general coding competence but also raises reliability concerns (e.g., self-verification failures) not specific to these four tasks. Missing for 10: independent hands-on verification specifically for lint-fixing, merge-conflict resolution, and dependency updates rather than general coding tasks.

          • [claimed-docs] writing tests for untested code, fixing lint errors across a project, resolving merge conflicts, updating dependencies, and writing release …
          • [claimed-docs] Claude Code handles the tedious tasks that eat up your day: writing tests for untested code, fixing lint errors across a project, resolving …
          • [claimed-docs] Claude Code maps and explains entire codebases in a few seconds. It uses agentic search to understand project structure and dependencies wit…
          • [claimed-docs] Claude Code integrates with GitHub, GitLab, and your command line tools to handle the entire workflow—reading issues, writing code, running …
          • [community] Claude is significantly better than other models at code assistant tasks, or at least in the way I use it.
          Google Antigravitypartialcommunity6/10

          Antigravity's docs describe general-purpose coding agents/subagents that can run tests, edit code, and operate across editor/terminal/browser (docs-18, docs-46, docs-40, docs-7), which implicitly covers writing tests and dependency/code edits, but there is no explicit documentation calling out lint-error fixing, merge-conflict resolution, or dependency updates as named capabilities. Community evidence is mixed on general quality/reliability but does not concretely refute these specific tasks. Missing for 10: explicit first-party documentation or hands-on examples of lint-fixing, merge-conflict resolution, and dependency-update workflows specifically.

          • [claimed-docs] an agent can delegate tasks—such as running tests or performing extensive codebase searches—to dedicated subagents.
          • [claimed-docs] an agent can delegate tasks—such as running tests or performing extensive codebase searches—to dedicated subagents
          • [claimed-docs] Build AI agents that autonomously read files, run commands, edit code, and more.
          • [claimed-docs] Able to autonomously operate across your editor, terminal, and browser.
          • [claimed-docs] The Agent SDK gives you the same tools, agent loop, and context management that power Google Antigravity, programmable in Python.
          • [community] It's not even good, honestly. I was using it for couple weeks before dropping that 2 months ago. The model was not good and slow, the harnes…

        Multimodal generation

        1. ai-native userGenerate a working app from a sketch, image, or PDF design

          weight 2 · round to Google Antigravity
          Claude Codenone0/10

          The evidence pack describes Claude Code's general coding, git, MCP, and automation capabilities but never mentions accepting a sketch, image, or PDF as design input to scaffold or generate an app. The closest reference (claude-code-docs-23) only describes updating an email template from Figma designs shared in Slack, not app generation from visual designs. Missing for 10: any documentation or example of image/PDF/sketch-to-code app generation, multimodal input support in the CLI, or a demonstrated workflow turning a design mockup into a working application.

            Google Antigravitypartialclaimed4/10

            Antigravity supports passing images, PDFs and other multimedia attachments to the agent as part of prompts (docs-15, docs-32), which implies it could take a sketch/image/PDF as design input for code generation, but there is no explicit documentation or example of a 'sketch-to-app' or 'design-to-code' workflow, nor any hands-on report of this being used successfully. missing for 10: dedicated design-to-app feature/workflow documentation, an example or case study of generating an app from an image/PDF, and independent verification that this works in practice.

            • [claimed-docs] Pass rich multimedia file attachments (images, videos, audio, and documents) to the agent alongside textual instruction prompt lists.
            • [claimed-docs] External files such as Google Drive links, PDFs, and Office documents now appear in their own Documents section in the sidebar above Artifac…

          Codebase understanding — how deeply the tool maps your repo — cross-file context, architecture awareness, historyCodebase understanding

          How deeply the tool maps your repo — cross-file context, architecture awareness, history

          Codebase mapping

          1. developerUnderstand how a codebase fits together to find where to start making changes

            weight 3 · round to Claude Code
            Claude Codefullclaimed7/10

            Docs explicitly claim Claude Code 'maps and explains entire codebases in a few seconds' using agentic search to understand project structure and dependencies without manual context selection, and separately states it 'understands your entire codebase' across files; CLAUDE.md further lets teams encode architecture decisions for onboarding. Missing for 10: independent/hands-on corroboration specifically validating codebase-mapping accuracy, and no benchmark or case study showing it correctly locates the right starting point in a large real-world repo.

            • [claimed-docs] Claude Code maps and explains entire codebases in a few seconds. It uses agentic search to understand project structure and dependencies wit…
            • [claimed-docs] It understands your entire codebase and can work across multiple files and tools to get things done.
            • [claimed-docs] CLAUDE.md is a markdown file you add to your project root that Claude Code reads at the start of every session.
            • [claimed-docs] CLAUDE.md is a markdown file you add to your project root that Claude Code reads at the start of every session. Use it to set coding standar…
            Google Antigravitypartialclaimed6/10

            Antigravity provides contextual codebase understanding indirectly: Projects give agents full context across multiple folders/repos, subagents can perform 'extensive codebase searches', and Artifacts can include architecture diagrams and implementation plans that map out how a change fits into the codebase. However, there's no dedicated codebase-mapping/explanation feature, and no independent/hands-on evidence confirming how well the agent actually explains codebase structure. Missing for 10: a first-class 'explain/visualize codebase architecture' feature, independent hands-on validation of comprehension quality on real repos.

            • [claimed-docs] a project can work with one folder or multiple folders (e.g., a frontend and a backend repo), providing your agents with all of the context …
            • [claimed-docs] an agent can delegate tasks—such as running tests or performing extensive codebase searches—to dedicated subagents.
            • [claimed-docs] Artifacts include rich markdown plans (Implementation Plans), code diffs, architecture diagrams, images, and browser recordings.
            • [claimed-docs] Agents work within Projects, which define the boundaries of the folders and repositories they can access.
            • [claimed-docs] an agent can delegate tasks—such as running tests or performing extensive codebase searches—to dedicated subagents
            • [claimed-docs] Google Antigravity's Editor view offers tab autocompletion, natural language code commands, and a configurable, and context-aware configurab…
          2. developerHave the agent map and explain an entire unfamiliar codebase without manually selecting context files

            weight 3 · round to Claude Code
            Claude Codefullclaimed7/10

            Claude Code's own product page explicitly states it 'maps and explains entire codebases in a few seconds' using 'agentic search to understand project structure and dependencies without you having to manually select context files,' directly matching the story, and other docs reinforce that it 'understands your entire codebase' across multiple files. Missing for 10: independent/hands-on evidence specifically corroborating the automatic codebase-mapping claim (community evidence covers general coding quality/trust issues but not this specific feature).

            • [claimed-docs] Claude Code maps and explains entire codebases in a few seconds. It uses agentic search to understand project structure and dependencies wit…
            • [claimed-docs] It understands your entire codebase and can work across multiple files and tools to get things done.
            • [claimed-docs] CLAUDE.md is a markdown file you add to your project root that Claude Code reads at the start of every session.
            Google Antigravitypartialclaimed5/10

            Docs indicate agents automatically get full-project context (docs-17, docs-35) and can delegate to subagents that perform 'extensive codebase searches' (docs-18/46), suggesting the agent can explore an unfamiliar repo without manual file selection. However, there is no explicit feature or example describing whole-codebase mapping/explanation, and no independent/hands-on evidence confirming this works well in practice. missing for 10: a dedicated 'explain codebase' or repo-mapping feature description, and independent verification of this on an unfamiliar large codebase.

            • [claimed-docs] a project can work with one folder or multiple folders (e.g., a frontend and a backend repo), providing your agents with all of the context …
            • [claimed-docs] an agent can delegate tasks—such as running tests or performing extensive codebase searches—to dedicated subagents.
            • [claimed-docs] Agents work within Projects, which define the boundaries of the folders and repositories they can access.
            • [claimed-docs] an agent can delegate tasks—such as running tests or performing extensive codebase searches—to dedicated subagents

          Context management

          1. developerHave the agent build and recall memory automatically across sessions

            weight 2 · round to Claude Code
            Claude Codepartialclaimed4/10

            Claude Code supports persistent project context via CLAUDE.md, which it reads at the start of every session, giving some continuity of 'memory' across sessions, and the VS Code extension keeps conversation history in-editor. However, this is a manually authored/maintained file, not an automatically built or recalled memory system that captures learnings from prior sessions without user intervention. Missing for 10: evidence of automatic memory formation/summarization from past sessions, automatic recall of prior task context without a manually maintained file, and any documentation of a persistent 'agent memory' feature beyond CLAUDE.md.

            • [claimed-docs] CLAUDE.md is a markdown file you add to your project root that Claude Code reads at the start of every session.
            • [claimed-docs] CLAUDE.md is a markdown file you add to your project root that Claude Code reads at the start of every session. Use it to set coding standar…
            • [claimed-docs] Review diffs visually, run multiple sessions side by side, schedule recurring tasks, and kick off cloud sessions.
            Google Antigravitynone0/10

            The evidence describes Projects, Rules, Artifacts, and context management, but none of these describe an automatic memory system that builds and recalls information across sessions without user re-specification; Rules are explicitly manual, and Projects only scope folders/permissions, not persistent learned memory. No documentation or community evidence confirms automatic cross-session memory recall.

            • [claimed-docs] a project can work with one folder or multiple folders (e.g., a frontend and a backend repo), providing your agents with all of the context …
            • [claimed-docs] At the rule level you can define how a rule should be activated: Manual... Always On... Model Decision... Glob
            • [claimed-docs] Agents work within Projects, which define the boundaries of the folders and repositories they can access.
            • [claimed-docs] Rules are manually defined constraints for the Agent to follow, at both the local and global levels.
          2. developerInclude multiple project directories in a single session for broader context

            weight 2 · round to Google Antigravity
            Claude Codenone0/10

            The evidence pack describes Claude Code understanding a single project's entire codebase and working across multiple files within it, but there is no mention of including multiple separate project directories in one session (e.g., an --add-dir style flag or multi-root workspace support).

              Google Antigravityfullclaimed8/10

              Docs explicitly state Projects can span multiple folders (e.g., a frontend and backend repo) giving agents full codebase context, with Projects defining folder/repo access boundaries and worktree support for isolated background folders. Missing for 10: independent/hands-on corroboration of multi-folder session use in practice.

              • [claimed-docs] Group your conversations into Projects, which can span multiple folders and support custom settings and scoped permissions.
              • [claimed-docs] a project can work with one folder or multiple folders (e.g., a frontend and a backend repo), providing your agents with all of the context …
              • [claimed-docs] Agents work within Projects, which define the boundaries of the folders and repositories they can access.
              • [claimed-docs] Worktree support: Projects natively support Git worktrees, allowing agents to operate in isolated background folders.
            • developerAdd a project instructions file to set coding standards and conventions the agent follows

              weight 3 · round to Claude Code
              Claude Codefullcommunity9/10

              First-party docs explicitly describe CLAUDE.md as a project-root markdown file read at every session start, used to set coding standards, architecture decisions, preferred libraries, and review checklists (claude-code-docs-5, claude-code-docs-22). Community evidence (claude-code-comm-15) independently confirms real-world use of CLAUDE.md files for guiding the agent, corroborating the feature exists and is actively used. Missing for 10: broader independent/hands-on documentation of best practices or examples beyond a single community mention.

              • [claimed-docs] CLAUDE.md is a markdown file you add to your project root that Claude Code reads at the start of every session.
              • [claimed-docs] CLAUDE.md is a markdown file you add to your project root that Claude Code reads at the start of every session. Use it to set coding standar…
              • [community] I've found that I have to add more and more CLAUDE.md guide rails, and my CLAUDE.md files have been exploding since around mid-March... I've…
              Google Antigravityfullclaimed7/10

              Antigravity supports Rules (manually defined constraints for the agent at local and global levels, with activation modes like Always On/Glob) which serve as a project instructions file for coding standards and conventions, and Projects scope these settings per folder/repo. missing for 10: no independent/hands-on confirmation of rules file format or behavior, and no evidence of a specific standardized file name (e.g. AGENTS.md-equivalent) or examples of it being used in practice.

              • [claimed-docs] At the rule level you can define how a rule should be activated: Manual... Always On... Model Decision... Glob
              • [claimed-docs] Rules are manually defined constraints for the Agent to follow, at both the local and global levels.
              • [claimed-docs] a project can work with one folder or multiple folders (e.g., a frontend and a backend repo), providing your agents with all of the context …
              • [claimed-docs] Agents work within Projects, which define the boundaries of the folders and repositories they can access.

            Issue diagnosis

            1. developerReproduce issues, narrow down root causes, and verify fixes

              weight 3 · round to Google Antigravity
              Claude Codedisputedcontradicted5/10

              Docs claim Claude Code can debug live apps, plan fixes, and 'verifies it works' across multi-file changes (claude-code-docs-2/17/28/51), supporting reproduce/root-cause/verify workflows, but hands-on community reports give a concrete counter-example where self-verification failed (a replace_all bug silently corrupted a constant, 'You absolutely can't trust it to self-verify') and describe recurring low-quality 'simplest fix' patches that break things (claude-code-comm-16, claude-code-comm-17). missing for 10: independent benchmark/case study specifically on bug reproduction and root-cause isolation, and resolution of the self-verification reliability concerns raised by users.

              • [claimed-docs] Describe what you want in plain language. Claude Code plans the approach, writes the code across multiple files, and verifies it works.
              • [claimed-docs] Debug live web applications | Chrome
              • [claimed-docs] Claude Code plans the approach, writes the code across multiple files, and verifies it works.
              • [claimed-docs] It understands your entire codebase and can work across multiple files and tools to get things done.
              • [community] I've been using Claude Code daily for months on a project with Elixir, Rust, and Python. The worst failure mode is when it does a replace_al…
              • [community] Whenever the phrase 'simplest fix' appears, it's time to pull the emergency break. This has gotten much worse over the past few weeks. It wi…
              Google Antigravitypartialcommunity6/10

              Antigravity's agents can run terminal commands and browser sessions, delegate to subagents that run tests or search the codebase (docs-18, docs-46, docs-7), and produce Artifacts with diffs, plans and browser recordings that could serve as reproduction/verification evidence (docs-30, docs-51). Headless/CI mode (docs-50) also supports automated verification loops. However there is no explicit documented workflow for issue reproduction or root-cause narrowing, and community reports show real-world reliability problems (deleted directories, exfiltration bugs) rather than confirmation that debugging workflows work well. Missing for 10: a dedicated debugging/root-cause-analysis feature, explicit test-verification-of-fix workflow, and independent hands-on validation that this works well in practice.

              • [claimed-docs] Able to autonomously operate across your editor, terminal, and browser.
              • [claimed-docs] an agent can delegate tasks—such as running tests or performing extensive codebase searches—to dedicated subagents.
              • [claimed-docs] Artifacts include rich markdown plans (Implementation Plans), code diffs, architecture diagrams, images, and browser recordings.
              • [claimed-docs] an agent can delegate tasks—such as running tests or performing extensive codebase searches—to dedicated subagents
              • [claimed-docs] Run Antigravity CLI non-interactively to script agent tasks, integrate with CI pipelines, and capture machine-readable output.
              • [claimed-docs] An Artifact is a structured deliverable created by the agent to accomplish its task and communicate its progress and thinking to the human u…
              • [community] Google Antigravity just deleted the contents of whole drive - came down to commanding a deletion of a 'directory with space in the name' wit…

            Ecosystem — integrations, plugins, and third-party ecosystem storiesEcosystem

            Integrations, plugins, and third-party ecosystem stories

            Marketplace

            1. developerEquip the agent with custom skills to perform specialized tasks

              weight 1 · round drawn
              Claude Codefullclaimed8/10

              Claude Code explicitly supports custom Skills ('Create skills to package repeatable workflows your team can share, like /review-pr or /deploy-staging') plus a scaffolding plugin (mcp-server-dev) for building custom tool integrations, giving developers a documented mechanism to equip the agent with specialized, shareable capabilities. Missing for 10: independent hands-on validation of the skills system's reliability/quality beyond first-party docs.

              • [claimed-docs] Create skills to package repeatable workflows your team can share, like `/review-pr` or `/deploy-staging`.
              • [claimed-docs] You can also have Claude scaffold a server for you with the official mcp-server-dev plugin
              • [claimed-docs] Claude Code can connect to hundreds of external tools and data sources through the Model Context Protocol (MCP)
              Google Antigravityfullclaimed8/10

              Antigravity has a dedicated Skills system: SKILL.md-based reusable packages of knowledge/instructions the agent follows for specific tasks, creatable/downloadable and accessible via slash commands, plus composable with plugins that bundle skills, rules, MCP servers, and hooks. This is documented across product and docs pages consistently, though no independent/community hands-on verification of custom skills specifically was found. Missing for 10: independent/hands-on corroboration of custom skill creation working in practice, and a marketplace/registry of shareable skills.

              • [claimed-docs] Create or download fully customizable skills to further your agent’s autonomy and transform how you get work done.
              • [claimed-docs] A skill is a folder containing a `SKILL.md` file with instructions that the agent can follow when working on specific tasks.
              • [claimed-docs] Skills are reusable packages of knowledge that extend what the agent can do.
              • [claimed-docs] Access plugins, MCP, skills, and hooks configurations instantly via slash commands, quickly enhancing your workflow.
              • [claimed-docs] Layer custom Python callables, Model Context Protocol (MCP) servers, and reusable agent skills over our built-in filesystem and terminal too…
              • [claimed-docs] Plugins are namespaced bundles that allow you to extend Antigravity’s capabilities by grouping skills, rules, MCP servers, and hooks into a …
            2. engineering-leadIntegrate third-party partner-built agent apps into my workflows

              weight 1 · round to Claude Code
              Claude Codepartialclaimed6/10

              Claude Code supports MCP integration with third-party tools/servers (Notion, Jira, Slack, Google Drive, custom servers) and can be extended via the Agent SDK, plugins, and Slack/GitHub integrations, enabling integration of partner-built apps into workflows. However, there's no explicit evidence of a curated marketplace or formal partner-app ecosystem comparable to a dedicated app store, and integration relies mainly on generic MCP connectors rather than pre-built 'partner agent apps.' Missing for 10: a documented partner/marketplace program for third-party agent apps, independent verification of partner integrations working reliably, and case studies of engineering teams integrating named partner-built agents.

              • [claimed-docs] With MCP, Claude Code can read your design docs in Google Drive, update tickets in Jira, pull data from Slack, or use your own custom toolin…
              • [claimed-docs] Claude Code can connect to hundreds of external tools and data sources through the Model Context Protocol (MCP)
              • [claimed-docs] claude mcp add --transport http notion https://mcp.notion.com/mcp
              • [claimed-docs] You can also have Claude scaffold a server for you with the official mcp-server-dev plugin
              • [claimed-docs] the Agent SDK lets you build your own agents powered by Claude Code's tools and capabilities, with full control over orchestration, tool acc…
              • [claimed-docs] Route tasks from team chat: mention @Claude in Slack with a bug report and get a pull request back
              Google Antigravitydisputedcontradicted4/10

              Vendor docs describe an extensibility layer (MCP servers, plugins, skills, hooks) that in principle lets teams plug in third-party building blocks (docs-14, docs-28, docs-12), suggesting an ecosystem for integrating outside agent capabilities. However, hands-on community reports directly contradict the notion of freely integrating partner-built agent apps: using a third-party agent ('Pi agent') alongside Antigravity triggered a Google account ban under Antigravity's TOS restricting 3rd-party usage, and users discovered unofficial vs official extensions causing confusion (comm-17, comm-18). Missing for 10: an official partner/marketplace program for third-party agent apps, clear TOS allowance for such integrations, and independent confirmation that such integrations work without account risk.

              • [claimed-docs] Layer custom Python callables, Model Context Protocol (MCP) servers, and reusable agent skills over our built-in filesystem and terminal too…
              • [claimed-docs] Plugins are namespaced bundles that allow you to extend Antigravity’s capabilities by grouping skills, rules, MCP servers, and hooks into a …
              • [claimed-docs] Access plugins, MCP, skills, and hooks configurations instantly via slash commands, quickly enhancing your workflow.
              • [community] Google Antigravity TOS: 3rd party usage can get Google account suspended. My friend got a ban by using Pi agent with Antigravity. They un-ba…
              • [community] The VSCode Antigravity extension I was using turns out to be a 3rd-party one. I found out only today that there's an official extension too,…

            Team knowledge

            1. engineering-leadCreate a shared workspace from my docs and repos as a common source of truth for the team

              weight 1 · round to Claude Code
              Claude Codepartialclaimed5/10

              CLAUDE.md gives teams a shared, repo-committed markdown file for coding standards, architecture decisions, and review checklists that Claude reads every session, and shareable Skills (e.g. /review-pr, /deploy-staging) let a lead codify team workflows; MCP integrations let Claude also pull in Google Drive docs, Jira tickets, and Slack data as additional context sources. However, this is scattered configuration/context-injection tooling rather than a dedicated 'workspace' or knowledge-base product that unifies docs and repos into one queryable source of truth for the whole team. Missing for 10: a purpose-built shared workspace/knowledge-base UI, cross-repo aggregation, and evidence of team-wide adoption/governance beyond per-repo CLAUDE.md files.

              • [claimed-docs] CLAUDE.md is a markdown file you add to your project root that Claude Code reads at the start of every session.
              • [claimed-docs] CLAUDE.md is a markdown file you add to your project root that Claude Code reads at the start of every session. Use it to set coding standar…
              • [claimed-docs] Create skills to package repeatable workflows your team can share, like `/review-pr` or `/deploy-staging`.
              • [claimed-docs] With MCP, Claude Code can read your design docs in Google Drive, update tickets in Jira, pull data from Slack, or use your own custom toolin…
              • [claimed-docs] Claude Code can read your design docs in Google Drive, update tickets in Jira, pull data from Slack, or use your own custom tooling.
              Google Antigravitynone0/10

              Antigravity's 'Projects' concept groups folders/repos for a single agent session's context (docs-5, docs-17, docs-35, docs-38) and can surface Docs/Drive links (docs-32), but there is no evidence of a multi-user, team-shared workspace or collaborative source-of-truth that an engineering-lead could set up for a whole team — Projects appear to be individually scoped, local constructs rather than shared team assets.

              • [claimed-docs] Group your conversations into Projects, which can span multiple folders and support custom settings and scoped permissions.
              • [claimed-docs] a project can work with one folder or multiple folders (e.g., a frontend and a backend repo), providing your agents with all of the context …
              • [claimed-docs] Agents work within Projects, which define the boundaries of the folders and repositories they can access.
              • [claimed-docs] Worktree support: Projects natively support Git worktrees, allowing agents to operate in isolated background folders.
              • [claimed-docs] External files such as Google Drive links, PDFs, and Office documents now appear in their own Documents section in the sidebar above Artifac…

            Tool integration

            1. developerConnect the agent to workflow tools like Jira, Slack, and Google Drive to extend its context

              weight 3 · round to Claude Code
              Claude Codefullclaimed9/10

              Docs explicitly state Claude Code can connect via MCP to Jira, Slack, Google Drive, and other custom tooling, with concrete examples (updating Jira tickets, pulling Slack data, Notion MCP server add command) and multiple transport options. Missing for 10: independent/hands-on third-party confirmation of these specific integrations working in practice beyond vendor docs.

              • [claimed-docs] Claude Code can read your design docs in Google Drive, update tickets in Jira, pull data from Slack, or use your own custom tooling.
              • [claimed-docs] With MCP, Claude Code can read your design docs in Google Drive, update tickets in Jira, pull data from Slack, or use your own custom toolin…
              • [claimed-docs] Update our standard email template based on the new Figma designs that were posted in Slack
              • [claimed-docs] Claude Code can connect to hundreds of external tools and data sources through the Model Context Protocol (MCP)
              • [claimed-docs] Implement features from issue trackers: "Add the feature described in JIRA issue ENG-4521 and create a PR on GitHub."
              • [claimed-docs] claude mcp add --transport http notion https://mcp.notion.com/mcp
              • [claimed-docs] Add the feature described in JIRA issue ENG-4521 and create a PR on GitHub.
              Google Antigravitypartialclaimed5/10

              Antigravity documents generic MCP support for connecting to 'local developer tools, databases, file parsers, and external remote APIs' and explicitly shows Google Drive links/files appearing in its sidebar Documents section, giving a plausible path to hook in workflow tools. However, there is no explicit documentation of Jira or Slack connectors/integrations, and no first-party or community evidence of anyone actually wiring these specific tools in via MCP. Missing for 10: explicit Jira/Slack connector docs or MCP server examples, and independent confirmation of successful workflow-tool integrations beyond Drive.

              • [claimed-docs] MCP lets Antigravity fetch structured context directly or execute safe actions on your behalf when needed.
              • [claimed-docs] lets AI agents and editors securely connect to local developer tools, databases, file parsers, and external remote APIs
              • [claimed-docs] External files such as Google Drive links, PDFs, and Office documents now appear in their own Documents section in the sidebar above Artifac…
              • [claimed-docs] Layer custom Python callables, Model Context Protocol (MCP) servers, and reusable agent skills over our built-in filesystem and terminal too…
            2. developerKick off agent tasks directly from GitHub, GitLab, Linear, or Slack

              weight 2 · round to Claude Code
              Claude Codepartialclaimed7/10

              Docs confirm task kickoff from GitHub (@claude mentions, GitHub Code Review, GitHub Actions) and Slack (@Claude mention returns a PR), plus GitLab CI/CD integration, but there is no evidence of Linear integration or a Linear-triggered agent workflow. missing for 10: explicit Linear integration/trigger support, independent/hands-on confirmation of cross-platform task kickoff.

              • [github] Use it in your terminal, IDE, or tag @claude on Github.
              • [claimed-docs] Route tasks from team chat: mention @Claude in Slack with a bug report and get a pull request back
              • [claimed-docs] Get automatic code review on every PR | GitHub Code Review
              • [claimed-docs] Claude Code integrates with GitHub, GitLab, and your command line tools to handle the entire workflow—reading issues, writing code, running …
              • [claimed-docs] In CI, you can automate code review and issue triage with GitHub Actions or GitLab CI/CD.
              Google Antigravitynone0/10

              No evidence that Antigravity supports kicking off agent tasks from GitHub, GitLab, Linear, or Slack; documentation covers IDE, CLI, SDK, scheduled tasks, and MCP but no mention of triggers from these external issue-tracker/chat platforms.

              Ide terminal integration — meeting you in the IDE and terminal — extensions, inline flows, contextIde terminal integration

              Meeting you in the IDE and terminal — extensions, inline flows, context

              Cross device continuity

              1. developerStart a task on one device and continue it later from another device or browser

                weight 2 · round to Claude Code
                Claude Codefullclaimed8/10

                Docs explicitly describe cross-device continuity: 'Remote Control' lets you continue work from phone/browser (docs-9), and 'claude --teleport' lets you start a task on web/mobile and pull it into your terminal later (docs-10), backed by browser/cloud session support (docs-13, docs-14, docs-26). missing for 10: independent/hands-on confirmation of teleport and remote-control reliability across devices

                • [claimed-docs] Step away from your desk and keep working from your phone or any browser with Remote Control
                • [claimed-docs] Kick off a long-running task on the web or the Claude mobile app, then pull it into your terminal with `claude --teleport`.
                • [claimed-docs] Review diffs visually, run multiple sessions side by side, schedule recurring tasks, and kick off cloud sessions.
                • [claimed-docs] Kick off long-running tasks and check back when they're done, work on repos you don't have locally, or run multiple tasks in parallel.
                • [claimed-docs] Run Claude Code in your browser with no local setup. Kick off long-running tasks and check back when they're done, work on repos you don't h…
                Google Antigravityfullclaimed7/10

                Antigravity Remote Control explicitly lets users securely connect to and drive their desktop Antigravity sessions from any web browser, directly enabling continuing a task started on one device from another device/browser, and scheduled/background tasks further support async continuation across sessions. Missing for 10: independent hands-on verification of cross-device continuity, details on session/state sync fidelity, and any community confirmation of this specific feature working in practice.

                • [claimed-docs] Antigravity Remote Control allows you to securely connect to and drive your Antigravity 2.0 desktop sessions running across your machines fr…
                • [claimed-docs] users can schedule messages to be sent to their agents while they’re away
                • [claimed-docs] Antigravity 2.0 serves as your AI agents’ central command center, providing a unified platform to launch, monitor, and orchestrate their act…

              Ide integration

              1. developerView interactive diffs and share selected code as context from within my JetBrains IDE

                weight 1 · round to Claude Code
                Claude Codefullclaimed8/10

                Docs explicitly describe a JetBrains plugin (IntelliJ IDEA, PyCharm, WebStorm, etc.) with interactive diff viewing and selection context sharing, directly matching the story. Missing for 10: independent/hands-on corroboration of the JetBrains plugin specifically (community evidence only covers CLI/terminal experience, not the IDE plugin).

                • [claimed-docs] A plugin for IntelliJ IDEA, PyCharm, WebStorm, and other JetBrains IDEs with interactive diff viewing and selection context sharing.
                Google Antigravitynone0/10

                Antigravity is documented as a standalone VSCode-fork IDE with its own Editor view, Artifacts diff viewer, and CLI/SDK — there is no mention anywhere in the docs, changelog, or community threads of a JetBrains plugin or JetBrains-specific integration for diffs or context sharing.

                • [claimed-docs] Google Antigravity's Editor view offers tab autocompletion, natural language code commands, and a configurable, and context-aware configurab…
                • [claimed-docs] Artifacts include rich markdown plans (Implementation Plans), code diffs, architecture diagrams, images, and browser recordings.
                • [claimed-docs] Added a "Hide Whitespace Changes" option to the Review Changes overflow menu and file diff viewers to filter out whitespace-only edits.
                • [claimed-docs] Code and data artifacts like SQL and JSONL files now open in a virtualized viewer with syntax highlighting and line numbers
                • [community] I went ahead and downloaded it, it looks to be a VSCode fork very similar to Cursor, with support for Gemini 3 Pro, Claude Sonnet 4.5, and G…
              2. developerChat with the coding assistant directly inside my IDE for contextual help

                weight 3 · round to Claude Code
                Claude Codefullclaimed8/10

                Official docs confirm dedicated IDE integrations (VS Code extension with inline diffs, @-mentions, plan review, conversation history; JetBrains plugin with diff viewing and selection context sharing), plus terminal-based chat usable from within an IDE, and GitHub explicitly states 'Use it in your terminal, IDE, or tag @claude on Github.' Missing for 10: independent hands-on validation specifically of the IDE chat experience (community evidence is mostly about CLI/terminal use and general quality, not IDE-embedded chat specifically).

                • [claimed-docs] The VS Code extension provides inline diffs, @-mentions, plan review, and conversation history directly in your editor.
                • [claimed-docs] A plugin for IntelliJ IDEA, PyCharm, WebStorm, and other JetBrains IDEs with interactive diff viewing and selection context sharing.
                • [github] Use it in your terminal, IDE, or tag @claude on Github.
                • [claimed-docs] Work with Claude directly in your codebase. Build, debug, and ship from your terminal, IDE, Slack, web, and more.
                Google Antigravityfullcommunity7/10

                Antigravity is a VSCode-fork IDE with an editor view offering tab autocompletion, natural language code commands, and a context-aware conversational agent, confirmed by community hands-on reports of using it like Cursor. This directly supports in-IDE chat for contextual help. missing for 10: independent review specifically praising chat UX/quality (community notes mixed quality/performance complaints), and no detailed walkthrough of the chat interface itself beyond high-level docs.

                • [claimed-docs] Google Antigravity's Editor view offers tab autocompletion, natural language code commands, and a configurable, and context-aware configurab…
                • [claimed-docs] offers tab autocompletion, natural language code commands, and a configurable, and context-aware configurable agent
                • [community] I went ahead and downloaded it, it looks to be a VSCode fork very similar to Cursor, with support for Gemini 3 Pro, Claude Sonnet 4.5, and G…
                • [claimed-docs] Users can select which reasoning model they want to use within the model selector drop-down under the conversation prompt box

              Session management

              1. developerReview diffs visually and run multiple sessions side by side in a desktop app

                weight 2 · round to Google Antigravity
                Claude Codepartialclaimed6/10

                First-party docs explicitly state the capability ('Review diffs visually, run multiple sessions side by side, schedule recurring tasks, and kick off cloud sessions'), closely matching the story, and related IDE integrations (VS Code inline diffs, JetBrains interactive diff viewer) support visual diff review, but this appears to describe a web/desktop companion app rather than a fully detailed, screenshot-documented desktop client, and no independent or hands-on evidence corroborates the side-by-side multi-session desktop UI. Missing for 10: independent/hands-on confirmation of the desktop app's diff viewer and multi-session UI, and richer first-party documentation (screenshots, feature depth) beyond a single summary line.

                • [claimed-docs] Review diffs visually, run multiple sessions side by side, schedule recurring tasks, and kick off cloud sessions.
                • [claimed-docs] A plugin for IntelliJ IDEA, PyCharm, WebStorm, and other JetBrains IDEs with interactive diff viewing and selection context sharing.
                • [claimed-docs] The VS Code extension provides inline diffs, @-mentions, plan review, and conversation history directly in your editor.
                • [claimed-docs] Available for macOS, Linux, and Windows.
                Google Antigravityfullcommunity8/10

                Antigravity's desktop app (confirmed as a VSCode-style editor) ships a dedicated 'Review Changes' diff viewer with whitespace filtering and syntax-highlighted artifacts (antigravity-docs-33, -34, -30), plus explicit support for running multiple agents/sessions in parallel across independent projects and worktrees from one command center (antigravity-docs-1, -37, -38, -10). Missing for 10: independent hands-on confirmation of the side-by-side multi-session UI specifically (community evidence mostly discusses general bugs/instability rather than this feature directly).

                • [claimed-docs] Artifacts include rich markdown plans (Implementation Plans), code diffs, architecture diagrams, images, and browser recordings.
                • [claimed-docs] Added a "Hide Whitespace Changes" option to the Review Changes overflow menu and file diff viewers to filter out whitespace-only edits.
                • [claimed-docs] Code and data artifacts like SQL and JSONL files now open in a virtualized viewer with syntax highlighting and line numbers
                • [claimed-docs] Planning Mode: The agent plans thoroughly before executing tasks... produces structured implementation plans called Artifacts
                • [claimed-docs] The agent always halts and requests your explicit approval before proceeding with proposed changes.
                • [claimed-docs] Orchestrate multiple autonomous agents working in parallel across independent projects.
                • [claimed-docs] Antigravity 2.0 serves as your AI agents’ central command center, providing a unified platform to launch, monitor, and orchestrate their act…
                • [claimed-docs] Worktree support: Projects natively support Git worktrees, allowing agents to operate in isolated background folders.
                • [community] I went ahead and downloaded it, it looks to be a VSCode fork very similar to Cursor, with support for Gemini 3 Pro, Claude Sonnet 4.5, and G…
              2. engineering-leadManage multiple agent-driven coding sessions from one unified workspace

                weight 2 · round to Google Antigravity
                Claude Codefullclaimed7/10

                Docs describe running multiple sessions side by side, kicking off parallel/cloud sessions from a browser, and spawning multiple coordinated sub-agents under a lead agent, which directly support a lead managing several agent sessions from one workspace (claude-code-docs-13, -14, -26, -6, -34, -44). Missing for 10: independent/hands-on confirmation of the 'unified workspace' UX (no community reports specifically validate multi-session management) and no detail on session-level access control across a team for the lead-agent view.

                • [claimed-docs] Review diffs visually, run multiple sessions side by side, schedule recurring tasks, and kick off cloud sessions.
                • [claimed-docs] Kick off long-running tasks and check back when they're done, work on repos you don't have locally, or run multiple tasks in parallel.
                • [claimed-docs] Run Claude Code in your browser with no local setup. Kick off long-running tasks and check back when they're done, work on repos you don't h…
                • [claimed-docs] Spawn multiple Claude Code agents that work on different parts of a task simultaneously. A lead agent coordinates the work, assigns subtasks…
                • [claimed-docs] Spawn multiple Claude Code agents that work on different parts of a task simultaneously.
                • [claimed-docs] Spawn multiple Claude Code agents that work on different parts of a task simultaneously. A lead agent coordin
                Google Antigravityfullcommunity8/10

                Antigravity's docs describe a unified 'command center' (antigravity-docs-37) that lets a lead orchestrate multiple autonomous agents in parallel across projects (antigravity-docs-1, antigravity-docs-10), grouped into Projects spanning folders/repos with scoped permissions (antigravity-docs-5, antigravity-docs-17, antigravity-docs-35), plus worktree isolation (antigravity-docs-38), scheduled/background tasks (antigravity-docs-2, antigravity-docs-39), subagent delegation (antigravity-docs-18/19), and even remote browser-based control of running sessions (antigravity-docs-26). This directly matches the engineering-lead's need to manage many concurrent agent sessions from one place. Missing for 10: independent verification of managing many simultaneous sessions at scale, and community reports note real stability/reliability issues (antigravity-comm-6, antigravity-comm-8, antigravity-comm-9) that temper confidence though they don't specifically contradict the multi-session orchestration claim.

                • [claimed-docs] Antigravity 2.0 serves as your AI agents’ central command center, providing a unified platform to launch, monitor, and orchestrate their act…
                • [claimed-docs] Orchestrate multiple autonomous agents working in parallel across independent projects.
                • [claimed-docs] Have multiple agents working in parallel, so larger tasks get tackled faster.
                • [claimed-docs] Group your conversations into Projects, which can span multiple folders and support custom settings and scoped permissions.
                • [claimed-docs] a project can work with one folder or multiple folders (e.g., a frontend and a backend repo), providing your agents with all of the context …
                • [claimed-docs] Worktree support: Projects natively support Git worktrees, allowing agents to operate in isolated background folders.
                • [claimed-docs] Automate routine checks with Scheduled Tasks, simply define a cron schedule and the agents start and run autonomously in the background.
                • [claimed-docs] Antigravity Remote Control allows you to securely connect to and drive your Antigravity 2.0 desktop sessions running across your machines fr…
                • [community] Google made its lack of interest in Antigravity IDE obvious from very early. Updates were few and far between and app-breaking bugs stuck ar…
                • [community] It's not even good, honestly. I was using it for couple weeks before dropping that 2 months ago. The model was not good and slow, the harnes…

              Terminal workflow

              1. developerRun a coding agent locally from my terminal

                weight 3 · round to Claude Code
                Claude Codefullcommunity9/10

                Claude Code is explicitly documented as a terminal-native coding agent: install via curl script, run with cd your-project && claude, available on macOS/Linux/Windows, and GitHub README confirms 'Use it in your terminal, IDE, or tag @claude on Github.' Community posts corroborate hands-on terminal use, noting it's 'implemented as a bash tool and not an editor replacement.' Missing for 10: broader independent benchmark or third-party review confirming consistent reliability of local terminal operation beyond a few anecdotal community posts.

                • [claimed-docs] cd your-project claude
                • [claimed-docs] curl -fsSL https://claude.ai/install.sh | bash
                • [claimed-docs] Available for macOS, Linux, and Windows.
                • [github] Use it in your terminal, IDE, or tag @claude on Github.
                • [community] The cost is absurd (compared to other LLM providers these days). I asked 3 questions and the cost was ~0.77c. I do like how this is implemen…
                Google Antigravityfullprobed8/10

                Google Antigravity ships an official CLI product (antigravity-cli) with terminal-native features like slash commands, headless/non-interactive mode for scripting, sandboxing, vim-mode editing, and config management, explicitly designed to run agents locally from the terminal, and a community comment confirms using 'Antigravity CLI with vscode' works fine. Missing for 10: deeper independent hands-on reviews specifically of the CLI (most community feedback focuses on the IDE, not the terminal tool) and no third-party benchmarks of terminal performance/reliability.

                • [claimed-docs] Edit, orchestrate, and build all in natural language. Tell your agents what you need, and they’ll work on getting it done.
                • [claimed-docs] Have multiple agents working in parallel, so larger tasks get tackled faster.
                • [claimed-docs] Navigate your entire workflow via standard terminal shortcuts: adjust permissions, themes, and preferences via /config and type /keybindings…
                • [claimed-docs] Access plugins, MCP, skills, and hooks configurations instantly via slash commands, quickly enhancing your workflow.
                • [claimed-docs] Run Antigravity CLI non-interactively to script agent tasks, integrate with CI pipelines, and capture machine-readable output.
                • [claimed-docs] Sensitive files like ~/.ssh and .env are blocked, anything not explicitly mounted is invisible inside the sandbox
                • [claimed-docs] Vim editor mode replaces the editing model in every multi-line input surface of the CLI
                • [probe] official CLI documented at https://antigravity.google/product/antigravity-cli
                • [community] I much prefer using Gemini CLI in combination with vscode. It works like a charm. Now, I'll do the same with Antigravity CLI and vscode. It …
              2. developerRun the agent non-interactively in scripts for workflow automation

                weight 2 · round to Claude Code
                Claude Codefullclaimed8/10

                Docs explicitly describe non-interactive automation: piping logs, running in CI, scheduling recurring tasks, GitHub Actions/GitLab CI/CD integration for automated code review and issue triage, and headless-style scripting per Unix philosophy. missing for 10: no explicit mention of a documented --print/non-interactive flag or exit-code behavior, and no independent/hands-on report confirming scripted CI usage works as described.

                • [claimed-docs] Claude Code is composable and follows the Unix philosophy. Pipe logs into it, run it in CI, or chain it with other tools
                • [claimed-docs] Run Claude on a schedule to automate work that repeats: morning PR reviews, overnight CI failure analysis, weekly dependency audits, or sync…
                • [claimed-docs] In CI, you can automate code review and issue triage with GitHub Actions or GitLab CI/CD.
                • [claimed-docs] Get automatic code review on every PR | GitHub Code Review
                Google Antigravityfullclaimed7/10

                Official docs explicitly describe a headless mode: 'Run Antigravity CLI non-interactively to script agent tasks, integrate with CI pipelines, and capture machine-readable output,' directly matching the workflow-automation story. Missing for 10: independent/hands-on confirmation of headless CI usage and details on machine-readable output format/exit codes.

                • [claimed-docs] Run Antigravity CLI non-interactively to script agent tasks, integrate with CI pipelines, and capture machine-readable output.
                • [claimed-docs] Edit, orchestrate, and build all in natural language. Tell your agents what you need, and they’ll work on getting it done.
                • [claimed-docs] Create or download fully customizable skills to further your agent’s autonomy and transform how you get work done.

              Openness — open source, data portability, and self-hosting storiesOpenness

              Open source, data portability, and self-hosting stories

              1. ai-native userDo everything through the API that I can do in the UI

                weight 2 · round to Claude Code
                Claude Codepartialclaimed5/10

                Claude Code exposes an Agent SDK for building custom agents with 'full control over orchestration, tool access, and permissions' (docs-18) and supports direct API-key access and CI/headless automation (docs-36, docs-39/40), suggesting core coding capabilities are programmatically accessible. However, evidence doesn't confirm parity for UI-specific features like Remote Control, teleport, mobile app, or Slack routing being fully reachable via the API/SDK. Missing for 10: explicit documentation that all UI-surfaced features (remote control, teleport, IDE-specific interactions) are equally available through the API/SDK, and independent confirmation of this parity.

                • [claimed-docs] the Agent SDK lets you build your own agents powered by Claude Code's tools and capabilities, with full control over orchestration, tool acc…
                • [claimed-docs] In CI, you can automate code review and issue triage with GitHub Actions or GitLab CI/CD.
                • [claimed-docs] If you've set the ANTHROPIC_API_KEY environment variable, Claude Code skips the login prompt and asks you to approve the key instead.
                • [claimed-docs] ANTHROPIC_API_KEY environment variable. Sent as the X-Api-Key header. Use this for direct Anthropic API access with a key from the Claude Co…
                • [claimed-docs] Use Claude Code as an MCP server. You can use Claude Code itself as an MCP server that other applications can connect to: claude mcp serve (…
                Google Antigravitypartialprobed4/10

                Antigravity offers an Agent SDK (Python programmable) and a headless/non-interactive CLI mode for scripting agent tasks, plugins, MCP, and hooks, suggesting substantial programmatic access to agent capabilities. However, there is no documented public REST/HTTP API or OpenAPI spec (probe explicitly found all openapi.json candidate paths 404'd), and no evidence that UI-only features like Remote Control, Editor tab-autocompletion, artifact review UI, or scheduled task UI are fully exposed via API parity. missing for 10: a documented public API/OpenAPI spec, confirmation that all UI features (remote control, artifact review, scheduling UI) have API equivalents, and independent verification of API-UI parity.

                • [claimed-docs] The Agent SDK gives you the same tools, agent loop, and context management that power Google Antigravity, programmable in Python.
                • [claimed-docs] Run Antigravity CLI non-interactively to script agent tasks, integrate with CI pipelines, and capture machine-readable output.
                • [claimed-docs] Layer custom Python callables, Model Context Protocol (MCP) servers, and reusable agent skills over our built-in filesystem and terminal too…
                • [probe] PROBE openapi: all candidate paths 404 (https://antigravity.google/openapi.json, https://antigravity.google/swagger.json, https://antigravit…
              2. ai-native userExport all of my data in open formats and leave

                weight 3 · round drawn
                Claude Codenone0/10

                The evidence pack contains no mention of a data export feature, session/conversation history export, or open-format portability guarantees for Claude Code — nothing addresses a user's ability to extract all their data and leave the platform. While Claude Code operates on local files (inherently open), there is no documented mechanism for exporting session logs, configs, or account data in open formats, so this applicable axis is unsupported.

                  Google Antigravitynone0/10

                  No evidence in the pack mentions data export, open-format portability, or account/data deletion features for Antigravity; docs cover projects, agents, artifacts, and CLI but never data portability or export-and-leave capability.

                  • ai-native userRead the product's source under an open license

                    weight 2 · round drawn
                    Claude Codenone0/10

                    No evidence Claude Code's source is available under an open license; in fact community discussion explicitly contrasts it with an open-source competitor, noting 'Codex CLI is FOSS, unlike Claude Code' — confirming it is closed-source.

                    • [community] Codex CLI is FOSS, unlike Claude Code, so Codex is less likely to do things like that, and it's one more reason to avoid Claude Code and Cla…
                    Google Antigravitynone0/10

                    No evidence of any open-source license or public source repository for Antigravity; it appears closed-source (VSCode fork distributed as binary download, third-party unofficial extensions noted). Nothing in the docs or community reports references source availability or a license.

                    • ai-native userSelf-host the core product

                      weight 3 · round drawn
                      Claude Codenone0/10

                      Claude Code is a closed-source CLI that requires an Anthropic API key or Claude.ai/Console login to function (docs-37, docs-39, docs-55) — there is no evidence of a self-hostable core model or backend. Community evidence explicitly notes it is not open source, unlike alternatives (comm-4), confirming the product cannot be self-hosted.

                      • [claimed-docs] Claude Pro or Max subscription: log in with your Claude.ai account.
                      • [claimed-docs] If you've set the ANTHROPIC_API_KEY environment variable, Claude Code skips the login prompt and asks you to approve the key instead.
                      • [claimed-docs] Individual users can log in with a claude.ai account, while teams can use Claude for Teams or Enterprise, the Claude Console, or a cloud pro…
                      • [community] Codex CLI is FOSS, unlike Claude Code, so Codex is less likely to do things like that, and it's one more reason to avoid Claude Code and Cla…
                      Google Antigravitynone0/10

                      No evidence anywhere in the pack indicates Antigravity can be self-hosted; it is described only as a downloadable desktop app/IDE/CLI/SDK connecting to Google's cloud-hosted models, with account/entitlement gating and TOS restrictions mentioned in community reports, but no self-hosted server or on-prem deployment option is documented.

                      • [claimed-docs] Visit antigravity.google/download to download Google Antigravity 2.0. Select your operating system below
                      • [claimed-docs] Antigravity 2.0 serves as your AI agents’ central command center, providing a unified platform to launch, monitor, and orchestrate their act…
                      • [community] don't buy Google AI subscription before you confirm you have 'Antigravity entitlement'... You can have verified account, bank card added, ac…
                      • [community] Google Antigravity TOS: 3rd party usage can get Google account suspended. My friend got a ban by using Pi agent with Antigravity. They un-ba…

                    Pricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limits

                    Free-tier ceilings, usage caps, and rate limits before you have to pay

                    Authentication

                    1. developerAuthenticate with an API key instead of an account login

                      weight 2 · round to Claude Code
                      Claude Codefullclaimed9/10

                      Docs explicitly confirm ANTHROPIC_API_KEY env var authentication bypasses the account login prompt, using it for direct API access via X-Api-Key header, as an alternative to Claude.ai account login. missing for 10: independent/hands-on community confirmation of this specific auth flow (only first-party docs cited).

                      • [claimed-docs] If you've set the ANTHROPIC_API_KEY environment variable, Claude Code skips the login prompt and asks you to approve the key instead.
                      • [claimed-docs] ANTHROPIC_API_KEY environment variable. Sent as the X-Api-Key header. Use this for direct Anthropic API access with a key from the Claude Co…
                      • [claimed-docs] Claude Pro or Max subscription: log in with your Claude.ai account.
                      • [claimed-docs] Individual users can log in with a claude.ai account, while teams can use Claude for Teams or Enterprise, the Claude Console, or a cloud pro…
                      Google Antigravitynone0/10

                      Evidence shows Antigravity requires a Google account/login and even ties usage to 'Antigravity entitlement' on that account (comm-19), with account-level bans possible (comm-17, comm-20); no docs or CLI reference mention an API-key authentication mode as an alternative to account login.

                      • [community] Google Antigravity TOS: 3rd party usage can get Google account suspended. My friend got a ban by using Pi agent with Antigravity. They un-ba…
                      • [community] don't buy Google AI subscription before you confirm you have 'Antigravity entitlement'... You can have verified account, bank card added, ac…
                      • [community] Banning the entire account rather than AI access is wildly user hostile... And then you get to fight the support bots and eventually go to t…
                    2. engineering-leadAuthenticate through an enterprise identity or cloud platform for compliance and scalability

                      weight 2 · round to Claude Code
                      Claude Codefullclaimed8/10

                      Claude Code documents enterprise authentication via SSO/SAML, domain capture, role-based permissions, compliance API, and managed policy settings under Claude for Enterprise, plus flexible auth options (Console API key, Claude.ai account, Teams/Enterprise, cloud provider) for scaling across org structures. missing for 10: independent/hands-on corroboration of SSO setup working in practice, and no explicit mention of cloud IAM integration (e.g., AWS/GCP native identity federation) beyond 'cloud provider' mention.

                      • [claimed-docs] Claude for Enterprise: adds SSO, domain capture, role-based permissions, compliance API, and managed policy settings for organization-wide C…
                      • [claimed-docs] Single sign-on (SSO/SAML) and domain capture
                      • [claimed-docs] Individual users can log in with a claude.ai account, while teams can use Claude for Teams or Enterprise, the Claude Console, or a cloud pro…
                      • [claimed-docs] Claude Pro or Max subscription: log in with your Claude.ai account.
                      • [claimed-docs] You can sign in to your Console account without creating an API key, even when your organization doesn't let developers create them.
                      Google Antigravitynone0/10

                      No evidence of SSO/SAML/OIDC, Google Workspace/Cloud IAM enterprise login, or any enterprise identity federation for Antigravity; docs mention only Google account sign-in and entitlement issues, with community reports of account suspensions rather than enterprise auth support. Missing for 10: SSO/SAML/OIDC support, Google Cloud IAM or Workspace admin console integration, enterprise provisioning/SCIM documentation.

                      • [community] don't buy Google AI subscription before you confirm you have 'Antigravity entitlement'... You can have verified account, bank card added, ac…
                      • [community] Banning the entire account rather than AI access is wildly user hostile... And then you get to fight the support bots and eventually go to t…
                    3. developerSign in with my existing product subscription plan to use the coding agent

                      weight 2 · round to Claude Code
                      Claude Codefullclaimed9/10

                      Docs explicitly confirm developers can log in with their existing Claude Pro or Max subscription (claude.ai account) instead of needing a separate API key, with API key as an alternative for direct API access. Missing for 10: independent/hands-on confirmation of the subscription login flow working smoothly in practice (community evidence focuses on other topics, not this login flow specifically).

                      • [claimed-docs] Claude Pro or Max subscription: log in with your Claude.ai account.
                      • [claimed-docs] Individual users can log in with a claude.ai account, while teams can use Claude for Teams or Enterprise, the Claude Console, or a cloud pro…
                      • [claimed-docs] If you've set the ANTHROPIC_API_KEY environment variable, Claude Code skips the login prompt and asks you to approve the key instead.
                      • [claimed-docs] ANTHROPIC_API_KEY environment variable. Sent as the X-Api-Key header. Use this for direct Anthropic API access with a key from the Claude Co…
                      Google Antigravitynone0/10

                      No documentation describes signing in with an existing Google AI/Gemini subscription plan to unlock Antigravity access, and community reports directly state that users with an active AI Pro subscription still could not use even the free tier without a separate 'Antigravity entitlement.' missing for 10: any first-party docs describing subscription-based sign-in, evidence of successful subscription-linked access, and confirmation that paid Google AI plans map directly to Antigravity usage.

                      • [community] don't buy Google AI subscription before you confirm you have 'Antigravity entitlement'... You can have verified account, bank card added, ac…
                      • [community] Google Antigravity TOS: 3rd party usage can get Google account suspended. My friend got a ban by using Pi agent with Antigravity. They un-ba…
                      • [community] On the pricing page it says free individual plan with 'generous rate limits'. I gave it an HTML file and 2 minutes later got: 'Model quota l…
                    4. developerSign in with a personal account to get free-tier access without managing API keys

                      weight 1 · round to Claude Code
                      Claude Codepartialclaimed6/10

                      Docs confirm individual developers can log in with a personal claude.ai account (Pro/Max subscription) instead of managing an API key, and that API-key auth is optional/alternate. However, evidence only references Pro/Max subscription login, not an explicit free tier for Claude Code — missing for 10: explicit confirmation that a free/no-cost claude.ai account grants Claude Code access, and independent user corroboration of free-tier login flow.

                      • [claimed-docs] Claude Pro or Max subscription: log in with your Claude.ai account.
                      • [claimed-docs] If you've set the ANTHROPIC_API_KEY environment variable, Claude Code skips the login prompt and asks you to approve the key instead.
                      • [claimed-docs] ANTHROPIC_API_KEY environment variable. Sent as the X-Api-Key header. Use this for direct Anthropic API access with a key from the Claude Co…
                      • [claimed-docs] Individual users can log in with a claude.ai account, while teams can use Claude for Teams or Enterprise, the Claude Console, or a cloud pro…
                      Google Antigravitydisputedcontradicted3/10

                      Vendor pages advertise a free individual plan with no mention of API key management, implying sign-in-with-personal-account access, but community hands-on reports directly contradict the free-tier promise — one user got a 'Model quota limit exceeded' error within minutes despite the 'generous rate limits' claim, and another describes being locked out of even the free tier due to an 'Antigravity entitlement' gate despite having an active subscription. missing for 10: first-party documentation explicitly describing the personal-account sign-in flow and free-tier terms, and independent confirmation that free-tier access works reliably without unexpected quota/entitlement blocks.

                      • [community] On the pricing page it says free individual plan with 'generous rate limits'. I gave it an HTML file and 2 minutes later got: 'Model quota l…
                      • [community] don't buy Google AI subscription before you confirm you have 'Antigravity entitlement'... You can have verified account, bank card added, ac…
                      • [claimed-docs] Visit antigravity.google/download to download Google Antigravity 2.0. Select your operating system below

                    Model choice

                    1. developerLet the tool automatically pick the best model for each task

                      weight 1 · round drawn
                      Claude Codenone0/10

                      No evidence in the pack describes automatic model selection or routing per task; users manually choose models (e.g., Sonnet vs Opus per comm-19) and there's no mention of an auto-select feature. Missing for 10: any docs describing automatic model routing/selection logic based on task complexity or cost.

                        Google Antigravitynone0/10

                        Docs describe a manual model selector dropdown where users choose the reasoning model themselves (antigravity-docs-16), not an automatic 'best model per task' selection mechanism; no evidence anywhere of automatic model routing or task-based model optimization.

                        • [claimed-docs] Users can select which reasoning model they want to use within the model selector drop-down under the conversation prompt box
                      • developerChoose which underlying AI model powers my session from multiple providers

                        weight 2 · round to Google Antigravity
                        Claude Codenone0/10

                        Evidence shows Claude Code authentication routes (Claude.ai login, API key, Console, Enterprise SSO) are all tied to Anthropic's own Claude models; there is no mention of selecting GPT, Gemini, or other third-party model providers to power a session. Since comparable coding tools do offer multi-provider model selection, this axis applies but is unevidenced here.

                        • [claimed-docs] Claude Pro or Max subscription: log in with your Claude.ai account.
                        • [claimed-docs] If you've set the ANTHROPIC_API_KEY environment variable, Claude Code skips the login prompt and asks you to approve the key instead.
                        • [claimed-docs] ANTHROPIC_API_KEY environment variable. Sent as the X-Api-Key header. Use this for direct Anthropic API access with a key from the Claude Co…
                        • [claimed-docs] Individual users can log in with a claude.ai account, while teams can use Claude for Teams or Enterprise, the Claude Console, or a cloud pro…
                        Google Antigravityfullcommunity8/10

                        Docs explicitly describe a model selector dropdown for choosing reasoning models, and community hands-on evidence confirms multiple providers (Gemini 3 Pro, Claude Sonnet 4.5, GPT-OSS 120B) are selectable, not locked to Gemini only. Missing for 10: pricing/tier restrictions per model and independent benchmarking of model-switching quality across providers.

                        • [claimed-docs] Users can select which reasoning model they want to use within the model selector drop-down under the conversation prompt box
                        • [community] I went ahead and downloaded it, it looks to be a VSCode fork very similar to Cursor, with support for Gemini 3 Pro, Claude Sonnet 4.5, and G…
                        • [community] Nice to see that it's not locked to just Gemini models.

                      Privacy posture — data-handling and privacy storiesPrivacy posture

                      Data-handling and privacy stories

                      1. ai-native userChoose where my data is stored (region/residency)

                        weight 2 · round drawn
                        Claude Codenone0/10

                        No evidence pack items mention data residency, regional data storage options, or geographic controls for where Claude Code data is processed/stored; only SSO/domain capture/compliance API for enterprise IAM are mentioned. Missing for 10: any documentation of region selection, data residency guarantees, or geo-specific storage controls.

                          Google Antigravitynone0/10

                          No evidence pack item mentions data residency, region selection, or storage location controls; only a telemetry on/off toggle is documented, which does not address data residency. Missing for 10: any documentation of regional data storage options, residency guarantees, or enterprise data-location controls.

                          • ai-native userPrevent my data from being used to train AI models

                            weight 3 · round to Google Antigravity
                            Claude Codenone0/10

                            The evidence pack includes enterprise/compliance features (SSO, compliance API, managed policies) but contains no mention of any training-data opt-out, data-usage policy, or explicit statement that user code/conversations are excluded from model training.

                              Google Antigravitypartialclaimed5/10

                              Antigravity's settings docs mention a Telemetry toggle to enable/disable sharing interaction logs 'to improve models,' which is the only evidence addressing training-data opt-out; there's no further detail on scope, default state, or enterprise data-processing guarantees. Missing for 10: independent verification the toggle actually excludes data from training, clarity on default setting, and any enterprise/DPA-level documentation of data usage.

                              • [claimed-docs] toggle Telemetry (enable/disable sharing interaction logs to improve models)
                            • ai-native userControl data retention and deletion

                              weight 2 · round to Google Antigravity
                              Claude Codenone0/10

                              The evidence pack shows enterprise features like SSO, domain capture, and a vague 'compliance API' but nothing describing user-controllable data retention settings or deletion of stored conversation/code data. No documentation addresses how users can view, export, or delete retained data.

                                Google Antigravitypartialclaimed3/10

                                Docs mention a Telemetry toggle to enable/disable sharing interaction logs, which is a privacy-related control, but there is no documented mechanism for viewing, exporting, or deleting stored data/history, nor any stated retention policy. Missing for 10: explicit data deletion controls, data export/retention policy documentation, and independent confirmation these settings work as described.

                                • [claimed-docs] toggle Telemetry (enable/disable sharing interaction logs to improve models)
                              • ai-native userOpt out of telemetry and usage tracking

                                weight 2 · round to Google Antigravity
                                Claude Codenone0/10

                                The evidence pack contains no documentation or reference to a telemetry/usage-tracking opt-out setting (e.g., no mention of a DISABLE_TELEMETRY flag, privacy settings page, or opt-out toggle) for Claude Code. Community commentary touches on unrelated trust/security concerns (anti-distillation fake tools, undercover mode) but none confirm or deny a telemetry opt-out mechanism.

                                  Google Antigravityfullclaimed7/10

                                  Docs confirm a settings toggle to enable/disable telemetry ('sharing interaction logs to improve models'), giving users a direct opt-out. Missing for 10: independent/hands-on confirmation that the toggle fully stops all data collection, and no detail on what telemetry remains even when disabled.

                                  • [claimed-docs] toggle Telemetry (enable/disable sharing interaction logs to improve models)

                                Review safety — keeping generated changes safe — diffs, approvals, guardrailsReview safety

                                Keeping generated changes safe — diffs, approvals, guardrails

                                Data governance

                                1. engineering-leadOpt out of having my code and prompts used for AI model training

                                  weight 1 · round to Google Antigravity
                                  Claude Codenone0/10

                                  The evidence pack contains no documentation or statements about Claude Code's data usage or model-training policies, nor any opt-out mechanism for code/prompt data. Enterprise features mentioned (SSO, compliance API, RBAC) do not address training data usage, and community items are unrelated to this specific concern.

                                    Google Antigravitypartialclaimed5/10

                                    Docs mention a Telemetry toggle to 'enable/disable sharing interaction logs to improve models,' which functions as an opt-out from data being used for model improvement, but there is no explicit documentation framing this as a training opt-out for enterprise/engineering-lead governance needs (e.g., no data-processing agreement, no distinction between prompts/code vs telemetry, no enterprise admin-level control). Missing for 10: explicit statement that code/prompts are excluded from training, org-wide/admin-level enforcement of the opt-out, and independent confirmation the toggle actually stops training use.

                                    • [claimed-docs] toggle Telemetry (enable/disable sharing interaction logs to improve models)

                                  Pr review

                                  1. developerHave the agent stage changes, write commit messages, create branches, and open pull requests

                                    weight 3 · round to Claude Code
                                    Claude Codefullclaimed8/10

                                    First-party docs explicitly state Claude Code 'stages changes, writes commit messages, creates branches, and opens pull requests' and integrates with GitHub/GitLab to handle the entire workflow including submitting PRs, corroborated by the GitHub repo description mentioning it 'handles git workflows'. Missing for 10: independent hands-on verification of a full stage-commit-branch-PR flow (community evidence discusses code quality/trust issues but not this specific git workflow failing).

                                    • [claimed-docs] Claude Code works directly with git. It stages changes, writes commit messages, creates branches, and opens pull requests.
                                    • [claimed-docs] Claude Code integrates with GitHub, GitLab, and your command line tools to handle the entire workflow—reading issues, writing code, running …
                                    • [github] helps you code faster by executing routine tasks, explaining complex code, and handling git workflows -- all through natural language comman…
                                    Google Antigravitypartialclaimed3/10

                                    Antigravity's agents can operate the terminal and natively support Git worktrees, which implies they could run git commands like staging, committing, and branching, but no documentation explicitly describes agent-driven commit message generation, branch creation, or PR opening (e.g., GitHub integration). Missing for 10: explicit docs on commit-message authoring, branch creation workflow, and pull-request creation/integration with GitHub/GitLab.

                                    • [claimed-docs] Worktree support: Projects natively support Git worktrees, allowing agents to operate in isolated background folders.
                                    • [claimed-docs] Able to autonomously operate across your editor, terminal, and browser.
                                    • [claimed-docs] Build AI agents that autonomously read files, run commands, edit code, and more.
                                  2. developerGet automatic code review with contextual feedback on every pull request

                                    weight 3 · round to Claude Code
                                    Claude Codefullcommunity7/10

                                    Docs explicitly advertise 'Get automatic code review on every PR | GitHub Code Review' plus CI-based automated code review/issue triage and enterprise security code review, and CLAUDE.md can encode review checklists; community evidence even notes Claude performs well specifically as a reviewer. missing for 10: independent hands-on validation of the GitHub Code Review integration itself and detail on how contextual feedback is generated/delivered on PRs.

                                    • [claimed-docs] Get automatic code review on every PR | GitHub Code Review
                                    • [claimed-docs] In CI, you can automate code review and issue triage with GitHub Actions or GitLab CI/CD.
                                    • [claimed-docs] Claude helps security teams and developers by reviewing code for security issues, drafts patches, and explains the risk in language your who…
                                    • [claimed-docs] CLAUDE.md is a markdown file you add to your project root that Claude Code reads at the start of every session. Use it to set coding standar…
                                    • [community] I have found that Claude Opus 4.6 is a better reviewer than it is an implementer. When Codex implements and Claude reviews, it's usually jus…
                                    Google Antigravitynone0/10

                                    Antigravity offers in-editor 'Review Changes' diff viewing and Artifact-based plan review, but there is no evidence of a GitHub/GitLab pull-request bot or CI-integrated review that automatically posts contextual feedback on every PR. The CLI headless mode allows scripting into CI, but no docs describe an automated PR-review workflow.

                                    • [claimed-docs] Planning Mode: The agent plans thoroughly before executing tasks... produces structured implementation plans called Artifacts
                                    • [claimed-docs] The agent always halts and requests your explicit approval before proceeding with proposed changes.
                                    • [claimed-docs] Added a "Hide Whitespace Changes" option to the Review Changes overflow menu and file diff viewers to filter out whitespace-only edits.
                                    • [claimed-docs] Run Antigravity CLI non-interactively to script agent tasks, integrate with CI pipelines, and capture machine-readable output.
                                  3. developerInspect diffs and run checks to catch problems before merging

                                    weight 3 · round drawn
                                    Claude Codepartialcommunity6/10

                                    Claude Code supports diff inspection (inline diffs in VS Code/JetBrains, visual diff review in web/desktop UI) and can run tests, lint, and CI checks as part of its workflow, plus automatic PR code review via GitHub integration. However, the story's 'inspect diffs and run checks before merging' as a cohesive reviewer workflow is only partially evidenced — there's no dedicated diff/lint/test-gate UI walkthrough, and community reports raise self-verification concerns (e.g., replace_all bugs going undetected). missing for 10: a dedicated pre-merge review workflow with integrated check-gating (not just individual features), independent hands-on validation of diff-review accuracy, and evidence addressing the self-verification skepticism raised in community reports.

                                    • [claimed-docs] Review diffs visually, run multiple sessions side by side, schedule recurring tasks, and kick off cloud sessions.
                                    • [claimed-docs] The VS Code extension provides inline diffs, @-mentions, plan review, and conversation history directly in your editor.
                                    • [claimed-docs] A plugin for IntelliJ IDEA, PyCharm, WebStorm, and other JetBrains IDEs with interactive diff viewing and selection context sharing.
                                    • [claimed-docs] Get automatic code review on every PR | GitHub Code Review
                                    • [claimed-docs] Hooks let you run shell commands before or after Claude Code actions, like auto-formatting after every file edit or running lint before a co…
                                    • [claimed-docs] In CI, you can automate code review and issue triage with GitHub Actions or GitLab CI/CD.
                                    • [community] I've been using Claude Code daily for months on a project with Elixir, Rust, and Python. The worst failure mode is when it does a replace_al…
                                    • [community] I have found that Claude Opus 4.6 is a better reviewer than it is an implementer. When Codex implements and Claude reviews, it's usually jus…
                                    Google Antigravitypartialcommunity6/10

                                    Docs describe a Review Changes/diff viewer (with whitespace filtering, syntax highlighting) and Artifacts containing code diffs, plus mandatory human approval before changes are applied, and subagents/CI headless mode that can run tests. This covers diff inspection and pre-merge gating, but there's no dedicated 'run checks' feature (e.g., integrated linting/test-run summary) beyond subagent test delegation, and no independent hands-on confirmation that this workflow reliably catches problems — community reports instead highlight safety failures (accidental deletion, data exfiltration) that occurred despite review/approval mechanisms. Missing for 10: independent verification that diff review + checks actually catch bugs pre-merge, and a dedicated automated check/test-report feature beyond ad-hoc subagent delegation.

                                    • [claimed-docs] Planning Mode: The agent plans thoroughly before executing tasks... produces structured implementation plans called Artifacts
                                    • [claimed-docs] The agent always halts and requests your explicit approval before proceeding with proposed changes.
                                    • [claimed-docs] Artifacts include rich markdown plans (Implementation Plans), code diffs, architecture diagrams, images, and browser recordings.
                                    • [claimed-docs] Added a "Hide Whitespace Changes" option to the Review Changes overflow menu and file diff viewers to filter out whitespace-only edits.
                                    • [claimed-docs] Code and data artifacts like SQL and JSONL files now open in a virtualized viewer with syntax highlighting and line numbers
                                    • [claimed-docs] an agent can delegate tasks—such as running tests or performing extensive codebase searches—to dedicated subagents.
                                    • [claimed-docs] Run Antigravity CLI non-interactively to script agent tasks, integrate with CI pipelines, and capture machine-readable output.
                                    • [community] Google Antigravity just deleted the contents of whole drive - came down to commanding a deletion of a 'directory with space in the name' wit…
                                    • [community] absolutely no sympathy for someone running Antigravity in Turbo mode (this is not the default and it clearly states that Antigravity auto-ex…

                                  Safe execution

                                  1. engineering-leadControl which external tools and integrations the agent is allowed to access

                                    weight 2 · round to Claude Code
                                    Claude Codepartialclaimed6/10

                                    Claude Code supports MCP server allow-listing via config (claude mcp add), sandboxed Bash tool with filesystem/network domain controls, and Enterprise-tier managed policy settings/SSO/role-based permissions that let an engineering lead govern tool and integration access. However, evidence doesn't show granular per-tool allow/deny lists at a team-policy level outside Enterprise, nor independent confirmation these controls reliably block unauthorized MCP/tool use in practice. missing for 10: fine-grained non-enterprise tool permission controls, independent/hands-on verification that access restrictions are enforced, and centralized audit/reporting of which integrations were actually used.

                                    • [claimed-docs] Claude Code can connect to hundreds of external tools and data sources through the Model Context Protocol (MCP)
                                    • [claimed-docs] claude mcp add --transport http notion https://mcp.notion.com/mcp
                                    • [claimed-docs] Stdio servers run as local processes on your machine. They're ideal for tools that need direct system access or custom scripts.
                                    • [claimed-docs] Use Claude Code as an MCP server. You can use Claude Code itself as an MCP server that other applications can connect to: claude mcp serve (…
                                    • [claimed-docs] Learn how Claude Code's sandboxed Bash tool provides filesystem and network isolation for safer, more autonomous agent execution. The Bash s…
                                    • [claimed-docs] Claude for Enterprise: adds SSO, domain capture, role-based permissions, compliance API, and managed policy settings for organization-wide C…
                                    • [claimed-docs] Single sign-on (SSO/SAML) and domain capture
                                    Google Antigravitydisputedcontradicted5/10

                                    Antigravity docs describe granular controls—Deny/Ask/Allow permission lists, MCP server configuration, plugins bundling MCP servers, and sandboxing that blocks sensitive files—giving engineering leads levers to restrict tool/integration access (antigravity-docs-23, antigravity-docs-42, antigravity-docs-12, antigravity-docs-28, antigravity-docs-48). However, independent security reports document that these controls were bypassed in practice: Gemini accessed .env files despite being configured not to, and the default Allowlist shipped with webhook.site, which was used as a live exfiltration vector—directly contradicting the claim that admins can reliably restrict external access (antigravity-comm-11, antigravity-comm-12, antigravity-comm-13). Missing for 10: evidence of a fix/patch to these bypasses, and no first-party acknowledgment/remediation documentation confirming the control now holds as designed.

                                    • [claimed-docs] Permissions are evaluated across three distinct access lists: Deny... Ask... Allow
                                    • [claimed-docs] Permissions are evaluated across three distinct access lists: Deny...Ask...Allow
                                    • [claimed-docs] Access plugins, MCP, skills, and hooks configurations instantly via slash commands, quickly enhancing your workflow.
                                    • [claimed-docs] Plugins are namespaced bundles that allow you to extend Antigravity’s capabilities by grouping skills, rules, MCP servers, and hooks into a …
                                    • [claimed-docs] Sensitive files like ~/.ssh and .env are blocked, anything not explicitly mounted is invisible inside the sandbox
                                    • [community] Google Antigravity exfiltrates data via indirect prompt injection attack: Gemini is not supposed to have access to .env files with default s…
                                    • [community] The default Allowlist provided with Antigravity includes 'webhook.site', which was used as an exfiltration vector for secrets.
                                    • [community] Antigravity was also vulnerable to the classic Markdown image exfiltration bug, reported a few days prior and flagged as 'intended behavior'…
                                  2. engineering-leadHave the agent operate inside a sandbox when interacting with code, tools, and network resources

                                    weight 2 · round to Claude Code
                                    Claude Codefullclaimed8/10

                                    Claude Code documents a dedicated sandboxed Bash tool that enforces filesystem and network isolation via OS-level boundaries, letting the agent run commands autonomously within defined limits rather than requiring per-command approval. missing for 10: independent/hands-on verification of sandbox robustness, and detail on sandboxing coverage for non-Bash tool calls (e.g., MCP tool network access).

                                    • [claimed-docs] Learn how Claude Code's sandboxed Bash tool provides filesystem and network isolation for safer, more autonomous agent execution. The Bash s…
                                    Google Antigravitydisputedcontradicted4/10

                                    Antigravity CLI docs describe a real sandbox mechanism (sensitive files like ~/.ssh and .env blocked, unmounted paths invisible) plus a permission allow/ask/deny system, suggesting sandboxed tool/network access is a documented feature. However, independent reports concretely contradict this: Gemini bypassed its own .env protection to exfiltrate secrets via indirect prompt injection using an allow-listed exfiltration endpoint, a known markdown-image exfiltration bug was dismissed as 'intended behavior,' and unrestrained terminal auto-execution led to a user's entire drive being deleted — showing the sandbox/permission boundary is not reliably enforced in practice. Missing for 10: consistent enforcement of sandbox boundaries against prompt-injection/exfiltration, first-party acknowledgment/fix of these incidents, and independent verification that the CLI's stated sandbox extends to the IDE agent's file/network access.

                                    • [claimed-docs] Sensitive files like ~/.ssh and .env are blocked, anything not explicitly mounted is invisible inside the sandbox
                                    • [claimed-docs] Permissions are evaluated across three distinct access lists: Deny... Ask... Allow
                                    • [claimed-docs] MCP lets Antigravity fetch structured context directly or execute safe actions on your behalf when needed.
                                    • [community] Google Antigravity exfiltrates data via indirect prompt injection attack: Gemini is not supposed to have access to .env files with default s…
                                    • [community] The default Allowlist provided with Antigravity includes 'webhook.site', which was used as an exfiltration vector for secrets.
                                    • [community] Antigravity was also vulnerable to the classic Markdown image exfiltration bug, reported a few days prior and flagged as 'intended behavior'…
                                    • [community] Google Antigravity just deleted the contents of whole drive - came down to commanding a deletion of a 'directory with space in the name' wit…
                                    • [community] The most useful suggestion from the Reddit thread: turn off 'Terminal Command Auto Execution' via File > Preferences > Antigravity Settings …
                                    • [community] absolutely no sympathy for someone running Antigravity in Turbo mode (this is not the default and it clearly states that Antigravity auto-ex…

                                  Security checks

                                  1. engineering-leadSee license and public-code matching references for AI-suggested code

                                    weight 1 · round drawn
                                    Claude Codenone0/10

                                    No evidence anywhere in the pack of license detection, public-code/OSS match references, or provenance attribution for AI-suggested code; Claude Code's documented features focus on code generation, review, MCP integrations, and workflow automation, not license/plagiarism matching.

                                      Google Antigravitynone0/10

                                      No evidence in the pack mentions license compliance checks, public-code/OSS matching, provenance detection, or any similar review-safety feature for AI-suggested code; the docs focus on agents, artifacts, permissions, and workflow tooling with no mention of license scanning.

                                      • developerGet contextual explanations and automatic fixes for security vulnerabilities

                                        weight 2 · round to Claude Code
                                        Claude Codefullclaimed7/10

                                        Anthropic's enterprise docs explicitly state Claude Code reviews code for security issues, drafts patches, and explains risk in plain language, directly matching the story's contextual-explanation-plus-fix pattern, and this is reinforced by automatic PR code review integration. Missing for 10: independent/hands-on evidence confirming automatic vulnerability fixes work reliably in practice, and more detail on the security-specific workflow beyond a single marketing mention.

                                        • [claimed-docs] Claude helps security teams and developers by reviewing code for security issues, drafts patches, and explains the risk in language your who…
                                        • [claimed-docs] Get automatic code review on every PR | GitHub Code Review
                                        • [claimed-docs] In CI, you can automate code review and issue triage with GitHub Actions or GitLab CI/CD.
                                        Google Antigravitynone0/10

                                        No evidence that Antigravity provides security-vulnerability-specific explanations or automatic fixes; the docs describe general agentic coding, planning, and review features but never mention vulnerability scanning or security remediation. Community evidence instead highlights security *problems* in Antigravity itself (prompt injection exfiltration), not a vulnerability-fixing capability for users' code.

                                        Not comparable on these axes

                                        1. ai-native userConnect an agent via an official MCP server

                                          weight 3 · not comparable
                                          Claude Codefullclaimed9/10

                                          Claude Code documents `claude mcp serve` to run itself as a stdio MCP server that other applications can connect to, in addition to being an MCP client that connects to hundreds of external servers. missing for 10: independent/hands-on third-party confirmation of the `claude mcp serve` server mode in actual use.

                                          • [claimed-docs] Use Claude Code as an MCP server. You can use Claude Code itself as an MCP server that other applications can connect to: claude mcp serve (…
                                          • [claimed-docs] Claude Code can connect to hundreds of external tools and data sources through the Model Context Protocol (MCP)
                                          • [claimed-docs] Stdio servers run as local processes on your machine. They're ideal for tools that need direct system access or custom scripts.
                                          • [claimed-docs] claude mcp add --transport http notion https://mcp.notion.com/mcp
                                          Google Antigravityn/a

                                          Antigravity is itself an agentic coding product (IDE/CLI/SDK) that acts as an MCP client—connecting to external MCP servers for tools/context (antigravity-docs-12, antigravity-docs-14, antigravity-docs-22, antigravity-docs-43)—rather than exposing itself as an MCP server for other agents to connect to. Per the agent-role exception, this axis (serving an official MCP server) does not apply to a product that is itself the agent/client.

                                          • [claimed-docs] Access plugins, MCP, skills, and hooks configurations instantly via slash commands, quickly enhancing your workflow.
                                          • [claimed-docs] Layer custom Python callables, Model Context Protocol (MCP) servers, and reusable agent skills over our built-in filesystem and terminal too…
                                          • [claimed-docs] MCP lets Antigravity fetch structured context directly or execute safe actions on your behalf when needed.
                                          • [claimed-docs] lets AI agents and editors securely connect to local developer tools, databases, file parsers, and external remote APIs
                                        2. ai-native userExplore an interactive API reference with runnable examples

                                          weight 2 · not comparable
                                          Claude Codenone0/10

                                          The evidence pack shows standard documentation pages and an Agent SDK reference, but nothing describing an interactive API reference with runnable/executable code examples (e.g., an in-browser sandbox or live API explorer). No such capability is evidenced anywhere in the docs, GitHub, or community items.

                                            Google Antigravityn/a

                                            Antigravity is an agentic coding IDE/CLI/SDK product, not an API/SaaS service exposing a public API surface meant for interactive exploration; the probe explicitly found no OpenAPI spec. An interactive API reference with runnable examples is not a fair axis for this kind of developer tool.

                                            • [probe] PROBE openapi: all candidate paths 404 (https://antigravity.google/openapi.json, https://antigravity.google/swagger.json, https://antigravit…
                                          • developerConfigure a reproducible cloud environment with the dependencies and setup steps my repository needs

                                            weight 2 · not comparable
                                            Claude Codepartialclaimed3/10

                                            Docs mention running Claude Code in the cloud/browser with no local setup and working on repos you don't have locally, implying some environment is provisioned, but there's no documentation of configuring a reproducible environment (e.g., setup scripts, dependency installation, devcontainer-style config) for cloud sessions. missing for 10: explicit environment/config file for cloud sandboxes, dependency installation steps, reproducibility guarantees across runs.

                                            • [claimed-docs] Review diffs visually, run multiple sessions side by side, schedule recurring tasks, and kick off cloud sessions.
                                            • [claimed-docs] Kick off long-running tasks and check back when they're done, work on repos you don't have locally, or run multiple tasks in parallel.
                                            • [claimed-docs] Run Claude Code in your browser with no local setup. Kick off long-running tasks and check back when they're done, work on repos you don't h…
                                            Google Antigravityn/a

                                            Antigravity is a local IDE/CLI/agent orchestration tool operating on a developer's own machine (or remote desktop sessions), not a cloud environment provisioning/dev-container service; there is no evidence of configuring reproducible cloud sandboxes with dependency/setup steps tied to a repo. This axis is a category error for this product type.