Skip to content

Rank #3 of 7 in Model Gateways & Routers

Vercel AI Gateway logo

Vercel AI Gateway

Vercel Inc. · commercial

npm 17.4M/wknpm/wk -4.7M

Showcase

Vercel AI Gateway homepage screenshot
homepage · captured Sep 2026 · view live ↗
Vercel AI Gateway docs screenshot
docs · captured Sep 2026 · view live ↗

Vercel ships more than one product — each judged line competes in its own arena on the same stories as everyone else.

LineArenaRankPA Score
VercelEdge & App Platforms#4/625/100
v0Vibe-Coding App Builders#4/621/100
Vercel SandboxAgent Sandboxes & Code Execution#5/825/100
Vercel AI Gatewaythis pageModel Gateways & Routers#3/729/100
skills.shAgent Skills & Extensions#5/518/100

Verified integrations

Connections to other tracked products — hover a chip for the verbatim evidence quote behind it.

By theme — the product's score on each story themeBy theme

Agenticness — how well agents can access and operate the productAgenticnessevidence →

How well agents can access and operate the product

39.5/100

Automation depth — how much of the product can run unattendedAutomation depthevidence →

How much of the product can run unattended

14.4/100

Caching performance — stories about caching performance in this arenaCaching performanceevidence →

Stories about caching performance in this arena

30.0/100

Cost controls — stories about cost controls in this arenaCost controlsevidence →

Stories about cost controls in this arena

69.0/100

Key management — stories about key management in this arenaKey managementevidence →

Stories about key management in this arena

30.0/100

Observability — seeing what the system is doing — logs, metrics, traces, alertsObservabilityevidence →

Seeing what the system is doing — logs, metrics, traces, alerts

67.5/100

Openness — open source, data portability, and self-hosting storiesOpennessevidence →

Open source, data portability, and self-hosting stories

16.2/100

Privacy posture — data-handling and privacy storiesPrivacy postureevidence →

Data-handling and privacy stories

0.0/100

Routing resilience — stories about routing resilience in this arenaRouting resilienceevidence →

Stories about routing resilience in this arena

56.3/100

Streaming tools — stories about streaming tools in this arenaStreaming toolsevidence →

Stories about streaming tools in this arena

47.1/100

Unified api — stories about unified api in this arenaUnified apievidence →

Stories about unified api in this arena

90.0/100

Story verdicts — every judged story with its evidenceStory verdicts

?

Sorted by importance (agentic first) (high → low) · 50/50 stories · click a row’s chevron for the rationale and evidence

Drive the product through a documented public API G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3full9/10T

Connect an agent via an official MCP server G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3none0/10

Delegate tasks to a built-in AI assistant inside the product G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness3n/auntestednone yet

Plug MCP servers into this product so it can use their tools G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3noneuntestednone yet

Point an agent at llms.txt or agent-oriented docs G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full9/10T

Build against official SDKs G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full8/10T

Run the product headlessly / in CI for automation G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full8/10T

Download a machine-readable API spec (OpenAPI or equivalent) G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full7/10T

Explore an interactive API reference with runnable examples G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial5/10T

Issue scoped/least-privilege API credentials for an agent G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial5/10C

Use an official CLI G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial3/10C

Rely on versioned APIs with a documented deprecation policy G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Get AI-generated insights and suggestions from my data inside the product G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2n/auntestednone yet

Operate the product with natural-language commands G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2n/auntestednone yet

Set up automations that run autonomously in the background G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2n/auntestednone yet

Subscribe to events via webhooks G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2noneuntestednone yet

Test against a sandbox environment without touching production data G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness1n/auntestednone yet

Call many model providers through one consistent API C

One endpoint

developerUnified api — stories about unified api in this arenaUnified api3full9/10T

Configure automatic fallback to another model or provider when one fails C

Fallbacks

platform engineerRouting resilience — stories about routing resilience in this arenaRouting resilience3full9/10C

Point existing OpenAI-compatible code at the gateway by changing only the base URL and key C

Compatibility

developerUnified api — stories about unified api in this arenaUnified api3full9/10T

Inspect logged requests and responses with latency, token counts, and cost attached C

Logs

platform engineerObservability — seeing what the system is doing — logs, metrics, traces, alertsObservability3full8/10C

Make tool and function calls across different providers with a consistent schema C

Tool calling

developerStreaming tools — stories about streaming tools in this arenaStreaming tools3full8/10T

Set hard budgets and spend limits per key, team, or user C

Budgets

platform engineerCost controls — stories about cost controls in this arenaCost controls3full8/10C

Track spend per model, key, team, or user across all providers in one place C

Spend tracking

platform engineerCost controls — stories about cost controls in this arenaCost controls3full8/10C

Mint gateway-managed keys for teams and apps without exposing raw provider keys C

Virtual keys

platform engineerKey management — stories about key management in this arenaKey management3full7/10C

My agent can switch models mid-task by policy — cost, capability, or availability — through gateway routing rules C

Policy routing

ai-native userRouting resilience — stories about routing resilience in this arenaRouting resilience3partial7/10X

Export all of my data in open formats and leave G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness3partial5/10X

Stream token-by-token responses through the gateway from any provider C

Streaming

developerStreaming tools — stories about streaming tools in this arenaStreaming tools3partial5/10C

Define rules that trigger actions automatically on events G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth3partial4/10C

Provision gateways, keys, and budgets programmatically through an admin API C

Programmatic admin

ai-native userKey management — stories about key management in this arenaKey management3none0/10

Prevent my data from being used to train AI models G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture3noneuntestednone yet

Self-host the core product G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness3noneuntestednone yet

Browse or query a catalog of available models with pricing and context-window metadata G

Catalog

developerUnified api — stories about unified api in this arenaUnified api2full9/10T

Set automatic retry policies for transient provider errors C

Retries

platform engineerRouting resilience — stories about routing resilience in this arenaRouting resilience2full8/10C

Cache responses at the gateway to cut cost and latency on repeated requests C

Caching

developerCaching performance — stories about caching performance in this arenaCaching performance2partial6/10X

Do everything through the API that I can do in the UI G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness2partial6/10T

Give an autonomous agent its own key with budget and rate guardrails so it cannot run away on spend C

Agent guardrails

ai-native userCost controls — stories about cost controls in this arenaCost controls2partial6/10X

Load-balance traffic across providers, deployments, or keys by weight, latency, or cost C

Load balancing

platform engineerRouting resilience — stories about routing resilience in this arenaRouting resilience2partial6/10C

Bring my own provider API keys and have the gateway use them for my traffic G

Byok

developerKey management — stories about key management in this arenaKey management2disputed5/10D

Run traffic through gateway infrastructure that adds minimal latency overhead to provider calls C

Latency

platform engineerCaching performance — stories about caching performance in this arenaCaching performance2partial4/10C

Smooth provider rate limits by spreading traffic across keys and queuing or throttling requests C

Rate limits

platform engineerRouting resilience — stories about routing resilience in this arenaRouting resilience2partial4/10C

Choose where my data is stored (region/residency) G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2noneuntestednone yet

Control data retention and deletion G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2noneuntestednone yet

Opt out of telemetry and usage tracking G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2noneuntestednone yet

Perform bulk operations across many items at once G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth2noneuntestednone yet

Read the product's source under an open license G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness2noneuntestednone yet

Schedule recurring jobs or workflows G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth2n/auntestednone yet

Export gateway logs and traces to my own observability stack G

Integrations

developerObservability — seeing what the system is doing — logs, metrics, traces, alertsObservability1partial5/10C

Request structured JSON-schema outputs across providers C

Tool calling

developerStreaming tools — stories about streaming tools in this arenaStreaming tools1noneuntestednone yet

Version, review, and roll back my automations G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth1n/auntestednone yet

Opportunities — the stories that would move this product's scores, from its own judged verdictsOpportunitiestop 8 of 27 stories with headroom

What would move Vercel AI Gateway’s scores — derived from its own judged verdicts, biggest headroom first. Each line quotes what the judge found missing; shipping it (or evidencing it publicly) is the fix.

  1. Agenticness — how well agents can access and operate the productPlug MCP servers into this product so it can use their tools

    nonemoves agent-readyimpact 45

    The evidence describes Vercel AI Gateway as a model-routing/proxy layer (provider switching, fallbacks, spend monitoring, caching, coding-agent connections) but contains no mention of MCP server integration or tool-use via MCP anywhere in the docs pack.

  2. Agenticness — how well agents can access and operate the productConnect an agent via an official MCP server

    nonemoves agent-readyimpact 45

    Missing: explicit MCP server documentation, MCP endpoint URL, and demonstration of an agent connecting via MCP.

  3. Openness — open source, data portability, and self-hosting storiesSelf-host the core product

    nonemoves PA Scoreimpact 30

    Vercel AI Gateway is a hosted cloud service with no evidence of a self-hostable core product or open-source release; nothing in the docs suggests an on-prem or self-managed deployment option.

  4. Privacy posture — data-handling and privacy storiesPrevent my data from being used to train AI models

    nonemoves PA Scoreimpact 30

    Missing: any privacy policy or training-opt-out statement, zero-data-retention terms, provider-level data-use guarantees.

  5. Key management — stories about key management in this arenaProvision gateways, keys, and budgets programmatically through an admin API

    nonemoves PA Scoreimpact 30

    Missing: explicit admin/API endpoints for creating gateways, generating keys, or setting budgets programmatically, and any docs or examples showing this workflow.

  6. Agenticness — how well agents can access and operate the productSubscribe to events via webhooks

    nonemoves agent-readyimpact 30

    No evidence anywhere in the pack of webhook subscriptions or event-driven notifications from AI Gateway; the product exposes REST APIs, logs, and observability dashboards, but nothing about webhooks for events like request completion, budget alerts, or failures.

  7. Agenticness — how well agents can access and operate the productRely on versioned APIs with a documented deprecation policy

    nonemoves API qualityimpact 30

    Missing: explicit versioning strategy documentation, deprecation/sunset timelines, migration guidance for breaking changes.

  8. Agenticness — how well agents can access and operate the productUse an official CLI

    partialq3/10moves agent-readyimpact 21

    Missing: full CLI command reference/docs, examples of CLI usage for model routing/config, independent confirmation of CLI functionality.

Showing the top 8 of 27 — every none/partial verdict in the story verdicts table is headroom.

Think a verdict is wrong? Every verdicts-table row has a Flag link — see the methodology.

Coverage map — which docs area, API section, or community source covers which judged storiesCoverage map6 surfaces · 30 covered stories

Where the cited evidence behind each covered verdict came from — the same citations the verdicts table shows, no extra judging.

docs30 stories

Claims vs evidence — vendor claims reconciled against independent verdictsClaims vs evidence

5 of 14 testable claims verified · 3 contradictedintegrity 0/100

25 distinct capability claims found in Vercel AI Gateway’s own claimed-docs/GitHub materials, reconciled against our judge’s independent verdicts.

5

Verified

6

Unverified

3

Contradicted

18

Undersold

Verified (8)
Unverified (10)
Contradicted (4)
Undersold (18)
Claims outside our story set (3)

Real capability claims found in Vercel AI Gateway’s own materials, but no story in this arena’s taxonomy covers them yet — that’s feedback on the taxonomy, not a mark against the product.

  • Generate vector embeddings for search, retrieval, and other tasks

    source ↗
  • Migrate existing provider integrations to AI Gateway

    source ↗
  • Routes requests, manages fallbacks and budgets, monitors usage, and connects supported coding agents

    source ↗
Suggest a story for these →

Pricing signals

  • 0%gateway feepay-as-you-goVercel AI Gateway itself charges no markup/platform fee on top of provider token pricing; per-model per-token rates are only listed on separate Models page.source ↗as of 2026-09-07
  • freegateway feefree tierEvery team gets $5/month included credit usable on Free Tier eligible models, with lower rate limits than paid tier.source ↗as of 2026-09-07

Extracted verbatim from the vendor’s own pricing page — hover a figure for the exact quote.

Business model

usage-basedfree-tier

Pay per token at provider list prices with no gateway markup; included monthly usage on Vercel plans, then usage-based billing.

pricing ↗

Score trend

How this product’s scores have moved as evidence and verdicts are re-derived — a point per change, not per day.

PA Score20 (Sep 4 '26)29 (Sep 4 '26)
Agent-ready38 (Sep 4 '26)41 (Sep 4 '26)

Flag

⚑ Flag a verdict

Think a verdict is wrong? Opens a prefilled GitHub issue — or use the ⚑ next to any verdict above.

Badge

Embed this product's score badge →

Hotlinked SVG — always shows the live current score.

For agents

Data

Agent surface uptime llms.txt 100% · openapi.json 100% (30d, checked every 6h since Sep 8 '26)