GPU Clouds arenaGPU Clouds
GPU cloud providers renting raw accelerator compute — on-demand instances, marketplace rentals, and multi-node clusters — judged on provisioning speed, spot/interruptible economics, cluster scale, storage and data movement, availability and quota transparency, per-second billing truth, and how completely an AI agent can discover pricing and provision GPUs through APIs, CLIs, and MCP. Raw GPU rental only: providers serving models over OpenAI-compatible APIs (Together, Fireworks) are judged in AI Inference Providers, and agent-oriented code-execution clouds (Modal, E2B) in Agent Sandboxes.
54 user stories · 270 judged cells · updated 2026-09-05 · Evidence as of 2026-09-05
Leaderboard — every product ranked by evidenceLeaderboard
| 1 | usage-based vs Runpod ↗ | 63/100 | 29/100 | 38/100 | 33/100 | 18/100 | ★ 220▲ 29/yrpypi 29k/wk | unclear | 22/36 verified · 2 disputed | 42/100 integrity | ||
| 2 | usage-based vs Vast.ai ↗ | 62/100 | 30/100 | 34/100 | 17/100 | 0/100 | pypi 149.7k/wk | $0.27/ GPU-hour | 21/31 verified | 50/100 integrity | ||
| 3 | usage-based vs Vast.ai ↗ | 27/100 | 9/100 | 31/100 | 36/100 | 9/100 | $0.79/ GPU-hour | 17/27 verified · 1 disputed | 50/100 integrity | |||
| 4 | usage-based vs Vast.ai ↗ | 47/100 | 0/100 | 0/100 | 25/100 | 9/100 | $2.7/ GPU-hour | 10/21 verified | 0/100 integrity | |||
| 5 | usage-based vs Vast.ai ↗ | 34/100 | 12/100 | 0/100 | 29/100 | 0/100 | $0.76/ GPU-hour | 12/21 verified · 3 disputed | 0/100 integrity |
Best by user type — persona-weighted winnersBest by user type
Per persona, the product with the highest persona-weighted coverage over just that persona's stories — not the same ranking as the overall PA Score leaderboard above.
Best for platform-engineer
CoreWeave
39/100
Runner-up:
Lambda (18/100)
6 platform-engineer stories scored
Story matrix — every product × every judged storyStory matrix
Access connectivity — stories about access connectivity in this arenaAccess connectivity
Ide
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Access connectivity — stories about access connectivity in this arenaOpen Jupyter or connect my IDE (VS Code/Cursor) to the instance in one step | developer | fullX 8/10 | partialC 6/10 | none 0/10 | partialC 5/10 | disputedD 4/10 |
Networking
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Access connectivity — stories about access connectivity in this arenaExpose ports to serve applications from my instance and connect instances over private networking | developer | fullC 7/10 | partialC 5/10 | none 0/10 | fullC 7/10 | partialC 5/10 |
Ssh
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Access connectivity — stories about access connectivity in this arenaSSH into my GPU instance with my own keys and get root-level control of the environment | developer | fullX 7/10 | fullX 8/10 | none 0/10 | partialC 6/10 | disputedD 5/10 |
Agenticness — how well agents can access and operate the productAgenticness
Agent access
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Agenticness — how well agents can access and operate the productPoint an agent at llms.txt or agent-oriented docs | ai-native | fullT 9/10 | none 0/10 | fullT 9/10 | fullT 9/10 | partialT 5/10 |
| Agenticness — how well agents can access and operate the productRun the product headlessly / in CI for automation | ai-native | fullT 8/10 | fullT 7/10 | fullT⚿ 7/10 | fullT 8/10 | partialT 6/10 |
| Agenticness — how well agents can access and operate the productPlug MCP servers into this product so it can use their tools | ai-native | n/a | n/a | n/a | n/a | n/a |
| Agenticness — how well agents can access and operate the productConnect an agent via an official MCP server | ai-native | fullT⚿ 9/10 | none 0/10 | partialT 6/10 | none 0/10 | none 0/10 |
| Agenticness — how well agents can access and operate the productUse an official CLI | ai-native | fullT 9/10 | none 0/10 | none 0/10 | fullT 9/10 | fullT 8/10 |
| Agenticness — how well agents can access and operate the productDrive the product through a documented public API | ai-native | fullT 9/10 | fullT 9/10 | fullT⚿ 9/10 | fullT 9/10 | fullT 8/10 |
| Agenticness — how well agents can access and operate the productIssue scoped/least-privilege API credentials for an agent | ai-native | none⚿ 0/10 | none 0/10 | none 0/10 | fullC 8/10 | none 0/10 |
| Agenticness — how well agents can access and operate the productBuild against official SDKs | ai-native | partialT 5/10 | partialT 3/10 | fullT⚿ 7/10 | fullT 9/10 | partialT 6/10 |
| Agenticness — how well agents can access and operate the productSubscribe to events via webhooks | ai-native | none 0/10 | partialC 3/10 | none 0/10 | none 0/10 | none 0/10 |
Agentic features
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Agenticness — how well agents can access and operate the productGet AI-generated insights and suggestions from my data inside the product | ai-native | n/a | n/a | n/a | n/a | n/a |
| Agenticness — how well agents can access and operate the productSet up automations that run autonomously in the background | ai-native | partialT 4/10 | none 0/10 | none 0/10 | partialT 5/10 | partialC 4/10 |
| Agenticness — how well agents can access and operate the productDelegate tasks to a built-in AI assistant inside the product | ai-native | none 0/10 | n/a | n/a | none 0/10 | n/a |
| Agenticness — how well agents can access and operate the productOperate the product with natural-language commands | ai-native | fullT⚿ 8/10 | partialT 3/10 | none 0/10 | fullT 7/10 | none 0/10 |
Api quality
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Agenticness — how well agents can access and operate the productExplore an interactive API reference with runnable examples | ai-native | partialT 5/10 | none 0/10 | none 0/10 | partialT 4/10 | none 0/10 |
| Agenticness — how well agents can access and operate the productDownload a machine-readable API spec (OpenAPI or equivalent) | ai-native | fullT 9/10 | fullT 9/10 | none 0/10 | fullT 9/10 | none 0/10 |
| Agenticness — how well agents can access and operate the productTest against a sandbox environment without touching production data | ai-native | none 0/10 | none 0/10 | none 0/10 | n/a | none 0/10 |
| Agenticness — how well agents can access and operate the productRely on versioned APIs with a documented deprecation policy | ai-native | none 0/10 | partialT 3/10 | none 0/10 | none 0/10 | none 0/10 |
Automation depth — how much of the product can run unattendedAutomation depth
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Automation depth — how much of the product can run unattendedPerform bulk operations across many items at once | ai-native | none 0/10 | partialT 5/10 | partialC 5/10 | partialT 6/10 | none 0/10 |
| Automation depth — how much of the product can run unattendedDefine rules that trigger actions automatically on events | ai-native | none 0/10 | none 0/10 | none 0/10 | partialC 4/10 | none 0/10 |
| Automation depth — how much of the product can run unattendedSchedule recurring jobs or workflows | ai-native | none 0/10 | none 0/10 | none 0/10 | none 0/10 | none 0/10 |
| Automation depth — how much of the product can run unattendedVersion, review, and roll back my automations | ai-native | n/a | n/a | n/a | none 0/10 | n/a |
Capacity availability — stories about capacity availability in this arenaCapacity availability
Availability
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Capacity availability — stories about capacity availability in this arenaSee real-time GPU availability by type and region before I try to provision, instead of discovering stockouts by failure | ml-engineer | none 0/10 | none 0/10 | partialC 5/10 | fullT 8/10 | none 0/10 |
Hardware
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Capacity availability — stories about capacity availability in this arenaChoose from current-generation datacenter GPUs (H100/H200/B200 class) as well as cheaper previous-generation options | ml-engineer | partialX 4/10 | fullX 9/10 | none 0/10 | partialT 5/10 | partialC 6/10 |
Quotas
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Capacity availability — stories about capacity availability in this arenaSee documented quotas and instance limits and raise them through a defined process | platform-engineer | none 0/10 | none 0/10 | fullC 8/10 | none 0/10 | partialC 3/10 |
Clusters scale — stories about clusters scale in this arenaClusters scale
Clusters
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Clusters scale — stories about clusters scale in this arenaProvision a multi-node GPU cluster with fast interconnect for distributed training without a sales cycle | ml-engineer | fullC 7/10 | fullT 9/10 | partialT⚿ 6/10 | disputedD 5/10 | none 0/10 |
Openness — open source, data portability, and self-hosting storiesOpenness
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Openness — open source, data portability, and self-hosting storiesDo everything through the API that I can do in the UI | ai-native | partialT⚿ 7/10 | partialT 6/10 | partialT⚿ 6/10 | fullT 9/10 | partialT 6/10 |
| Openness — open source, data portability, and self-hosting storiesExport all of my data in open formats and leave | ai-native | partialC 5/10 | partialC 6/10 | partialC 3/10 | partialC 3/10 | partialT 4/10 |
| Openness — open source, data portability, and self-hosting storiesRead the product's source under an open license | ai-native | none 0/10 | n/a | n/a | none 0/10 | n/a |
| Openness — open source, data portability, and self-hosting storiesSelf-host the core product | ai-native | none 0/10 | n/a | n/a | n/a | n/a |
Pricing billing — stories about pricing billing in this arenaPricing billing
Billing
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Pricing billing — stories about pricing billing in this arenaPull usage and billing breakdowns programmatically to attribute GPU spend by team or workload | platform-engineer | none 0/10 | none 0/10 | partialC 4/10 | partialC 4/10 | none 0/10 |
| Pricing billing — stories about pricing billing in this arenaI am billed at per-second or per-minute granularity and only while my instance is actually running | ml-engineer | fullC 9/10 | fullX 8/10 | none 0/10 | none 0/10 | none 0/10 |
Discovery
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Pricing billing — stories about pricing billing in this arenaQuery the GPU catalog with live pricing and availability from a public or documented endpoint before committing any spend | ai-native | partialT 4/10 | partialT 6/10 | none 0/10 | fullT 9/10 | none 0/10 |
Pricing
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Pricing billing — stories about pricing billing in this arenaLock in reserved or committed-use discounts for sustained GPU capacity | platform-engineer | partialC 6/10 | partialC 4/10 | none 0/10 | fullC 7/10 | none 0/10 |
| Pricing billing — stories about pricing billing in this arenaSee the published per-GPU-hour price for every GPU type on a public pricing page without talking to sales | ml-engineer | partialC 5/10 | fullX 8/10 | none 0/10 | partialT 7/10 | none 0/10 |
Spot
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Pricing billing — stories about pricing billing in this arenaRent spot or interruptible GPU capacity at a deep discount with clearly documented preemption semantics | ml-engineer | none 0/10 | none 0/10 | partialC 4/10 | partialX 6/10 | none 0/10 |
Privacy posture — data-handling and privacy storiesPrivacy posture
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Privacy posture — data-handling and privacy storiesChoose where my data is stored (region/residency) | ai-native | partialC 3/10 | none 0/10 | none 0/10 | none 0/10 | partialC 4/10 |
| Privacy posture — data-handling and privacy storiesPrevent my data from being used to train AI models | ai-native | n/a | n/a | none 0/10 | n/a | n/a |
| Privacy posture — data-handling and privacy storiesControl data retention and deletion | ai-native | none 0/10 | none 0/10 | none 0/10 | partialT 3/10 | none 0/10 |
| Privacy posture — data-handling and privacy storiesOpt out of telemetry and usage tracking | ai-native | none 0/10 | n/a | none 0/10 | none 0/10 | none 0/10 |
Provisioning lifecycle — creating, updating, and tearing down resources across their lifecycleProvisioning lifecycle
Agent ops
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Provisioning lifecycle — creating, updating, and tearing down resources across their lifecycleProvision a GPU, monitor it, run a workload, and tear it down end to end through documented APIs, CLI, or MCP without a human in the console | ai-native | fullT⚿ 8/10 | partialT 6/10 | partialT⚿ 6/10 | fullT 8/10 | partialT 5/10 |
Manage
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Provisioning lifecycle — creating, updating, and tearing down resources across their lifecycleSet auto-shutdown timers or spend limits so a forgotten instance can't silently run up a huge bill | ml-engineer | none 0/10 | none 0/10 | none 0/10 | none 0/10 | partialC 6/10 |
| Provisioning lifecycle — creating, updating, and tearing down resources across their lifecycleStart, stop, restart, and terminate instances programmatically and keep paying only for what is running | developer | fullT 9/10 | disputedD 6/10 | partialT⚿ 6/10 | fullT 8/10 | partialT 5/10 |
Provision
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Provisioning lifecycle — creating, updating, and tearing down resources across their lifecycleProvision an on-demand GPU instance from the console or API and be running code on it within minutes | developer | fullT 8/10 | fullT 8/10 | partialT⚿ 6/10 | fullT 7/10 | disputedD 5/10 |
Serverless endpoints — stories about serverless endpoints in this arenaServerless endpoints
Serverless
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Serverless endpoints — stories about serverless endpoints in this arenaDeploy code to autoscaling serverless GPU workers that scale to zero, instead of managing always-on instances | developer | fullC 8/10 | none 0/10 | none 0/10 | partialC 6/10 | none 0/10 |
Storage data — storing and moving data — persistence, formats, durabilityStorage data
Data movement
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Storage data — storing and moving data — persistence, formats, durabilityMove data in and out efficiently — S3-compatible endpoints, cloud-storage sync, or documented transfer tooling | developer | fullC 8/10 | fullT 8/10 | partialC 6/10 | partialX 5/10 | none 0/10 |
Storage
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Storage data — storing and moving data — persistence, formats, durabilityAttach persistent network storage that survives instance teardown, so datasets and checkpoints outlive any single GPU rental | developer | fullC 8/10 | fullC 8/10 | partialC 5/10 | partialX 6/10 | partialX 6/10 |
Templates images — stories about templates images in this arenaTemplates images
Images
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Templates images — stories about templates images in this arenaRun my own Docker image or custom machine template with my exact environment | developer | fullX 8/10 | partialC 6/10 | partialC 3/10 | fullC 8/10 | fullX 7/10 |
Templates
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Templates images — stories about templates images in this arenaLaunch from pre-built ML templates (PyTorch, CUDA, vLLM, ComfyUI) instead of assembling an environment from scratch | developer | fullX 9/10 | partialC 5/10 | none 0/10 | fullC 8/10 | partialX 5/10 |
Trust governance — stories about trust governance in this arenaTrust governance
Compliance
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Trust governance — stories about trust governance in this arenaVerify the provider's security and compliance posture (SOC 2, data handling, datacenter tiers) before putting proprietary models on it | platform-engineer | partialX 5/10 | none 0/10 | none 0/10 | disputedD 4/10 | none 0/10 |
Governance
| Story | Persona | |||||
|---|---|---|---|---|---|---|
| Trust governance — stories about trust governance in this arenaManage team members with roles and scoped API keys so credentials and spend stay controlled | platform-engineer | none 0/10 | none 0/10 | partialC 4/10 | partialC 4/10 | none 0/10 |
Adjacent arenas — categories often shopped togetherAdjacent arenas
Shopping this category often means shopping these too.
Model Gateways & Routers arenaModel Gateways & Routers
7 products · leader: LiteLLM
AI Inference Providers arenaAI Inference Providers
7 products · leader: Groq
Local LLM Runtimes arenaLocal LLM Runtimes
7 products · leader: LocalAI
Robotics Software Platforms arenaRobotics Software Platforms
5 products · leader: ROS 2