Try itExperimental
See what an agent can do with Context.dev before you ever sign up. Pick a story: recorded sessions replay real probe-harness transcripts; commands tagged live-capable can re-run against the real endpoint from our edge, right now (▶ run live — the exact same request, live and recorded lines always labeled); the live MCP handshake runs real requests from our edge, right now — including, where the server allows it, one real read-only tool call (bring your own key for auth-gated servers); sandboxed self-drive sessions are designed and gated (docs/TRY-IT.md).
$curl -si https://api.context.dev/v1/web/scrape/markdownrecorded session — replayed, not liveVerified integrations
No integration evidence found in our corpus for this product yet — that means none was found, never that it doesn’t integrate.
By theme — the product's score on each story themeBy theme
Agenticness — how well agents can access and operate the productAgenticnessevidence →
How well agents can access and operate the product
Anti bot — getting past bot defenses — CAPTCHAs, fingerprinting, blocksAnti botevidence →
Getting past bot defenses — CAPTCHAs, fingerprinting, blocks
Automation depth — how much of the product can run unattendedAutomation depthevidence →
How much of the product can run unattended
Dev experience — day-to-day developer experience — setup friction, docs, debugging, iteration speedDev experienceevidence →
Day-to-day developer experience — setup friction, docs, debugging, iteration speed
Extraction quality — how faithfully content is extracted — structure, fidelity, edge casesExtraction qualityevidence →
How faithfully content is extracted — structure, fidelity, edge cases
Js rendering — handling JavaScript-heavy pages — rendering, waiting, dynamic contentJs renderingevidence →
Handling JavaScript-heavy pages — rendering, waiting, dynamic content
Openness — open source, data portability, and self-hosting storiesOpennessevidence →
Open source, data portability, and self-hosting stories
Output formats — stories about output formats in this arenaOutput formatsevidence →
Stories about output formats in this arena
Pricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limitsevidence →
Free-tier ceilings, usage caps, and rate limits before you have to pay
Privacy posture — data-handling and privacy storiesPrivacy postureevidence →
Data-handling and privacy stories
Scale reliability — behavior under load — scaling limits, uptime, failure handlingScale reliabilityevidence →
Behavior under load — scaling limits, uptime, failure handling
Story verdicts — every judged story with its evidenceStory verdicts
Follow the green: where the map greys out is where Context.dev stops today. ✓ full · ~ partial · ! disputed · — none · n/a not applicable.
Agenticness — how well agents can access and operate the productAgenticness
How well agents can access and operate the product
API surface
Drive the product through a documented public API
✓9/10
unlocks → Official SDKs · Versioning policy · API sandbox · Apply a preset configuration tuned for research agents that returns structured, citable output
Subscribe to events via webhooks
~5/10
Build against official SDKs
—0/10
Issue scoped/least-privilege API credentials for an agent
✓7/10
Connect an agent via an official MCP server
✓8/10
Download a machine-readable API spec (OpenAPI or equivalent)
✓9/10
unlocks → Interactive API docs · Official SDKs
Rely on versioned APIs with a documented deprecation policy
—–
Test against a sandbox environment without touching production data
—–
Explore an interactive API reference with runnable examples
—0/10
Agentic features
Delegate tasks to a built-in AI assistant inside the product
—–
Operate the product with natural-language commands
~6/10
Plug MCP servers into this product so it can use their tools
n/an/a
Get AI-generated insights and suggestions from my data inside the product
—0/10
Set up automations that run autonomously in the background
~6/10
Apply a preset configuration tuned for research agents that returns structured, citable output
—–
The documented rate limit (requests per second or minute) enforced on my API key before throttling kicks in
~4/10
Anti bot — getting past bot defenses — CAPTCHAs, fingerprinting, blocksAnti bot
Getting past bot defenses — CAPTCHAs, fingerprinting, blocks
Block evasion
Proxy rotation
Automation depth — how much of the product can run unattendedAutomation depth
How much of the product can run unattended
Dev experience — day-to-day developer experience — setup friction, docs, debugging, iteration speedDev experience
Day-to-day developer experience — setup friction, docs, debugging, iteration speed
Share scrapers with teammates and manage organizations and role-based permissions
—0/10
Deployment flexibility
Connect the scraping API to no-code automation platforms like n8n or Zapier through a prebuilt connector
—–
Build scrapers using popular open-source automation libraries like Playwright, Puppeteer, Selenium, or Scrapy
—–
Export my scraped data and job configurations in a portable format to migrate to another provider without lock-in
—–
Extraction quality — how faithfully content is extracted — structure, fidelity, edge casesExtraction quality
How faithfully content is extracted — structure, fidelity, edge cases
Ai extraction
Scrape a web page with a single API call and get its raw HTML back
~5/10
Automatically detect and filter personally identifiable information out of scraped content before it reaches storage
—–
Extract text content from PDFs, Word, Excel, and PowerPoint files without hosting them myself
✓8/10
Get automatic captions for images on a page so a text-only model can reason about visual content
—–
Search the web and get full page content from results in a single call instead of just links and snippets
—0/10
Extract specific fields from a page using CSS or XPath selector rules
—–
Extract data from very large tables using intelligent chunking so it fits within processing limits
—–
Js rendering — handling JavaScript-heavy pages — rendering, waiting, dynamic contentJs rendering
Handling JavaScript-heavy pages — rendering, waiting, dynamic content
Headless rendering
Interactive automation
Control the browser viewport width and height when rendering a page
~3/10
Openness — open source, data portability, and self-hosting storiesOpenness
Open source, data portability, and self-hosting stories
Output formats — stories about output formats in this arenaOutput formats
Stories about output formats in this arena
Pricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limits
Free-tier ceilings, usage caps, and rate limits before you have to pay
Cost optimization
Cost transparency
Trade off latency against completeness by controlling exactly when content is returned
~6/10
The maximum concurrent sessions or requests allowed on my pricing tier and the cost to raise that cap
—0/10
Privacy posture — data-handling and privacy storiesPrivacy posture
Data-handling and privacy stories
Scale reliability — behavior under load — scaling limits, uptime, failure handlingScale reliability
Behavior under load — scaling limits, uptime, failure handling
Rely on adaptive crawling that automatically stops once enough information has been gathered to answer my query
—0/10
Batch processing
Spin up many concurrent scraping sessions to gather data at scale
~5/10
Configure the crawler to respect robots.txt rules and target-site rate limits automatically
—0/10
Resume a crashed deep crawl from a saved checkpoint instead of restarting from scratch
—0/10
Check a public status page showing uptime history and past incident postmortems before committing to the service
—–
Scheduling monitoring
Sorted by importance (agentic first) (high → low) · 96/96 stories · click a row’s chevron for the rationale and evidence
Drive the product through a documented public API G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 3 | full | 9/10 | Tprobed | |
Connect an agent via an official MCP server G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 3 | full | 8/10 | Tprobed | |
The documented rate limit (requests per second or minute) enforced on my API key before throttling kicks in G Api quality | data-engineer | Agenticness — how well agents can access and operate the productAgenticness | 3 | partial | 4/10 | Cclaimed | |
Plug MCP servers into this product so it can use their tools G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 3 | n/a | 0/10 | ||
Delegate tasks to a built-in AI assistant inside the product G Agentic features | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 3 | none | untested | none yet | |
Download a machine-readable API spec (OpenAPI or equivalent) G Api quality | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | full | 9/10 | Tprobed | |
Point an agent at llms.txt or agent-oriented docs G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | full | 9/10 | Tprobed | |
Run the product headlessly / in CI for automation G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | full | 8/10 | Tprobed | |
Issue scoped/least-privilege API credentials for an agent G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | full | 7/10 | Cclaimed | |
Operate the product with natural-language commands G Agentic features | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | partial | 6/10 | Tprobed | |
Set up automations that run autonomously in the background G Agentic features | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | partial | 6/10 | Cclaimed | |
Use an official CLI G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | partial | 6/10 | Tprobed | |
Subscribe to events via webhooks G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | partial | 5/10 | Cclaimed | |
Build against official SDKs G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | none | 0/10 | ||
Explore an interactive API reference with runnable examples G Api quality | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | none | 0/10 | ||
Get AI-generated insights and suggestions from my data inside the product G Agentic features | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | none | 0/10 | ||
Apply a preset configuration tuned for research agents that returns structured, citable output C Agentic features | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | none | untested | none yet | |
Rely on versioned APIs with a documented deprecation policy G Api quality | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | none | untested | none yet | |
Test against a sandbox environment without touching production data G Api quality | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 1 | none | untested | none yet | |
Crawl an entire website and get content from all its pages with one request C Site crawling | developer | Scale reliability — behavior under load — scaling limits, uptime, failure handlingScale reliability | 3 | full | 8/10 | Cclaimed | |
Get clean LLM-ready text directly instead of dealing with blocking, rendering, and messy HTML myself C Llm ready output | ai-native user | Output formats — stories about output formats in this arenaOutput formats | 3 | full | 8/10 | Cclaimed | |
Receive scraped content as clean markdown instead of raw HTML C Content formats | developer | Output formats — stories about output formats in this arenaOutput formats | 3 | full | 8/10 | Cclaimed | |
Receive scraped content as structured JSON C Content formats | developer | Output formats — stories about output formats in this arenaOutput formats | 3 | full | 8/10 | Tprobed | |
Script page interactions like clicking, filling inputs, and scrolling before content is returned C Interactive automation | developer | Js rendering — handling JavaScript-heavy pages — rendering, waiting, dynamic contentJs rendering | 3 | full | 8/10 | Cclaimed | |
Batch scrape thousands of URLs asynchronously C Batch processing | data-engineer | Scale reliability — behavior under load — scaling limits, uptime, failure handlingScale reliability | 3 | partial | 7/10 | Cclaimed | |
Extract structured data from a page using natural language instructions instead of writing selectors C Ai extraction | developer | Extraction quality — how faithfully content is extracted — structure, fidelity, edge casesExtraction quality | 3 | full | 7/10 | Cclaimed | |
Render JavaScript-heavy single-page applications and get the fully rendered HTML C Headless rendering | developer | Js rendering — handling JavaScript-heavy pages — rendering, waiting, dynamic contentJs rendering | 3 | partial | 6/10 | Cclaimed | |
Define rules that trigger actions automatically on events G | ai-native user | Automation depth — how much of the product can run unattendedAutomation depth | 3 | partial | 5/10 | Cclaimed | |
Scrape a web page with a single API call and get its raw HTML back C Basic scraping | developer | Extraction quality — how faithfully content is extracted — structure, fidelity, edge casesExtraction quality | 3 | partial | 5/10 | Cclaimed | |
Spin up many concurrent scraping sessions to gather data at scale C Concurrency | data-engineer | Scale reliability — behavior under load — scaling limits, uptime, failure handlingScale reliability | 3 | partial | 5/10 | Xcommunity | |
Export all of my data in open formats and leave G | ai-native user | Openness — open source, data portability, and self-hosting storiesOpenness | 3 | partial | 4/10 | Cclaimed | |
Route requests through a rotating pool of proxy IPs to avoid blocks C Proxy rotation | developer | Anti bot — getting past bot defenses — CAPTCHAs, fingerprinting, blocksAnti bot | 3 | none | 0/10 | ||
Search the web and get full page content from results in a single call instead of just links and snippets C Search integration | developer | Extraction quality — how faithfully content is extracted — structure, fidelity, edge casesExtraction quality | 3 | none | 0/10 | ||
Use an undetected browser mode to bypass sophisticated bot detection systems C Block evasion | developer | Anti bot — getting past bot defenses — CAPTCHAs, fingerprinting, blocksAnti bot | 3 | none | 0/10 | ||
Use premium residential or datacenter proxies to bypass sites that are hard to scrape C Proxy rotation | developer | Anti bot — getting past bot defenses — CAPTCHAs, fingerprinting, blocksAnti bot | 3 | none | 0/10 | ||
Export my scraped data and job configurations in a portable format to migrate to another provider without lock-in C Migration lock in | developer | Dev experience — day-to-day developer experience — setup friction, docs, debugging, iteration speedDev experience | 3 | none | untested | none yet | |
Extract specific fields from a page using CSS or XPath selector rules C Selector extraction | developer | Extraction quality — how faithfully content is extracted — structure, fidelity, edge casesExtraction quality | 3 | none | untested | none yet | |
Prevent my data from being used to train AI models G | ai-native user | Privacy posture — data-handling and privacy storiesPrivacy posture | 3 | none | untested | none yet | |
Self-host the core product G | ai-native user | Openness — open source, data portability, and self-hosting storiesOpenness | 3 | none | untested | none yet | |
Set a spending cap or usage alert so proxy/credit consumption doesn't silently blow past my budget G Cost transparency | developer | Pricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limits | 3 | none | untested | none yet | |
Whether exceeding my plan's monthly credit or request quota triggers overage charges or a hard cutoff G Cost transparency | developer | Pricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limits | 3 | none | untested | none yet | |
Capture a screenshot of a full page or a specific selected area C Visual capture | developer | Output formats — stories about output formats in this arenaOutput formats | 2 | full | 8/10 | Cclaimed | |
Do everything through the API that I can do in the UI G | ai-native user | Openness — open source, data portability, and self-hosting storiesOpenness | 2 | full | 8/10 | Tprobed | |
Extract text content from PDFs, Word, Excel, and PowerPoint files without hosting them myself C Document extraction | data-engineer | Extraction quality — how faithfully content is extracted — structure, fidelity, edge casesExtraction quality | 2 | full | 8/10 | Cclaimed | |
Instantly discover all URLs on a website without fully crawling it C Site crawling | developer | Scale reliability — behavior under load — scaling limits, uptime, failure handlingScale reliability | 2 | full | 8/10 | Cclaimed | |
Pass a JSON schema so the API returns structured data matching that schema C Ai extraction | developer | Extraction quality — how faithfully content is extracted — structure, fidelity, edge casesExtraction quality | 2 | full | 8/10 | Tprobed | |
Choose exactly which output format is returned, such as markdown, HTML, text, or frontmatter C Content formats | developer | Output formats — stories about output formats in this arenaOutput formats | 2 | partial | 6/10 | Cclaimed | |
Have an LLM read a page and decide what structured fields to pull out without pre-written selectors C Ai extraction | ai-native user | Extraction quality — how faithfully content is extracted — structure, fidelity, edge casesExtraction quality | 2 | partial | 6/10 | Cclaimed | |
Have the API wait for a specific selector to appear before returning the rendered page C Headless rendering | developer | Js rendering — handling JavaScript-heavy pages — rendering, waiting, dynamic contentJs rendering | 2 | partial | 6/10 | Cclaimed | |
Monitor target pages for content changes, such as price or listing updates, and get notified as they happen C Scheduling monitoring | data-engineer | Scale reliability — behavior under load — scaling limits, uptime, failure handlingScale reliability | 2 | partial | 6/10 | Cclaimed | |
Perform bulk operations across many items at once G | ai-native user | Automation depth — how much of the product can run unattendedAutomation depth | 2 | partial | 6/10 | Cclaimed | |
Schedule recurring jobs or workflows G | ai-native user | Automation depth — how much of the product can run unattendedAutomation depth | 2 | partial | 5/10 | Cclaimed | |
Schedule scraping jobs to run automatically at specific times C Scheduling monitoring | developer | Scale reliability — behavior under load — scaling limits, uptime, failure handlingScale reliability | 2 | partial | 5/10 | Cclaimed | |
Whether failed, blocked, or empty-result requests still consume my billing quota G Cost transparency | developer | Pricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limits | 2 | partial | 5/10 | Cclaimed | |
Keep interacting with an already-scraped page, clicking and filling forms to reach content behind a login wall C Interactive automation | developer | Js rendering — handling JavaScript-heavy pages — rendering, waiting, dynamic contentJs rendering | 2 | partial | 4/10 | Cclaimed | |
Monitor job performance, validate data quality, and receive alerts when something fails C Scheduling monitoring | data-engineer | Scale reliability — behavior under load — scaling limits, uptime, failure handlingScale reliability | 2 | partial | 4/10 | Cclaimed | |
Run a deep crawl using a breadth-first strategy with a configurable maximum page limit C Site crawling | data-engineer | Scale reliability — behavior under load — scaling limits, uptime, failure handlingScale reliability | 2 | partial | 4/10 | Cclaimed | |
Automatically retry through a chain of different proxies when anti-bot detection blocks a request C Block evasion | data-engineer | Anti bot — getting past bot defenses — CAPTCHAs, fingerprinting, blocksAnti bot | 2 | none | 0/10 | ||
Configure the crawler to respect robots.txt rules and target-site rate limits automatically C Crawl compliance | data-engineer | Scale reliability — behavior under load — scaling limits, uptime, failure handlingScale reliability | 2 | none | 0/10 | ||
Have an agent automatically get past a CAPTCHA, login, or form wall without my manual intervention C Block evasion | ai-native user | Anti bot — getting past bot defenses — CAPTCHAs, fingerprinting, blocksAnti bot | 2 | none | 0/10 | ||
Rely on adaptive crawling that automatically stops once enough information has been gathered to answer my query C Ai driven crawling | ai-native user | Scale reliability — behavior under load — scaling limits, uptime, failure handlingScale reliability | 2 | none | 0/10 | ||
Request a proxy from a specific country to get geolocation-appropriate content C Proxy rotation | developer | Anti bot — getting past bot defenses — CAPTCHAs, fingerprinting, blocksAnti bot | 2 | none | 0/10 | ||
Request semantically chunked output instead of one large content blob, so it feeds cleanly into a retrieval pipeline C Llm ready output | ai-native user | Output formats — stories about output formats in this arenaOutput formats | 2 | none | 0/10 | ||
Resume a crashed deep crawl from a saved checkpoint instead of restarting from scratch C Fault tolerance | data-engineer | Scale reliability — behavior under load — scaling limits, uptime, failure handlingScale reliability | 2 | none | 0/10 | ||
Reuse a persistent browser profile with saved cookies and login state across multiple requests C Session persistence | developer | Js rendering — handling JavaScript-heavy pages — rendering, waiting, dynamic contentJs rendering | 2 | none | 0/10 | ||
Route multiple requests through the same proxy IP using a session identifier to maintain a consistent identity P Proxy rotation | developer | Anti bot — getting past bot defenses — CAPTCHAs, fingerprinting, blocksAnti bot | 2 | none | 0/10 | ||
Run a ready-made scraper from a marketplace instead of building one from scratch C Quickstart | developer | Dev experience — day-to-day developer experience — setup friction, docs, debugging, iteration speedDev experience | 2 | none | 0/10 | ||
Share scrapers with teammates and manage organizations and role-based permissions P Collaboration | developer | Dev experience — day-to-day developer experience — setup friction, docs, debugging, iteration speedDev experience | 2 | none | 0/10 | ||
The maximum concurrent sessions or requests allowed on my pricing tier and the cost to raise that cap G Plan scale limits | data-engineer | Pricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limits | 2 | none | 0/10 | ||
Access a managed remote browser sandbox for interactive, manual browsing workflows C Interactive automation | developer | Js rendering — handling JavaScript-heavy pages — rendering, waiting, dynamic contentJs rendering | 2 | none | untested | none yet | |
Automatically detect and filter personally identifiable information out of scraped content before it reaches storage C Data safety | data-engineer | Extraction quality — how faithfully content is extracted — structure, fidelity, edge casesExtraction quality | 2 | none | untested | none yet | |
Build and deploy custom serverless scraping scripts on the platform without managing my own infrastructure C Deployment flexibility | developer | Dev experience — day-to-day developer experience — setup friction, docs, debugging, iteration speedDev experience | 2 | none | untested | none yet | |
Build scrapers using popular open-source automation libraries like Playwright, Puppeteer, Selenium, or Scrapy C Library compatibility | developer | Dev experience — day-to-day developer experience — setup friction, docs, debugging, iteration speedDev experience | 2 | none | untested | none yet | |
Check a public status page showing uptime history and past incident postmortems before committing to the service G Operational transparency | data-engineer | Scale reliability — behavior under load — scaling limits, uptime, failure handlingScale reliability | 2 | none | untested | none yet | |
Choose where my data is stored (region/residency) G | ai-native user | Privacy posture — data-handling and privacy storiesPrivacy posture | 2 | none | untested | none yet | |
Connect the scraping API to no-code automation platforms like n8n or Zapier through a prebuilt connector C Integrations | developer | Dev experience — day-to-day developer experience — setup friction, docs, debugging, iteration speedDev experience | 2 | none | untested | none yet | |
Control data retention and deletion G | ai-native user | Privacy posture — data-handling and privacy storiesPrivacy posture | 2 | none | untested | none yet | |
Deploy the scraping service via a Docker container for production use C Deployment flexibility | developer | Dev experience — day-to-day developer experience — setup friction, docs, debugging, iteration speedDev experience | 2 | none | untested | none yet | |
Get automatic captions for images on a page so a text-only model can reason about visual content C Multimodal extraction | ai-native user | Extraction quality — how faithfully content is extracted — structure, fidelity, edge casesExtraction quality | 2 | none | untested | none yet | |
Let the API automatically pick the cheapest configuration that still succeeds C Cost optimization | developer | Pricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limits | 2 | none | untested | none yet | |
Opt out of telemetry and usage tracking G | ai-native user | Privacy posture — data-handling and privacy storiesPrivacy posture | 2 | none | untested | none yet | |
Pass my own session cookies so the API fetches pages requiring authentication C Session persistence | developer | Js rendering — handling JavaScript-heavy pages — rendering, waiting, dynamic contentJs rendering | 2 | none | untested | none yet | |
Plug in a local or self-hosted LLM as the extraction backend instead of a cloud-only model C Ai extraction | developer | Extraction quality — how faithfully content is extracted — structure, fidelity, edge casesExtraction quality | 2 | none | untested | none yet | |
Read the product's source under an open license G | ai-native user | Openness — open source, data portability, and self-hosting storiesOpenness | 2 | none | untested | none yet | |
Self-host an open-source version of the scraper instead of relying on a hosted cloud service C Deployment flexibility | developer | Dev experience — day-to-day developer experience — setup friction, docs, debugging, iteration speedDev experience | 2 | none | untested | none yet | |
Set how much reasoning effort an autonomous agent spends on a data-gathering task (low, medium, high) C Cost optimization | ai-native user | Pricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limits | 2 | none | untested | none yet | |
Trade off latency against completeness by controlling exactly when content is returned C Performance tuning | developer | Pricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limits | 1 | partial | 6/10 | Cclaimed | |
Control the browser viewport width and height when rendering a page C Render configuration | developer | Js rendering — handling JavaScript-heavy pages — rendering, waiting, dynamic contentJs rendering | 1 | partial | 3/10 | Cclaimed | |
Apply different crawl configurations to different URL patterns within a single batch job C Batch processing | developer | Scale reliability — behavior under load — scaling limits, uptime, failure handlingScale reliability | 1 | none | 0/10 | ||
Block ads on the target page to speed up scraping requests C Cost optimization | developer | Pricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limits | 1 | none | untested | none yet | |
Block images and CSS resources by default to reduce bandwidth and speed up requests C Cost optimization | developer | Pricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limits | 1 | none | untested | none yet | |
Extract data from very large tables using intelligent chunking so it fits within processing limits C Structured data handling | data-engineer | Extraction quality — how faithfully content is extracted — structure, fidelity, edge casesExtraction quality | 1 | none | untested | none yet | |
Monitor live system metrics and worker/browser pool status through a real-time dashboard C Scheduling monitoring | developer | Scale reliability — behavior under load — scaling limits, uptime, failure handlingScale reliability | 1 | none | untested | none yet | |
Publish my custom scraper to a public marketplace and earn revenue when others use it P Quickstart | developer | Dev experience — day-to-day developer experience — setup friction, docs, debugging, iteration speedDev experience | 1 | none | untested | none yet | |
Start building immediately using a library of ready-made project templates C Quickstart | developer | Dev experience — day-to-day developer experience — setup friction, docs, debugging, iteration speedDev experience | 1 | none | untested | none yet | |
Version, review, and roll back my automations G | ai-native user | Automation depth — how much of the product can run unattendedAutomation depth | 1 | n/a | untested | none yet |
Opportunities — the stories that would move this product's scores, from its own judged verdictsOpportunitiestop 8 of 77 stories with headroom
What would move Context.dev’s scores — derived from its own judged verdicts, biggest headroom first. Each line quotes what the judge found missing; shipping it (or evidencing it publicly) is the fix.
Agenticness — how well agents can access and operate the productDelegate tasks to a built-in AI assistant inside the product
nonemoves Built-in AIimpact 45
The axis applies to this product kind (peer products hold positive or none verdicts on this story), so lack of evidence for an applicable capability is "none", never "na".
Pricing limits — free-tier ceilings, usage caps, and rate limits before you have to payWhether exceeding my plan's monthly credit or request quota triggers overage charges or a hard cutoff
nonemoves PA Scoreimpact 30
No evidence pack item discusses what happens when a monthly credit or request quota is exceeded—no mention of overage billing or hard cutoffs; only per-minute rate-limit headers and timeout behavior are documented, which are unrelated to plan quota exhaustion.
Extraction quality — how faithfully content is extracted — structure, fidelity, edge casesExtract specific fields from a page using CSS or XPath selector rules
nonemoves PA Scoreimpact 30
Context.dev's extraction is schema-based (JSON Schema-driven structured extraction) with no evidence of CSS or XPath selector-based field extraction rules; docs mention Markdown conversion, crawling, and JSON-schema extraction but never selector syntax.
Dev experience — day-to-day developer experience — setup friction, docs, debugging, iteration speedExport my scraped data and job configurations in a portable format to migrate to another provider without lock-in
nonemoves PA Scoreimpact 30
No evidence of an export feature for scraped data or job configurations in a portable format, nor any migration/lock-in-avoidance tooling; data is returned via API responses (Markdown/JSON) but no mention of bulk export or config portability to another provider.
Openness — open source, data portability, and self-hosting storiesSelf-host the core product
nonemoves PA Scoreimpact 30
Context.dev is a hosted API/SaaS product (web scraping, extraction, MCP, CLI) with no evidence of an open-source core or self-hosting option; all evidence points to a cloud-only API service.
Anti bot — getting past bot defenses — CAPTCHAs, fingerprinting, blocksUse premium residential or datacenter proxies to bypass sites that are hard to scrape
nonemoves PA Scoreimpact 30
No documentation or product page mentions residential/datacenter proxies, IP rotation, or anti-bot bypass infrastructure; community comments explicitly note the absence of any proxy mention and question whether the product can handle high-value/anti-scraping targets like LinkedIn.
Privacy posture — data-handling and privacy storiesPrevent my data from being used to train AI models
nonemoves PA Scoreimpact 30
No evidence pack item addresses data-training opt-out, privacy policy on model training, or any commitment about customer data usage for AI training; the pack only covers scraping/crawling features, API key restrictions, and community pricing/proxy debates.
Anti bot — getting past bot defenses — CAPTCHAs, fingerprinting, blocksRoute requests through a rotating pool of proxy IPs to avoid blocks
nonemoves PA Scoreimpact 30
No documentation or product page mentions proxy IP rotation, residential proxies, or anti-blocking infrastructure; a community comment on Hacker News explicitly notes the homepage never mentions 'ip' and questions whether rotating/residential proxies are used at all.
Showing the top 8 of 77 — every none/partial verdict in the story verdicts table is headroom.
Think a verdict is wrong? Every verdicts-table row has a Flag link — see the methodology.
Coverage map — which docs area, API section, or community source covers which judged storiesCoverage map13 surfaces · 41 covered stories
Where the cited evidence behind each covered verdict came from — the same citations the verdicts table shows, no extra judging.
Guides docs34 stories
- Run the product headlessly / in CI for automation
- Drive the product through a documented public API
- Issue scoped/least-privilege API credentials for an agent
- Subscribe to events via webhooks
- Set up automations that run autonomously in the background
- Perform bulk operations across many items at once
- Define rules that trigger actions automatically on events
- Schedule recurring jobs or workflows
- Extract structured data from a page using natural language instructions instead of writing selectors
- Pass a JSON schema so the API returns structured data matching that schema
- Have an LLM read a page and decide what structured fields to pull out without pre-written selectors
- Scrape a web page with a single API call and get its raw HTML back
- Extract text content from PDFs, Word, Excel, and PowerPoint files without hosting them myself
- Render JavaScript-heavy single-page applications and get the fully rendered HTML
- Have the API wait for a specific selector to appear before returning the rendered page
- Keep interacting with an already-scraped page, clicking and filling forms to reach content behind a login wall
- Script page interactions like clicking, filling inputs, and scrolling before content is returned
- Control the browser viewport width and height when rendering a page
- Do everything through the API that I can do in the UI
- Export all of my data in open formats and leave
- Receive scraped content as clean markdown instead of raw HTML
- Choose exactly which output format is returned, such as markdown, HTML, text, or frontmatter
- Receive scraped content as structured JSON
- Get clean LLM-ready text directly instead of dealing with blocking, rendering, and messy HTML myself
- Capture a screenshot of a full page or a specific selected area
- Trade off latency against completeness by controlling exactly when content is returned
- Batch scrape thousands of URLs asynchronously
- Spin up many concurrent scraping sessions to gather data at scale
- Monitor target pages for content changes, such as price or listing updates, and get notified as they happen
- Monitor job performance, validate data quality, and receive alerts when something fails
- Schedule scraping jobs to run automatically at specific times
- Run a deep crawl using a breadth-first strategy with a configurable maximum page limit
- Crawl an entire website and get content from all its pages with one request
- Instantly discover all URLs on a website without fully crawling it
docs.context.dev12 stories
- Drive the product through a documented public API
- Download a machine-readable API spec (OpenAPI or equivalent)
- Extract structured data from a page using natural language instructions instead of writing selectors
- Have an LLM read a page and decide what structured fields to pull out without pre-written selectors
- Scrape a web page with a single API call and get its raw HTML back
- Do everything through the API that I can do in the UI
- Export all of my data in open formats and leave
- Receive scraped content as clean markdown instead of raw HTML
- Choose exactly which output format is returned, such as markdown, HTML, text, or frontmatter
- Receive scraped content as structured JSON
- Get clean LLM-ready text directly instead of dealing with blocking, rendering, and messy HTML myself
- Crawl an entire website and get content from all its pages with one request
Optimization docs8 stories
- Run the product headlessly / in CI for automation
- Drive the product through a documented public API
- Whether failed, blocked, or empty-result requests still consume my billing quota
- Trade off latency against completeness by controlling exactly when content is returned
- Batch scrape thousands of URLs asynchronously
- Spin up many concurrent scraping sessions to gather data at scale
- The documented rate limit (requests per second or minute) enforced on my API key before throttling kicks in
- Monitor job performance, validate data quality, and receive alerts when something fails
Install CLI docs7 stories
- Point an agent at llms.txt or agent-oriented docs
- Run the product headlessly / in CI for automation
- Use an official CLI
- Drive the product through a documented public API
- Operate the product with natural-language commands
- Do everything through the API that I can do in the UI
- Receive scraped content as structured JSON
OpenAPI spec6 stories
- Run the product headlessly / in CI for automation
- Drive the product through a documented public API
- Download a machine-readable API spec (OpenAPI or equivalent)
- Pass a JSON schema so the API returns structured data matching that schema
- Do everything through the API that I can do in the UI
- Receive scraped content as structured JSON
MCP docs3 stories
context.dev3 stories
Install MCP docs3 stories
Auth docs2 stories
Install skill docs2 stories
Probe proofs — replayable recordings from the probe harnessProbe proofs
Replayable recordings from our probe harness — see the Prove-It protocol to submit one.
$curl -si https://api.context.dev/v1/web/scrape/markdownreproduced$ curl -si https://api.context.dev/v1/web/scrape/markdown
HTTP/2 401
date: Tue, 15 Sep 2026 17:38:02 GMT
content-type: application/json; charset=utf-8
content-length: 293
cf-ray: a3b9671f7b30b8b6-SJC
cf-cache-status: DYNAMIC
server: cloudflare
vary: Origin, Accept-Encoding
access-control-expose-headers: X-Context-ZDR,X-Request-Id
alt-svc: h3=":443"; ma=86400
x-cloud-trace-context: 111bcfbe29fdcf2212fcc1c92940a30c;o=1
x-request-id: de9a3c50-1ef8-42b5-8df7-895e4f606578
speculation-rules: "/cdn-cgi/speculation"
report-to: {"group":"cf-nel","max_age":604800,"endpoints":[{"url":"https://a.nel.cloudflare.com/report/v4?s=f3tbh1Zv0c2OQ2ottuZbvZnl9X5T8jDWfAet86EYtKVhe6eCNb%2BQ7ocs9owtj1pLR3YgbdPAnJA%2F77k0T5avADHP5RUOBJnYKpSIGee4rPxy%2B5gs3ayK2%2BCUZuCvriBKwg%3D%3D"}]}
nel: {"report_to":"cf-nel","success_fraction":0.0,"max_age":604800}
{"message":"Unauthorized: No API [redacted] provided, please create one at context.dev and include it in the request header under [[redacted] <your-api-[redacted]>] (no brackets). You can read more at docs.context.dev","error_code":"NOT_FOUND","request_id":"de9a3c50-1ef8-42b5-8df7-895e4f606578"}
$curl -si -X POST https://mcp.context.dev/mcp -H 'Content-Type: application/json' -d '<jsonrpc initialize>'reproduced$ curl -si -X POST https://mcp.context.dev/mcp -H 'Content-Type: application/json' -d '<jsonrpc initialize>'
HTTP/2 401
date: Tue, 15 Sep 2026 17:38:02 GMT
content-type: application/json
cf-ray: a3b9671c59241ed2-SJC
cf-cache-status: DYNAMIC
access-control-allow-origin: *
server: cloudflare
vary: accept-encoding
www-authenticate: Bearer error="unauthorized", error_description="Authorization needed", resource_metadata="https://mcp.context.dev/.well-known/oauth-protected-resource/mcp"
access-control-expose-headers: *
{"error":"Missing Authorization header"}
$curl -s https://www.context.dev/openapi.json | head -c 300reproduced$ curl -s https://www.context.dev/openapi.json | head -c 300
{"openapi":"3.1.0","info":{"title":"Context API","description":"API for retrieving context data from any website","version":"1.0.0"},"servers":[{"url":"https://api.context.dev/v1"}],"tags":[{"name":"Batch","description":"Scrape many pages or crawl a site asynchronously."},{"name":"Monitors","descrip
Claims vs evidence — vendor claims reconciled against independent verdictsClaims vs evidence
5 of 17 testable claims verified · 0 contradicted → integrity 29/100
18 distinct capability claims found in Context.dev’s own claimed-docs/GitHub materials, reconciled against our judge’s independent verdicts.
5
Verified
12
Unverified
0
Contradicted
24
Undersold
Verified (5)
“Scrapes websites into Markdown and extracts JSON for AI agents/apps”
“Crawls relevant pages and returns data matching a custom JSON Schema with grounding/coverage/freshness controls”
Pass a JSON schema so the API returns structured data matching that schemafullproof ↗
“Lets an AI client connect to Context.dev's tools for live web and company data”
“Provides an official CLI to call Context.dev from the terminal with JSON output usable in scripts/CI”
“Provides agent-oriented documentation teaching a coding agent how to use the API”
Point an agent at llms.txt or agent-oriented docsfullproof ↗
Unverified (12)
“Scrapes websites into Markdown and extracts JSON for AI agents/apps”
Receive scraped content as clean markdown instead of raw HTMLfullproof ↗
“Crawls a small site section (up to 500 pages) and returns Markdown in one synchronous response”
Crawl an entire website and get content from all its pages with one requestfullproof ↗
“Runs large background crawls of up to 25,000 pages with progress tracking and Markdown/HTML retrieval on completion”
“Reads a site's public sitemaps to return a filtered URL list without rendering pages”
Instantly discover all URLs on a website without fully crawling itfullproof ↗
“Renders a URL and returns viewport, full-page, or offset PNG screenshots”
Capture a screenshot of a full page or a specific selected areafullproof ↗
“Supports scripted page interactions (click, wait, scroll) before scraping or extraction, with success checks”
Script page interactions like clicking, filling inputs, and scrolling before content is returnedfullproof ↗
“Converts PDFs, Office documents, and spreadsheets to Markdown, with OCR recovery for scanned PDFs”
Extract text content from PDFs, Word, Excel, and PowerPoint files without hosting them myselffullproof ↗
“Watches a page, sitemap, or dataset on a schedule and delivers signed change events”
Monitor target pages for content changes, such as price or listing updates, and get notified as they happenpartialproof ↗
“Watches a page, sitemap, or dataset on a schedule and delivers signed change events”
“Supports restricted API keys scoped to only selected operations”
Issue scoped/least-privilege API credentials for an agentfullproof ↗
“Supports a 'return-partial' mode returning completed work with a marker, or failing without charge if nothing usable exists”
Whether failed, blocked, or empty-result requests still consume my billing quotapartialproof ↗
“Exposes rate-limit headers on authenticated API responses when a per-minute limit applies”
The documented rate limit (requests per second or minute) enforced on my API key before throttling kicks inpartialproof ↗
Undersold (24)
Run the product headlessly / in CI for automationfullproof ↗
Drive the product through a documented public APIfullproof ↗
Set up automations that run autonomously in the backgroundpartialproof ↗
Operate the product with natural-language commandspartialproof ↗
Download a machine-readable API spec (OpenAPI or equivalent)fullproof ↗
Perform bulk operations across many items at oncepartialproof ↗
Define rules that trigger actions automatically on eventspartialproof ↗
Extract structured data from a page using natural language instructions instead of writing selectorsfullproof ↗
Have an LLM read a page and decide what structured fields to pull out without pre-written selectorspartialproof ↗
Scrape a web page with a single API call and get its raw HTML backpartialproof ↗
Render JavaScript-heavy single-page applications and get the fully rendered HTMLpartialproof ↗
Have the API wait for a specific selector to appear before returning the rendered pagepartialproof ↗
Keep interacting with an already-scraped page, clicking and filling forms to reach content behind a login wallpartialproof ↗
Control the browser viewport width and height when rendering a pagepartialproof ↗
Do everything through the API that I can do in the UIfullproof ↗
Export all of my data in open formats and leavepartialproof ↗
Choose exactly which output format is returned, such as markdown, HTML, text, or frontmatterpartialproof ↗
Get clean LLM-ready text directly instead of dealing with blocking, rendering, and messy HTML myselffullproof ↗
Trade off latency against completeness by controlling exactly when content is returnedpartialproof ↗
Spin up many concurrent scraping sessions to gather data at scalepartialproof ↗
Monitor job performance, validate data quality, and receive alerts when something failspartialproof ↗
Schedule scraping jobs to run automatically at specific timespartialproof ↗
Run a deep crawl using a breadth-first strategy with a configurable maximum page limitpartialproof ↗
Claims outside our story set (3)
Real capability claims found in Context.dev’s own materials, but no story in this arena’s taxonomy covers them yet — that’s feedback on the taxonomy, not a mark against the product.
“Retrieves brand profile data (logos, colors, description, social links) via the API”
source ↗“Brand data extraction includes fonts, styleguide, and address in addition to logos/colors/socials”
source ↗“Supports an OAuth-like device flow: discover, register, deliver setup link/code, user claims in browser, poll for access token, then call API”
source ↗
Business model
Free monthly credits on sign-up (work-email refills), then credit-based usage pricing per API call across scrape, crawl, extract, and brand endpoints.
pricing ↗Score trend
How this product’s scores have moved as evidence and verdicts are re-derived — a point per change, not per day.
Try Experimental
Run it in the microterminal →Recorded agent sessions — and a live MCP handshake where the vendor ships one.
Flag
⚑ Flag a verdictThink a verdict is wrong? Opens a prefilled GitHub issue — or use the ⚑ next to any verdict above.
For agents
