Web content scraping
Scrape a URL into clean markdown for LLM pipelines and content processing. The pricing page also distinguishes HTML and sitemap retrieval as separate web extraction endpoints.
Context.dev is a web scraping and crawl API for AI agents and LLM applications. Pull live page content, crawl sites, and extract structured data with one REST API.
Context.dev is a web scraping and crawl API for AI agents and LLM applications. It provides a single REST API for pulling live web content, crawling sites, and extracting structured data from pages and websites.
The product is positioned as a way to replace internal scraping infrastructure or multiple vendors with one service. The site highlights outputs such as markdown, rendered HTML, sitemaps, screenshots, brand data, and JSON Schema-based extraction, plus an agentic setup flow that can be completed in the browser.
Scrape a URL into clean markdown for LLM pipelines and content processing. The pricing page also distinguishes HTML and sitemap retrieval as separate web extraction endpoints.
Crawl a website and extract structured data into a JSON Schema you define, which is useful when you need typed outputs instead of raw page text.
Crawl a site sitemap to discover page URLs across a domain, helping teams index or inventory site content before further processing.
Retrieve brand details by domain, company name, email, or stock ticker, including logos and other metadata referenced on the site.
Generate viewport or full-page screenshots, with a separate endpoint for image retrieval when enrichment is enabled.
Use the agentic setup flow to register through `auth.md`, then finish claim in the browser before polling for access.
Feed current page content into an AI agent or RAG pipeline as markdown so the model works from live web data instead of a frozen snapshot.
Crawl a website into a JSON Schema-defined output when you need structured records for a database, workflow, or downstream automation.
Inventory pages, documents, or product sections by crawling sitemaps and following site structure across a domain.
Pull brand metadata such as logos and identifiers for onboarding, enrichment, or profile display in product experiences.
Generate screenshots or scraped page assets for review, QA, or content capture workflows.
Yes. Context.dev offers a free tier. The pricing page says new users can get 500 API credits with a work email, or 250 credits with a free email provider, along with different request rates.
API credits are consumed when you call endpoints. Web scraping endpoints cost 1 credit per call, while brand-related endpoints such as logo retrieval, color extraction, font detection, and brand profiles cost 10 credits per call. Logo Link uses a separate quota.
Context.dev provides official SDKs for TypeScript, Python, Ruby, Go, and PHP.
Yes. The FAQ says subdomains are fully supported.
The service supports agentic registration. An agent can register using a user's email, deliver the setup link and code, and then the user completes the claim ceremony in the browser before the agent polls for access.
ByteAsk is a terminal-first AI coding agent for C and C++ that edits repos and verifies changes with compilers, debuggers, sanitizers, and tests.
hob is an independent workspace for coding agents, with local control over sessions, terminals, history, routing, and follow-up work.
Ably Chat is a chat API platform for custom realtime chat apps, with rooms, typing indicators, presence, reactions, message updates and usage-based pricing.
Manta AI is an autonomous web app testing tool that maps app behavior, catches regressions, and generates tests from a URL, no scripts or selectors needed.
SonOf connects to your repo and PM tool, audits your codebase, and turns approved work into shipped tickets with senior engineering review.
Ghost is a terminal-based AI assistant for chatting, code generation, and CLI tasks. Includes free models, supports Linux, macOS, Windows, and is open source.