---
source_url: "https://www.context.dev/data/scrape-and-crawl"
title: "Scrape & Crawl - Web Scraping & Crawling APIs | Context.dev"
mirrored_at: 2026-08-15T15:38:49.685Z
host: www.context.dev
cited_in_42a: true
mirror_canonical: "https://index.42a.ai/www.context.dev/data/scrape-and-crawl"
---

> **Original source:** https://www.context.dev/data/scrape-and-crawl

[

Web Scraping API

markdown, HTML, sitemap, search, full-site crawls



](https://www.context.dev/web-scraping-api)[

Extract

structured data from any site via JSON schema



](https://www.context.dev/data/extract)[

Brand Data

logos, colors, fonts, styleguide, description, socials, address



](https://www.context.dev/data/brand-data)[

Logo Link

logo CDN



](https://www.context.dev/use-cases/logo-link)[

Pull Images

images, logos, and screenshots from any URL



](https://www.context.dev/data/pull-images)[

Classification

NAICS, SIC, transaction identification



](https://www.context.dev/data/classification)

Scrape & Crawl

## Web scraping & crawling APIs

Scrape any URL as Markdown or HTML, extract sitemaps, search the web, and crawl entire sites — with JavaScript rendering and automatic proxy escalation built in. Beyond web pages, the scrapers parse XML, PDF, DOCX, and DOC content into clean, LLM-ready output.

## Scraping & Crawling APIs

6

[

### Web Scraping API — Overview

Start here: what the web scraping API includes — output formats, JavaScript rendering, stealth, pricing, and code examples in every SDK.

](https://www.context.dev/web-scraping-api)[

### URL to Markdown API

Convert any URL to clean GitHub Flavored Markdown. Preserve or strip links and images. Ideal for LLMs and RAG pipelines.

](https://www.context.dev/data/web-scrape-markdown-api)[

### Web Scrape HTML API

Scrape raw HTML from any URL with a single call. Automatic proxy escalation handles blocked sites and returns clean output.

](https://www.context.dev/data/web-scrape-html-api)[

### Sitemap Extractor API

Extract page URLs from any website sitemap. Supports sitemap index files, parallel fetching, deduplication, and non-page resource filtering.

](https://www.context.dev/data/web-scrape-sitemap-api)[

### Website Crawler API

Crawl any site and extract page content as Markdown. Configure depth, page limits, URL filters, and subdomain following.

](https://www.context.dev/data/crawl-website-api)[

### Web Search API

Search the web and optionally scrape each result to Markdown in one round-trip. 1 credit per result.

](https://docs.context.dev/api-reference/web-scraping/search)

## One key. Every endpoint.

A single Context.dev account unlocks every API on this page. Start free and scale up only when you need to.

[Get API Access](https://www.context.dev/signup)

## Ship an agent that actually knows things.

Free tier, 10-minute integration, and the same API powering agents at Mintlify, daily.dev, and Propane. No credit card to start.