Context.dev – Web Scraping API for AI Agents
A YC S26 API that scrapes any URL to clean markdown or structured data, extracts brand logos and colors, and monitors sites for changes.
Tag
14 posts tagged #web-scraping
Browse 14 posts tagged Web Scraping, including practical setup notes, reviews, comparisons, and workflow patterns for engineers working with AI tools.
A YC S26 API that scrapes any URL to clean markdown or structured data, extracts brand logos and colors, and monitors sites for changes.
One REST API to scrape, crawl, extract structured data, screenshot, and identify brands from any URL. Free tier included. Built for AI agents.
A fast REST API for screenshots, PDFs, web scraping, and content extraction built with Fastify and Playwright. Free tier: 200 requests/month with Go SDK and MCP server.
Context.dev (YC S26) converts any URL to clean markdown, extracts brand data, and crawls full sites via one REST API with TypeScript and Python SDKs.
A desktop studio for browser automation and web scraping. Build flows visually, watch them run live, and export plain Playwright code you own forever.
Simplex provides API-first browser automation infrastructure — headless browsers, proxies, and captcha solving without managing your own browser fleet. YC S24.
Open-source web scraping API that converts any URL into clean Markdown or structured data for AI agents. Supports search, scrape, crawl, and map endpoints with LLM-ready output.
Crawlee is an Apify-built open-source web scraping library that handles bot detection, proxy rotation and headless browsers so you focus on data extraction.
Spidra is a no-code AI web scraping platform. Point at any URL, describe what you want in plain text, and get structured data back - no CSS selectors, no proxy management.
Intuned Agent converts natural language prompts into production-ready Playwright code, fixes flows when sites change, and deploys with built-in stealth, scheduling, and auto-scaling.
MrScraper is visual web scraping platform with built-in proxies, headless browsers, and AI extraction. A no-code workflow for getting clean data from any site at scale.
API Parrot automatically reverse engineers HTTP APIs by tracing data correlations between requests, building visual flow diagrams, and exporting runnable.
Hyperbrowser spins up hundreds of headless browser sessions in secure isolated environments with sub-second launch times, captcha solving, and residential.
Cloud browsers built for AI agents that need to navigate, scrape, and interact with the web at scale. Handles proxies, stealth detection, and concurrent.