Tag
#web-scraping
8 open-source projects filed under this tag.
agent-reach
Agent Reach gives your AI coding agent internet access — read Twitter, YouTube, Reddit, GitHub, and 10+ platforms with zero API costs.
bigset
BigSet turns plain English into structured datasets from the live web. AI agents research, verify, and auto-refresh your data. Next.js 16 + Fastify + Mastra.
browser-harness
Self-healing browser harness that connects LLMs directly to Chrome via CDP — no middleware, no abstractions, just raw browser control for AI agents.
cloakbrowser
Stealth Chromium browser with 58 C++ source-level fingerprint patches — drop-in Playwright/Puppeteer replacement that scores 0.9 on reCAPTCHA v3 and passes Cloudflare Turnstile.
invisible-playwright
Stealth Firefox that passes every bot detection test. Drop-in Playwright replacement with C++ level fingerprint patching and 0.90 reCAPTCHA v3 score.
obscura
Obscura is a Rust-based headless browser for AI agents and web scraping — 7x less memory than Chrome with built-in anti-detection and MCP support.
scrapling
Adaptive Python web scraping framework with anti-bot bypass, smart element tracking, and full crawl capabilities for modern web scraping.
shardbrowser
Free open-source anti-detect browser with engine-level fingerprint spoofing, 170+ device profiles, MCP server, and multi-language SDKs for web scraping and testing.