Category · 95 repos
Browser Automation & Scraping
Scrapers, crawlers, headless browsers and agents that use the web for you. Ranked by star velocity over the last 24 hours.
Showing 51–95 of 95
Fastest and cheapest web agent
Professional browser automation bookmarklets with a sleek interactive gallery and automated deployment.
The headless browser for AI agents and web scraping
Two methods to collect real Google SERP data—a free scraper for basic use and the enterprise-grade Bright Data API for high-volume demands.
A powerful, type-safe web scraping library for TypeScript.
AI-powered product monitoring and self-healing scraper platform using Bright Data.
The fastest browser for AI agents to run browser automation, built for sharing your logged-in browser state with your AI agents, like Codex or Claude Code, without disturbing you. Zero cost, zero config.
Machine Learning (AI) Scraper helper service for changedetection.io
Jev-powered model routing, memory, compaction, skill selection, computer and browser use for Hermes agents (also Claude Code and Codex)
A focused Playwright helper for setting up one xAI account and creating one API key.
A Docker-deployed AV metadata web scraper that aggregates data from multiple sites
Key Technologies Python | FastAPI | MariaDB | NLP | FinBERT | Sentence Transformers | LLM APIs | Qdrant | Streamlit | BeautifulSoup | Playwright | Git | Docker | REST APIs
Local-first cross-platform desktop workspace for Claude Code / agents: multi-agent, Git worktrees, code diffs, skill marketplace, multi-model, Computer Use, task-aware desktop pets, with WeChat, Feishu, DingTalk, Telegram, WhatsApp and H5 access.
Xiaohongshu crawler and data collection, Xiaohongshu reverse engineering, direct messages, livestreams, and an end-to-end Xiaohongshu operations solution
Stealth headless browser for AI agents — bypass Cloudflare, bot detection, and anti-scraping. Drop-in Puppeteer/Playwright replacement.
Pure standalone, zero-browser Meta Threads intelligence engine & CLI. Cloud API & SaaS at https://skelepel.id
BuildDraft.gr — Compare computer component prices in Greece from Skroutz, BestPrice, Shopflix, Snif, and e-shop.gr, with price history and PC assembly. © 2026 Teo Ampatzis
Aliyun CAPTCHA 2.0 V3 (INPAINTING) automatic slider solver — jsdom environment spoofing + multimodal AI image recognition to locate the missing piece
Find way more from the Wayback Machine, Common Crawl, Alien Vault OTX, URLScan, VirusTotal, GhostArchive & Intelligence X!
A Python open-source project for “the road to learning programming on your own,” with a comprehensive tutorial: AI lab, valuable videos, data structures, learning guides, hands-on machine learning, hands-on deep learning, web scraping, big-tech interview experiences, programmer life, and resource sharing.
A multi-task, real-time/scheduled monitoring and intelligent analysis system for Xianyu, built with Playwright and AI, with a full-featured admin UI. Helps users find products they want among Xianyu's vast listings.
privacy-first, open-source and self-hosted CAPTCHA alternative with PoW and instrumentation challenges.
RowsX is a Chrome extension that performs simple web scraping tasks for business users. It loads data from website tables into spreadsheets. Developed by Rows.com.
Daidai Didi general-purpose CAPTCHA recognition OCR, PyPI version
Chrome DevTools for coding agents
Local Google Maps lead extraction, operated by Claude. Clean CSVs with phones, emails, websites and socials. No API keys, runs on your machine.
Open-source stealth Chromium with engine-level fingerprint spoofing - de-Googled, drop-in Playwright, fully buildable and verifiable from source.
A web scraper in python to scrape data from the NUCES Lahore faculty webpage.
The complete load testing platform. Everything you need for production-grade load tests. Serverless & distributed. Load test with Playwright. Load test HTTP APIs, GraphQL, WebSocket, and more. Use any Node.js module.
The fastest, cheapest browser agent (with Jev), with deep reasoning when it matters
Open-source, local-first social media research agent for Instagram, TikTok, and LinkedIn. Jev routes read-only steps; socai CLI captures cited browser evidence.
Bulk file downloader for datanodes.to and fuckingfast.co — real-Chrome extraction over CDP, pure-HTTP extraction with a Chrome TLS fingerprint, aiohttp streaming, WebView2 GUI and a headless CLI. Windows, Python 3.10+, MIT.
Playwright Model Context Protocol Server - Tool to automate Browsers and APIs in Claude Desktop, Cline, Cursor IDE and More 🔌
scrape data from the itunes app store
HasData MCP server for web scraping and search. Google Search, Maps, Amazon, Zillow, Airbnb, YouTube and 20 more sites as structured JSON, hosted over streamable HTTP.
Self-hosted visual web scraper: record a browser session, pick the elements you want, and replay it on a schedule with Playwright. Sessions export as editable JSON flow files.
DeepSeek Harness plugin: give your agent a browser with a persistent identity - engine-level fingerprint spoofing, unlimited free local profiles, Android device emulation, passkeys that survive, and residential proxy egress.
A simple bot that uses selenium to farm Microsoft Rewards written in Python
Web fetch, search, and crawl for AI agents. Built from scratch in Rust. No keys, no accounts. AGPL v3.
A Claude Code plugin: an E2E browser-test skill pipeline (author, refine, compile, and run Playwright tests via Playwright MCP).
Dungeon Crawl: Stone Soup official repository
TikTok Scraper. Download video posts, collect user/trend/hashtag/music feed metadata, sign URL and etc.
Score a browser against itself - cross-check its JavaScript fingerprint against what the network layer actually saw. CLI + packages behind liarjs.dev.
源码级 Chromium 指纹内核,检测站按普通 Chrome 评分。Source-level Chromium fingerprint kernel that passes bot detection tests. Not JS injection — C++ patches compiled into the binary.
Anime Garden mirror site | Anime BT resource aggregation site | Open API for anime BT resources