Public catalog
Browse tools
48 shown. Search ranks against problem language; browsing defaults to stronger public signals.
Scrapling
turnstile captcha blocking scraper
Verified 2026-08-10 · 68,957 stars
Playwright
browser automation click fill forms
Verified 2026-08-10 · 20,110 stars
puppeteer
Node Chrome automation via DevTools Protocol
Verified 2026-08-10 · 8,691 stars
browser-use
agent needs to click through a website and fill forms
Verified 2026-08-10 · 6,311 stars
undetected-chromedriver
undetected chromedriver for bot walls
Verified 2026-08-10 · 5,372 stars
lightpanda
use lightpanda for scraping
Verified 2026-08-10 · 4,292 stars
playwright-mcp
Playwright MCP for agent browser control
Verified 2026-08-10 · 3,770 stars
playwright-python
browser automation and scraping in Python with Playwright
Verified 2026-08-10 · 3,509 stars
httpx / requests
async HTTP client for scrapers with timeouts and HTTP/2
Verified 2026-08-10 · 3,038 stars
appium
use appium for browser automation
Verified 2026-08-10 · 2,832 stars
Trafilatura
extract main article text and metadata from news HTML for RAG
Verified 2026-08-10 · 2,538 stars
newspaper3k
extract news article title body and authors from article URLs
Verified 2026-08-10 · 1,713 stars
puppeteer-extra
stealth puppeteer plugins for logged-in scrapes
Verified 2026-08-10 · 1,649 stars
stagehand
AI native browser act extract observe
Verified 2026-08-10 · 1,553 stars
botasaurus
all-in-one anti-bot web scraping framework
Verified 2026-08-10 · 1,300 stars
requests-html
requests-like session that can render JS with pyppeteer for simple pages
Verified 2026-08-10 · 926 stars
nodriver
async undetected chrome driver for scrapers
Verified 2026-08-10 · 913 stars
midscene
use midscene for browser automation
Verified 2026-08-10 · 837 stars
Pyppeteer
control headless Chromium from async Python for JS-rendered pages
Verified 2026-08-10 · 833 stars
firecrawl-mcp
Firecrawl as an MCP tool for agents to crawl URLs
Verified 2026-08-10 · 493 stars
lxml
parse HTML/XML at scale with lxml
Verified 2026-08-10 · 401 stars
Splash
render JavaScript pages via a lightweight HTTP API for scrapy
Verified 2026-08-10 · 297 stars
Parsel
CSS and XPath extraction on HTML responses in scrapy-style code
Verified 2026-08-10 · 251 stars
mcp-playwright
MCP server so Claude can drive Playwright
Verified 2026-08-10 · 232 stars
zendriver
undetected chrome automation fork
Verified 2026-08-10 · 166 stars
html5lib
parse broken real-world HTML with a standards HTML5 parser
Verified 2026-08-10 · 117 stars
agent-browser
An AI agent needs to drive a real Chrome browser from shell commands, reading pages via accessibility-tree snapshots and clicking elements by stable refs instead of brittle CSS selectors
Verified 2026-08-10
alpine-chrome
use alpine-chrome for browser automation
Verified 2026-08-10
Beautiful Soup
parse static HTML into a navigable tree for scrapers
Verified 2026-08-10
browser-tools-mcp
use browser-tools-mcp for browser automation
Verified 2026-08-10
Browserbase
hosted headless browsers for scraping or agents
Verified 2026-08-10
Browserless
hosted Chrome CDP for scraping and PDFs
Verified 2026-08-10
computer-use-ootb
desktop UI control beyond pure web scrapers
Verified 2026-08-10
Crawl4AI
LLM-friendly web crawling and extraction
Verified 2026-08-10
Crawlee
production crawlers with browser and HTTP queues
Verified 2026-08-10
Crawlee for Python
Your Python scraper keeps dying on transient errors and blocks and you need built-in retries, proxy rotation, and session management
Verified 2026-08-10
cua
computer use agent stack
Verified 2026-08-10
DrissionPage
Python browser automation combining requests and browser modes
Verified 2026-08-10
Firecrawl
API crawl that returns markdown from any URL
Verified 2026-08-10
Jina Reader
turn a public web page into clean Markdown with one HTTP request
Verified 2026-08-10
MechanicalSoup
automate form login and multi-page HTML flows without a full browser
Verified 2026-08-10
open-codex-computer-use
computer-use style desktop control for coding agents
Verified 2026-08-10
Oxylabs Browser Agent
Your scraping scripts keep breaking because CSS selectors rot after every site redesign and you would rather describe the steps in plain English
Verified 2026-08-10
patchright
playwright patched for bot detection
Verified 2026-08-10
Playwright for Go
A Go service must scrape or exercise a JavaScript-rendered site and plain HTTP fetches return empty or skeleton markup
Verified 2026-08-10
selenium
classic WebDriver browser automation across browsers
Verified 2026-08-10
selenium-driverless
driverless selenium against bot detection
Verified 2026-08-10
WaterCrawl
self-host asynchronous website crawling with a queue and API
Verified 2026-08-10