GitHub repo leaderboard by stars, growth rate and activity.
JavaScript API for Chrome and Firefox
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.
A high-level browser automation library.
🖥 Chrome automation made simple. Runs locally or headless on AWS Lambda.
High-performance browser automation bridge and multi-instance orchestrator with advanced stealth injection and real-time dashboard.
Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Parsel, BeautifulSoup, Playwright, and raw HTTP. Both headful and headless mode. With proxy rotation.
💯 Teach puppeteer new tricks through plugins.
Web page PDF/PNG rendering done right. Self-hosted service for rendering receipts, invoices, or any content.
A Headless Chrome rendering solution
Distributed crawler powered by Headless Chrome
| # | Repo | Language | Stars | 30-day trend | Last updated |
|---|---|---|---|---|---|
| 1 | JavaScript API for Chrome and Firefox | TypeScript | 95,564 | last pushed Yesterday | |
| 2 | Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation. | TypeScript | 25,713 | last pushed Yesterday | |
| 3 | A high-level browser automation library. | JavaScript | 19,768 | last pushed 2 years ago | |
| 4 | 🖥 Chrome automation made simple. Runs locally or headless on AWS Lambda. | TypeScript | 13,217 | last pushed 8 years ago | |
| 5 | High-performance browser automation bridge and multi-instance orchestrator with advanced stealth injection and real-time dashboard. | Go | 10,243 | last pushed 2 days ago | |
| 6 | Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Parsel, BeautifulSoup, Playwright, and raw HTTP. Both headful and headless mode. With proxy rotation. | Python | 9,500 | last pushed 2 days ago | |
| 7 | 💯 Teach puppeteer new tricks through plugins. | JavaScript | 7,397 | last pushed 2 years ago | |
| 8 | Web page PDF/PNG rendering done right. Self-hosted service for rendering receipts, invoices, or any content. | HTML | 7,101 | last pushed 3 years ago | |
| 9 | A Headless Chrome rendering solution | TypeScript | 5,952 | last pushed 4 years ago | |
| 10 | Distributed crawler powered by Headless Chrome | JavaScript | 5,635 | last pushed 3 years ago |
All · 11,474