GitHub repo leaderboard by stars, growth rate and activity.
Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Parsel, BeautifulSoup, Playwright, and raw HTTP. Both headful and headless mode. With proxy rotation.
dude uncomplicated data extraction: A simple framework for writing web scrapers using Python decorators
| # | Repo | Language | Stars | 30-day trend | Last updated |
|---|---|---|---|---|---|
| 1 | Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Parsel, BeautifulSoup, Playwright, and raw HTTP. Both headful and headless mode. With proxy rotation. | Python | 9,508 | last pushed 22 hours ago | |
| 2 | dude uncomplicated data extraction: A simple framework for writing web scrapers using Python decorators | Python | 426 | last pushed 1 year ago |
All · 11,658