Bảng xếp hạng repo GitHub theo sao, tốc độ tăng trưởng và mức độ hoạt động.
The context API to search, scrape, and interact with the web at scale. 🔥
🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!
Scrapy, a fast high-level web crawling & scraping framework for Python.
A visual no-code/code-free web crawler/spider易采集:一个可视化浏览器自动化测试/数据采集/网页爬虫软件,可以无代码图形化的设计和执行爬虫任务。别名:ServiceWrapper面向Web应用的智能化服务封装系统。
👾 Fast and simple video download library and CLI tool written in Go
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.
🚀「Douyin_TikTok_Download_API」是一个开箱即用的高性能异步抖音、快手、TikTok、Bilibili数据爬取工具,支持API调用,在线批量解析及下载。
| # | Repo | Ngôn ngữ | Sao | Xu hướng 30 ngày | Cập nhật lần cuối |
|---|---|---|---|---|---|
| 1 | The context API to search, scrape, and interact with the web at scale. 🔥 | TypeScript | 178.457 | push cuối Hôm qua | |
| 2 | 🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl! | Python | 79.739 | push cuối 7 ngày trước | |
| 3 | Scrapy, a fast high-level web crawling & scraping framework for Python. | Python | 64.268 | push cuối 2 ngày trước | |
| 4 | A visual no-code/code-free web crawler/spider易采集:一个可视化浏览器自动化测试/数据采集/网页爬虫软件,可以无代码图形化的设计和执行爬虫任务。别名:ServiceWrapper面向Web应用的智能化服务封装系统。 | JavaScript | 44.521 | push cuối Hôm qua | |
| 5 | 👾 Fast and simple video download library and CLI tool written in Go | Go | 31.670 | push cuối 6 tháng trước | |
| 6 | Python scraper based on AI | Python | 30.783 | push cuối 3 ngày trước | |
| 7 | Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation. | TypeScript | 25.713 | push cuối Hôm qua | |
| 8 | Elegant Scraper and Crawler Framework for Golang | Go | 25.504 | push cuối 1 tuần trước | |
| 9 | Python ProxyPool for web spider | Python | 23.686 | push cuối 3 tháng trước | |
| 10 | 🚀「Douyin_TikTok_Download_API」是一个开箱即用的高性能异步抖音、快手、TikTok、Bilibili数据爬取工具,支持API调用,在线批量解析及下载。 | Python | 20.035 | push cuối 11 tháng trước |
Tất cả · 11.474