GitHub repo leaderboard by stars, growth rate and activity.
🗃 Open source self-hosted web archiving. Takes URLs/browser history/bookmarks/Pocket/Pinboard/etc., saves HTML, JS, PDFs, media, and more...
Browser extension for viewing archived and cached versions of web pages, available for Chrome, Edge and Safari
A collection of special paths linked to common sensitive APIs, devops internals, frameworks conf, known misconfigurations, juicy APIs ..etc. It could be used as a part of web content discovery, to scan passively for high-quality endpoints and quick-wins.
复刻网站的 Agent Skill:抓只读镜像、从压缩代码逐行还原、自动比对验收。An agent skill that mirrors a website, rebuilds it from the minified code, and verifies the result with automated diffs.
Wayback Machine API interface & a command-line tool
A command-line utility and Scrapy middleware for scraping time series data from Archive.org's Wayback Machine.
A lightweight tool for scraping current and historic Google Analytics data
A Scrapy middleware for scraping time series data from Archive.org's Wayback Machine.
| # | Repo | Language | Stars | 30-day trend | Last updated |
|---|---|---|---|---|---|
| 1 | 🗃 Open source self-hosted web archiving. Takes URLs/browser history/bookmarks/Pocket/Pinboard/etc., saves HTML, JS, PDFs, media, and more... | Python | 28,267 | last pushed 4 days ago | |
| 2 | Browser extension for viewing archived and cached versions of web pages, available for Chrome, Edge and Safari | JavaScript | 1,593 | last pushed 3 months ago | |
| 3 | A collection of special paths linked to common sensitive APIs, devops internals, frameworks conf, known misconfigurations, juicy APIs ..etc. It could be used as a part of web content discovery, to scan passively for high-quality endpoints and quick-wins. | — | 1,193 | last pushed 5 months ago | |
| 4 | 复刻网站的 Agent Skill:抓只读镜像、从压缩代码逐行还原、自动比对验收。An agent skill that mirrors a website, rebuilds it from the minified code, and verifies the result with automated diffs. | JavaScript | 1,028 | last pushed 4 days ago | |
| 5 | Wayback Machine API interface & a command-line tool | Python | 604 | last pushed 3 years ago | |
| 6 | A command-line utility and Scrapy middleware for scraping time series data from Archive.org's Wayback Machine. | Python | 483 | last pushed 3 years ago | |
| 7 | A lightweight tool for scraping current and historic Google Analytics data | Python | 239 | last pushed 2 years ago | |
| 8 | A Scrapy middleware for scraping time series data from Archive.org's Wayback Machine. | Python | 124 | last pushed 3 years ago |
All · 11,658