Awesome Web Archiving is a code repository or project page with source code, releases, documentation, or setup notes.
Directory
Search results
Published directory entries matching your search.
Awesome Web Scraping is an open source network utility for diagnostics, connectivity checks, or internet testing.
Dev Scraping APIs is a GitHub repository or organization with source code, releases, documentation, or project resources.
brozzler is an open source project with code, documentation, and setup resources for web archiving and scraping.
datahoarder-website-to-markdown is an open source project with code, documentation, and setup resources for web archiving and scraping.
DownloadNet (dn) is an open source project with code, documentation, and setup resources for web archiving and scraping.
Heritrix3 is an open source project with code, documentation, and setup resources for web archiving and scraping.
Httrack is an open source project with code, documentation, and setup resources for web archiving and scraping.
MarkSnip is an open source project with code, documentation, and setup resources for web archiving and scraping.
Resurrect Pages Fork is an open source project with code, documentation, and setup resources for web archiving and scraping.
Scoop is an open source project with code, documentation, and setup resources for web archiving and scraping.
SuckIT is an open source project with code, documentation, and setup resources for web archiving and scraping.
Trawl is an open source project with code, documentation, and setup resources for web archiving and scraping.
Waymore is an open source project with code, documentation, and setup resources for web archiving and scraping.
Self-hosted bookmark manager for saving links, archiving pages, and organizing web resources.
Browser extension and tool for saving complete web pages into a single HTML file.
wayback machine spn scripts is an open source project for saving, replaying, or working with archived web pages.
Dorks Eye Google Hacking Dork Scraping and Searching Script. Dorks Eye is a script I made in python 3. With this tool, you can easily find Google Dorks. Dork Eye collects potentially vulnerable web pages and applications on the Internet or other awesome info that is picked up by Google's search bots. Author: Jolanda de Koff
grab-site is an open source web scraping or crawling project for collecting website data.
Scrapling is an open source web scraping or crawling project for collecting website data.
Spider Suite is an open source web scraping or crawling project for collecting website data.
Twitch VOD and Live Stream archiving platform. Includes a rendered and real-time chat for each archive.
Kotlin/Java library and cli tool for scraping posts and media from various sources with neither authorization nor full page rendering (Facebook, Instagram, Twitter, Youtube, Tiktok, Telegram, Twitch, Reddit, 9GAG, Pinterest, Flickr, Tumblr, Coub, Vimeo, IFunny, VK, Odnoklassniki, Pikabu)
An archiving tool with an IM-style interface that prioritizes privacy and accessibility, integrated with various archival services including Internet Archive, archive.today, Ghostarchive, IPFS, Telegraph, and file systems.