Awesome Web Archiving is a code repository or project page with source code, releases, documentation, or setup notes.
Directory
Search results
Published directory entries matching your search.
Self-hosted web archiving tool for saving pages, media, PDFs, and snapshots of online content.
Browser-based web archiving tool for capturing pages and saving them as web archive files.
Data Hoarding is a web archiving or scraping tool for saving pages, data, or online content.
Web archiving project implemented by National Diet Library Japan.
Wget2 is an open source project with code, documentation, and setup resources for web archiving and scraping.
80legs is a web archiving or scraping tool for saving pages, data, or online content.
Web archiving platform for preserving, browsing, and searching curated digital collections.
Webpage archiving service for saving snapshots of pages and preserving online evidence.
Archivematica helps save, browse, replay, or preserve web pages and online resources.
ArchiveTeam helps save, browse, replay, or preserve web pages and online resources.
Browse research datasets and datasets related to economics.
Auto Load is an open source project for saving, replaying, or working with archived web pages.
brozzler is an open source project with code, documentation, and setup resources for web archiving and scraping.
CachedView is a web archiving or scraping tool for saving pages, data, or online content.
Commands is a web archiving or scraping tool for saving pages, data, or online content.
Service offering for the long-term archiving of cultural heritage.
Find data collections, study records and related research documentation.
datahoarder-website-to-markdown is an open source project with code, documentation, and setup resources for web archiving and scraping.
Solution for digital long-term archiving.
DownloadNet (dn) is an open source project with code, documentation, and setup resources for web archiving and scraping.
Twitch VOD and Live Stream archiving platform. Includes a rendered and real-time chat for each archive.
Heritrix is a web archiving or scraping tool for saving pages, data, or online content.
Heritrix3 is an open source project with code, documentation, and setup resources for web archiving and scraping.