Awesome Web Scraping is an open source network utility for diagnostics, connectivity checks, or internet testing.
Directory
Search results
Published directory entries matching your search.
Web Scraping FYI helps crawl, scrape, copy, or extract data and pages from websites.
Dev Scraping APIs is a GitHub repository or organization with source code, releases, documentation, or project resources.
Data Hoarding is a web archiving or scraping tool for saving pages, data, or online content.
Wget2 is an open source project with code, documentation, and setup resources for web archiving and scraping.
Python tool for scraping public posts and profiles from social networks without using official platform APIs.
Airtable-based OSINT resource list for simple scraping workflows and Twitter-focused research links.
web.scraper.workers.dev helps crawl, scrape, copy, or extract data and pages from websites.
Import.io helps businesses scale web scraping and data extraction for real-time pricing, competitor tracking, and high-quality web data.
Python tool for collecting email addresses from search engines and public web sources during authorized research.
80legs is a web archiving or scraping tool for saving pages, data, or online content.
One API key, 63 AI agent tools: web search, scraping, AI generation, OCR, crypto data. LangChain + MCP ready. Pay per call with credits or x402/USDC.
brozzler is an open source project with code, documentation, and setup resources for web archiving and scraping.
CachedView is a web archiving or scraping tool for saving pages, data, or online content.
Commands is a web archiving or scraping tool for saving pages, data, or online content.
CopySite helps crawl, scrape, copy, or extract data and pages from websites.
Cyotek WebCopy helps crawl, scrape, copy, or extract data and pages from websites.
datahoarder-website-to-markdown is an open source project with code, documentation, and setup resources for web archiving and scraping.
Dorks Eye Google Hacking Dork Scraping and Searching Script. Dorks Eye is a script I made in python 3. With this tool, you can easily find Google Dorks. Dork Eye collects potentially vulnerable web pages and applications on the Internet or other awesome info that is picked up by Google's search bots. Author: Jolanda de Koff
DownloadNet (dn) is an open source project with code, documentation, and setup resources for web archiving and scraping.
grab-site is an open source web scraping or crawling project for collecting website data.
Heritrix is a web archiving or scraping tool for saving pages, data, or online content.
Heritrix3 is an open source project with code, documentation, and setup resources for web archiving and scraping.
Httrack is an open source project with code, documentation, and setup resources for web archiving and scraping.