Directory

Search results

Published directory entries matching your search.

Search results

6,406 listings
github.com

Scoop is an open source project with code, documentation, and setup resources for web archiving and scraping.

Web Archives 330
github.com

Scrapling is an open source web scraping or crawling project for collecting website data.

Internet Utilities 211
github.com

Spider Suite is an open source web scraping or crawling project for collecting website data.

Internet Utilities 163
spidersuite.io

SpiderSuite helps crawl, scrape, copy, or extract data and pages from websites.

Internet Utilities 294
github.com

SuckIT is an open source project with code, documentation, and setup resources for web archiving and scraping.

Web Archives 331
tools.digitalmethods.net

Digital Methods tool for scraping search engine results and collecting web research data.

Security Tools & Toolkits 266
github.com

Trawl is an open source project with code, documentation, and setup resources for web archiving and scraping.

Internet Utilities 316
vegetableman.github.io

Vandal is a web archiving or scraping tool for saving pages, data, or online content.

Web Archives 242
machawk1.github.io

WAIL is a web archiving or scraping tool for saving pages, data, or online content.

Web Archives 328
github.com

Waymore is an open source project with code, documentation, and setup resources for web archiving and scraping.

Internet Utilities 180
webscrapeai.com

Webscrape AI is the perfect tool for collecting data from the web without the hassle of manual scraping. No coding skills required.

AI Coding & Development 210
practicalbetterments.com

Wiki DL Guide is a web archiving or scraping tool for saving pages, data, or online content.

Web Archives 327
pypi.org

Python OSINT tool for collecting public tweets, followers, likes, and profile data from Twitter.

X / Twitter 328
alltheplaces.xyz

Open source data scraping software project.

Dataset & Data Search 348
crummy.com

Beautiful Soup: a library designed for screen-scraping HTML and XML.

AI Data & Analytics 293
github.com

Proxy service that helps scraping tools handle Cloudflare and browser-based anti-bot challenges.

Internet Utilities 343
github.com

A Playwright-based Node.js tool that bypasses search engine anti-scraping mechanisms to execute Google searches. Local alternative to SERP APIs with MCP server integration.

Google Communities & Tools 252
github.com

Project Eyes On is a high-speed, multi-threaded surveillance tool by Y0oshi (@rde0) for locating open IP cameras worldwide. Unifies Google Dorking and Directory Scraping into a single OSINT engine.

Google Communities & Tools 268
github.com

Kotlin/Java library and cli tool for scraping posts and media from various sources with neither authorization nor full page rendering (Facebook, Instagram, Twitter, Youtube, Tiktok, Telegram, Twitch, Reddit, 9GAG, Pinterest, Flickr, Tumblr, Coub, Vimeo, IFunny, VK, Odnoklassniki, Pikabu)

Twitch 205
github.com

全自动短视频搬运工具,支持自动下载、去重、AI生成标题+标签、上传,可二开扩展至多平台,例如:TikTok->视频号/抖音/小红书、抖音->TikTok/视频号/小红书......video-processing, automation, tiktok, selenium, pyqt5, ffmpeg, bot, data-scraping, video-deduplication.

TikTok 330
aprilfoolsdayontheweb.com

April Fools Day On The Web is a miscellaneous web resource for discovery, directories, archives, web culture, or useful tools.

Fun & Interesting 277
theuselessweb.com

The Useless Web is a miscellaneous web resource for discovery, directories, archives, web culture, or useful tools.

Fun & Interesting 162
webneko.net

Web Neko is a miscellaneous web resource for discovery, directories, archives, web culture, or useful tools.

Fun & Interesting 279
web-rewind.com

Web Rewind is a miscellaneous web resource for discovery, directories, archives, web culture, or useful tools.

Fun & Interesting 337