DataBackup is a code project with source, releases, documentation, or setup notes.
Directory
Search results
Published directory entries matching your search.
datahoarder-website-to-markdown is an open source project with code, documentation, and setup resources for web archiving and scraping.
A pure-python graphics and GUI library built on PyQt4 / PySide and NumPy.
Open-source Python toolkit that tracks whether ChatGPT, Claude, and Perplexity cite your site. MIT license, runs locally on your credentials.
Curated GitHub list of public datasets across science, government, finance, and research topics.
A dashboard library for interactive visualizations using flask socketio and react.
ML powered analytics engine for outlier/anomaly detection and root cause analysis.
Analytical Web Apps for Python, R, Julia, and Jupyter.
dataforseo-claude helps research keywords, search demand, related terms, and content opportunities for SEO.
DataMonitor is a code project with source, releases, documentation, or setup notes.
A listener that streams your spark events logs to delight, a free and improved spark UI.
Storage layer that brings scalable, ACID transactions to Apache Spark and other engines.
A Julia package for probability distributions and associated functions.
SQL database that you can fork, clone, branch, merge, push and pull just like a git repository.
Python library for decision tree visualization and model interpretation.
A large Dataset of synchronised Audio, LyrIcs and vocal notes.
References for Go were mostly cut-and-pasted from.
A dataset format for creating, storing, and collaborating on AI datasets of any size.
Distributed POSIX file system built on top of Redis and S3.
Unified metadata exploration API service for Hive, RDS, Teradata, Redshift, S3 and Cassandra.
Open source ML model versioning, metadata, and experiment management.
A high-quality tool for convert PDF to Markdown and JSON.
Fast DataFrame library for Rust and Python, designed as a faster alternative to Pandas.
Massively parallel self-organizing maps: accelerate training on multicore CPUs, GPUs, and clusters, has python API.