Unified metadata exploration API service for Hive, RDS, Teradata, Redshift, S3 and Cassandra.
Directory
Search results
Published directory entries matching your search.
Analyzes structured data and helps produce queries, insights or visualizations.
Open source ML model versioning, metadata, and experiment management.
Algorithm capable of fully capturing the impact of data drift on performance.
Notebook experience in your Clojure namespace.
Distributed, masterless, high performance, fault tolerant data processing. Written entirely in Clojure.
A high-quality tool for convert PDF to Markdown and JSON.
Bring multiple data streams into one dashboard.
Cleansing, pre-processing, feature engineering, exploratory data analysis and easy ML with PySpark backend.
Pachyderm is a version control system for data.
Create HTML profiling reports from pandas DataFrame objects.
Clojure API wrapping Python's Pandas library.
Analyzes structured data and helps produce queries, insights or visualizations.
A Clojure/Clojurescript notebook application/-library based on Gorilla-REPL.
CLI tool that allows you to build data profiles and write assertion tests for easily evaluating and tracking your data's reliability over time.
Fast DataFrame library for Rust and Python, designed as a faster alternative to Pandas.
Portable annotation tool for creating labeled datasets.
pprof is a tool for visualization and analysis of profiling data
Materials and IPython notebooks for "Python for Data Analysis" by Wes McKinney, published by O'Reilly Media.
Simple plotting for Python. Wrapper for D3xterjs; easily render charts in-browser.
Markov Chain Monte Carlo sampling toolkit.
A pure-python graphics and GUI library built on PyQt4 / PySide and NumPy.
A python framework to transform natural language questions to queries in a database query language.
A self-organizing data hub with S3 support.