Collect, aggregate, and visualize a data ecosystem's metadata.
Directory
Search results
Published directory entries matching your search.
A tensor-based framework for large-scale data computation which is often regarded as a parallel and distributed version of NumPy.
MeTA: ModErn Text Analysis is a C++ Data Sciences Toolkit that facilitates mining big text data.
Unified metadata exploration API service for Hive, RDS, Teradata, Redshift, S3 and Cassandra.
Analyzes structured data and helps produce queries, insights or visualizations.
Open source ML model versioning, metadata, and experiment management.
Algorithm capable of fully capturing the impact of data drift on performance.
Notebook experience in your Clojure namespace.
Distributed, masterless, high performance, fault tolerant data processing. Written entirely in Clojure.
A high-quality tool for convert PDF to Markdown and JSON.
Bring multiple data streams into one dashboard.
Cleansing, pre-processing, feature engineering, exploratory data analysis and easy ML with PySpark backend.
Pachyderm is a version control system for data.
Create HTML profiling reports from pandas DataFrame objects.
Clojure API wrapping Python's Pandas library.
Analyzes structured data and helps produce queries, insights or visualizations.
A Clojure/Clojurescript notebook application/-library based on Gorilla-REPL.
CLI tool that allows you to build data profiles and write assertion tests for easily evaluating and tracking your data's reliability over time.
Fast DataFrame library for Rust and Python, designed as a faster alternative to Pandas.
Portable annotation tool for creating labeled datasets.
pprof is a tool for visualization and analysis of profiling data
Materials and IPython notebooks for "Python for Data Analysis" by Wes McKinney, published by O'Reilly Media.
Simple plotting for Python. Wrapper for D3xterjs; easily render charts in-browser.
Markov Chain Monte Carlo sampling toolkit.