A library to compare Pandas, Polars, and Spark data frames. It provides stats and lets users adjust for match accuracy.
Directory
Search results
Published directory entries matching your search.
Reproducible data setup for reproducible science.
Library for working with tabular data in Julia.
A lightweight framework for data analysis in JavaScript.
Collect, clean and visualize your data in Python.
Open-source package for validating ML models & data, with various checks and suites.
Drop-in replacement for Jupyter and an AI-native workspace for modern data teams.
Library of SAS Enterprise Miner process flow diagrams to help you learn by example about specific data mining topics.
Tools for exploratory data analysis in Python.
Functions and data dependencies for loading various word embeddings.
Clojure Data Visualisation library, based on Statistiker and D3.
Datawrapper An open source data visualization platform helping everyone to create simple, correct and embeddable charts. Also at.
Analyzes structured data and helps produce queries, insights or visualizations.
Analyzes structured data and helps produce queries, insights or visualizations.
Knowledge sharing platform for data scientists and other technical professions.
Repeatable, atomic and versioned data lake on top of object storage.
Modern columnar data format for ML implemented in Rust.
Pretrain computer vision models on unlabeled data for industrial applications.
Collect, aggregate, and visualize a data ecosystem's metadata.
A tensor-based framework for large-scale data computation which is often regarded as a parallel and distributed version of NumPy.
Analyzes structured data and helps produce queries, insights or visualizations.
Algorithm capable of fully capturing the impact of data drift on performance.
Distributed, masterless, high performance, fault tolerant data processing. Written entirely in Clojure.
Bring multiple data streams into one dashboard.