A library to compare Pandas, Polars, and Spark data frames. It provides stats and lets users adjust for match accuracy.
Directory
Search results
Published directory entries matching your search.
Reproducible data setup for reproducible science.
Library for working with tabular data in Julia.
A lightweight framework for data analysis in JavaScript.
Collect, clean and visualize your data in Python.
A GitHub Repository Where you can Learn Datavisualizatoin Basics to Intermediate level.
Open-source package for validating ML models & data, with various checks and suites.
Drop-in replacement for Jupyter and an AI-native workspace for modern data teams.
A listener that streams your spark events logs to delight, a free and improved spark UI.
Storage layer that brings scalable, ACID transactions to Apache Spark and other engines.
A Julia package for probability distributions and associated functions.
Library of SAS Enterprise Miner process flow diagrams to help you learn by example about specific data mining topics.
SQL database that you can fork, clone, branch, merge, push and pull just like a git repository.
Tools for exploratory data analysis in Python.
In-process SQL OLAP database management system.
Data Version Control - Git for Data & Models - ML Experiments Management.
Framework to create ChatGPT like bots over your dataset.
Functions and data dependencies for loading various word embeddings.
The Python ensemble sampling toolkit for affine-invariant MCMC.
Clojure Data Visualisation library, based on Statistiker and D3.
Visualizations for understanding and analyzing machine learning datasets.
A large Dataset of synchronised Audio, LyrIcs and vocal notes.
Crafty statistical graphics for Julia.
A Clojure dataframe library that runs on Apache Spark.