CLI tool that allows you to build data profiles and write assertion tests for easily evaluating and tracking your data's reliability over time.
Directory
Search results
Published directory entries matching your search.
Fast DataFrame library for Rust and Python, designed as a faster alternative to Pandas.
Portable annotation tool for creating labeled datasets.
pprof is a tool for visualization and analysis of profiling data
AI-powered tool for automated PR analysis, feedback, suggestions, and more.
AI-powered tool for automated PR analysis, feedback, suggestions and more.
Simple plotting for Python. Wrapper for D3xterjs; easily render charts in-browser.
Markov Chain Monte Carlo sampling toolkit.
A pure-python graphics and GUI library built on PyQt4 / PySide and NumPy.
A python framework to transform natural language questions to queries in a database query language.
A self-organizing data hub with S3 support.
Julia package for loading many of the data sets available in R.
Python project for creating a persona-style summary from public Reddit account activity.
A javascript library containing a collection of least squares fitting methods for finding a trend in a set of data.
Basic sampling algorithms for Julia.
Analyzes structured data and helps produce queries, insights or visualizations.
A beautiful graphing toolkit for Ruby.
High performance distributed data processing in NodeJS.
A system for quickly generating training data with weak supervision.
AI-generated visualization prototyping and editing platform, support 2D, 3D models, combined with LLM(Large Language Model) for quick editing.
Massively parallel self-organizing maps: accelerate training on multicore CPUs, GPUs, and clusters, has python API.
Spark is a fast and general engine for large-scale data processing.
Analyzes structured data and helps produce queries, insights or visualizations.
Statistical modelling and econometrics in Python.