Algorithm capable of fully capturing the impact of data drift on performance.
Directory
Search results
Published directory entries matching your search.
Open source ML model versioning, metadata, and experiment management.
Analyzes structured data and helps produce queries, insights or visualizations.
Unified metadata exploration API service for Hive, RDS, Teradata, Redshift, S3 and Cassandra.
MeTA: ModErn Text Analysis is a C++ Data Sciences Toolkit that facilitates mining big text data.
A tensor-based framework for large-scale data computation which is often regarded as a parallel and distributed version of NumPy.
Collect, aggregate, and visualize a data ecosystem's metadata.
A federated, open-source data catalog for all your big data and small data.
Fast and easy data exploration by automating the visualization and data analysis process.
Pretrain computer vision models on unlabeled data for industrial applications.
Modern columnar data format for ML implemented in Rust.
Repeatable, atomic and versioned data lake on top of object storage.
Pandas API on Apache Spark. Makes data scientists more productive when interacting with big data.
Knowledge sharing platform for data scientists and other technical professions.
Rendering beautiful SVG maps in Python.
Jupyter Notebooks as Markdown Documents, Julia, Python or R scripts.
Distributed POSIX file system built on top of Redis and S3.
Analyzes structured data and helps produce queries, insights or visualizations.
A dataset format for creating, storing, and collaborating on AI datasets of any size.
Implementation of the hdbscan algorithm in Python - used for clustering.
Open-source project that applies AI to data analysis, extraction, visualization, or business intelligence.
Grid studio is a web-based spreadsheet application with full integration of the Python programming language.
Graph layout algorithms in pure Julia.
References for Go were mostly cut-and-pasted from.