Grid studio is a web-based spreadsheet application with full integration of the Python programming language.
Directory
Search results
Published directory entries matching your search.
Implementation of the hdbscan algorithm in Python - used for clustering.
A dataset format for creating, storing, and collaborating on AI datasets of any size.
Analyzes structured data and helps produce queries, insights or visualizations.
Analyzes structured data and helps produce queries, insights or visualizations.
A lightweight set of tools for loading and sharing data in data science projects.
The power of Chart.js in Jupyter Notebook.
Machine learning toolkit with classification and clustering for Node.js; supports visualization (see visualml.io).
Distributed POSIX file system built on top of Redis and S3.
Jupyter Notebooks as Markdown Documents, Julia, Python or R scripts.
Rendering beautiful SVG maps in Python.
Knowledge sharing platform for data scientists and other technical professions.
Pandas API on Apache Spark. Makes data scientists more productive when interacting with big data.
Repeatable, atomic and versioned data lake on top of object storage.
Modern columnar data format for ML implemented in Rust.
Pretrain computer vision models on unlabeled data for industrial applications.
A federated, open-source data catalog for all your big data and small data.
Collect, aggregate, and visualize a data ecosystem's metadata.
A tensor-based framework for large-scale data computation which is often regarded as a parallel and distributed version of NumPy.
MeTA: ModErn Text Analysis is a C++ Data Sciences Toolkit that facilitates mining big text data.
Unified metadata exploration API service for Hive, RDS, Teradata, Redshift, S3 and Cassandra.
Analyzes structured data and helps produce queries, insights or visualizations.
Open source ML model versioning, metadata, and experiment management.
Algorithm capable of fully capturing the impact of data drift on performance.