Jupyter Notebooks as Markdown Documents, Julia, Python or R scripts.
Directory
Search results
Published directory entries matching your search.
Rendering beautiful SVG maps in Python.
Knowledge sharing platform for data scientists and other technical professions.
Pandas API on Apache Spark. Makes data scientists more productive when interacting with big data.
Repeatable, atomic and versioned data lake on top of object storage.
Modern columnar data format for ML implemented in Rust.
Pretrain computer vision models on unlabeled data for industrial applications.
A federated, open-source data catalog for all your big data and small data.
Collect, aggregate, and visualize a data ecosystem's metadata.
A tensor-based framework for large-scale data computation which is often regarded as a parallel and distributed version of NumPy.
MeTA: ModErn Text Analysis is a C++ Data Sciences Toolkit that facilitates mining big text data.
Unified metadata exploration API service for Hive, RDS, Teradata, Redshift, S3 and Cassandra.
Analyzes structured data and helps produce queries, insights or visualizations.
Open source ML model versioning, metadata, and experiment management.
Algorithm capable of fully capturing the impact of data drift on performance.
Notebook experience in your Clojure namespace.
Distributed, masterless, high performance, fault tolerant data processing. Written entirely in Clojure.
A high-quality tool for convert PDF to Markdown and JSON.
Bring multiple data streams into one dashboard.
Cleansing, pre-processing, feature engineering, exploratory data analysis and easy ML with PySpark backend.
Pachyderm is a version control system for data.
Create HTML profiling reports from pandas DataFrame objects.
Clojure API wrapping Python's Pandas library.
Analyzes structured data and helps produce queries, insights or visualizations.