Portable annotation tool for creating labeled datasets.
Directory
Search results
Published directory entries matching your search.
Fast DataFrame library for Rust and Python, designed as a faster alternative to Pandas.
CLI tool that allows you to build data profiles and write assertion tests for easily evaluating and tracking your data's reliability over time.
A Clojure/Clojurescript notebook application/-library based on Gorilla-REPL.
Simple, realtime visualization of neural network training performance.
Clojure API wrapping Python's Pandas library.
Create HTML profiling reports from pandas DataFrame objects.
Open-source project that applies AI to data analysis, extraction, visualization, or business intelligence.
Cleansing, pre-processing, feature engineering, exploratory data analysis and easy ML with PySpark backend.
Bring multiple data streams into one dashboard.
Distributed, masterless, high performance, fault tolerant data processing. Written entirely in Clojure.
Unified metadata exploration API service for Hive, RDS, Teradata, Redshift, S3 and Cassandra.
A tensor-based framework for large-scale data computation which is often regarded as a parallel and distributed version of NumPy.
Fast and easy data exploration by automating the visualization and data analysis process.
Pandas API on Apache Spark. Makes data scientists more productive when interacting with big data.
Jupyter Notebooks as Markdown Documents, Julia, Python or R scripts.
A dataset format for creating, storing, and collaborating on AI datasets of any size.
Grid studio is a web-based spreadsheet application with full integration of the Python programming language.
References for Go were mostly cut-and-pasted from.
A large Dataset of synchronised Audio, LyrIcs and vocal notes.
Clojure Data Visualisation library, based on Statistiker and D3.
Framework to create ChatGPT like bots over your dataset.
In-process SQL OLAP database management system.
SQL database that you can fork, clone, branch, merge, push and pull just like a git repository.