Spark is a GitHub repository or organization with source code, releases, documentation, or project resources.
Directory
Search results
Published directory entries matching your search.
SparkleShare helps upload, sync, or manage files across cloud storage, remote drives, or shared storage services.
A listener that streams your spark events logs to delight, a free and improved spark UI.
A distributed machine learning framework Apache Spark.
Spark is a fast and general engine for large-scale data processing.
A genomics processing engine and specialized file format built using Apache Avro, Apache Spark and Parquet. Apache 2 licensed.
A library to compare Pandas, Polars, and Spark data frames. It provides stats and lets users adjust for match accuracy.
Storage layer that brings scalable, ACID transactions to Apache Spark and other engines.
A Clojure dataframe library that runs on Apache Spark.
ML engine that supports distributed learning on Hadoop, Spark or your laptop via APIs in R, Python, Scala, REST/JSON.
Open-source project that provides AI-assisted learning, courses, tutorials, or study support.
A service for deployment Apache Spark MLLib machine learning models as realtime, batch or reactive web services.
Pandas API on Apache Spark. Makes data scientists more productive when interacting with big data.
Lambda Architecture Framework using Apache Spark and Apache Kafka with a specialization for real-time large-scale machine learning.
Analyzes structured data and helps produce queries, insights or visualizations.