apache/hudi
Upserts, Deletes And Incremental Processing on Big Data.
Apache Spark - A unified analytics engine for large-scale data processing
Quick read
Latest capture 2026-08-02 03:10
11 observed captures since 2026-05-22. Observed captures are shown by default.
Stars from first capture +443
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
Upserts, Deletes And Incremental Processing on Big Data.
Simple and Distributed Machine Learning
PySpark + Scikit-learn = Sparkit-learn
Apache Maven core
Spark plugin to retrieve metrics from a variety of cluster resources
Project SnappyData - memory optimized analytics database, based on Apache Spark™ and Apache Geode™. Stream, Transact, Analyze, Predict in one cluster