internetarchive/Sparkling
Internet Archive's Sparkling Data Processing Library
A pure Python implementation of Apache Spark's RDD and DStream interfaces.
Appears on
Quick read
Latest capture 2026-08-31 03:02
0 paths
Agent instructions and tool configuration found in this repository.
No config files detected.
8 observed captures since 2026-05-22. Observed captures are shown by default.
Stars from first capture 0
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
Internet Archive's Sparkling Data Processing Library
PySpark + Scikit-learn = Sparkit-learn
Apache Spark - A unified analytics engine for large-scale data processing
A Data Streaming Library for Efficient Neural Network Training
Project SnappyData - memory optimized analytics database, based on Apache Spark™ and Apache Geode™. Stream, Transact, Analyze, Predict in one cluster
Simple and Distributed Machine Learning