github Actively maintained

archivesunleashed/twut

An open-source toolkit for analyzing line-oriented JSON Twitter archives with Apache Spark.

1 awesome list

Quick read

Stars
10
Forks
2
Open issues
1
Commits
41

Activity and growth

Latest capture 2026-09-03 03:03

Stars · last 7 days
0 0.0%
Commits · last 7 days
0 0.0%
Stars since tracking
0
Stored snapshots
7

Classification

Metadata

Language
Scala
License
Apache-2.0
Default branch
main
Created
2019-11-29
First commit
2019-11-29
Last pushed
2026-03-17
GitHub updated
2025-10-16
Last synced
2026-09-03 03:03
Stack scanned
2026-09-03 03:03
Archived
No

AI development signals

0 paths

Agent instructions and tool configuration found in this repository.

No config files detected.

Growth history

Tracked growth

7 observed captures since 2026-05-23. Observed captures are shown by default.

Stars from first capture 0

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

archivesunleashed/aut

The Archives Unleashed Toolkit is an open-source toolkit for analyzing web archives.

158 stars
Scala 1 awesome list

helgeho/ArchiveSpark

An Apache Spark framework for easy data processing, extraction as well as derivation for web archives and archival collections, developed at Internet Archive.

162 stars
Scala 1 awesome list

DocNow/twarc

A command line tool (and Python library) for archiving Twitter JSON

1,394 stars
Python 1 awesome list

twitter/twemproxy

A fast, light-weight proxy for memcached and redis

12,337 stars
C 1 awesome list