github Actively maintained

turicas/crau

Easy-to-use Web archiver

1 awesome list

Quick read

Stars
64
Forks
10
Open issues
11
Commits
76

Activity and growth

Latest capture 2026-07-31 03:05

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
0
Stored snapshots
6

Classification

Technology stack

Metadata

Language
Python
License
LGPL-3.0
Default branch
develop
Created
2019-10-26
First commit
2019-10-26
Last pushed
2026-04-13
GitHub updated
2025-12-05
Last synced
2026-07-31 03:05
Stack scanned
2026-07-31 03:05
Archived
No

AI development signals

0 paths

Agent instructions and tool configuration found in this repository.

No config files detected.

Growth history

Tracked growth

6 observed captures since 2026-05-23. Observed captures are shown by default.

Stars from first capture 0

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

karust/gogetcrawl

Extract web archive data using Wayback Machine and Common Crawl

183 stars
Go 1 awesome list

chfoo/warcat

Tool and library for handling Web ARChive (WARC) files.

165 stars
Python 1 awesome list

oduwsdl/archivenow

A Tool To Push Web Resources Into Web Archives

434 stars
Python 1 awesome list

ArchiveTeam/grab-site

The archivist's web crawler: WARC output, dashboard for all crawls, dynamic ignore patterns

1,602 stars
Python 1 awesome list

webrecorder/warcio

Streaming WARC/ARC library for fast web archive IO

471 stars
Python 1 awesome list