github Actively maintained

webrecorder/browsertrix-crawler

Run a high-fidelity browser-based web archiving crawler in a single Docker container

1 awesome list

Quick read

Stars
1,127
Forks
151
Open issues
144
Commits
752

Activity and growth

Latest capture 2026-09-03 03:04

Stars · last 7 days
0 0.0%
Commits · last 7 days
0 0.0%
Stars since tracking
+85
Stored snapshots
7

Classification

Metadata

Language
TypeScript
License
AGPL-3.0
Default branch
main
Created
2020-11-02
First commit
2020-10-31
Last pushed
2026-08-27
GitHub updated
2026-09-02
Last synced
2026-09-03 03:04
Stack scanned
2026-09-03 03:04
Archived
No

AI development signals

0 paths

Agent instructions and tool configuration found in this repository.

No config files detected.

Growth history

Tracked growth

7 observed captures since 2026-05-23. Observed captures are shown by default.

Stars from first capture +85

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

webrecorder/browsertrix

Browsertrix is the hosted, high-fidelity, browser-based crawling service from Webrecorder designed to make web archiving easier and more accessible for all!

468 stars
TypeScript 1 awesome list

openzim/zimit

Make a ZIM file from any Web site and surf offline!

827 stars
Python 1 awesome list

internetarchive/heritrix3

Heritrix is the Internet Archive's open-source, extensible, web-scale, archival-quality web crawler project.

3,313 stars
Java 1 awesome list

trickstercache/trickster

Open Source HTTP Reverse Proxy Cache and Time Series Dashboard Accelerator

2,085 stars
Go 1 awesome list

retracedhq/retraced

🔥 A fully open source audit logs service and embeddable UI easily deployed to your own Kubernetes cluster. Brought to you by replicated.com and boxyhq.com 🚀

449 stars
TypeScript 1 awesome list