Awesome

GitHub projects from awesome lists

Search awesome repositories

Search names, descriptions, topics, tags, and stacks, then tune results by ecosystem, freshness, health, and cross-list signal.

Repos indexed
18,423
Awesome lists tracked
132
Current results
21

Find repositories

Start broad, then narrow by ecosystem, freshness, health, and growth.

Clear 1 refinement
Search mode
Tune results
More filters Topics, generated tags, stack, files, age, archive status, and growth.
Ecosystem
Files

Choose a suggestion or use commas to require multiple files.

Health

Uses known first-commit dates.

Momentum
Filters by observed commit-count growth over the repository's latest 7-day capture window. Repositories without a recent baseline are excluded.
Filters by observed GitHub star growth over the repository's latest 7-day capture window. Repositories without a recent baseline are excluded.
Reset filters
21 repos shown
Highlighted

Open highlighted repo slot

Put your repository first

Promote a GitHub repo at the top of Awesome repository list views for 7 days.

ArchiveBox/ArchiveBox

🗃 Open source self-hosted web archiving. Takes URLs/browser history/bookmarks/Pocket/Pinboard/etc., saves HTML, JS, PDFs, media, and more...

AI dev
Updated
2026-09-11
Lists
4 list mentions
First commit
2017-05-05
History
46 history points
License
MIT
Issues
159 open
Forks
1,572
Commits
6,802 commits
Star growth, last 7 days
0 0.0%
Commit velocity, last 7 days
0 0.0%
internetarchive/heritrix3

Heritrix is the Internet Archive's open-source, extensible, web-scale, archival-quality web crawler project.

Updated
2026-09-02
Lists
1 list mention
First commit
2009-05-11
History
7 history points
License
NOASSERTION
Issues
37 open
Forks
792
Commits
3,022 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
webrecorder/browsertrix-crawler

Run a high-fidelity browser-based web archiving crawler in a single Docker container

Updated
2026-08-27
Lists
1 list mention
First commit
2020-10-31
History
7 history points
License
AGPL-3.0
Issues
144 open
Forks
151
Commits
752 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
webrecorder/browsertrix

Browsertrix is the hosted, high-fidelity, browser-based crawling service from Webrecorder designed to make web archiving easier and more accessible for all!

Updated
2026-09-04
Lists
1 list mention
First commit
2021-06-28
History
7 history points
License
AGPL-3.0
Issues
322 open
Forks
77
Commits
2,172 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
helgeho/ArchiveSpark

An Apache Spark framework for easy data processing, extraction as well as derivation for web archives and archival collections, developed at Internet Archive.

Updated
2025-10-08
Lists
1 list mention
First commit
2015-08-06
History
7 history points
License
MIT
Issues
5 open
Forks
19
Commits
154 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
archivesunleashed/warclight

A Rails engine supporting the discovery of web archives.

Archived
Updated
2023-06-13
Lists
1 list mention
First commit
2017-08-03
History
7 history points
License
NOASSERTION
Issues
8 open
Forks
9
Commits
301 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history