github Actively maintained

apify/crawlee

Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.

1 awesome list

Quick read

Stars
25,646
Forks
1,658
Open issues
146
Commits
5,689

Activity and growth

Latest capture 2026-09-04 10:54

Stars · last 7 days
0 0.0%
Commits · last 7 days
0 0.0%
Stars since tracking
0
Stored snapshots
1

Classification

Metadata

Language
TypeScript
License
Apache-2.0
Default branch
master
Created
2016-08-26
First commit
2016-08-26
Last pushed
2026-09-04
GitHub updated
2026-09-04
Last synced
2026-09-04 10:54
Stack scanned
2026-09-04 10:54
Archived
No

AI development signals

2 paths

Agent instructions and tool configuration found in this repository.

Claude Code 2

View paths

Growth history

Tracked growth

1 observed capture since 2026-09-04. Observed captures are shown by default.

Stars from first capture 0

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

apify/crawlee-python

Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Parsel, BeautifulSoup, Playwright, and raw HTTP. Both headful and headless mode. With proxy rotation.

9,393 stars
Python 2 awesome lists

any4ai/AnyCrawl

AnyCrawl 🚀: A Node.js/TypeScript crawler that turns websites into LLM-ready data and extracts structured SERP results from Google/Bing/Baidu/etc. Native multi-threading for bulk processing.

3,447 stars
TypeScript 1 awesome list

s0rg/crawley

The unix-way web crawler

341 stars
Go 2 awesome lists

amantus-ai/llm-codes

Transform developer documentation to clean Markdown

345 stars
TypeScript 0 awesome lists

firecrawl/firecrawl

The context API to search, scrape, and interact with the web at scale. 🔥

176,202 stars
TypeScript 2 awesome lists

unclecode/crawl4ai

🚀🤖 Crawl4AI: Open-source LLM Friendly Web Crawler & Scraper. Don't be shy, join here: https://discord.gg/jP8KfhDhyN

81,229 stars
Python 3 awesome lists