github Actively maintained

D4Vinci/Scrapling

🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!

1 awesome list

Quick read

Stars
73,436
Forks
7,333
Open issues
0
Commits
1,583

Activity and growth

Latest capture 2026-08-11 04:36

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
+15021
Stored snapshots
60

Classification

Metadata

Language
Python
License
BSD-3-Clause
Default branch
main
Created
2024-10-13
First commit
2024-10-13
Last pushed
2026-08-10
GitHub updated
2026-08-11
Last synced
2026-08-11 04:36
Stack scanned
2026-08-11 04:36
Archived
No

AI development signals

0 paths

Agent instructions and tool configuration found in this repository.

No config files detected.

Growth history

Tracked growth

60 observed captures since 2026-06-02. Observed captures are shown by default.

Stars from first capture +15021

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

firecrawl/firecrawl

The context API to search, scrape, and interact with the web at scale. 🔥

176,202 stars
TypeScript 2 awesome lists

unclecode/crawl4ai

🚀🤖 Crawl4AI: Open-source LLM Friendly Web Crawler & Scraper. Don't be shy, join here: https://discord.gg/jP8KfhDhyN

81,229 stars
Python 3 awesome lists

Pyx-Corp/spectrawl

The unified web layer for AI agents. Search (8 engines), stealth browse, auth, and act on 24 platforms. One npm install, self-hosted.

28 stars
JavaScript 1 awesome list

bighippoman/intercept-mcp

MCP server with multi-tier fallback chain for fetching web content as clean markdown

8 stars
TypeScript 1 awesome list

apify/crawlee-python

Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Parsel, BeautifulSoup, Playwright, and raw HTTP. Both headful and headless mode. With proxy rotation.

9,393 stars
Python 2 awesome lists

lightpanda-io/browser

Lightpanda: the headless browser designed for AI and automation

34,410 stars
Zig 3 awesome lists