s0rg/crawley
The unix-way web crawler
Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Parsel, BeautifulSoup, Playwright, and raw HTTP. Both headful and headless mode. With proxy rotation.
Appears on
Quick read
Latest capture 2026-08-03 04:38
42 observed captures since 2026-05-27. Observed captures are shown by default.
Stars from first capture +288
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
The unix-way web crawler
🚀🤖 Crawl4AI: Open-source LLM Friendly Web Crawler & Scraper. Don't be shy, join here: https://discord.gg/jP8KfhDhyN
The context API to search, scrape, and interact with the web at scale. 🔥
Transform developer documentation to clean Markdown
Fast Python web crawler for RAG and AI ingestion. Extracts clean Markdown from any site for LLMs and vector stores.
🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!