← Explore
TOPIC

#crawling

Open source repositories tagged with #crawling, ranked by health score.

apify
apify/crawlee
TypeScript
89
health

Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.

25.5k
dondai44423
dondai44423/donsetch
Rust
89
health

Web fetch, search, and crawl for AI agents. Built from scratch in Rust. No keys, no accounts. AGPL v3.

472
hardkoded
hardkoded/puppeteer-sharp
C#
89
health

Headless Chrome .NET API

3.9k