< cd ../gallery
[DevTools]
crawlee
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.
@apify
maintainer
★ 24.7k stars
# README
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation. Maintained by apify on GitHub, where it has earned 24,703 stars from the community.
It's actively developed around apify, automation, crawler, and is a solid reference for anyone building with these tools.
# tags
# install
npm install crawleelanguages
TypeScript57.632%
MDX33.553%
JavaScript6.898%
CSS1.065%
Dockerfile0.826%
Python0.024%
HTML0.002%
last commit1 week ago
licenseApache-2.0
more Express repos