Skip to content
Advertisement

crawlee Open source

Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScri

Collection & Scraping

Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.

Open-source data tool (Apache-2.0 license, 24,902★). Source: GitHub (apify/crawlee).

Text
Advertisement