apify

apify/crawlee

Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.

3 good first / help-wanted issues · TypeScript · last activity Aug 18, 2026

25.4K stars 1.6K forks 25.4K watchers TypeScript Apache License 2.0
apify automation crawler crawling headless headless-chrome javascript nodejs npm playwright puppeteer scraper scraping typescript web-crawler web-crawling web-scraping
3 Open Issues Need Help Last updated: Aug 18, 2026

Open Issues Need Help

View All on GitHub
apify/crawlee
25.4K

Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.

TypeScript
#apify#automation#crawler#crawling#headless#headless-chrome#javascript#nodejs#npm#playwright#puppeteer#scraper#scraping#typescript#web-crawler#web-crawling#web-scraping
Monitor mode 5 months ago
feature help wanted t-tooling hacktoberfest
apify/crawlee
25.4K

Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.

TypeScript
#apify#automation#crawler#crawling#headless#headless-chrome#javascript#nodejs#npm#playwright#puppeteer#scraper#scraping#typescript#web-crawler#web-crawling#web-scraping
bug good first issue t-tooling
apify/crawlee
25.4K

Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.

TypeScript
#apify#automation#crawler#crawling#headless#headless-chrome#javascript#nodejs#npm#playwright#puppeteer#scraper#scraping#typescript#web-crawler#web-crawling#web-scraping