C
crawlee
apify/crawlee
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.
★25.2kstars
TypeScript
Apache-2.0
Updated: Today
📋 Project at a Glance
Tap to expand
What's this?A , built with TypeScript open-source project in the Agents category, core strengths: apify/automation
Who made it?Maintained by apify team, 25.2K⭐ on GitHub, #80 out of 2166 in Agents
Why does it exist?The apify team recognized that existing apify tools in Agents were hard to use. crawlee was designed to make automation more accessible.
What can it do?Key use cases: crawler, crawling, headless
How to install with AI?Use an AI coding assistant to automatically run npm/pnpm install and configure the project per the README.
🔗 github.com/apify/crawlee | 官网 https://crawlee.dev
🔗 github.com/apify/crawlee | 官网 https://crawlee.dev
Topics
apifyautomationcrawlercrawlingheadlessheadless-chromejavascriptnodejsnpmplaywrightpuppeteerscraperscrapingtypescriptweb-crawlerweb-crawlingweb-scraping