GitHub 仓库
apify/crawlee ↗Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.
2.6wStars
1.7kForks
140Issues
135Watchers
● TypeScript Apache-2.0 最近更新 5 小时前 创建于 2016-08-26
apifyautomationcrawlercrawlingheadlessheadless-chromejavascriptnodejs
相似推荐
进入对比 →