← 返回专题广场
web-crawling
4 个项目 · ⭐ 35.0k1
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.
TypeScript
⭐ 25.5k
⑂ 1.6k
Apache-2.0
· 1 天前推送
1 天前
最近推送
2
Declarative data automation language and Go runtime for structured extraction workflows.
Go
⭐ 6.0k
⑂ 323
Apache-2.0
· 5 小时前推送
5 小时前
最近推送
3
A powerful Model Context Protocol (MCP) server that provides an all-in-one solution for public web access.
JavaScript
⭐ 2.6k
⑂ 323
MIT
· 10 天前推送
10 天前
最近推送
4
Best headless browser for AI agents. Lite, Fast, High-Compatibility. Built in Rust
Rust
⭐ 913
⑂ 55
Apache-2.0
· 3 小时前推送
3 小时前
最近推送