web-crawler
17 个项目 · ⭐ 275.2kThe context API to search, scrape, and interact with the web at scale. 🔥
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.
Distributed web crawler admin platform for spiders management regardless of languages and frameworks. 分布式爬虫管理平台,支持任何语言和框架
The go-to web for your AI coding agent — local-first search, fetch, crawl & research over MCP. No API keys, no cloud, $0/query. Public beta.
Cross Platform C# web crawler framework built for speed and flexibility. Please star this project! +1.
Best headless browser for AI agents. Lite, Fast, High-Compatibility. Built in Rust
🔥 This repository contains complete application examples, including websites and other projects, developed using Firecrawl.
Open source web infrastructure for AI. Scrape, crawl, and automate the web, clean markdown, browser sessions, ready for your agents.