Back to List
Notice:This resource is provided by a third-party author. Please review the code with AI tools or manually before use to ensure security and compatibility.
TypeScriptwatercrawl/WaterCrawl

WaterCrawl

Transform Web Content into LLM-Ready Data

77.1/100
2.2KForks: 278
View on GitHubHomepage →
Loading report...

Similar Projects

firecrawl

93

The context API to search, scrape, and interact with the web at scale. 🔥

TypeScript177.7K

crawlee

94

Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.

TypeScript25.7K

maxun

91

🔥 The open-source no-code platform for web scraping, crawling, search and AI data extraction • Turn websites into structured APIs in minutes 🔥

TypeScript17.4K

llm-scraper

62

Turn any webpage into structured data using LLMs

TypeScript6.9K
Back to List