AnyCrawl
AnyCrawl 🚀: A Node.js/TypeScript crawler that turns websites into LLM-ready data and extracts structured SERP results from Google/Bing/Baidu/etc. Native multi-threading for bulk processing.
- Type: Framework
- Imported from GitHub
- Popularity: 3.5K GitHub stars
- License: MIT
- Source: https://github.com/any4ai/AnyCrawl
- Repository: https://github.com/any4ai/anycrawl
- Tags: ai-scraping, aitools, crawl, data, html-to-markdown, rag, scrape, scraping, serp, webscraper
- Updated: 2026-10-01
More frameworks
- open-webui — User-friendly AI Interface (Supports Ollama, OpenAI API, ...)
- langchain — The agent engineering platform.
- awesome-llm-apps — 100+ AI Agents, Agent Skills and RAG Apps - Free and Open Source.
- graphify — Turn any codebase, with its docs, SQL schemas, configs, and PDFs, into a queryable knowledge graph. A /graphify skill for Claude Code, Curs…
- ragflow — RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with Agent capabilities to create…
- PaddleOCR — Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PD…
- crawl4ai — Open-source web crawler and scraper for LLMs and AI agents: any website into clean, LLM-ready Markdown. Run it yourself, or use Crawl4AI Cl…
- hello-agents — 📚 《从零开始构建智能体》——从零开始的智能体原理与实践教程