Firecrawl
π₯ Turn entire websites into LLM-ready markdown or structured data using an efficient API. Easily scrape, crawl, and extract data.
There are 4 open-source alternatives to Scrapy in this directory. The most popular is Firecrawl with 43.7k GitHub stars, followed by Scrapling(6.4k stars). All are open source, free to use, and many can be self-hosted.
| # | Tool | Stars | Language | License |
|---|---|---|---|---|
| 1 | Firecrawl π₯ Turn entire websites into LLM-ready markdown or structured data using an efficient API. Easily scrape, crawl, and extract data. | 43.7k | TypeScript | GNU Affero General Public License v3.0 |
| 2 | Scrapling π·οΈ An undetectable, powerful, flexible, high-performance Python library to make Web Scraping Easy and Effortless as it should be! | 6.4k | Python | BSD 3-Clause "New" or "Revised" License |
| 3 | Pipet A Swiss-army tool for scraping and extracting data from online assets, designed for hackers and data enthusiasts. | 4.7k | Go | MIT License |
| 4 | AnyCrawl AnyCrawl is a Node.js and TypeScript-powered web crawler that transforms websites into data suitable for large language models (LLMs) and extracts structured SERP results from search engines like Google, Bing, and Baidu. It features native multi-threading for efficient, bulk-scale processing. | 2.5k | TypeScript | MIT License |
π₯ Turn entire websites into LLM-ready markdown or structured data using an efficient API. Easily scrape, crawl, and extract data.
π·οΈ An undetectable, powerful, flexible, high-performance Python library to make Web Scraping Easy and Effortless as it should be!
A Swiss-army tool for scraping and extracting data from online assets, designed for hackers and data enthusiasts.
AnyCrawl is a Node.js and TypeScript-powered web crawler that transforms websites into data suitable for large language models (LLMs) and extracts structured SERP results from search engines like Google, Bing, and Baidu. It features native multi-threading for efficient, bulk-scale processing.
The top three are Firecrawl(43.7k stars), Scrapling(6.4k stars), Pipet(4.7k stars). All are open source and free to use.
Yes. Every tool listed here is open source and free to use in your own projects. Many can be self-hosted on your own infrastructure, which means no subscription fees and full control over your data.
Many of the alternatives listed are self-hostable. Each tool's page lists hosting details, system requirements, and licensing terms.
Get notified about new tools and updates to existing ones.