Firecrawl is an API and open server for search, crawling, and turning web pages into clean Markdown or structured data.
Open Source: #web-scraping
Catalog projects marked with #web-scraping. Tags work as dedicated landing pages, so related tools are easier to find and connect.
This collection holds 6 projects with a combined 419,990 GitHub stars. Main languages: Python, TypeScript, Ruby.
Repositories
Found: 6
firecrawl/firecrawl
Firecrawl
unclecode/crawl4ai
Crawl4AI
Crawl4AI is a Python crawler and web-data extraction tool that prepares pages for LLM, RAG, and agent workflows.
D4Vinci/Scrapling
Scrapling
Scrapling is a Python web scraping framework for requests, crawling, and adaptive data extraction.
scrapy/scrapy
Scrapy
Scrapy is a Python framework for web crawlers and structured data extraction from websites.
huginn/huginn
Huginn
Huginn is a system for personal automation: agents read the web, watch events, and run actions on your own server.
NaiboWang/EasySpider
EasySpider
EasySpider is a visual tool for browser automation and web data collection.