Open source crawling projects
Every project in the registry tagged crawling, ranked by real GitHub adoption.
🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl! Don't be shy, join here: https://discord.gg/EMgGbDce
Scrapy, a fast high-level web crawling & scraping framework for Python.
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or G
Elegant Scraper and Crawler Framework for Golang
No-code web scraping, crawling, and extraction platform
Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, P
A Chrome DevTools Protocol driver for web automation and scraping.
Headless Chrome .NET API
Open source SEO audit tool.
Related tags
Frequently asked questions
How many open source crawling projects are there?
This registry tracks 9 projects tagged crawling, with 236,358 GitHub stars between them. The most-adopted is Scrapling at 81,798 stars.
Are these crawling projects free to use?
Yes — 9 of the 9 carry an explicit open-source licence across 4 distinct licences, so there is no licence fee. Where a project also sells a hosted or enterprise version, the self-hosted path remains free.
Which crawling project should I choose?
The list above is ranked by GitHub stars, but stars measure attention rather than fit. Check three things on each card: the licence (permissive versus copyleft), the language it is written in, and the last-push date — a high-star project that has not been pushed in a year is a liability.
Are these crawling projects still maintained?
8 of the 9 were pushed in the last 90 days, and every card shows its exact last-push date so you can see the rest. Sort your shortlist by that date before committing to a migration.