Open source scraper projects
Every project in the registry tagged scraper, ranked by real GitHub adoption.
Create agents that monitor and act on your behalf. Your agents are standing by!
👾 Fast and simple video download library and CLI tool written in Go
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or G
Elegant Scraper and Crawler Framework for Golang
🚀 Self-hosted TikTok & Douyin scraper and no-watermark video downloader — async REST API, MCP server, CLI and web console for posts, profiles, comments and pla
Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, P
Related tags
Frequently asked questions
How many open source scraper projects are there?
This registry tracks 7 projects tagged scraper, with 344,322 GitHub stars between them. The most-adopted is firecrawl at 181,627 stars.
Are these scraper projects free to use?
Yes — 7 of the 7 carry an explicit open-source licence across 3 distinct licences, so there is no licence fee. Where a project also sells a hosted or enterprise version, the self-hosted path remains free.
Which scraper project should I choose?
The list above is ranked by GitHub stars, but stars measure attention rather than fit. Check three things on each card: the licence (permissive versus copyleft), the language it is written in, and the last-push date — a high-star project that has not been pushed in a year is a liability.
Are these scraper projects still maintained?
6 of the 7 were pushed in the last 90 days, and every card shows its exact last-push date so you can see the rest. Sort your shortlist by that date before committing to a migration.