Open source ai-scraping projects

Every project in the registry tagged ai-scraping, ranked by real GitHub adoption.

projects 4 combined stars ★ 296K refresh nightly
01 firecrawl ★ 182K

last push7 hours ago languageTypeScript licenseAGPL-3.0
02 Scrapling ★ 82K

🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl! Don't be shy, join here: https://discord.gg/EMgGbDce

last push3 days ago languagePython licenseBSD-3-Clause
03 Scrapegraph-ai ★ 31K

Python scraper based on AI

last push11 days ago languagePython licenseMIT
04 parsera ★ 1.4K

Lightweight library for scraping web-sites with LLMs

last push9 months ago languagePython licenseGPL-2.0

Related tags

← all tags

Frequently asked questions

How many open source ai-scraping projects are there?

This registry tracks 4 projects tagged ai-scraping, with 295,837 GitHub stars between them. The most-adopted is firecrawl at 181,627 stars.

Are these ai-scraping projects free to use?

Yes — 4 of the 4 carry an explicit open-source licence across 4 distinct licences, so there is no licence fee. Where a project also sells a hosted or enterprise version, the self-hosted path remains free.

Which ai-scraping project should I choose?

The list above is ranked by GitHub stars, but stars measure attention rather than fit. Check three things on each card: the licence (permissive versus copyleft), the language it is written in, and the last-push date — a high-star project that has not been pushed in a year is a liability.

Are these ai-scraping projects still maintained?

3 of the 4 were pushed in the last 90 days, and every card shows its exact last-push date so you can see the rest. Sort your shortlist by that date before committing to a migration.