crawlee-python

crawlee-python

Alternative to ScrapingBee

Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Parsel, BeautifulSoup, Playwright, and raw HTTP. Both headful and headless mode. With proxy rotation.

9.5k 8032 HN points
Apache-2.0
last commit 2026-06-11
Website Source
Share:

About

Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Parsel, BeautifulSoup, Playwright, and raw HTTP. Both headful and headless mode. With proxy rotation.

Languages

Contributors30

Free alternative to ScrapingBee

crawlee-python is a free, open-source alternative to ScrapingBee. Here's why:

  • 🕸️ Advanced Web Scraping
  • 🤖 Browser Automation
  • 🧠 AI Data Extraction
  • 📄 Multi-Format Downloads
  • 🔄 Proxy Rotation
See all open-source alternatives to ScrapingBee
Comments Theme
slug: crawlee-python