GitHub 项目简介: Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files…
README: Crawlee is available as [`crawlee`](https://pypi.org/project/crawlee/) package on PyPI. This package includes the core functionality, while additional features are available as op…
README: Your crawlers will appear almost human-like and fly under the radar of modern bot protections even with the default configuration.
README: Unified interface for **HTTP & headless browser** crawling. - Automatic **parallel crawling** based on available system resources. - Written in Python with **type hints** - enhanc…