Scrapy, a fast high-level web crawling & scraping framework for Python.
Go to file
Adrian Chaves 1b917950f5 Remove unneeded comment 2026-04-29 16:40:41 +02:00
.github Add CI timeouts while at it 2026-04-29 12:45:52 +02:00
docs sphinx-scrapy: 0.8.4 → 0.8.5 (#7472) 2026-04-29 14:17:26 +02:00
extras Fix and ban the top-level twisted.internet.reactor imports. (#6835) 2025-05-28 15:53:52 +05:00
scrapy Remove unneeded comment 2026-04-29 16:40:41 +02:00
sep Add sphinx-lint. (#6920) 2025-06-28 01:37:20 +02:00
tests Address typing issue 2026-04-29 16:23:23 +02:00
tests_typing Move to mypy --strict with exceptions. (#7300) 2026-03-02 15:47:23 +05:00
.git-blame-ignore-revs Update tool versions (#7127) 2025-10-27 14:11:31 +01:00
.gitattributes
…
.gitignore Add llms.txt and llms-full.txt generation (#7380) 2026-04-06 10:24:21 +02:00
.pre-commit-config.yaml sphinx-scrapy: 0.8.4 → 0.8.5 (#7472) 2026-04-29 14:17:26 +02:00
.readthedocs.yml Use sphinx-scrapy 0.7.1 (#7406) 2026-04-06 15:29:59 +02:00
AUTHORS
…
CODE_OF_CONDUCT.md
…
CONTRIBUTING.md
…
INSTALL.md
…
LICENSE
…
NEWS
…
README.rst Remove Python 3.9 support (#7121) 2025-10-27 12:37:49 +01:00
SECURITY.md Bump version: 2.14.2 → 2.15.0 2026-04-09 17:00:40 +05:00
codecov.yml
…
conftest.py Enable in-process HTTP tests without a reactor. (#7254) 2026-02-13 19:08:06 +01:00
pyproject.toml Bump version: 2.15.1 → 2.15.2 2026-04-28 15:28:51 +02:00
tox.ini sphinx-scrapy: 0.8.4 → 0.8.5 (#7472) 2026-04-29 14:17:26 +02:00

README.rst

<html xmlns="http://www.w3.org/1999/xhtml" xml:lang="en" lang="en"> <head> </head>

Scrapy is a web scraping framework to extract structured data from websites. It is cross-platform, and requires Python 3.10+. It is maintained by Zyte (formerly Scrapinghub) and many other contributors.

Install with:

System Message: WARNING/2 (<stdin>, line 52)

Cannot analyze code. Pygments package not found.

.. code:: bash

    pip install scrapy

And follow the documentation to learn how to use it.

If you wish to contribute, see Contributing.

</html>