Scrapy, a fast high-level web crawling & scraping framework for Python.
Go to file
D00399830 fcf3d8e0a0 Updated the documentation for developer tools to have JavaScript instead of Javascript, as JavaScript is the more correct way to write it 2022-03-21 14:09:31 -06:00
.github CI: stop using tox-pip-version (#5389) 2022-02-04 12:27:39 +01:00
artwork Fixed artwork/README formatting 2020-01-15 08:54:25 +04:00
docs Updated the documentation for developer tools to have JavaScript instead of Javascript, as JavaScript is the more correct way to write it 2022-03-21 14:09:31 -06:00
extras Fix typos 2021-10-11 22:32:42 +08:00
scrapy Bump version: 2.6.0 → 2.6.1 2022-03-01 13:48:40 +01:00
sep Fix typos 2021-10-11 22:32:42 +08:00
tests Suggest installing the brotli package instead of brotlipy (#4267) 2022-03-17 05:39:54 +01:00
.bandit.yml bandit: allow-list B324 for the time being 2022-03-01 13:01:20 +01:00
.bumpversion.cfg Bump version: 2.6.0 → 2.6.1 2022-03-01 13:48:40 +01:00
.coveragerc Remove deprecated xlib module 2019-09-13 14:32:05 -03:00
.flake8 Reducing amount of warnings during test run (#5162) 2021-05-28 14:45:06 +05:00
.gitattributes Maybe the problem is not in the code after all 2020-08-13 06:35:09 +02:00
.gitignore Reducing amount of warnings during test run (#5162) 2021-05-28 14:45:06 +05:00
.readthedocs.yml Make Python 3.10 support official (#5265) 2021-10-18 22:09:17 +02:00
AUTHORS Scrapinghub → Zyte 2021-02-02 15:03:20 +01:00
CODE_OF_CONDUCT.md Add FAQ to code of Conduct (#5177) 2021-06-11 12:49:41 +05:00
CONTRIBUTING.md Be consistent with domain used for links to documentation website 2019-01-31 01:28:53 -03:00
INSTALL Be consistent with domain used for links to documentation website 2019-01-31 01:28:53 -03:00
LICENSE added oxford commas to LICENSE 2018-06-01 21:48:43 -03:00
MANIFEST.in Include additional files in sdists 2018-11-16 13:38:19 -05:00
NEWS added NEWS file pointing to docs/news.rst 2012-04-28 23:32:51 -03:00
README.rst Update Logo in README.rst (#5258) 2022-02-08 15:36:25 +01:00
codecov.yml codecov config: disable project check, tweak PR comments 2017-05-19 00:01:27 +05:00
conftest.py Add Deferred-to-Future helpers (#5288) 2021-10-22 18:46:01 +02:00
pylintrc Fix and pin pylint. 2021-11-26 12:25:45 +05:00
pytest.ini Copy resource classes from twisted.web.test.test_webclient. 2022-02-08 21:01:16 +05:00
setup.cfg Engine tests: fix item class spider, add minimal type hints 2021-04-09 13:09:47 -03:00
setup.py Merge pull request from GHSA-mfjm-vh54-3f96 2022-03-01 12:38:19 +01:00
tox.ini Freeze and upgrade CI packages (#5429) 2022-03-01 17:29:08 +01:00

README.rst

<html xmlns="http://www.w3.org/1999/xhtml" xml:lang="en" lang="en"> <head> </head>
https://scrapy.org/img/scrapylogo.png

Scrapy

Overview

Scrapy is a fast high-level web crawling and web scraping framework, used to crawl websites and extract structured data from their pages. It can be used for a wide range of purposes, from data mining to monitoring and automated testing.

Scrapy is maintained by Zyte (formerly Scrapinghub) and many other contributors.

Check the Scrapy homepage at https://scrapy.org for more information, including a list of features.

Requirements

  • Python 3.6+
  • Works on Linux, Windows, macOS, BSD

Install

The quick way:

pip install scrapy

See the install section in the documentation at https://docs.scrapy.org/en/latest/intro/install.html for more details.

Documentation

Documentation is available online at https://docs.scrapy.org/ and in the docs directory.

Releases

You can check https://docs.scrapy.org/en/latest/news.html for the release notes.

Community (blog, twitter, mail list, IRC)

See https://scrapy.org/community/ for details.

Contributing

See https://docs.scrapy.org/en/master/contributing.html for details.

Code of Conduct

Please note that this project is released with a Contributor Code of Conduct (see https://github.com/scrapy/scrapy/blob/master/CODE_OF_CONDUCT.md).

By participating in this project you agree to abide by its terms. Please report unacceptable behavior to opensource@zyte.com.

Companies using Scrapy

See https://scrapy.org/companies/ for a list.

Commercial Support

See https://scrapy.org/support/ for details.

</html>