Scrapy, a fast high-level web crawling & scraping framework for Python.
Go to file
Eugenio Lacuesta 55ae2109c9
Remove deprecated TextResponse.body_as_unicode
2022-02-05 13:11:13 -03:00
.github CI: stop using tox-pip-version (#5389) 2022-02-04 12:27:39 +01:00
artwork Fixed artwork/README formatting 2020-01-15 08:54:25 +04:00
docs Fix FEED_URI_PARAMS: custom params throws KeyError (#4966) 2022-01-28 14:30:30 -03:00
extras Fix typos 2021-10-11 22:32:42 +08:00
scrapy Remove deprecated TextResponse.body_as_unicode 2022-02-05 13:11:13 -03:00
sep Fix typos 2021-10-11 22:32:42 +08:00
tests Remove deprecated TextResponse.body_as_unicode 2022-02-05 13:11:13 -03:00
.bandit.yml Mark bandit’s 402 check as addressed by #4180 (#4181) 2019-12-05 14:48:31 +01:00
.bumpversion.cfg Bump version: 2.4.1 → 2.5.0 2021-04-06 19:13:32 +05:00
.coveragerc Remove deprecated xlib module 2019-09-13 14:32:05 -03:00
.flake8 Reducing amount of warnings during test run (#5162) 2021-05-28 14:45:06 +05:00
.gitattributes Maybe the problem is not in the code after all 2020-08-13 06:35:09 +02:00
.gitignore Reducing amount of warnings during test run (#5162) 2021-05-28 14:45:06 +05:00
.readthedocs.yml Make Python 3.10 support official (#5265) 2021-10-18 22:09:17 +02:00
AUTHORS Scrapinghub → Zyte 2021-02-02 15:03:20 +01:00
CODE_OF_CONDUCT.md Add FAQ to code of Conduct (#5177) 2021-06-11 12:49:41 +05:00
CONTRIBUTING.md Be consistent with domain used for links to documentation website 2019-01-31 01:28:53 -03:00
INSTALL Be consistent with domain used for links to documentation website 2019-01-31 01:28:53 -03:00
LICENSE added oxford commas to LICENSE 2018-06-01 21:48:43 -03:00
MANIFEST.in Include additional files in sdists 2018-11-16 13:38:19 -05:00
NEWS added NEWS file pointing to docs/news.rst 2012-04-28 23:32:51 -03:00
README.rst Using Logo Scrapy in Readme.md 2021-10-03 13:26:20 +07:00
codecov.yml codecov config: disable project check, tweak PR comments 2017-05-19 00:01:27 +05:00
conftest.py Add Deferred-to-Future helpers (#5288) 2021-10-22 18:46:01 +02:00
pylintrc Fix and pin pylint. 2021-11-26 12:25:45 +05:00
pytest.ini Add Deferred-to-Future helpers (#5288) 2021-10-22 18:46:01 +02:00
setup.cfg Engine tests: fix item class spider, add minimal type hints 2021-04-09 13:09:47 -03:00
setup.py Make Python 3.10 support official (#5265) 2021-10-18 22:09:17 +02:00
tox.ini Fix and pin pylint. 2021-11-26 12:25:45 +05:00

README.rst

<html xmlns="http://www.w3.org/1999/xhtml" xml:lang="en" lang="en"> <head> </head>
/artwork/scrapy-logo.jpg

Scrapy

Overview

Scrapy is a fast high-level web crawling and web scraping framework, used to crawl websites and extract structured data from their pages. It can be used for a wide range of purposes, from data mining to monitoring and automated testing.

Scrapy is maintained by Zyte (formerly Scrapinghub) and many other contributors.

Check the Scrapy homepage at https://scrapy.org for more information, including a list of features.

Requirements

  • Python 3.6+
  • Works on Linux, Windows, macOS, BSD

Install

The quick way:

pip install scrapy

See the install section in the documentation at https://docs.scrapy.org/en/latest/intro/install.html for more details.

Documentation

Documentation is available online at https://docs.scrapy.org/ and in the docs directory.

Releases

You can check https://docs.scrapy.org/en/latest/news.html for the release notes.

Community (blog, twitter, mail list, IRC)

See https://scrapy.org/community/ for details.

Contributing

See https://docs.scrapy.org/en/master/contributing.html for details.

Code of Conduct

Please note that this project is released with a Contributor Code of Conduct (see https://github.com/scrapy/scrapy/blob/master/CODE_OF_CONDUCT.md).

By participating in this project you agree to abide by its terms. Please report unacceptable behavior to opensource@zyte.com.

Companies using Scrapy

See https://scrapy.org/companies/ for a list.

Commercial Support

See https://scrapy.org/support/ for details.

</html>