Scrapy, a fast high-level web crawling & scraping framework for Python.
Go to file
nyov cf50561b86
Allow passing classes directly in Settings (#3873)
Co-authored-by: Adrián Chaves <adrian@chaves.io>
2020-08-26 13:08:14 +02:00
.github/ISSUE_TEMPLATE
…
artwork Fixed artwork/README formatting 2020-01-15 08:54:25 +04:00
docs Allow passing classes directly in Settings (#3873) 2020-08-26 13:08:14 +02:00
extras Change super syntax (#4707) 2020-08-04 20:42:01 +02:00
scrapy Allow passing classes directly in Settings (#3873) 2020-08-26 13:08:14 +02:00
sep Fix a spelling error: ie. → i.e. (#4338) 2020-02-18 17:58:31 +01:00
tests Allow passing classes directly in Settings (#3873) 2020-08-26 13:08:14 +02:00
.bandit.yml
…
.bumpversion.cfg Bump version: 2.2.0 → 2.3.0 2020-08-04 20:07:02 +02:00
.coveragerc
…
.gitattributes Maybe the problem is not in the code after all 2020-08-13 06:35:09 +02:00
.gitignore Typing: Tox env, CI job 2020-06-18 13:56:07 -03:00
.readthedocs.yml Allow doc to be downloadable on readthedocs.org 2020-05-19 02:17:11 +02:00
.travis.yml Remove Python 3.5 from CI (#4743) 2020-08-22 09:33:35 +02:00
AUTHORS
…
CODE_OF_CONDUCT.md
…
CONTRIBUTING.md
…
INSTALL
…
LICENSE
…
MANIFEST.in
…
NEWS
…
README.rst Bump minimum Python version to 3.5.2 (#4615) 2020-06-11 14:53:59 +02:00
azure-pipelines.yml Remove Python 3.5 from CI (#4743) 2020-08-22 09:33:35 +02:00
codecov.yml
…
conftest.py Merge branch 'master' into response_ip_address 2020-02-23 18:13:52 -03:00
pylintrc Skip checks introduced in Pylint 2.6.0 2020-08-21 17:06:54 +02:00
pytest.ini Update pytest.ini 2020-07-10 18:22:43 +05:30
setup.cfg Conditional request attribute binding for responses (#4632) 2020-08-17 10:39:59 +02:00
setup.py Bitbucket no longer supports Mercurial repositories (#4738) 2020-08-19 17:45:24 +02:00
tox.ini Remove Python 3.5 from CI (#4743) 2020-08-22 09:33:35 +02:00

README.rst

<html xmlns="http://www.w3.org/1999/xhtml" xml:lang="en" lang="en"> <head> </head>

Scrapy

Overview

Scrapy is a fast high-level web crawling and web scraping framework, used to crawl websites and extract structured data from their pages. It can be used for a wide range of purposes, from data mining to monitoring and automated testing.

Check the Scrapy homepage at https://scrapy.org for more information, including a list of features.

Requirements

  • Python 3.5.2+
  • Works on Linux, Windows, macOS, BSD

Install

The quick way:

pip install scrapy

See the install section in the documentation at https://docs.scrapy.org/en/latest/intro/install.html for more details.

Documentation

Documentation is available online at https://docs.scrapy.org/ and in the docs directory.

Releases

You can check https://docs.scrapy.org/en/latest/news.html for the release notes.

Community (blog, twitter, mail list, IRC)

See https://scrapy.org/community/ for details.

Contributing

See https://docs.scrapy.org/en/master/contributing.html for details.

Code of Conduct

Please note that this project is released with a Contributor Code of Conduct (see https://github.com/scrapy/scrapy/blob/master/CODE_OF_CONDUCT.md).

By participating in this project you agree to abide by its terms. Please report unacceptable behavior to opensource@scrapinghub.com.

Companies using Scrapy

See https://scrapy.org/companies/ for a list.

Commercial Support

See https://scrapy.org/support/ for details.

</html>