Scrapy, a fast high-level web crawling & scraping framework for Python.
Go to file
Lucy Wang acec4260d2 catch CertificateError in tls verification 2018-06-21 11:07:00 -03:00
artwork changing README to README.rst 2017-01-22 19:23:44 -05:00
debian Merge pull request #934 from Dineshs91/zsh-support 2015-07-31 03:18:19 +05:00
docs update scrapinghub.com urls to use https 2017-08-24 16:03:36 -03:00
extras xrange() --> range() for Python 3 2017-07-24 22:06:17 +02:00
scrapy catch CertificateError in tls verification 2018-06-21 11:07:00 -03:00
sep changing README to README.rst 2017-01-22 19:23:44 -05:00
tests catch CertificateError in tls verification 2018-06-21 11:07:00 -03:00
.bumpversion.cfg Bump version: 1.3.2 → 1.4.0 2017-05-18 23:01:05 +02:00
.coveragerc Add coverage report trough codecov.io 2015-08-13 13:56:24 -03:00
.gitignore add a couple more lines to gitignore 2017-02-13 18:40:53 +05:00
.travis.yml Remove duplicate PyPy toxenv from Travis config 2017-06-19 17:45:28 +03:00
AUTHORS added Nicolas Ramirez to AUTHORS 2013-03-14 12:44:39 -03:00
CODE_OF_CONDUCT.md update code of conduct http://contributor-covenant.org/version/1/4 2016-12-27 11:32:15 -02:00
CONTRIBUTING.md Put a blurb about support channels in CONTRIBUTING 2015-07-24 01:48:43 +00:00
INSTALL fix link to online installation instructions 2012-10-02 12:26:14 +01:00
LICENSE [PEDANTIC] FIX trailing whitespaces in LICENSE. 2017-04-22 00:24:18 +02:00
MANIFEST.in Ignore explicitly compiled python files. 2016-11-08 20:52:32 -03:00
Makefile.buildbot Generated version as pep440 and dpkg compatible 2015-06-16 00:16:09 +00:00
NEWS added NEWS file pointing to docs/news.rst 2012-04-28 23:32:51 -03:00
README.rst DOC change "releases" section content 2017-05-30 00:50:32 +05:00
codecov.yml codecov config: disable project check, tweak PR comments 2017-05-19 00:01:27 +05:00
conftest.py Simplify if statement 2016-01-18 07:45:36 +01:00
pytest.ini Don't collect tests by their class name 2015-05-04 18:10:04 -03:00
requirements-py3.txt require w3lib 1.17+ 2017-02-15 00:32:44 +05:00
requirements.txt require w3lib 1.17+ 2017-02-15 00:32:44 +05:00
setup.cfg Build universal wheels 2016-03-01 11:00:20 +01:00
setup.py Use environment markers for custom PyPy requirements 2017-06-19 19:16:50 +03:00
tox.ini Jessi toxenv: Add cryptography as per https://packages.debian.org/jessie/python-cryptography 2017-07-24 19:30:08 +02:00

README.rst

<html xmlns="http://www.w3.org/1999/xhtml" xml:lang="en" lang="en"> <head> </head>

Scrapy

Overview

Scrapy is a fast high-level web crawling and web scraping framework, used to crawl websites and extract structured data from their pages. It can be used for a wide range of purposes, from data mining to monitoring and automated testing.

For more information including a list of features check the Scrapy homepage at: http://scrapy.org

Requirements

  • Python 2.7 or Python 3.3+
  • Works on Linux, Windows, Mac OSX, BSD

Install

The quick way:

pip install scrapy

For more details see the install section in the documentation: http://doc.scrapy.org/en/latest/intro/install.html

Documentation

Documentation is available online at http://doc.scrapy.org/ and in the docs directory.

Releases

You can find release notes at https://doc.scrapy.org/en/latest/news.html

Community (blog, twitter, mail list, IRC)

See http://scrapy.org/community/

Contributing

See http://doc.scrapy.org/en/master/contributing.html

Code of Conduct

Please note that this project is released with a Contributor Code of Conduct (see https://github.com/scrapy/scrapy/blob/master/CODE_OF_CONDUCT.md).

By participating in this project you agree to abide by its terms. Please report unacceptable behavior to opensource@scrapinghub.com.

Companies using Scrapy

See http://scrapy.org/companies/

Commercial Support

See http://scrapy.org/support/

</html>