Scrapy, a fast high-level web crawling & scraping framework for Python.
Go to file
Konstantin Lopuhin b4eb60e527 Install PyPyDispatcher for PyPy tests
Using https://github.com/lopuhin/pydispatcher, pypy branch.
This is executed as a separate step to avoid changing
default requirements.txt and setup.py. If just added to "deps"
in tox, this install command will be executed as one command
and PyPyDispatcher will not override PyDispatcher.
2017-06-15 13:34:58 +03:00
artwork changing README to README.rst 2017-01-22 19:23:44 -05:00
debian Merge pull request #934 from Dineshs91/zsh-support 2015-07-31 03:18:19 +05:00
docs Merge pull request #2781 from crasker/patch-1 2017-06-14 03:29:45 +05:00
extras Update scrapy.1 2017-01-25 11:28:20 +01:00
scrapy Move garbage_collect to scrapy.utils.python 2017-06-15 13:14:31 +03:00
sep changing README to README.rst 2017-01-22 19:23:44 -05:00
tests Fix get_func_args tests under PyPy 2017-06-15 13:07:59 +03:00
.bumpversion.cfg Bump version: 1.3.2 → 1.4.0 2017-05-18 23:01:05 +02:00
.coveragerc Add coverage report trough codecov.io 2015-08-13 13:56:24 -03:00
.gitignore add a couple more lines to gitignore 2017-02-13 18:40:53 +05:00
.travis.yml Travis CI: use portable pypy for Linux 2017-04-12 16:32:21 +02:00
AUTHORS added Nicolas Ramirez to AUTHORS 2013-03-14 12:44:39 -03:00
CODE_OF_CONDUCT.md update code of conduct http://contributor-covenant.org/version/1/4 2016-12-27 11:32:15 -02:00
CONTRIBUTING.md Put a blurb about support channels in CONTRIBUTING 2015-07-24 01:48:43 +00:00
INSTALL fix link to online installation instructions 2012-10-02 12:26:14 +01:00
LICENSE [PEDANTIC] FIX trailing whitespaces in LICENSE. 2017-04-22 00:24:18 +02:00
MANIFEST.in Ignore explicitly compiled python files. 2016-11-08 20:52:32 -03:00
Makefile.buildbot Generated version as pep440 and dpkg compatible 2015-06-16 00:16:09 +00:00
NEWS added NEWS file pointing to docs/news.rst 2012-04-28 23:32:51 -03:00
README.rst DOC remove “Python 3 progress” badge 2017-02-16 04:22:19 +05:00
codecov.yml codecov config: disable project check, tweak PR comments 2017-05-19 00:01:27 +05:00
conftest.py Simplify if statement 2016-01-18 07:45:36 +01:00
pytest.ini Don't collect tests by their class name 2015-05-04 18:10:04 -03:00
requirements-py3.txt require w3lib 1.17+ 2017-02-15 00:32:44 +05:00
requirements.txt require w3lib 1.17+ 2017-02-15 00:32:44 +05:00
setup.cfg Build universal wheels 2016-03-01 11:00:20 +01:00
setup.py require w3lib 1.17+ 2017-02-15 00:32:44 +05:00
tox.ini Install PyPyDispatcher for PyPy tests 2017-06-15 13:34:58 +03:00

README.rst

<html xmlns="http://www.w3.org/1999/xhtml" xml:lang="en" lang="en"> <head> </head>

Scrapy

Overview

Scrapy is a fast high-level web crawling and web scraping framework, used to crawl websites and extract structured data from their pages. It can be used for a wide range of purposes, from data mining to monitoring and automated testing.

For more information including a list of features check the Scrapy homepage at: http://scrapy.org

Requirements

  • Python 2.7 or Python 3.3+
  • Works on Linux, Windows, Mac OSX, BSD

Install

The quick way:

pip install scrapy

For more details see the install section in the documentation: http://doc.scrapy.org/en/latest/intro/install.html

Releases

You can download the latest stable and development releases from: http://scrapy.org/download/

Documentation

Documentation is available online at http://doc.scrapy.org/ and in the docs directory.

Community (blog, twitter, mail list, IRC)

See http://scrapy.org/community/

Contributing

See http://doc.scrapy.org/en/master/contributing.html

Code of Conduct

Please note that this project is released with a Contributor Code of Conduct (see https://github.com/scrapy/scrapy/blob/master/CODE_OF_CONDUCT.md).

By participating in this project you agree to abide by its terms. Please report unacceptable behavior to opensource@scrapinghub.com.

Companies using Scrapy

See http://scrapy.org/companies/

Commercial Support

See http://scrapy.org/support/

</html>