Scrapy, a fast high-level web crawling & scraping framework for Python.
Go to file
Paul Tremberth fe3a2fcce9 Build universal wheels 2016-03-01 14:18:14 +01:00
artwork added artwork files properly now 2012-03-20 10:46:45 -03:00
debian Merge pull request #934 from Dineshs91/zsh-support 2015-07-31 03:18:19 +05:00
docs Update release notes about change of default S3 ACL policy to "private" 2016-02-29 13:15:50 +01:00
extras Merge pull request #934 from Dineshs91/zsh-support 2015-07-31 03:18:19 +05:00
scrapy Bump version: 1.1.0rc1 → 1.1.0rc2 2016-02-29 13:15:54 +01:00
sep Spelling fixes 2015-12-13 19:39:48 -08:00
tests Merge pull request #1809 from scrapy/backport-1.1-1787 2016-02-24 17:54:56 +01:00
.bumpversion.cfg Bump version: 1.1.0rc1 → 1.1.0rc2 2016-02-29 13:15:54 +01:00
.coveragerc Add coverage report trough codecov.io 2015-08-13 13:56:24 -03:00
.gitignore add coverage files to gitignore 2015-08-26 01:58:33 +05:00
.travis.yml Enable travis builds on tag patterns 2016-02-03 23:11:17 -03:00
AUTHORS added Nicolas Ramirez to AUTHORS 2013-03-14 12:44:39 -03:00
CODE_OF_CONDUCT.md Add Code of Conduct Version 1.3.0 from http://contributor-covenant.org/ 2016-01-15 18:01:04 +01:00
CONTRIBUTING.md Put a blurb about support channels in CONTRIBUTING 2015-07-24 01:48:43 +00:00
INSTALL fix link to online installation instructions 2012-10-02 12:26:14 +01:00
LICENSE mv scrapy/trunk to root as part of svn2hg migration 2009-05-06 15:55:17 -03:00
MANIFEST.in ENH: include tests/ to source distribution in MANIFEST.in 2015-06-25 23:00:00 -04:00
Makefile.buildbot Generated version as pep440 and dpkg compatible 2015-06-16 00:16:09 +00:00
NEWS added NEWS file pointing to docs/news.rst 2012-04-28 23:32:51 -03:00
README.rst Add link to CoC mardown file on Github 2016-01-27 13:04:08 +01:00
conftest.py Simplify if statement 2016-01-18 07:45:36 +01:00
pytest.ini Don't collect tests by their class name 2015-05-04 18:10:04 -03:00
requirements-py3.txt raise minimal twisted version for py3 2016-01-19 17:25:50 +03:00
requirements.txt upgrade parsel and add shim for deprecated selectorlist methods 2015-08-11 15:20:33 -03:00
setup.cfg Build universal wheels 2016-03-01 14:18:14 +01:00
setup.py upgrade parsel and use its function to instantiate root for finding form 2015-08-11 14:09:34 -03:00
tox.ini [backport][1.1.x] Py3 S3 botocore 2016-02-18 22:48:37 +01:00

README.rst

<html xmlns="http://www.w3.org/1999/xhtml" xml:lang="en" lang="en"> <head> </head>

Scrapy

Overview

Scrapy is a fast high-level web crawling and web scraping framework, used to crawl websites and extract structured data from their pages. It can be used for a wide range of purposes, from data mining to monitoring and automated testing.

For more information including a list of features check the Scrapy homepage at: http://scrapy.org

Requirements

  • Python 2.7
  • Works on Linux, Windows, Mac OSX, BSD

Install

The quick way:

pip install scrapy

For more details see the install section in the documentation: http://doc.scrapy.org/en/latest/intro/install.html

Releases

You can download the latest stable and development releases from: http://scrapy.org/download/

Documentation

Documentation is available online at http://doc.scrapy.org/ and in the docs directory.

Community (blog, twitter, mail list, IRC)

See http://scrapy.org/community/

Contributing

Please note that this project is released with a Contributor Code of Conduct (see https://github.com/scrapy/scrapy/blob/master/CODE_OF_CONDUCT.md).

By participating in this project you agree to abide by its terms. Please report unacceptable behavior to opensource@scrapinghub.com.

See http://doc.scrapy.org/en/master/contributing.html

Companies using Scrapy

See http://scrapy.org/companies/

Commercial Support

See http://scrapy.org/support/

</html>