Scrapy, a fast high-level web crawling & scraping framework for Python.
Go to file
Patrick Connolly 9c29215bef Place brackets on own lines with JsonItemExporter.
Placing the opening and closing brackets on their own lines makes it slightly easier to sort lines after the `spider_closed` signal is fired.
2016-04-25 15:28:41 +02:00
artwork added artwork files properly now 2012-03-20 10:46:45 -03:00
debian Merge pull request #934 from Dineshs91/zsh-support 2015-07-31 03:18:19 +05:00
docs Reference StackOverflow's "minimal, complete, and verifiable example" guide 2016-04-12 16:13:05 +02:00
extras Merge pull request #934 from Dineshs91/zsh-support 2015-07-31 03:18:19 +05:00
scrapy Place brackets on own lines with JsonItemExporter. 2016-04-25 15:28:41 +02:00
sep Spelling fixes 2015-12-13 19:39:48 -08:00
tests Use newer w3lib.url.safe_url_string() and re-enable HTTP request tests 2016-04-20 17:36:58 +02:00
.bumpversion.cfg Allow more pre-releases with bumpversion 2016-04-21 16:53:17 +02:00
.coveragerc Add coverage report trough codecov.io 2015-08-13 13:56:24 -03:00
.gitignore add coverage files to gitignore 2015-08-26 01:58:33 +05:00
.travis.yml Enable travis builds on tag patterns 2016-02-03 23:11:17 -03:00
AUTHORS added Nicolas Ramirez to AUTHORS 2013-03-14 12:44:39 -03:00
CODE_OF_CONDUCT.md Add Code of Conduct Version 1.3.0 from http://contributor-covenant.org/ 2016-01-15 18:01:04 +01:00
CONTRIBUTING.md Put a blurb about support channels in CONTRIBUTING 2015-07-24 01:48:43 +00:00
INSTALL fix link to online installation instructions 2012-10-02 12:26:14 +01:00
LICENSE mv scrapy/trunk to root as part of svn2hg migration 2009-05-06 15:55:17 -03:00
MANIFEST.in ENH: include tests/ to source distribution in MANIFEST.in 2015-06-25 23:00:00 -04:00
Makefile.buildbot Generated version as pep440 and dpkg compatible 2015-06-16 00:16:09 +00:00
NEWS added NEWS file pointing to docs/news.rst 2012-04-28 23:32:51 -03:00
README.rst Add link to CoC mardown file on Github 2016-01-27 13:04:08 +01:00
conftest.py Simplify if statement 2016-01-18 07:45:36 +01:00
pytest.ini Don't collect tests by their class name 2015-05-04 18:10:04 -03:00
requirements-py3.txt raise minimal twisted version for py3 2016-01-19 17:25:50 +03:00
requirements.txt Bump up w3lib requirement to v1.14.2 2016-04-20 17:37:16 +02:00
setup.cfg Build universal wheels 2016-03-01 14:18:14 +01:00
setup.py Remove duplicate code now handled by newer w3lib 2016-04-11 17:39:17 +02:00
tox.ini Add support for Sphinx 1.4 2016-03-30 17:49:14 +02:00

README.rst

<html xmlns="http://www.w3.org/1999/xhtml" xml:lang="en" lang="en"> <head> </head>

Scrapy

Overview

Scrapy is a fast high-level web crawling and web scraping framework, used to crawl websites and extract structured data from their pages. It can be used for a wide range of purposes, from data mining to monitoring and automated testing.

For more information including a list of features check the Scrapy homepage at: http://scrapy.org

Requirements

  • Python 2.7
  • Works on Linux, Windows, Mac OSX, BSD

Install

The quick way:

pip install scrapy

For more details see the install section in the documentation: http://doc.scrapy.org/en/latest/intro/install.html

Releases

You can download the latest stable and development releases from: http://scrapy.org/download/

Documentation

Documentation is available online at http://doc.scrapy.org/ and in the docs directory.

Community (blog, twitter, mail list, IRC)

See http://scrapy.org/community/

Contributing

Please note that this project is released with a Contributor Code of Conduct (see https://github.com/scrapy/scrapy/blob/master/CODE_OF_CONDUCT.md).

By participating in this project you agree to abide by its terms. Please report unacceptable behavior to opensource@scrapinghub.com.

See http://doc.scrapy.org/en/master/contributing.html

Companies using Scrapy

See http://scrapy.org/companies/

Commercial Support

See http://scrapy.org/support/

</html>