Scrapy, a fast high-level web crawling & scraping framework for Python.
Go to file
Paul Tremberth 600e7bbd75 Bump version: 1.0.6 → 1.0.7 2017-03-03 19:19:12 +01:00
artwork added artwork files properly now 2012-03-20 10:46:45 -03:00
debian Generated version as pep440 and dpkg compatible 2015-06-16 00:13:30 +00:00
docs Set release date for 1.0.7 2017-03-03 19:18:09 +01:00
extras removed SUFFIX from scrapy name package 2015-06-15 18:04:30 -03:00
scrapy Bump version: 1.0.6 → 1.0.7 2017-03-03 19:19:12 +01:00
sep Spelling fixes 2015-12-30 15:23:11 -03:00
tests Fix tests on URL path encoding for links from latin1 document 2016-04-07 22:43:47 +02:00
.bumpversion.cfg Bump version: 1.0.6 → 1.0.7 2017-03-03 19:19:12 +01:00
.coveragerc Added rules to Makefile.buildbot for generating coverage reports 2010-12-15 11:13:45 -02:00
.gitignore minor corrections in documentation. 2015-04-19 18:58:15 +04:00
.travis.yml Yet another try to build 1.0.4 tag 2015-12-30 16:21:06 -03:00
AUTHORS added Nicolas Ramirez to AUTHORS 2013-03-14 12:44:39 -03:00
CONTRIBUTING.md Put a blurb about support channels in CONTRIBUTING 2015-08-06 17:54:13 -03:00
INSTALL fix link to online installation instructions 2012-10-02 12:26:14 +01:00
LICENSE mv scrapy/trunk to root as part of svn2hg migration 2009-05-06 15:55:17 -03:00
MANIFEST.in ENH: include tests/ to source distribution in MANIFEST.in 2015-07-01 01:40:39 -03:00
Makefile.buildbot Changed buildbot makefile to use 'pytest' 2016-01-19 09:03:01 +02:00
NEWS added NEWS file pointing to docs/news.rst 2012-04-28 23:32:51 -03:00
README.rst Add PyPI download stats badge 2015-12-30 15:02:18 -03:00
conftest.py Ignoring xlib/tx folder, depending on Twisted version. 2015-12-30 15:31:10 -03:00
pytest.ini Don't collect tests by their class name 2015-05-04 18:10:04 -03:00
requirements.txt Require at least lxml 2.3 2017-03-02 16:32:10 +01:00
setup.cfg remove no longer existent examples from doc_files used in bdist_rpm. closes GH-417 2013-10-08 15:18:45 -02:00
setup.py Require at least lxml 2.3 2017-03-02 16:32:10 +01:00
tox.ini Add support for Sphinx 1.4 2016-03-30 17:55:20 +02:00

README.rst

<html xmlns="http://www.w3.org/1999/xhtml" xml:lang="en" lang="en"> <head> </head>

Scrapy

Overview

Scrapy is a fast high-level web crawling and web scraping framework, used to crawl websites and extract structured data from their pages. It can be used for a wide range of purposes, from data mining to monitoring and automated testing.

For more information including a list of features check the Scrapy homepage at: http://scrapy.org

Requirements

  • Python 2.7
  • Works on Linux, Windows, Mac OSX, BSD

Install

The quick way:

pip install scrapy

For more details see the install section in the documentation: http://doc.scrapy.org/en/latest/intro/install.html

Releases

You can download the latest stable and development releases from: http://scrapy.org/download/

Documentation

Documentation is available online at http://doc.scrapy.org/ and in the docs directory.

Community (blog, twitter, mail list, IRC)

See http://scrapy.org/community/

Companies using Scrapy

See http://scrapy.org/companies/

Commercial Support

See http://scrapy.org/support/

</html>