Scrapy, a fast high-level web crawling & scraping framework for Python.
Go to file
Denys Butenko 8e0b2bd343 Resolved issue #546. Output format parsing from filename extension. 2014-03-19 19:00:55 +02:00
artwork
…
bin
…
debian Added "six>=1.5.2" to requirements 2014-01-15 13:05:00 +06:00
docs add message when raise IngoreReques; fix item_scraped document 2014-03-18 15:23:25 +08:00
extras remove references to deprecated scrapy-developers list 2014-02-16 21:44:49 -02:00
scrapy Resolved issue #546. Output format parsing from filename extension. 2014-03-19 19:00:55 +02:00
sep sep 14 for #629 2014-03-07 18:05:21 -05:00
.coveragerc
…
.gitignore Added request_fingerprint method to dupefilter classes so they could be easily subclassed without need to override entire request_seen method. 2014-01-15 13:02:05 +02:00
.travis.yml reindent travis.yml as per travis gem defaults 2013-12-24 11:26:48 -02:00
AUTHORS
…
CONTRIBUTING.md
…
INSTALL
…
LICENSE
…
MANIFEST.in
…
Makefile.buildbot
…
NEWS
…
README.rst Drop Python 2.6 support 2013-10-29 13:44:00 -02:00
requirements.txt test_command_deploy, test_contrib_linkextractors 2014-01-11 14:30:27 +06:00
setup.cfg
…
setup.py Added "six>=1.5.2" to requirements 2014-01-15 13:05:00 +06:00
tests-requirements.txt testing PIL dependency is removed because there is a new mitmproxy version 2014-02-06 00:49:26 +06:00
tox.ini TST Improved twisted installation in tox.ini for Python 3.3 2014-02-17 17:21:38 +06:00

README.rst

<html xmlns="http://www.w3.org/1999/xhtml" xml:lang="en" lang="en"> <head> </head>

Scrapy

https://badge.fury.io/py/Scrapy.png https://secure.travis-ci.org/scrapy/scrapy.png?branch=master

Overview

Scrapy is a fast high-level screen scraping and web crawling framework, used to crawl websites and extract structured data from their pages. It can be used for a wide range of purposes, from data mining to monitoring and automated testing.

For more information including a list of features check the Scrapy homepage at: http://scrapy.org

Requirements

  • Python 2.7
  • Works on Linux, Windows, Mac OSX, BSD

Install

The quick way:

pip install scrapy

For more details see the install section in the documentation: http://doc.scrapy.org/en/latest/intro/install.html

Releases

You can download the latest stable and development releases from: http://scrapy.org/download/

Documentation

Documentation is available online at http://doc.scrapy.org/ and in the docs directory.

Community (blog, twitter, mail list, IRC)

See http://scrapy.org/community/

Companies using Scrapy

See http://scrapy.org/companies/

Commercial Support

See http://scrapy.org/support/

</html>