Scrapy, a fast high-level web crawling & scraping framework for Python.
Go to file
smirecki 8379bea7ed Typos corrections
I've made a few small corrections, some spelling changes and typo fixes.
I've tried to respect regional spelling differences and avoided proposing hyphenating compound words.

 Please enter the commit message for your changes. Lines starting
2015-10-02 23:48:27 -04:00
artwork added artwork files properly now 2012-03-20 10:46:45 -03:00
debian Merge pull request #934 from Dineshs91/zsh-support 2015-07-31 03:18:19 +05:00
docs Typos corrections 2015-10-02 23:48:27 -04:00
extras Merge pull request #934 from Dineshs91/zsh-support 2015-07-31 03:18:19 +05:00
scrapy Merge pull request #1498 from Preetwinder/scrapy-add_scheme 2015-09-30 21:38:58 -03:00
sep mark SEP-019 as Final 2014-10-02 01:17:31 -03:00
tests Restructure tests for add_http_if_no_scheme function 2015-09-24 17:57:25 +00:00
.bumpversion.cfg Bump version: 1.0.0rc1 → 1.1.0dev1 2015-05-22 13:24:50 -03:00
.coveragerc Add coverage report trough codecov.io 2015-08-13 13:56:24 -03:00
.gitignore add coverage files to gitignore 2015-08-26 01:58:33 +05:00
.travis.yml don't run tests twice on Travis if a PR is made from a scrapy/scrapy branch 2015-09-04 22:06:46 +05:00
AUTHORS added Nicolas Ramirez to AUTHORS 2013-03-14 12:44:39 -03:00
CONTRIBUTING.md Put a blurb about support channels in CONTRIBUTING 2015-07-24 01:48:43 +00:00
INSTALL fix link to online installation instructions 2012-10-02 12:26:14 +01:00
LICENSE mv scrapy/trunk to root as part of svn2hg migration 2009-05-06 15:55:17 -03:00
MANIFEST.in ENH: include tests/ to source distribution in MANIFEST.in 2015-06-25 23:00:00 -04:00
Makefile.buildbot Generated version as pep440 and dpkg compatible 2015-06-16 00:16:09 +00:00
NEWS added NEWS file pointing to docs/news.rst 2012-04-28 23:32:51 -03:00
README.rst Add PyPI download stats badge 2015-09-17 14:07:57 -03:00
conftest.py remove scrapy.utils.testsite from PY3 ignores 2015-08-01 00:43:13 +05:00
pytest.ini Don't collect tests by their class name 2015-05-04 18:10:04 -03:00
requirements-py3.txt add service_identity to scrapy install_requires 2015-08-11 20:14:22 +05:00
requirements.txt upgrade parsel and add shim for deprecated selectorlist methods 2015-08-11 15:20:33 -03:00
setup.cfg remove no longer existent examples from doc_files used in bdist_rpm. closes GH-417 2013-10-08 15:18:45 -02:00
setup.py upgrade parsel and use its function to instantiate root for finding form 2015-08-11 14:09:34 -03:00
tox.ini Do not be verbose with coverage report by default 2015-08-13 19:02:36 -03:00

README.rst

<html xmlns="http://www.w3.org/1999/xhtml" xml:lang="en" lang="en"> <head> </head>

Scrapy

Overview

Scrapy is a fast high-level web crawling and web scraping framework, used to crawl websites and extract structured data from their pages. It can be used for a wide range of purposes, from data mining to monitoring and automated testing.

For more information including a list of features check the Scrapy homepage at: http://scrapy.org

Requirements

  • Python 2.7
  • Works on Linux, Windows, Mac OSX, BSD

Install

The quick way:

pip install scrapy

For more details see the install section in the documentation: http://doc.scrapy.org/en/latest/intro/install.html

Releases

You can download the latest stable and development releases from: http://scrapy.org/download/

Documentation

Documentation is available online at http://doc.scrapy.org/ and in the docs directory.

Community (blog, twitter, mail list, IRC)

See http://scrapy.org/community/

Companies using Scrapy

See http://scrapy.org/companies/

Commercial Support

See http://scrapy.org/support/

</html>