mirror of https://github.com/scrapy/scrapy.git
(closes #1027, #1095, #1102) |
||
|---|---|---|
| artwork | ||
| debian | ||
| docs | ||
| extras | ||
| scrapy | ||
| sep | ||
| tests | ||
| .bumpversion.cfg | ||
| .coveragerc | ||
| .gitignore | ||
| .travis-workarounds.sh | ||
| .travis.yml | ||
| AUTHORS | ||
| CONTRIBUTING.md | ||
| INSTALL | ||
| LICENSE | ||
| MANIFEST.in | ||
| Makefile.buildbot | ||
| NEWS | ||
| README.rst | ||
| conftest.py | ||
| pytest.ini | ||
| requirements.txt | ||
| setup.cfg | ||
| setup.py | ||
| tox.ini | ||
README.rst
<html xmlns="http://www.w3.org/1999/xhtml" xml:lang="en" lang="en">
<head>
</head>
</html>
Scrapy
Overview
Scrapy is a fast high-level web crawling and screen scraping framework, used to crawl websites and extract structured data from their pages. It can be used for a wide range of purposes, from data mining to monitoring and automated testing.
For more information including a list of features check the Scrapy homepage at: http://scrapy.org
Requirements
- Python 2.7
- Works on Linux, Windows, Mac OSX, BSD
Install
The quick way:
pip install scrapy
For more details see the install section in the documentation: http://doc.scrapy.org/en/latest/intro/install.html
Releases
You can download the latest stable and development releases from: http://scrapy.org/download/
Documentation
Documentation is available online at http://doc.scrapy.org/ and in the docs directory.