Scrapy, a fast high-level web crawling & scraping framework for Python.
Go to file
Pablo Hoffman cae22930c8 Added ExecutionQueue class for feeding spiders and requests to scrape. This
class can (and is meant to) be subclassed by projects that want to use a custom
mechanism for feeding spiders to crawl. For example, a queue that pulls spiders
to scrape from Amazon SQS (an example will be added soon).

Also introduced a rather big core refactoring of Scrapy manager and Scrapy
engine.
2010-05-26 10:29:32 -03:00
bin added scrapy.service and scrapy.tac for running from twistd 2010-04-11 03:37:08 -03:00
docs ItemLoader: Update docs for {add,replace,get}_{value,xpath} 2010-05-18 17:54:25 +08:00
examples SEP12 implementation 2010-04-01 18:27:22 -03:00
extras removed old untested (and probably broken) code 2010-04-01 04:05:53 -03:00
profiling/priorityqueue mv scrapy/trunk to root as part of svn2hg migration 2009-05-06 15:55:17 -03:00
scrapy Added ExecutionQueue class for feeding spiders and requests to scrape. This 2010-05-26 10:29:32 -03:00
scripts removed python2.5 from rpm-install.sh script 2009-06-16 13:14:40 -03:00
.hgignore ignore docs/build 2009-07-25 15:21:22 -03:00
.hgtags Added tag 0.8 for changeset eef0b17d8752 2009-12-12 18:02:42 -02:00
AUTHORS simplified and improved AUTHORS file 2010-02-19 23:16:55 -02:00
INSTALL mv scrapy/trunk to root as part of svn2hg migration 2009-05-06 15:55:17 -03:00
LICENSE mv scrapy/trunk to root as part of svn2hg migration 2009-05-06 15:55:17 -03:00
MANIFEST.in added some missing file to MANIFEST.in 2009-09-28 23:55:00 -03:00
README mv scrapy/trunk to root as part of svn2hg migration 2009-05-06 15:55:17 -03:00
setup.cfg added bitmap for windows installer 2009-09-17 02:01:40 -03:00
setup.py title-cased project name in setup.py 2009-12-12 15:48:02 -02:00

README

This is Scrapy, an opensource screen scraping framework written in Python.

For more visit the project home page at http://scrapy.org