scrapy/scrapy/trunk/docs/faq.rst

1.2 KiB

<html xmlns="http://www.w3.org/1999/xhtml" xml:lang="en" lang="en"> <head> </head>

Frequently Asked Questions

How does Scrapy compare to BeatifulSoul or lxml?

BeautifulSoup and lxml are libraries for parsing HTML and XML. Scrapy is an application framework for writing web spiders that crawl web sites and extract data from it. Scrapy provides some mechanisms for extracting data (called selectors) but you can easily use BeautifulSoup or lxml if you feel more comfortable with them. After all, they're just parsing libraries which can be imported and used from any Python code.

To illustrate from another point of view, comparing BeautifulSoup or lxml to Scrapy is like comparing urllib or urlparse to Django (a popular Python web framework).

Does Scrapy work in Python 3.0?

No, and there are no ongoing plans to port Scrapy to Python 3.0 yet. At the moment Scrapy requires Python 2.5 or 2.6 to work.

</html>