Re: Extracting web data
Lennart Regebro <[email protected]>
| Newsgroups | gmane.comp.python.web |
|---|---|
| Message-ID | <[email protected]> |
On Tue, Feb 22, 2011 at 01:52, Aaron Watters <arw1961-/[email protected]> wrote: > BeautifulSoup is the standard response. > I think lxml will not work very well unless the > html is extremely nicely formatted, but I could > be wrong. > lxml handles broken HTML pretty well. Tere are Windows binaries here: http://pypi.python.org/pypi/lxml/2.2.8 //Lennart _______________________________________________ Web-SIG mailing list [email protected] Web SIG: http://www.python.org/sigs/web-sig Unsubscribe: http://mail.python.org/mailman/options/web-sig/gcpw-web-sig%40m.gmane.org