HTML Parser problem
"Shashank Kavishwar" <[email protected]> Wed, 9 Mar 2005 12:10:04 -0800
| Newsgroups | gmane.comp.lib.libwww |
|---|---|
| Message-ID | <[email protected]> |
I have a problem with parsing HTML data. If the HTML data contains a '<' with no ending '>' the rest of the HTML is not parsed. Eg: <html> <body> <p>This is a test paragraph </p> Normal text here <<<<<<< John Doe e-mail: [email protected] web site: www.johndoe.com <http://www.johndoe.com/> </body> </html> When I parse this HTML, using the HTMLToPlain() call, I only get until '. text here'. Everything after the '<<<<<<' is skipped. Am I missing something? Thanks, Shashank FrontBridge introduces Message Archive and Secure Email. Get leading Enterprise Message Security services from FrontBridge. www.frontbridge.com.