Re: List of pages to fix for HTML parsing errors

Laurence Rowe <[email protected]>
Newsgroups gmane.comp.web.zope.plone.documentation,gmane.comp.web.zope.plone.devel
Message-ID <[email protected]>
2009/5/17 Israel Saeta Pérez <[email protected]>:
> I've fixed a bunch of them related to documentation and striked them out in
> the spreadsheet. As you've said, most of them come from stx documents, where
> spurious <p> elements are inserted, even inside literal blocks. There are
> also some problems with URL encoding (with composite query strings) and
> entities encoding (&amp; &lt; &gt;).
>
> By the way, do we really need such a strict parsing that makes the page
> rendering blow up whenever the XHTML is not perfect? Can't the libxml2
> parser be 'patched' to accept valid enough XHTML?

The HTMLParser does try to be tolerant, but in SAX mode it seems to
break down when the errors are in an open tag near a chunk boundary
(zope serves data in 1024 byte chunks). If you have the C and SAX foo
to fix this, that would be great!

Laurence

------------------------------------------------------------------------------
Crystal Reports - New Free Runtime and 30 Day Trial
Check out the new simplified licensing option that enables 
unlimited royalty-free distribution of the report engine 
for externally facing server and web deployment. 
http://p.sf.net/sfu/businessobjects
lmpx.com only provides a reader for public news (NNTP) servers. It is not affiliated with the servers or forums shown here and is not responsible for the content of articles, which is written by their respective authors.