Problem with Locator2.getEncoding

Elliotte Rusty Harold <[email protected]> Mon, 26 Apr 2004 16:03:02 -0400
Newsgroups gmane.text.xml.sax.devel
Message-ID <p06010209bcb31ae119b0@[192.168.254.88]>
Currently the JavaDoc for Locator2.getEncoding says:

Note that some recent W3C specifications require that text in some 
encodings be normalized, using Unicode Normalization Form C, before 
processing. Such normalization must be performed by applications, and 
would normally be triggered based on the value returned by this 
method.


Is this really accurate? Specifically,

1. Does normalization depend on the encoding?
2. *Must* applications do it?

I think it's a should, not a must. I also think that the parser is 
supposed to do this when converting from a non-Unicode encoding, but 
not when converting from a Unicode encoding.

Unless this can be cleared up a lot, I suggest just deleting this paragraph.
-- 

   Elliotte Rusty Harold
   [email protected]
   Effective XML (Addison-Wesley, 2003)
   http://www.cafeconleche.org/books/effectivexml            
   http://www.amazon.com/exec/obidos/ISBN%3D0321150406/ref%3Dnosim/cafeaulaitA 


-------------------------------------------------------
This SF.net email is sponsored by: The Robotic Monkeys at ThinkGeek
For a limited time only, get FREE Ground shipping on all orders of $35
or more. Hurry up and shop folks, this offer expires April 30th!
http://www.thinkgeek.com/freeshipping/?cpg=12297
_______________________________________________
List: sax-devel, [email protected]
See:  http://www.saxproject.org/
https://lists.sourceforge.net/lists/listinfo/sax-devel