[Fwd]: [indic] pD7A..pD7E Malayalam chillu

"Mahesh T. Pai" <[email protected]> Fri, 27 Aug 2004 01:09:22 +0530
Newsgroups gmane.org.region.india.gnu-malayalam
Message-ID <[email protected]>
Comments welcome.

----- Forwarded message from Eric Muller <emuller-dv/[email protected]> -----

 > Date: Thu, 26 Aug 2004 10:48:05 -0700
 > From: Eric Muller <emuller-dv/[email protected]>
 > Subject: [indic] pD7A..pD7E Malayalam chillu
 > To: indic <[email protected]>
 > User-Agent: Mozilla/5.0 (Windows; U; Windows NT 5.1; en-US; rv:1.7)
 >  Gecko/20040616
 > Message-id: <412E2255.3020601-dv/[email protected]>
 > Organization: Adobe Systems Incorporated
 > 
 > pD7A MALAYALAM LETTER NN
 > pD7B MALAYALAM LETTER N
 > pD7C MALAYALAM LETTER RR
 > pD7D MALAYALAM LETTER L
 > pD7E MALAYALAM LETTER LL
 > 
 > From previous discussions on this list, I am almost convinced that 
 > those need to be encoded as characters (and there is even one more, not 
 > in modern used, based on KA). What is needed is a proposal that gives 
 > evidence of the constrast between, e.g., NA + visible virama and chillu N.
 > 
 > There is no question that only one of the two renderings is acceptable 
 > in any given circumstance, and that the proper rendering cannot be 
 > determined from the context; in other words, that there must be some 
 > difference in the representation of text. What is less clear is whether 
 > that difference of representation should be achieved by encoding new 
 > characters, or by using the current mechanism of following the virama by 
 > ZWJ (to get the chillu) or ZWNJ (to get the visible virama). [To be 
 > sure, the joiner mechanism is currently underspecified, since the 
 > rendering of a sequence that does not include a joiner is not defined; 
 > but let's set that aside for now.]
 > 
 > My understanding is that there are pairs of words, which differ only in 
 > having, e.g., a NA + virama or a chillu N. To take an example from 
 > English "last" and "lost" differ in using "a" or "o", and even if "o" 
 > was used only in this word, this would be a good enough reason to 
 > encoded an "o". What we need are actual examples of this situation, 
 > preferably a couple of examples for each of the six characters. I'll be 
 > happy to turn those examples into a proposal in the form preferred by 
 > the UTC and ISO (to be reviewed here, of course).
 > 
 > 
 > Eric.
 > 
 > 

----- End forwarded message -----

-- 
         Mahesh T. Pai    <<>>   http://paivakil.port5.com