Re: Global and national e-mail address

John C Klensin <[email protected]> Wed, 10 Dec 2003 09:11:39 -0500
Newsgroups gmane.ietf.imaa
Message-ID <[email protected]>
Dan,

Three observations (short this time)...

(i) Computer geeks and their possible preferences aside, people 
tend to not like transliterations (writing of a name that would 
normally be written in one character set in the characters of 
another).  Whether they like "better" transliterations more than 
"worse" transliterations is a cultural issue.

(ii) To reasonably transliterate names and languages, one needs 
not only a collection of the right phonemes, but an appropriate 
and accurate notation for tones.   You can't get those out of a 
small extension to Latin letters.  I'm told that one can get a 
reasonable approximation of all of the relevant phonemes and 
tones with IPA, but IPA not only uses some characters that are 
distinctly non-Latin-based, but also uses a rather complex 
collection of combining diacriticals.  And, at least unless one 
is a professional phonologist, learning IPA and how to use it 
accurately is _hard_ (having had people attempt to teach it to 
me twice, once when I was young enough to learn these things). 
For some hints in a reference that is easily accessible to most 
of us, see the discussion of IPA Characters in the Unicode 
definition (3.0 or 4.0, take your pick).

(iii) If one wants even an approximation to accurate 
transliteration, the symbol-overloading in Latin scripts is bad 
news.  E.g., the sound of "ö" (o with diaresis, U+00F6) is 
different in, e.g., Swedish and German.  I.e., they are 
different characters, even if they look the same and even if 
Unicode "unified" them.  If one is trying to transliterate, 
e.g., Arabic into Roman characters, does one pick a character on 
the basis of the Swedish phoneme or the German ones?  (Hint, as 
soon as you start down that path, you end up sliding toward IPA.)

It appears to me that you are proposing a very Euro-centric view 
of things, and it won't work all that well even for Europe.

regards,
    john


--On Wednesday, 10 December, 2003 10:32 +0100 Dan Oscarsson 
<[email protected]> wrote:

>
> I will here comment several comments frpm John C Klensin, Adan
> M Costello and others.
>
> With this topic I did not want to talk about replacing ASCII
> with ISO 8859-1 or about what character encoding to use.
> Instead I wanted to discuss what names could be suitable to
> use as a global fallback name.
>
> To be able to write a name you need, at least, to be able to
> have letters so you can write all phonemes used. ASCII only
> contains 26 letters and they cannot represent all phonemes
> very well. For example, Swedish have three vowals in addition
> to the ones available in ASCII. They are represented by the
> letters "åäö". These are letters, not an "a" or "o" with an
> accent above. Without those three letters you cannot write all
> Swedish names. Accents I can live without, but not the letters
> for our additional phonemes.
>
> To be able to write most names in the world I think you need
> to be able to write about 60 phonems. Not everybody need their
> own letter (English have about 45 phonems but the 26 letters
> are enough). So I would expect by adding not that many more
> letters to the ones in ASCII we could get a quite faire
> representation of all names in the world. That would be more
> acceptable  to have on a business card.
>
> And just like Adam my Swedish keyboard do not have any
> accented or diacritic letters. But I can still type quite a
> lot of them by using "compose" or "alt graph".
>
> ASCII will never be good enough for use as a "global" name for
> Swedish. But with a few more letters added it would be
> possible. I expect the same would work for most other
> languages. I would not be surprised if 35-45 letters would be
> enough to give quite good representation of the native names.
>
>    Dan
>
>