Re: get the sourcecode [of UTF-8]
A bughunter via Unicode <[email protected]>
| Newsgroups | gmane.text.unicode.general |
|---|---|
| Message-ID | <VbTP9pBTBW_BbwGR6JrCzg1ChpXVOxjwi8YMexRmTqARaY89zIFMfh7jW8C1-vPbIClez2pQtukEWJ4P75v4sUFvf0EPzMlQLZbbe81kTYM=@proton.me> |
-----BEGIN PGP SIGNED MESSAGE----- Hash: SHA512 My reply to Steven interspersed. Please do not reply further unless you intend to answer my Originating concise yet full though simple one line relevent ontopic Question. Where to get the sourcecode of relevent (version) UTF-8?: in order to checksum text against the specific encoding map (codepage). from [email protected] Sent with Proton Mail secure email. On Friday, November 8th, 2024 at 06:20, Steven R. Loomis <[email protected]> wrote: > > On Thu, Nov 7, 2024 at 10:35 PM A bughunter via Unicode <[email protected]> wrote: > > > @Julian says "There are no codepages in Unicode. (Or I suppose there is exactly > > one.)" ( https://corp.unicode.org/pipermail/unicode/2024-November/011116.html ) > > > > does conflict with > > > > @Otto says "UTF-8 is one method (of a handfull of standardized methods) to represent Unicode text at the bit level in order to conveniently transfer, or store, it." ( https://corp.unicode.org/pipermail/unicode/2024-November/011132.html ) > > > There's no conflict here. UTF-8 is a form of Unicode, it's not a codepage. > > > @Jim says he is putting up a post like a seeing eye dog by pasting from my GitHub ( https://www.github.com/freedom-foundation ) "I summarised what I understand of your project as a courtesy to my fellow unicode-list subscribers. " > > > It's helpful to understand what the purpose of your project is. You could checksum across UTF-8 code units, or UTF-32 code units, or UTF-16 code units, or 21-bit scalar values. It's up to you. > > > I chuse this example because it goes to show that Unicode consortium is disappointed. You will probably read in there somewhere that it intended to solve a problem of many codepages hower you see that it has become something more complicated than the problem it were to simplify to solve. Steven no, you have here listed 4 codepages which are all uncompatible there is no one codepage like Julian claimed. > History does not bear out your claim. In fact, Unicode has solved exactly what it set out to accomplish. The problem of many codepages now exists mostly due to older implementations and data, plus (very rarely) due to new implementations that choose a difficult path. > > -s -----BEGIN PGP SIGNATURE----- Version: ProtonMail wnUEARYKACcFgmctzXAJkKkWZTlQrvKZFiEEZlQIBcAycZ2lO9z2qRZlOVCu 8pkAAE8GAQDew0BFa3W22qr7iDGyqduXp6FqeQbff6WOl74t4v91bQD9EoXy 1E6kMS6E6Ouq7uy82sFm1EGgw0pO/BuRXcR0VgA= =XJs0 -----END PGP SIGNATURE-----
publickey - [email protected] - 0x66540805.asc
(application/pgp-keys, 653 B)
-----BEGIN PGP PUBLIC KEY BLOCK----- xjMEZu0X1xYJKwYBBAHaRw8BAQdAH0I47jDsPZ6gvb+YUGBnAx7Jyf14AVOH xa8y0+dmN5bNLUFfYnVnaHVudGVyQHByb3Rvbi5tZSA8QV9idWdodW50ZXJA cHJvdG9uLm1lPsKMBBAWCgA+BYJm7RfXBAsJBwgJkKkWZTlQrvKZAxUICgQW AAIBAhkBApsDAh4BFiEEZlQIBcAycZ2lO9z2qRZlOVCu8pkAAD9FAP9/ddT6 56Gka9NtMvmdoY5ktNgqbY5Xbd9fx6kPE5/4tQD/XiialKQHjmwAtbcSe1Q+ 3cxYLxNhjU7mynQspv9dxADOOARm7RfXEgorBgEEAZdVAQUBAQdAnfp/z2Fw RkpvUgf7mqYI9RKnTVadwGfgaQLhmwg3LxMDAQgHwngEGBYKACoFgmbtF9cJ kKkWZTlQrvKZApsMFiEEZlQIBcAycZ2lO9z2qRZlOVCu8pkAAJi8AQC+fnOm 4Vj9QmH4H0GVt7RuOQK+wOQ1PRvpymSjeyBJOwD9GYuvxOAVK8iAupJ+ppwM r36VukIe1pXuHo9RhjveAw0= =FQFw -----END PGP PUBLIC KEY BLOCK-----
publickey - [email protected] - 0x66540805.asc.sig
(application/pgp-signature, 119 B) - not displayed