Re: get the sourcecode [of UTF-8]

A bughunter via Unicode <[email protected]>
Newsgroups gmane.text.unicode.general
Message-ID <VbTP9pBTBW_BbwGR6JrCzg1ChpXVOxjwi8YMexRmTqARaY89zIFMfh7jW8C1-vPbIClez2pQtukEWJ4P75v4sUFvf0EPzMlQLZbbe81kTYM=@proton.me>
-----BEGIN PGP SIGNED MESSAGE-----
Hash: SHA512

My reply to Steven interspersed.
Please do not reply further unless you intend to answer my Originating concise yet full though simple one line relevent ontopic Question.

Where to get the sourcecode of relevent (version) UTF-8?: in order to checksum text against the specific encoding map (codepage).

from [email protected]

Sent with Proton Mail secure email.

On Friday, November 8th, 2024 at 06:20, Steven R. Loomis <[email protected]> wrote:

> 
> On Thu, Nov 7, 2024 at 10:35 PM A bughunter via Unicode <[email protected]> wrote:
> 
> > @Julian says "There are no codepages in Unicode. (Or I suppose there is exactly
> > one.)" ( https://corp.unicode.org/pipermail/unicode/2024-November/011116.html )
> > 
> > does conflict with
> > 
> > @Otto says "UTF-8 is one method (of a handfull of standardized methods) to represent Unicode text at the bit level in order to conveniently transfer, or store, it." ( https://corp.unicode.org/pipermail/unicode/2024-November/011132.html )
> 
> 
> There's no conflict here. UTF-8 is a form of Unicode, it's not a codepage.
> 
> > @Jim says he is putting up a post like a seeing eye dog by pasting from my GitHub ( https://www.github.com/freedom-foundation ) "I summarised what I understand of your project as a courtesy to my fellow unicode-list subscribers. "
> 
> 
> It's helpful to understand what the purpose of your project is. You could checksum across UTF-8 code units, or UTF-32 code units, or UTF-16 code units, or 21-bit scalar values. It's up to you.
> 
> > I chuse this example because it goes to show that Unicode consortium is disappointed. You will probably read in there somewhere that it intended to solve a problem of many codepages hower you see that it has become something more complicated than the problem it were to simplify to solve.

Steven no, you have here listed 4 codepages which are all uncompatible there is no one codepage like Julian claimed.
> History does not bear out your claim. In fact, Unicode has solved exactly what it set out to accomplish. The problem of many codepages now exists mostly due to older implementations and data, plus (very rarely) due to new implementations that choose a difficult path.
> 
> -s
-----BEGIN PGP SIGNATURE-----
Version: ProtonMail

wnUEARYKACcFgmctzXAJkKkWZTlQrvKZFiEEZlQIBcAycZ2lO9z2qRZlOVCu
8pkAAE8GAQDew0BFa3W22qr7iDGyqduXp6FqeQbff6WOl74t4v91bQD9EoXy
1E6kMS6E6Ouq7uy82sFm1EGgw0pO/BuRXcR0VgA=
=XJs0
-----END PGP SIGNATURE-----
publickey - [email protected] - 0x66540805.asc (application/pgp-keys, 653 B)
-----BEGIN PGP PUBLIC KEY BLOCK-----

xjMEZu0X1xYJKwYBBAHaRw8BAQdAH0I47jDsPZ6gvb+YUGBnAx7Jyf14AVOH
xa8y0+dmN5bNLUFfYnVnaHVudGVyQHByb3Rvbi5tZSA8QV9idWdodW50ZXJA
cHJvdG9uLm1lPsKMBBAWCgA+BYJm7RfXBAsJBwgJkKkWZTlQrvKZAxUICgQW
AAIBAhkBApsDAh4BFiEEZlQIBcAycZ2lO9z2qRZlOVCu8pkAAD9FAP9/ddT6
56Gka9NtMvmdoY5ktNgqbY5Xbd9fx6kPE5/4tQD/XiialKQHjmwAtbcSe1Q+
3cxYLxNhjU7mynQspv9dxADOOARm7RfXEgorBgEEAZdVAQUBAQdAnfp/z2Fw
RkpvUgf7mqYI9RKnTVadwGfgaQLhmwg3LxMDAQgHwngEGBYKACoFgmbtF9cJ
kKkWZTlQrvKZApsMFiEEZlQIBcAycZ2lO9z2qRZlOVCu8pkAAJi8AQC+fnOm
4Vj9QmH4H0GVt7RuOQK+wOQ1PRvpymSjeyBJOwD9GYuvxOAVK8iAupJ+ppwM
r36VukIe1pXuHo9RhjveAw0=
=FQFw
-----END PGP PUBLIC KEY BLOCK-----
publickey - [email protected] - 0x66540805.asc.sig (application/pgp-signature, 119 B) - not displayed
lmpx.com only provides a reader for public news (NNTP) servers. It is not affiliated with the servers or forums shown here and is not responsible for the content of articles, which is written by their respective authors.