Re: Interpretation of non-UTF8 strings

[email protected] (Dominic Mitchell)
Newsgroups perl.unicode
Organization Semantico
Message-ID <[email protected]>
Marcin 'Qrczak' Kowalczyk wrote:
> This leaves chr() ambiguous, so there should be some other function for
> making Unicode code points, as chr should probably be kept for
> compatibility to mean the default encoding.

In the past when I've needed to guarantee Unicode code points, I've used 
  unpack("U",300).  chr() violates the principle of least astonishment 
(for me anyway) by producing single byte output for input between 0x7f < 
n < 0x100.

-Dom

-- 
| Semantico: creators of major online resources          |
|       URL: http://www.semantico.com/                   |
|       Tel: +44 (1273) 722222 / Fax: +44 (1273) 723232  |
|   Address: 33 Bond St., Brighton, Sussex, BN1 1RD, UK. |
lmpx.com only provides a reader for public news (NNTP) servers. It is not affiliated with the servers or forums shown here and is not responsible for the content of articles, which is written by their respective authors.