Re: sed error message reports byte position instead of char position when program contains UTF-8

John Cowan <[email protected]>
Newsgroups gmane.comp.gnu.utils.bugs
Message-ID <[email protected]>
Eli Zaretskii scripsit:

> How do you expect Sed to know what character set is being used for the
> command line?  Are we again going to limit ourselves to the current
> locale's charset?

I think it is a reasonable assumption that the command line uses the
same encoding that is used for the input files.  Nothing else will
make common cases like "sed 's/föö*/bär/'" work correctly.

-- 
John Cowan    [email protected]    http://ccil.org/~cowan
Objective consideration of contemporary phenomena compel the conclusion
that optimum or inadequate performance in the trend of competitive
activities exhibits no tendency to be commensurate with innate capacity,
but that a considerable element of the unpredictable must invariably be
taken into account. --Ecclesiastes 9:11, Orwell/Brown version
lmpx.com only provides a reader for public news (NNTP) servers. It is not affiliated with the servers or forums shown here and is not responsible for the content of articles, which is written by their respective authors.