c/3804: Extended ASCII "wide" characters not behaving with UTF-8 locale
Neil Booth
neil@daikokuya.demon.co.uk
Wed Jul 25 14:07:00 GMT 2001
Alex Eulenberg wrote:-
> Second item: According to the Sun compiler documentation
> http://docs.sun.com/htmlcoll/coll.33.7/iso-8859-1/CUG/tguide.html#785 , for
> all ANSI/ISO Compilers, "When the compilation system encounters a wide
> character constant or wide string literal, each multibyte character is
> converted into a wide character, as if by calling the mbtowc() function."
and the sentence should finish "with an implementation-defined current
locale". We don't want to restrict ourselves to a single locale for a
translation unit, since a file can include headers where a different
locale might be appropriate.
We have not decided exactly what we will do, but it is looking likely
that we will assume each source file is UTF8 unless otherwise
indicated. "Otherwise indicated" might be with a magic comment at the
top of the file, like Emacs.
Neil.
More information about the Gcc-bugs
mailing list