From: John Joyce Date: 2008-07-10T10:25:27+09:00 Subject: Re: String#upcase/downcase with UTF-8 strings in Ruby 1.9 On Jul 9, 2008, at 8:17 PM, Stefan Schmidt wrote: >> The document for String#upcase says: > > Yes, sorry, I should have read the documentation > >> See "Note:". Tim Bray have persuaded me to do so, since case >> conversion outside of ASCII region is highly dependent on country, >> language, culture and script. > > So basically the Python guys are going down a wrong route ? > > # -*- coding: utf-8 -*- > import string > print string.upper(u"aoueäöüé") > print string.lower(u"AOUEÄÖÜÉ") > > works as expected. > > Cheers, Stefan > No. They're going down a different route. Seriously, the language handling is something that could easily be handled by extensions. It does not need to be a core part of the language. Even operating systems handle these things with proprietary and very sophisticated techniques based on the language in question. In most cases, what you are expecting to be the correct upper case characters may be 'correct' but it will ultimately depend on the language and the context.