From: Charlie Bowman Date: 2006-05-26T00:11:19+09:00 Subject: Re: How does one transform UTF-8 encoded characters to ASCII? --=-PYv7os84BCjfRBY81HZc Content-Type: text/plain Content-Transfer-Encoding: 7bit Sorry to just into this thread but I have the exact opposite problem. Open Office uses utf8 and I need to be able to copy text from the OO document into a web form. Is this possible without saving the document as a plain text file first? On Fri, 2006-05-26 at 00:03 +0900, Michal Suchanek wrote: > On 5/24/06, Wes Gamble wrote: > > I'm a little embarrassed about asking this, but here goes... > > > > I am using HTMLEntities.decode_entities on a string of HTML. > > > > I don't understand how to make my text, which now contains UTF-8 > > characters, display correctly in say, Notepad. All of the entities are > > preceded by the character A-circumflex. My guess is that Notepad > > doesn't know how to handle UTF-8, for example. > > At least on Windows XP the notepad can handle various encodings. The > problem is convincing it to use the right encoding. There is no > obvoius way. > One way that might work: > Save a text as utf-8 in notepad. Notepad inserts a mark at the > beginning of the text. You can then copy the mark to the beginning of > any of your texts and it would be readable in notepad. But it would no > longer parse as a valid HTML, ruby, or whatever. > > Or just drop notepad and use a text editor. > > HTH > > Michal Charlie Bowman http://www.recentrambles.com --=-PYv7os84BCjfRBY81HZc--