From: Nick Snels Date: 2006-01-31T20:08:05+09:00 Subject: Re: Premature end of regular expression with non-ascii chara Hi Axel, thanks for the reply. If I try your code, my characters with accents don't get translated to numbers, unfortunately. Do you know where these numbers come from, I looked on the net but \352 is not the octal, hexadecimal or UTF-8 representation of 棚 . Could you split the following sentence for me and let me know what the result is: a="Ils sont tr竪s 辿nerv辿 les regexps." splitted_text=a.split(/\s/) Not my best French. But if I try this, 'tr竪s 辿nerv辿 les' is still one part, eventhough I split it on the spaces. Maybe it is different with you and then I have to look deeper. Thanks for your help. If anybody is able to split is like 'tr竪s', '辿nerv辿', 'les' please let me know!! Kind regards, Nick -- Posted via http://www.ruby-forum.com/.