From: matz@... (Yukihiro Matsumoto) Date: 2002-03-25T15:10:59+09:00 Subject: Re: Unicode in Regexp followup Hi, In message "Re: Unicode in Regexp followup" on 02/03/25, Yukihiro Matsumoto writes: |By the definition of UTF-8, 0x80-0xEF at the first byte of multibyte |sequence are invalid, so Ruby treats them as if they are single byte |characters. Oops, it should be 0x80-0xBF. matz.