From: Stanley Xu Date: 2011-03-22T23:27:00+09:00 Subject: How could I make the Ruby 1.9 string ignore the invalid utf-8 byte sequence in split? --bcaec52e60174e8ca2049f13070b Content-Type: text/plain; charset=ISO-8859-1 Dear buddies, I am using ruby to run some map reduce job in hadoop streaming. Unfortunately, we have some dirty data which have invalid byte sequence as the input. So while running things like line.chomp.split("\t") I will get Best wishes, Stanley Xu --bcaec52e60174e8ca2049f13070b--