From: Arun Kumar Date: 2009-08-23T05:10:28+09:00 Subject: Re: Parsing pdf files --000e0cd35526f97fb50471c07ff7 Content-Type: text/plain; charset=UTF-8 Content-Transfer-Encoding: quoted-printable That's really very sad :( On Sat, Aug 22, 2009 at 10:33 PM, Gregory Brown wrote: > On Sat, Aug 22, 2009 at 12:33 PM, Arun Kumar > wrote: > > hello all, > > Does anyone know a good pdf parser that retains > formatting > > after its extracted text? I used PDF::Reader, but when you extract text > you > > just get a stream of characters that are not at all intelligible. When = I > > copy a pdf contents from a pdf reader to Gedit text editor in linux it > > retains its format. I'm looking for something like that. > > This doesn't exist in Ruby, unfortunately. > > -greg > > --=20 || =E0=A4=B6=E0=A5=8D=E0=A4=B0=E0=A5=80 =E0=A4=9C=E0=A4=BE=E0=A4=A8=E0=A4= =95=E0=A5=80=E0=A4=B0=E0=A4=98=E0=A5=81=E0=A4=A8=E0=A4=BE=E0=A4=A5=E0=A5=8B= =E0=A4=B5=E0=A4=BF=E0=A4=9C=E0=A4=AF=E0=A4=A4=E0=A5=87 || --000e0cd35526f97fb50471c07ff7--