From: "Иван Бишевац" Date: 2012-09-06T18:16:15+09:00 Subject: Re: Parsing through downloaded html --f46d042e0077700b1204c904f197 Content-Type: text/plain; charset=UTF-8 http://nokogiri.org/ is great for this. You need parsing html, look at tutorial on their site: http://nokogiri.org/tutorials/parsing_an_html_xml_document.html 2012/9/6 Sybren Kooistra > Hi all, > > I've collected a number of thousands of .hmtl documents and I need to > know how to parse through all these documents (that are in one folder) > automatically. > > So, I want to copy certain parts of all of these .html documents (for > example the header), but the websites are offline, on my hard disk, in > stead of online. > > What's the way to go? > > -- > Posted via http://www.ruby-forum.com/. > > --f46d042e0077700b1204c904f197 Content-Type: text/html; charset=UTF-8 Content-Transfer-Encoding: quoted-printable http://nokogiri.org/=C2=A0is great for= this. You need parsing html, look at tutorial on their site:=C2=A0http:/= /nokogiri.org/tutorials/parsing_an_html_xml_document.html

2012/9/6 Sybren Kooistra &= lt;lists@ruby-for= um.com>
Hi all,

I've collected a number of thousands of .hmtl documents and I need to know how to parse through all these documents (that are in one folder)
automatically.

So, I want to copy certain parts of all of these .html documents (for
example the header), but the websites are offline, on my hard disk, in
stead of online.

What's the way to go?

--
Posted via http://= www.ruby-forum.com/.


--f46d042e0077700b1204c904f197--