From: 7stud -- Date: 2012-09-05T10:40:55+09:00 Subject: Re: Parsing Newb Help I'm not at all clear what the *specific* things are that you want to extract from the website. In any case, you need to click on View/Source in your browser and examine the raw html to figure out what tags you need to extract and how to identify them. Look at the web page in your browser then use Find or Search to locate the same text in the raw html. Then read some basic xpath tutorials starting here: http://www.engineyard.com/blog/2010/getting-started-with-nokogiri/ Here is an example of how to get the names of the restaurants: require 'nokogiri' #require 'open-uri' #doc = Nokogiri::HTML(open("http://www.threescompany.com/")) html =< Stuff

Fishermen's Grotto

blah blah blah

Marnee Thai Restaurant

MY_HTML doc = Nokogiri::HTML(html) doc.xpath('//h3[@class="title fn org"]/a[1]').each do |node| puts node.text end --output:-- Fishermen's Grotto Marnee Thai Restaurant Parsing html requires a good understanding of html structure, e.g. parents, children, siblings, etc., and css, e.g. classes, ids, etc. As a beginner it is better to take baby steps, not jump in the deep end of the pool, so this project may be too hard for you. -- Posted via http://www.ruby-forum.com/.