From: "ilpuccio.febo@..." Date: 2008-07-09T17:01:15+09:00 Subject: Re: implementing a simple and efficient index system On 6 Lug, 20:22, Janus Bor wrote: > Hello everyone, > > I'm pretty new to Ruby and programming in general. Here's my problem: > > I'm writing a program that will automatically download protein sequences > from a server and write them into the corresponding file. Every single > sequence has a unique id and I have to eliminate duplicates. However, as > the number of sequences might exceed 50 000, I can't simply save all > sequences in a hash (with their id as key) and then write them to hd > after downloading has finished. So my idea is to write every sequence to > the corresponding file immediately, but first I have to check if it has > been processed already. You can use BioRuby+BioSQL, fetching data from a remote server and storing into the db.