PubMed HealthSearch

Biomedical subjects

F Eggenberger

Publications and source records attributed to F Eggenberger.

3 recordsLinked to original sources

FastAlert--an automatic search system to alert about new entries in biological sequence databanks.

This paper describes a new tool enabling awareness of new sequence databank entries of interest. The FastAlert system relieves the researcher from the burden of repeating FASTA searches in order to keep up with the rapidly growing amount of information found in biological sequence databanks. The query sequence can be submitted from any computer connected to the Internet. Upon registration, the databank, including the updates, is scanned at periodic intervals with the sequence provided. The results, so-called FastAlert reports, are delivered via electronic mail. The reports contain the FASTA best-scores list and the similarity statistics for each entry listed.

Amino Acid Sequence

List update processing (LUP)--solving the sequence database update problem.

Sequence databases of today require frequent updating. Mirror procedures to copy incrementally updated databases as cumulative sets are the preferred method and can be implemented by straightforward scripting. However, limited bandwidth of networks and the increase of data require more powerful paradigms to reduce the workload reliably. We suggest the List Update Processing (LUP) principle. The system has been implemented on an experimental basis to update the Swiss EMBnet Node (BioComputing Basel, CH) with data from the European Bioinformatics Institute (EMBL Outstation, Hinxton Hall, UK). The results obtained from the prototype suggest to expand the system to several sites.

CD-ROM

A compression mechanism for sequence databases to improve the efficiency of conventional tools.

This paper describes a method to compress molecular biology databases that are characterized by an increasing proportion of data derived from genome projects. The performance of our tool has been tested on various data files of the EMBL nucleotide sequence database. The best compression ratios were achieved on EST (Expressed Sequence Tags) data, typically derived from large-scale sequence projects. The compression of sequence database updates was tested in combination with the common Unix compression program 'compress'. Our tool improved the efficiency of 'compress' on average by 16%.

Base Sequence