| Title: | Fast and sensitive multiple alignment of large genomic sequences |
| Authors: | Brudno, Michael Chapman, Michael A Gottgens, Berthold Batzoglou, Serafim Morgenstern, Burkhard |
| Issue Date: | 23-Dec-2003 |
| Citation: | BMC Bioinformatics 2003, 4:66 |
| Abstract: | Abstract Background Genomic sequence alignment is a powerful method for genome analysis and annotation, as alignments are routinely used to identify functional sites such as genes or regulatory elements. With a growing number of partially or completely sequenced genomes, multiple alignment is playing an increasingly important role in these studies. In recent years, various tools for pair-wise and multiple genomic alignment have been proposed. Some of them are extremely fast, but often efficiency is achieved at the expense of sensitivity. One way of combining speed and sensitivity is to use an anchored-alignment approach. In a first step, a fast search program identifies a chain of strong local sequence similarities. In a second step, regions between these anchor points are aligned using a slower but more accurate method. Results Herein, we present CHAOS, a novel algorithm for rapid identification of chains of local pair-wise sequence similarities. Local alignments calculated by CHAOS are used as anchor points to improve the running time of DIALIGN, a slow but sensitive multiple-alignment tool. We show that this way, the running time of DIALIGN can be reduced by more than 95% for BAC-sized and longer sequences, without affecting the quality of the resulting alignments. We apply our approach to a set of five genomic sequences around the stem-cell-leukemia (SCL) gene and demonstrate that exons and small regulatory elements can be identified by our multiple-alignment procedure. Conclusion We conclude that the novel CHAOS local alignment tool is an effective way to significantly speed up global alignment tools such as DIALIGN without reducing the alignment quality. We likewise demonstrate that the DIALIGN/CHAOS combination is able to accurately align short regulatory sequences in distant orthologues. |
| Description: | RIGHTS : This article is licensed under the BioMed Central licence at http://www.biomedcentral.com/about/license which is similar to the 'Creative Commons Attribution Licence'. In brief you may : copy, distribute, and display the work; make derivative works; or make commercial use of the work - under the following conditions: the original author must be given credit; for any reuse or distribution, it must be made clear to others what the license terms of this work are. |
| URI: | http://www.dspace.cam.ac.uk/handle/1810/238125 http://dx.doi.org/10.1186/1471-2105-4-66 |
| Appears in Collections: | Scholarly works - Haematology |
Files in This Item:
|
| Additional resources for this item |
|---|
| search for alternative versions in eresources@cambridge |
| retrieve citation metadata in EndNote format |
This item has been accessed 336 times.
Items in DSpace are protected by copyright, with all rights reserved, unless otherwise indicated.

