Alignment uncertainty and genomic analysis.
نویسندگان
چکیده
The statistical methods applied to the analysis of genomic data do not account for uncertainty in the sequence alignment. Indeed, the alignment is treated as an observation, and all of the subsequent inferences depend on the alignment being correct. This may not have been too problematic for many phylogenetic studies, in which the gene is carefully chosen for, among other things, ease of alignment. However, in a comparative genomics study, the same statistical methods are applied repeatedly on thousands of genes, many of which will be difficult to align. Using genomic data from seven yeast species, we show that uncertainty in the alignment can lead to several problems, including different alignment methods resulting in different conclusions.
منابع مشابه
IT - Business Strategic Alignment and Organizational Agility: The Moderating Role of Environmental Uncertainty
This study investigates the effect of IT-business strategic alignment on organizational agility by considering the effects of IT flexibility and IT capability on strategic alignment. Also this study investigates the moderating role of environmental uncertainty on the relationship between strategic alignment and organizational agility. This research is an applied research based on purpose and de...
متن کاملUncertainty in homology inferences: assessing and improving genomic sequence alignment.
Sequence alignment underpins all of comparative genomics, yet it remains an incompletely solved problem. In particular, the statistical uncertainty within inferred alignments is often disregarded, while parametric or phylogenetic inferences are considered meaningless without confidence estimates. Here, we report on a theoretical and simulation study of pairwise alignments of genomic DNA at huma...
متن کاملA Novel Pseudo-Alignment Approach to Fast Genomic Sequence Comparison
Standard methods for sequence analysis and phylogeny reconstruction are based on (multiple) sequence alignments. These methods are known to be accurate but if larger genomic sequences are to be analysed they reach their limits. Consequently, faster but less precise alignment-free methods are increasingly used for genomic sequence analysis. In this work, a novel approach to fast genomic sequence...
متن کاملMolecular phylogeny of some avian species using Cytochrome b gene sequence analysis
Veritable identification and differentiation of avian species is a vital step in conservative, taxonomic, forensic, legal and other ornithological interventions. Therefore, this study involved the application of molecular approach to identify some avian species i.e. Chicken (Gallus gallus), Muskovy duck (Cairina moschata), Japanese quail (Coturnix japonica), Laughing dove (Streptopelia senegale...
متن کاملMultiple alignment of genomic sequences using CHAOS, DIALIGN and ABC
Comparative analysis of genomic sequences is a powerful approach to discover functional sites in these sequences. Herein, we present a WWW-based software system for multiple alignment of genomic sequences. We use the local alignment tool CHAOS to rapidly identify chains of pairwise similarities. These similarities are used as anchor points to speed up the DIALIGN multiple-alignment program. Fin...
متن کاملذخیره در منابع من
با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید
عنوان ژورنال:
- Science
دوره 319 5862 شماره
صفحات -
تاریخ انتشار 2008