Loading…

Ribosomal RNA as molecular barcodes: a simple correlation analysis without sequence alignment

Motivation: We explored the feasibility of using unaligned rRNA gene sequences as DNA barcodes, based on correlation analysis of composition vectors (CVs) derived from nucleotide strings. We tested this method with seven rRNA (including 12, 16, 18, 26 and 28S) datasets from a wide variety of organis...

Full description

Saved in:
Bibliographic Details
Published in:Bioinformatics 2006-07, Vol.22 (14), p.1690-1701
Main Authors: Chu, K. H., Li, C. P., Qi, J.
Format: Article
Language:English
Subjects:
Citations: Items that this one cites
Items that cite this one
Online Access:Request full text
Tags: Add Tag
No Tags, Be the first to tag this record!
Description
Summary:Motivation: We explored the feasibility of using unaligned rRNA gene sequences as DNA barcodes, based on correlation analysis of composition vectors (CVs) derived from nucleotide strings. We tested this method with seven rRNA (including 12, 16, 18, 26 and 28S) datasets from a wide variety of organisms (from archaea to tetrapods) at taxonomic levels ranging from class to species. Result: Our results indicate that grouping of taxa based on CV analysis is always in good agreement with the phylogenetic trees generated by traditional approaches, although in some cases the relationships among the higher systemic groups may differ. The effectiveness of our analysis might be related to the length and divergence among sequences in a dataset. Nevertheless, the correct grouping of sequences and accurate assignment of unknown taxa make our analysis a reliable and convenient approach in analyzing unaligned sequence datasets of various rRNAs for barcoding purposes. Availability: The newly designed software (CVTree 1.0) is publicly available at the Composition Vector Tree (CVTree) web server Contact:kahouchu@cuhk.edu.hk
ISSN:1367-4803
1460-2059
1367-4811
DOI:10.1093/bioinformatics/btl146