Skip to main page content
Access keys NCBI Homepage MyNCBI Homepage Main Content Main Navigation
. 2007 Jan;4(1):63-72.
doi: 10.1038/nmeth976. Epub 2006 Dec 10.

Accurate Phylogenetic Classification of Variable-Length DNA Fragments

Affiliations

Accurate Phylogenetic Classification of Variable-Length DNA Fragments

Alice Carolyn McHardy et al. Nat Methods. .

Abstract

Metagenome studies have retrieved vast amounts of sequence data from a variety of environments leading to new discoveries and insights into the uncultured microbial world. Except for very simple communities, the encountered diversity has made fragment assembly and the subsequent analysis a challenging problem. A taxonomic characterization of metagenomic fragments is required for a deeper understanding of shotgun-sequenced microbial communities, but success has mostly been limited to sequences containing phylogenetic marker genes. Here we present PhyloPythia, a composition-based classifier that combines higher-level generic clades from a set of 340 completed genomes with sample-derived population models. Extensive analyses on synthetic and real metagenome data sets showed that PhyloPythia allows the accurate classification of most sequence fragments across all considered taxonomic ranks, even for unknown organisms. The method requires no more than 100 kb of training sequence for the creation of accurate models of sample-specific populations and can assign fragments >or=1 kb with high specificity.

Similar articles

See all similar articles

Cited by 198 articles

See all "Cited by" articles

Publication types

LinkOut - more resources

Feedback