Movatterモバイル変換


[0]ホーム

URL:


Publisher LogoGenome Research

Skip to main page content

Advanced Search
  • AACR Annual Meeting

MEGAN analysis of metagenomic data

  1. Daniel H. Huson1,3,
  2. Alexander F. Auch1,
  3. Ji Qi2, and
  4. Stephan C. Schuster2,3
  1. 1 Center for Bioinformatics, Tübingen University, Sand 14, 72076 Tübingen, Germany;
  2. 2 Center for Comparative Genomics and Bioinformatics, Center for Infectious Disease Dynamics, Penn State University, University Park, Pennsylvania 16802, USA

Abstract

Metagenomics is the study of the genomic content of a sample of organisms obtained from a common habitat using targeted or random sequencing. Goals include understanding the extent and role of microbial diversity. The taxonomical content of such a sample is usually estimated by comparison against sequence databases of known sequences. Most published studies use the analysis of paired-end reads, complete sequences of environmental fosmid and BAC clones, or environmental assemblies. Emerging sequencing-by-synthesis technologies with very high throughput are paving the way to low-cost random “shotgun” approaches. This paper introduces MEGAN, a new computer program that allows laptop analysis of large metagenomic data sets. In a preprocessing step, the set of DNA sequences is compared against databases of known sequences using BLAST or another comparison tool. MEGAN is then used to compute and explore the taxonomical content of the data set, employing the NCBI taxonomy to summarize and order the results. A simple lowest common ancestor algorithm assigns reads to taxa such that the taxonomical level of the assigned taxon reflects the level of conservation of the sequence. The software allows large data sets to be dissected without the need for assembly or the targeting of specific phylogenetic markers. It provides graphical and statistical output for comparing different data sets. The approach is applied to several data sets, including the Sargasso Sea data set, a recently published metagenomic data set sampled from a mammoth bone, and several complete microbial genomes. Also, simulations that evaluate the performance of the approach for different read lengths are presented.

Footnotes

  • Add to CiteULikeCiteULike
  • Add to DeliciousDelicious
  • Add to DiggDigg
  • Add to FacebookFacebook
  • Add to Google+Google+
  • Add to RedditReddit
  • Add to TwitterTwitter

What's this?

« Previous |Next Article »Table of Contents
OPEN ACCESS ARTICLE

This Article

  1. Published in AdvanceJanuary 25, 2007, doi:10.1101/gr.5969107 Genome Res. 17: 377-386Copyright © 2007, Cold Spring Harbor Laboratory Press
  1. »AbstractFree
  2. Full TextFree
  3. Full Text (PDF)Free
  4. All Versions of this Article:
    1. gr.5969107v1
    2. 17/3/377most recent

Article Category

Services

  1. Alert me when this article is cited
  2. Alert me if a correction is posted
  3. Similar articles in this journal
  4. Similar articles in Web of Science
  5. Article Metrics
  6. Similar articles in PubMed
  7. Download to citation manager
  8. Permissions

Citing Articles

  1. Load citing article information
  2. Citing articles via Web of Science
  3. Citing articles via Google Scholar

Google Scholar

  1. Articles by Huson, D. H.
  2. Articles by Schuster, S. C.
  3. Search for related content

PubMed/NCBI

  1. PubMed citation
  2. Articles by Huson, D. H.
  3. Articles by Schuster, S. C.

Share

    • Add to CiteULikeCiteULike
    • Add to DeliciousDelicious
    • Add to DiggDigg
    • Add to FacebookFacebook
    • Add to Google+Google+
    • Add to RedditReddit
    • Add to TwitterTwitter

    What's this?

Current Issue

  1. March 2025, 35 (3)
  1. Current Issue
  1. Alert me to new issues of Genome Research

Copyright © 2025 by Cold Spring Harbor Laboratory Press

  • Print ISSN:1088-9051
  • Online ISSN:1549-5469

[8]ページ先頭

©2009-2025 Movatter.jp