Abstract

Surveys of microbial communities (microbiota), typically measured as relative abundance of species, have illustrated the importance of these communities in human health and disease. Yet, statistical artifacts commonly plague the analysis of relative abundance data. Here, we introduce the PhILR transform, which incorporates microbial evolutionary models with the isometric log-ratio transform to allow off-the-shelf statistical tools to be safely applied to microbiota surveys. We demonstrate that analyses of community-level structure can be applied to PhILR transformed data with performance on benchmarks rivaling or surpassing standard tools. Additionally, By decomposing distance in the PhILR transformed space, we identified neighboring clades that may have adapted to distinct human body sites. Decomposing variance revealed that covariation of bacterial clades within human body sites increases with phylogenetic relatedness. Together, these findings illustrate how the PhILR transform combines statistical and phylogenetic models to overcome compositional data challenges and enable evolutionary insights relevant to microbial communities.

Data availability

The following previously published data sets were used
    1. Human Microbiome Project Consortium
    (2010) Human Microbiome Project
    Publicly available at HMPDACC (v35 download of files 6, 9, and 10).
    1. Costello EK
    2. Lauber CL
    3. Hamady M
    4. Fierer N
    5. Gordon JI
    6. Knight R
    (2009) Costello Skin Sites
    Publicly available as part of the FEMS Benchmark dataset (2011) provided Dan Knights.

Article and author information

Author details

  1. Justin D Silverman

    Program in Computational Biology and Bioinformatics, Duke University, Durham, United States
    Competing interests
    The authors declare that no competing interests exist.
    ORCID icon "This ORCID iD identifies the author of this article:" 0000-0002-3063-2098
  2. Alex D Washburne

    Nicholas School of the Environment, Duke University, Durham, United States
    Competing interests
    The authors declare that no competing interests exist.
  3. Sayan Mukherjee

    Program in Computational Biology and Bioinformatics, Duke University, Durham, United States
    Competing interests
    The authors declare that no competing interests exist.
  4. Lawrence A David

    Program in Computational Biology and Bioinformatics, Duke University, Durham, United States
    For correspondence
    lawrence.david@duke.edu
    Competing interests
    The authors declare that no competing interests exist.
    ORCID icon "This ORCID iD identifies the author of this article:" 0000-0002-3570-4767

Funding

Global Probiotics Council (Young Investigator Grant for Probiotics Research)

  • Lawrence A David

Searle Scholars Program (15-SSP-184 Research Agreement)

  • Lawrence A David

Alfred P. Sloan Foundation (BR2014-003)

  • Lawrence A David

The funders had no role in study design, data collection and interpretation, or the decision to submit the work for publication.

Reviewing Editor

  1. Anthony Fodor, University of North Carolina at Charlotte

Version history

  1. Received: September 27, 2016
  2. Accepted: February 13, 2017
  3. Accepted Manuscript published: February 15, 2017 (version 1)
  4. Version of Record published: February 27, 2017 (version 2)

Copyright

© 2017, Silverman et al.

This article is distributed under the terms of the Creative Commons Attribution License permitting unrestricted use and redistribution provided that the original author and source are credited.

Metrics

  • 11,393
    views
  • 1,717
    downloads
  • 236
    citations

Views, downloads and citations are aggregated across all versions of this paper published by eLife.

Download links

A two-part list of links to download the article, or parts of the article, in various formats.

Downloads (link to download the article as PDF)

Open citations (links to open the citations from this article in various online reference manager services)

Cite this article (links to download the citations from this article in formats compatible with various reference manager tools)

  1. Justin D Silverman
  2. Alex D Washburne
  3. Sayan Mukherjee
  4. Lawrence A David
(2017)
A phylogenetic transform enhances analysis of compositional microbiota data
eLife 6:e21887.
https://doi.org/10.7554/eLife.21887

Share this article

https://doi.org/10.7554/eLife.21887

Further reading

    1. Genetics and Genomics
    2. Neuroscience
    Kenneth Chiou, Noah Snyder-Mackler
    Insight

    Single-cell RNA sequencing reveals the extent to which marmosets carry genetically distinct cells from their siblings.

    1. Genetics and Genomics
    Can Hu, Xue-Ting Zhu ... Jin-Qiu Zhou
    Research Article

    Telomeres, which are chromosomal end structures, play a crucial role in maintaining genome stability and integrity in eukaryotes. In the baker’s yeast Saccharomyces cerevisiae, the X- and Y’-elements are subtelomeric repetitive sequences found in all 32 and 17 telomeres, respectively. While the Y’-elements serve as a backup for telomere functions in cells lacking telomerase, the function of the X-elements remains unclear. This study utilized the S. cerevisiae strain SY12, which has three chromosomes and six telomeres, to investigate the role of X-elements (as well as Y’-elements) in telomere maintenance. Deletion of Y’-elements (SY12), X-elements (SY12XYΔ+Y), or both X- and Y’-elements (SY12XYΔ) did not impact the length of the terminal TG1-3 tracks or telomere silencing. However, inactivation of telomerase in SY12, SY12XYΔ+Y, and SY12XYΔ cells resulted in cellular senescence and the generation of survivors. These survivors either maintained their telomeres through homologous recombination-dependent TG1-3 track elongation or underwent microhomology-mediated intra-chromosomal end-to-end joining. Our findings indicate the non-essential role of subtelomeric X- and Y’-elements in telomere regulation in both telomerase-proficient and telomerase-null cells and suggest that these elements may represent remnants of S. cerevisiae genome evolution. Furthermore, strains with fewer or no subtelomeric elements exhibit more concise telomere structures and offer potential models for future studies in telomere biology.