Unknown

Dataset Information

0

Coda4microbiome: compositional data analysis for microbiome cross-sectional and longitudinal studies.


ABSTRACT:

Background

One of the main challenges of microbiome analysis is its compositional nature that if ignored can lead to spurious results. Addressing the compositional structure of microbiome data is particularly critical in longitudinal studies where abundances measured at different times can correspond to different sub-compositions.

Results

We developed coda4microbiome, a new R package for analyzing microbiome data within the Compositional Data Analysis (CoDA) framework in both, cross-sectional and longitudinal studies. The aim of coda4microbiome is prediction, more specifically, the method is designed to identify a model (microbial signature) containing the minimum number of features with the maximum predictive power. The algorithm relies on the analysis of log-ratios between pairs of components and variable selection is addressed through penalized regression on the "all-pairs log-ratio model", the model containing all possible pairwise log-ratios. For longitudinal data, the algorithm infers dynamic microbial signatures by performing penalized regression over the summary of the log-ratio trajectories (the area under these trajectories). In both, cross-sectional and longitudinal studies, the inferred microbial signature is expressed as the (weighted) balance between two groups of taxa, those that contribute positively to the microbial signature and those that contribute negatively. The package provides several graphical representations that facilitate the interpretation of the analysis and the identified microbial signatures. We illustrate the new method with data from a Crohn's disease study (cross-sectional data) and on the developing microbiome of infants (longitudinal data).

Conclusions

coda4microbiome is a new algorithm for identification of microbial signatures in both, cross-sectional and longitudinal studies. The algorithm is implemented as an R package that is available at CRAN ( https://cran.r-project.org/web/packages/coda4microbiome/ ) and is accompanied with a vignette with a detailed description of the functions. The website of the project contains several tutorials: https://malucalle.github.io/coda4microbiome/.

SUBMITTER: Calle ML 

PROVIDER: S-EPMC9990256 | biostudies-literature | 2023 Mar

REPOSITORIES: biostudies-literature

altmetric image

Publications

coda4microbiome: compositional data analysis for microbiome cross-sectional and longitudinal studies.

Calle M Luz ML   Pujolassos Meritxell M   Susin Antoni A  

BMC bioinformatics 20230306 1


<h4>Background</h4>One of the main challenges of microbiome analysis is its compositional nature that if ignored can lead to spurious results. Addressing the compositional structure of microbiome data is particularly critical in longitudinal studies where abundances measured at different times can correspond to different sub-compositions.<h4>Results</h4>We developed coda4microbiome, a new R package for analyzing microbiome data within the Compositional Data Analysis (CoDA) framework in both, cro  ...[more]

Similar Datasets

| S-EPMC7671404 | biostudies-literature
| S-EPMC6697339 | biostudies-literature
| S-EPMC8199728 | biostudies-literature
| S-EPMC11066923 | biostudies-literature
| S-EPMC9012043 | biostudies-literature
| S-EPMC7031076 | biostudies-literature
| S-EPMC9294433 | biostudies-literature
| S-EPMC7410344 | biostudies-literature
| S-EPMC7768662 | biostudies-literature
| S-EPMC4549296 | biostudies-literature