CORTEXA
← Browse
openalexFrontiers in Microbiology2026-07-23Cited by 0

Omics data in relative values are almost subcompositionally coherent

Marina Martínez-Álvaro, Michael Greenacre, A. Blasco

Introduction Omics data are compositional and often expressed as relative abundances after total sum scaling normalization. An important statistical issue with compositional data is the lack of subcompositional coherence, meaning that relative abundances change when data are re-normalized after removing or adding features. While this problem is well documented for small compositions, it has not been investigated in large Omics datasets, which typically contain hundreds or thousands of features and where subcompositions are ubiquitous. Subcompositions arise, for example, when using different reference datasets, sequencing depths or when filtering low-abundant features from the database. In such cases, the most abundant features are preferentially retained, whereas variation between original or full compositions and subcompositions is mainly driven by less abundant features. The standard solution to this problem is the use of logratio transformations, but these complicate interpretations and require handling zeros, which are frequent in Omics data and whose imputation introduces spurious variability. Methods Here, we evaluated subcompositional coherence in five representative Omics datasets: fecal 16S metagenomics, rumen metagenomics (taxonomic and functional levels), liver transcriptomics, and plasma metabolomics, considering both unsupervised and supervised learning contexts. We generated 100 random subcompositions comprising one-third of the original features under an abundance-weighted subcomposition scheme and compared their statistical outputs with those from the full composition. Results and discussion Raw Omics data showed near-perfect coherence: relative abundances, pairwise correlations and sample distances all exhibited very high (scaled) concordances (≥0.98–0.99). Outputs from commonly used supervised models (linear regression, PLS, random forest, and linear mixed models with a Gaussian kernel) were also highly subcompositionally coherent. We conclude that large Omics datasets expressed as relative abundances are almost subcompositionally coherent when considering a weighted subcomposition scheme, thereby challenging one of the criticisms of using relative data in the Omics field over logratio transformations.

View free PDFSource page

Related papers

openalexFrontiers in Microbiology2026-07-24

Flavor compounds regulation in rice-flavor Baijiu by Bacillus velezensis: a genetic and functional analysis

Hong Ren, Ren Huang, Y Q Xie, Qinghua Chen, Feng Chen, Runhua Huang, et al.

To achieve targeted enhancement of the flavor complexity of rice-flavor Baijiu, this study evaluated the impact of fortified fermentation with four saccharification-promoting Bacillus velezensis strains on the volatile profile of fermented grains (Jiupei), and further dissected t…

View free PDFSource page
openalexFrontiers in Microbiology2026-07-24

Geographic origin and wheat variety shape microbial and functional profiles of brewing wheat for Daqu fermentation

Hao Tang, Shouhang Zheng, L B Yao, Qin Wang, Ning Yang, Na Luo, et al.

Introduction Wheat is the primary raw material for traditional Baijiu Daqu fermentation, yet its role as a carrier of functional microbiota and its contribution to Daqu quality remain poorly understood. Methods A total of 135 wheat samples representing five geographic regions and…

View free PDFSource page
openalexFrontiers in Microbiology2026-07-24

Detection and colocalization of STEC genetic markers in wheat flour using whole cell digital droplet PCR

Sophie Butot, C Gaille, Lise Michot, Mathieu Seppey, Caroline Barretto, Solenn Pruvost, et al.

Introduction Shiga toxin-producing Escherichia coli (STEC) poses a significant food safety risk in wheat-based products. The detection of these pathogens in complex food matrices is challenging due to the labor-intensive nature of conventional methodologies and the variability in…

View free PDFSource page
openalexFrontiers in Microbiology2026-07-24

Whole-genome sequencing and molecular analysis of carbapenem-resistant Escherichia coli from clinical samples collected in the framework of the EURGen-Net CCRE survey, Italy 2019

Maria Giufrè, Giulia Errico, Maria Del Grosso, Sara Giancristofaro, Michela Pagnotta, Fabio D’Ambrosio, et al.

Introduction Carbapenem-resistant (CR)-Enterobacterales are a serious threat to public health worldwide. However, there are limited surveillance data on the molecular epidemiology of CR- Escherichia coli . To address this need, the European Centre for Disease Prevention and Contr…

View free PDFSource page
openalexFrontiers in Microbiology2026-07-24

Metabolic reprogramming is associated with symptomatic COVID-19: a serum proteomics and causal inference study identifying ALDOB and glycerol as candidate metabolic correlates

Nana Guo, Xianlei Zhou, Ziyi Pang, Yan Li, Caixiao Jiang, Minghao Geng, et al.

Introduction Coronavirus disease 2019 (COVID-19) shows prominent clinical heterogeneity, presenting two distinct clinical phenotypes including asymptomatic infection and severe symptomatic disease, yet the molecular mechanisms driving such phenotypic differences remain unclear. T…

View free PDFSource page
openalexFrontiers in Microbiology2026-07-24

Synthetic toehold biosensors for detecting Listeria monocytogenes in raw milk

Víctor M. Carballo-Uicab, Luz E. Casados-Vázquez, Drew Endy, J. E. Barboza-Corona

Introduction To our knowledge, direct detection of foodborne bacterial pathogens using toehold biosensors has not been demonstrated in complex food matrices such as raw milk, particularly for Listeria monocytogenes . Here, we report the design and evaluation of toehold biosensors…

View free PDFSource page