Powered by OpenAIRE graph
Found an issue? Give us feedback
image/svg+xml art designer at PLoS, modified by Wikipedia users Nina, Beao, JakobVoss, and AnonMoos Open Access logo, converted into svg, designed by PLoS. This version with transparent background. http://commons.wikimedia.org/wiki/File:Open_Access_logo_PLoS_white.svg art designer at PLoS, modified by Wikipedia users Nina, Beao, JakobVoss, and AnonMoos http://www.plos.org/ arXiv.org e-Print Ar...arrow_drop_down
image/svg+xml art designer at PLoS, modified by Wikipedia users Nina, Beao, JakobVoss, and AnonMoos Open Access logo, converted into svg, designed by PLoS. This version with transparent background. http://commons.wikimedia.org/wiki/File:Open_Access_logo_PLoS_white.svg art designer at PLoS, modified by Wikipedia users Nina, Beao, JakobVoss, and AnonMoos http://www.plos.org/
https://dx.doi.org/10.48550/ar...
Article . 2023
License: arXiv Non-Exclusive Distribution
Data sources: Datacite
versions View all 2 versions
addClaim

This Research product is the result of merged Research products in OpenAIRE.

You have already added 0 works in your ORCID record related to the merged Research product.

Sweetwater: An interpretable and adaptive autoencoder for efficient tissue deconvolution

Authors: de la Fuente, Jesus; Legarra, Naroa; Serrano, Guillermo; Marin-Goni, Irene; Diaz-Mazkiaran, Aintzane; Sendin, Markel Benito; Osta, Ana Garcia; +4 Authors

Sweetwater: An interpretable and adaptive autoencoder for efficient tissue deconvolution

Abstract

Single-cell RNA-sequencing (scRNA-seq) stands as a powerful tool for deciphering cellular heterogeneity and exploring gene expression profiles at high resolution. However, its high cost renders it impractical for extensive sample cohorts within routine clinical care, hindering its broader applicability. Hence, many methodologies have recently arised to estimate cell type proportions from bulk RNA-seq samples (known as deconvolution methods). However, they have several limitations: Many depend on selecting a robust scRNA-seq reference dataset, which is often challenging. Secondly, building reliable pseudobulk samples requires determining the optimal number of genes or cells involved in the simulated data generation process, which has not been studied in depth. Moreover, pseudobulk and bulk RNA-seq samples often exhibit distribution shifts. Finally, most modern deconvolution approaches behave as a black box, and the underlying mechanisms of the deconvolution task are still unknown, which can compromise the reliability of the results. In this work, we present Sweetwater, an adaptive and interpretable autoencoder able to efficiently deconvolve bulk RNA-seq and microarray samples leveraging multiple classes of reference data, such as scRNA-seq and single-nuclei RNA-seq. Moreover, it can be trained on a mixture of FACS-sorted FASTQ files, which we newly propose to use as this reduces platform-specific biases and may potentially outperform single-cell-based references. Also, we demonstrate that Sweetwater effectively uncovers biologically meaningful patterns during the training process, increasing the reliability of the results. Sweetwater is available at https://github.com/ubioinformat/Sweetwater, and we anticipate will facilitate and expedite the accurate examination of high-throughput clinical data across diverse applications.

Keywords

Genomics (q-bio.GN), FOS: Biological sciences, Quantitative Biology - Genomics, Quantitative Biology - Quantitative Methods, Quantitative Methods (q-bio.QM)

  • BIP!
    Impact byBIP!
    citations
    This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).
    0
    popularity
    This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.
    Average
    influence
    This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).
    Average
    impulse
    This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.
    Average
Powered by OpenAIRE graph
Found an issue? Give us feedback
citations
This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).
BIP!Citations provided by BIP!
popularity
This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.
BIP!Popularity provided by BIP!
influence
This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).
BIP!Influence provided by BIP!
impulse
This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.
BIP!Impulse provided by BIP!
0
Average
Average
Average
Green