Compound models and Pearson residuals for single-cell RNA-seq data without UMIs.

Jan Lause,Christoph Ziegenhain,Leonard Hartmanis,Philipp Berens,Dmitry Kobak

doi:10.1101/2023.08.02.551637

Compound models and Pearson residuals for single-cell RNA-seq data without UMIs.

Jan Lause, Christoph Ziegenhain + Show 3 more

Open Access

https://doi.org/10.1101/2023.08.02.551637

Copy DOI

Journal: bioRxiv : the preprint server for biology	Publication Date: Jul 25, 2024
License type: CC BY-NC-ND 4.0

Affiliation: Bernstein Center for Computational Neuroscience Tübingen, Hertie Institute for Clinical Brain Research, University of Tübingen, Karolinska Institutet

#Pearson Residuals #Broken Power Law + Show 8 more

Abstract
Full-Text
Similar Papers

Abstract

Recent work employed Pearson residuals from Poisson or negative binomial models to normalize UMI data. To extend this approach to non-UMI data, we model the additional amplification step with a compound distribution: we assume that sequenced RNA molecules follow a negative binomial distribution, and are then replicated following an amplification distribution. We show how this model leads to compound Pearson residuals, which yield meaningful gene selection and embeddings of Smart-seq2 datasets. Further, we suggest that amplification distributions across several sequencing protocols can be described by a broken power law. The resulting compound model captures previously unexplained overdispersion and zero-inflation patterns in non-UMI data.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

More From: bioRxiv : the preprint server for biology

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.