» Articles » PMID: 36536300

VirPool: Model-based Estimation of SARS-CoV-2 Variant Proportions in Wastewater Samples

Overview
Publisher Biomed Central
Specialty Biology
Date 2022 Dec 19
PMID 36536300
Authors
Affiliations
Soon will be listed here.
Abstract

Background: The genomes of SARS-CoV-2 are classified into variants, some of which are monitored as variants of concern (e.g. the Delta variant B.1.617.2 or Omicron variant B.1.1.529). Proportions of these variants circulating in a human population are typically estimated by large-scale sequencing of individual patient samples. Sequencing a mixture of SARS-CoV-2 RNA molecules from wastewater provides a cost-effective alternative, but requires methods for estimating variant proportions in a mixed sample.

Results: We propose a new method based on a probabilistic model of sequencing reads, capturing sequence diversity present within individual variants, as well as sequencing errors. The algorithm is implemented in an open source Python program called VirPool. We evaluate the accuracy of VirPool on several simulated and real sequencing data sets from both Illumina and nanopore sequencing platforms, including wastewater samples from Austria and France monitoring the onset of the Alpha variant.

Conclusions: VirPool is a versatile tool for wastewater and other mixed-sample analysis that can handle both short- and long-read sequencing data. Our approach does not require pre-selection of characteristic mutations for variant profiles, it is able to use the entire length of reads instead of just the most informative positions, and can also capture haplotype dependencies within a single read.

Citing Articles

SWAMPy: simulating SARS-CoV-2 wastewater amplicon metagenomes.

Boulton W, Fidan F, Denise H, De Maio N, Goldman N Bioinformatics. 2024; 40(9).

PMID: 39226177 PMC: 11401744. DOI: 10.1093/bioinformatics/btae532.


Impact of reference design on estimating SARS-CoV-2 lineage abundances from wastewater sequencing data.

Assmann E, Agrawal S, Orschler L, Bottcher S, Lackner S, Holzer M Gigascience. 2024; 13.

PMID: 39115959 PMC: 11308188. DOI: 10.1093/gigascience/giae051.


Nanopore sequencing technology and its applications.

Zheng P, Zhou C, Ding Y, Liu B, Lu L, Zhu F MedComm (2020). 2023; 4(4):e316.

PMID: 37441463 PMC: 10333861. DOI: 10.1002/mco2.316.

References
1.
Rios G, Lacoux C, Leclercq V, Diamant A, Lebrigand K, Lazuka A . Monitoring SARS-CoV-2 variants alterations in Nice neighborhoods by wastewater nanopore sequencing. Lancet Reg Health Eur. 2021; 10:100202. PMC: 8372489. DOI: 10.1016/j.lanepe.2021.100202. View

2.
Bibby K, Bivins A, Wu Z, North D . Making waves: Plausible lead time for wastewater based epidemiology as an early warning system for COVID-19. Water Res. 2021; 202:117438. PMC: 8274973. DOI: 10.1016/j.watres.2021.117438. View

3.
Hrudey S, Conant B . The devil is in the details: emerging insights on the relevance of wastewater surveillance for SARS-CoV-2 to public health. J Water Health. 2022; 20(1):246-270. DOI: 10.2166/wh.2021.186. View

4.
Virtanen P, Gommers R, Oliphant T, Haberland M, Reddy T, Cournapeau D . SciPy 1.0: fundamental algorithms for scientific computing in Python. Nat Methods. 2020; 17(3):261-272. PMC: 7056644. DOI: 10.1038/s41592-019-0686-2. View

5.
Crits-Christoph A, Kantor R, Olm M, Whitney O, Al-Shayeb B, Lou Y . Genome Sequencing of Sewage Detects Regionally Prevalent SARS-CoV-2 Variants. mBio. 2021; 12(1). PMC: 7845645. DOI: 10.1128/mBio.02703-20. View