Abstract
Pooling is a method of simultaneously testing multiple samples for the presence of pathogens. Pooling of SARS-CoV-2 tests is increasing in popularity, due to its high testing throughput. A popular pooling scheme is Dorfman pooling: test N individuals simultaneously, if the test is positive, each individual is then tested separately; otherwise, all are declared negative. Most analyses of the error rates of pooling schemes assume that including more than a single infected sample in a pooled test does not increase the probability of a positive outcome. We challenge this assumption with experimental data and suggest a novel and parsimonious probabilistic model for the outcomes of pooled tests. As an application, we analyse the false-negative rate (i.e. the probability of a negative result for an infected individual) of Dorfman pooling. We show that the false-negative rates under Dorfman pooling increase when the prevalence of infection decreases. However, low infection prevalence is exactly the condition when Dorfman pooling achieves highest throughput efficiency. We therefore urge the cautious use of pooling and development of pooling schemes that consider correctly accounting for tests’ error rates.
Keywords: pooling, probabilistic model, COVID-19, qRT-PCR
1. Introduction
Reverse transcription polymerase chain reaction (RT-PCR) testing is the backbone and gold standard of infection surveillance across the globe, and a key component in the fight against the COVID-19 pandemic. One major application of this technology is the rapid testing of contacts of positive individuals in order to cut transmission chains. Moreover, in some countries, RT-PCR testing has become a routine procedure of surveillance, testing incoming air passengers in airports [1,2], school students [3], or even the entire adult population [4].
As such, the need for large-scale testing has resulted in the application of techniques to increase the throughput of RT-PCR tests [5–10]. These techniques, often termed pooling, or group testing, have originated in the 1940s in the seminal work of Dorfman [5,11]. Under Dorfman pooling, one selects N individuals and performs a single RT-PCR test on their combined (pooled) samples. If the pooled test yields a positive result, then each individual is retested separately; otherwise, everyone is declared negative. The throughput efficiency of Dorfman pooling has been demonstrated empirically [5] and its error rates thoroughly investigated [12–14].
The study and practice of pooling has advanced considerably since Dorfman’s original paper, with several other pooling techniques being applied for SARS-CoV-2. For example, in recursive pooling [13,15], if the first pooled test is positive, the pool is split into two and the splitting process repeats. Otherwise, all pool members are declared negative. In Matrix pooling [8] a population of size N = mn is arranged in a matrix of dimensions m × n. Each row and column are then pooled and individuals in the intersection of positive rows and columns are tested separately. Other, more sophisticated pooling schemes exist [16], but studies investigating the underlying assumptions of pooling are surprisingly lacking.
Studies focused on pooling for SARS-CoV-2 are in a consensus that sample dilution effects [17,18] are not a concern, even for pools as large as 64 individuals [5,6,9,19,20]. Consequently, studies assumed that the probability of a true-positive (the test’s sensitivity) does not depend on the number of infected samples in the pool, but rather on the existence of at least one such sample. Thus, the probability of a positive result in a pooled test has been assumed identical for a pool with one sample from an infected individual and, e.g. five such samples. This assumption is common in the group testing literature [6,12,13], as well as in more specific, SARS-CoV-2 focused studies [14,21]. In this study we challenge this commonly made assumption and demonstrate how using a more accurate probabilistic model affects the estimation of false-negative rates for Dorfman pooling.
Our analysis has several implications: previously estimated optimal Dorfman pool size [12] should be revisited in general and specifically for the case of SARS-CoV-2, as it is influenced by the prevalence of infection. The error rates of other pooling schemes in use should also be revisited [10,13]. Moreover, studies that seek precise error models should be conducted for other pathogens as well, since the standard error model could prove flawed. Specifically, all assumptions should be scrutinized, including (e.g.) the independence of tests, an extremely widespread assumption in the group testing literature [12,13] which dates back to Dorfman’s original paper [11]. Finally, new methods should be developed in order to take advantage of more realistic models.
2. Methods
Formally, we consider a pool containing N individuals {1, …, N}. We denote the true infection state θ ∈ {0, 1}N, so individual i is infected if θi = 1. The RT-PCR test’s sensitivity (true-positive rate) is denoted Se, and the test’s specificity (true-negative rate) is denoted Sp. Pooled test result (data) is denoted d ∈ {0, 1}, where d = 0 if the test returned a negative result.
2.1. The common assumption
Previous studies of pooling schemes assumed that the false-negative probability does not depend on the number of infected samples, but merely on the existence of at least one such sample in a pool [12,13]. Current studies of pooling in the context of SARS-CoV-2 also employ a similar assumption [14,21]. Explicitly, these studies assume
| 2.1 |
Below, we refer to equation (2.1) as the common assumption.
2.2. Refuting the common assumption
We refute the common assumption with experimental data collected from [22], and summarized in table 1. There, the authors investigate Dorfman pooling and, regardless of the pooled test result, follow up and test each pool member separately. We focus on 128 pools for which at least one subsequent separate test was positive—of which 29 pooled tests were negative and 99 positive. In the data cited in [22], of the 29 negative pools, subsequent separate testing yielded a single positive result in 24. By contrast, of the 99 positive pools, 42 yielded a single positive test upon subsequent separate testing.
Table 1.
Contingency table of data from [22].
| negative pool | positive pool | |
|---|---|---|
| no. subsequent positives = 1 | 24 | 42 |
| no. subsequent positives ≥ 1 | 5 | 57 |
| total | 29 | 99 |
The data in table 1 allows us to test the following null hypothesis H0: The probability of a pooled false-negative is equal for pools with one subsequent positively tested member and pools with two or more such members. H0 is a direct consequence of the common assumption, and rejecting H0 implies the common assumption is not realistic, at least for SARS-CoV-2.
We apply Fisher’s exact test for the presence of more than one positive individual in correctly identified pools. Fisher’s test yields an increased odds ratio of 6.4, 95% CI (2.2, 23.4), with a p-value ≈10−4. Thus, we reject H0, refuting the common assumption.
2.3. Our model
Since the essence of the refuted common assumption is that amplification of all samples occurs only once, we assume a more realistic model: amplification of viral RNA succeeds or fails for each sample independently. Furthermore, according to [12–14,21], a false-positive does not depend on the number of negative samples in a pool either. For lack of data pointing otherwise, we incorporate this assumption into our model with a small modification. We do assume that a false amplification can occur only once per pool. However, we also assume false amplification is independent of any other correct amplification. Specifically, it is possible that every correct amplification fails and an erroneous one occurs simultaneously. This assumption is somewhat specific for the current application of screening for SARS-CoV-2 via RT-PCR. For example, cross-reactivity with other coronaviruses would have violated this assumption. However, cross-reactivity was ruled out in [23]. These assumptions lead to the following model, illustrated in figure 1 and summarized in equation (2.2).
| 2.2 |
Figure 1.
Illustration of the revised probabilistic model. A pool contains individuals {1, 2, 3} with state θ = (1, 1, 0) (i.e. individuals 1 and 2 are infected). A negative pooled test (d = 0) occurs when three detection paths fail: a false-negative occurs for individuals 1 and 2, each with probability (1 − Se). Additionally, no false-positive detection occurs, with probability Sp. Individual 3 is not infected and does not contribute to the probability of the pooled test result.
2.4. Application: scheme false-negative rate
We calculate the false-negative rate for a single infected individual, henceforth referred to as ‘Donald’, under our model and under the common assumption. We distinguish three types of false-negative events when performing pooling. A single test’s false-negative is the event of a negative result upon testing Donald separately, i.e. in an RT-PCR test without pooling. A pooled false-negative occurs when a pooled test containing Donald’s sample (and other samples) yields a negative result, i.e. the pooling fails to detect at least one positive result. Lastly, a scheme false-negative occurs when an entire pooling scheme fails to identify Donald as infected. Our first task is to calculate Dorfman’s scheme false-negative rate. Equivalently, we ask: what is the probability of not identifying Donald as infected under Dorfman pooling?
Denote the prevalence of infection in the (tested) population q. We denote Donald as individual 1, so that θ1 = 1. Then
| 2.3 |
If the pooled test yields a positive result, Donald is tested separately. Taking a conservative stand, it is assumed that such a simple procedure poses no risk of introducing contaminant RNA. Therefore, the separate test yields a positive result with probability Se.
We calculate the probability that Donald is mistakenly identified as not infected, henceforth referred to as the scheme’s false-negative rate and denoted Sfn. In order to correctly identify an infected individual as infected, both pooled and separate tests have to yield a positive result. Thus, the scheme’s false-negative rate is
| 2.4 |
2.4.1. Comparison metric
The single test false-negative rate 1 − Se and scheme false-negative rate Sfn are compared via
| 2.5 |
Eexact is the percentage increase in the pooling scheme false-negative rate, relative to the single test false-negative rate.
According to the common assumption, the scheme false-negative rate is . A straight forward calculation shows that this implies the percentage increase in scheme false-negative rate is .
3. Results
3.1. Scheme false-negative
We plot Eexact for varying prevalence q and sensitivity Se values, and make the comparison with Ecommon. As recommended by [5], we apply different pool sizes N, for different prevalence values. We consider a specificity of Sp = 0.95 [5] and a range of reasonable sensitivity and prevalence values [23–26]. We observe that the increase in the scheme false-negative rate (compared to 1 − Se), expressed by Eexact, is at least (figure 2).
Figure 2.
Relative increase in Dorfman pooling false-negative rates Eexact. (a) Colours represent Eexact, the relative percentage increase in the scheme false-negative rates relative to the single test false-negative rates equation (2.5). (b) Colours represent the difference between Ecommon and Eexact. The infection prevalence q is varied on the x-axis, while the test sensitivity is varied on the y-axis. Pool size N was chosen according to q as in [5]. Panel a shows that Eexact is largest for low prevalence values q. The difference between Ecommon and Eexact can be as large as 30%, as seen in b.
Interestingly, an increase in infection prevalence monotonically decreases the scheme false-negative rate, as can also be easily seen from equation (2.4). For the chosen parameter ranges, the increase in the single test false-negative rates increases the relative error Eexact. These effects can be seen in figure 2a, upon conditioning on pool size. Extending the range for Sp yields no qualitative differences. We further compare Eexact to Ecommon, showing the discrepancy changes as a function of both prevalence and the single test sensitivity (figure 2b).
4. Discussion
In this study, a novel probabilistic model is developed to capture realistic features of the outcomes of pooled RT-PCR tests, parameterized for SARS-CoV-2. Contrary to the common assumption, we assume, based on data (table 1), that multiple infected individuals increase the likelihood of a positive pooled test result. A direct consequence of our model is that low values of infection prevalence increase the false-negative rates of Dorfman pooling. Importantly, our results remain qualitatively similar under varying parameter values, in the observed ranges [23–25,27] (figure 2). Therefore, our results give rise to a conflict: low infection prevalence leads to high efficiency of Dorfman pooling [5] but also increases false-negative rates.
As the COVID-19 pandemic progresses, the infection prevalence in various tested populations undergoes frequent changes. Hence, as our results suggest, pooling schemes employed for mass testing should be used with caution in populations where infection rates are low. Such mass-tested populations often include air travel passengers [28], nursing homes [5], or presymptomatic and asymptomatic individuals [29], and can be crucial for controlling outbreaks [30].
Furthermore, an especially important consideration in designing pooling schemes is explicitly taking into account the intrinsic RT-PCR error rates. As illustrated above, analysing the expected error rates of pooling schemes under a realistic error model is readily achievable. Designing pooling schemes that optimize testing while accounting for error rates is, however, considerably more involved. Nonetheless, various algorithms explicitly incorporate error rates into their pooling optimization processes [12,13], including a recent study by our group which employed the herein described error model [31]. In our proposed algorithm, we have used an experimental design criterion to create an algorithm which iteratively performs pooling, updating its information based on previous results. In doing so, our pooling scheme shows potential in terms of both reducing false-negative and false-positive rates, as well as increasing efficiency, although it has not yet been implemented in real-world settings.
To conclude, pooling is an important technique that can increase testing throughput in a cost-effective manner. Nevertheless, care must be given to pooling schemes’ false-negative rates, especially under low infection prevalence settings.
Supplementary Material
Authors' contributions
Data and relevant code for this research work are stored in GitHub: https://github.com/yairdaon/errorates and have been archived within the Zenodo repository: https://doi.org/10.5281/zenodo.5510563 [32]. Y.D., A.H. and U.O. designed the study. Y.D. wrote code.
Competing interests
We declare we have no competing interests.
Funding
Y.D. was supported by a post-doctoral fellowship from the Tel Aviv University Center for Combating Pandemics and the Raymond and Beverly Sackler dean’s post-doctoral fellowship.
References
- 1.Dollard P et al. 2020. Risk assessment and management of COVID-19 among travelers arriving at designated us airports, January 17–September 13, 2020. Morb. Mortal. Wkly. Rep. 69, 1681. ( 10.15585/mmwr.mm6945a4) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 2.Norizuki M, Hachiya M, Motohashi A, Moriya A, Mezaki K, Kimura M, Sugiura W, Akashi H, Umeda T. 2021. Effective screening strategies for detection of asymptomatic COVID-19 travelers at airport quarantine stations: exploratory findings in Japan. Global Health Med. 3, 107-111. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 3.Sweeney-Reed CM, Wolff D, Niggel J, Kabesch M, Apfelbacher C. 2021. Pool testing as a strategy for prevention of SARS-CoV-2 outbreaks in schools: protocol for a feasibility study. JMIR Res. Protoc. 10, e28673. ( 10.2196/28673) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 4.Holt Ed. 2020. Slovakia to test all adults for SARS-CoV-2. Lancet 396, 1386-1387. ( 10.1016/S0140-6736(20)32261-3) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 5.Barak N et al. 2021. Lessons from applied large-scale pooling of 133, 816 SARS-CoV-2 RT-PCR tests. Sci. Transl. Med. 13, eabf2823. ( 10.1126/scitranslmed.abf2823) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 6.Grobe N, Cherif A, Wang X, Dong Z, Kotanko P. 2021. Sample pooling: burden or solution? Clin. Microbiol. Infect. 27(9) 1212-1220 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 7.Hanel R, Thurner S. 2020. Boosting test-efficiency by pooled testing for SARS-CoV-2—Formula for optimal pool size. PLoS ONE 15, e0240652. ( 10.1371/journal.pone.0240652) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 8.Kwiatkowski Jr TJ, Zoghbi HY, Ledbetter SA, Ellison KA, Chinault AC. 1990. Rapid identification of yeast artificial chromosome clones by matrix pooling and crude lysate PCR. Nucleic Acids Res. 18, 7191. ( 10.1093/nar/18.23.7191) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 9.Lohse S, Pfuhl T, Berkó-Göttel B, Rissland J, Geißler T, Gärtner B, Becker SL, Schneitler S, Smola S. 2020. Pooling of samples for testing for SARS-CoV-2 in asymptomatic people. Lancet Infect. Dis. 20, 1231-1232. ( 10.1016/S1473-3099(20)30362-5) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 10.Noriega R, Samore M. 2020. Increasing testing throughput and case detection with a pooled-sample Bayesian approach in the context of COVID-19. bioRxiv. ( 10.1101/2020.04.03.024216) [DOI]
- 11.Dorfman R. 1943. The detection of defective members of large populations. Ann. Math. Stat. 14, 436-440. ( 10.1214/aoms/1177731363) [DOI] [Google Scholar]
- 12.Aprahamian H, Bish DR, Bish EK. 2020. Optimal group testing: structural properties and robust solutions, with application to public health screening. INFORMS J. Comput. 32, 895-911. ( 10.1287/ijoc.2019.0942) [DOI] [Google Scholar]
- 13.Kim H-Y, Hudgens MG, Dreyfuss JM, Westreich DJ, Pilcher CD. 2007. Comparison of group testing algorithms for case identification in the presence of test errors. Biometrics 63, 1152-1163. ( 10.1111/j.1541-0420.2007.00817.x) [DOI] [PubMed] [Google Scholar]
- 14.Pikovski A, Bentele K. 2020. Pooling of coronavirus tests under unknown prevalence. Epidemiol. Infect. 148, e183. ( 10.1017/S0950268820001752) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 15.Eberhardt JN, Breuckmann NP, Eberhardt CS. 2020. Multi-stage group testing improves efficiency of large-scale COVID-19 screening. J. Clin. Virol. 128, 104382. ( 10.1016/j.jcv.2020.104382) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 16.Shental N et al. 2020. Efficient high-throughput SARS-CoV-2 testing to detect asymptomatic carriers. Sci. Adv. 6, eabc5961. ( 10.1126/sciadv.abc5961) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 17.Hwang FK. 1976. Group testing with a dilution effect. Biometrika 63, 671-680. ( 10.1093/biomet/63.3.671) [DOI] [Google Scholar]
- 18.Wein LM, Zenios SA. 1996. Pooled testing for HIV screening: capturing the dilution effect. Oper. Res. 44, 543-569. ( 10.1287/opre.44.4.543) [DOI] [Google Scholar]
- 19.Hirotsu Y et al. 2020. Pooling RT-qPCR testing for SARS-CoV-2 in 1000 individuals of healthy and infection-suspected patients. Sci. Rep. 10, 1-8. ( 10.1038/s41598-019-56847-4) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 20.Yelin I et al. 2020. Evaluation of COVID-19 RT-qPCR test in multi sample pools. Clin. Infect. Dis. 71, 2073-2078. ( 10.1093/cid/ciaa531) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 21.Cherif A, Grobe N, Wang X, Kotanko P. 2020. Simulation of pool testing to identify patients with coronavirus disease 2019 under conditions of limited test availability. JAMA Netw. Open 3, e2013 075-e2013 075. ( 10.1001/jamanetworkopen.2020.13075) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 22.de Salazar A et al. 2020. Sample pooling for SARS-CoV-2 RT-PCR screening. Clin. Microbiol. Infect. 26, 1687.e1-1687.e5. ( 10.1016/j.cmi.2020.09.008) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 23.van Kasteren PB, van Der Veer B, van den Brink S, Wijsman L, de Jonge J, van den Brandt A, Molenkamp R, Reusken CBEM, Meijer A. 2020. Comparison of seven commercial RT-PCR diagnostic kits for COVID-19. J. Clin. Virol. 128, 104412. ( 10.1016/j.jcv.2020.104412) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 24.Watson J, Whiting PF, Brush JE. 2020. Interpreting a COVID-19 test result. BMJ 369, m1808. [DOI] [PubMed] [Google Scholar]
- 25.Wikramaratna PS, Paton RS, Ghafari M, Lourenço J. 2020. Estimating the false-negative test probability of SARS-CoV-2 by RT-PCR. Eurosurveillance 25, 2000568. ( 10.2807/1560-7917.ES.2020.25.50.2000568) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 26.Woloshin S, Patel N, Kesselheim AS. 2020. False negative tests for SARS-CoV-2 infection-challenges and implications. N. Engl. J. Med. 383, e38. ( 10.1056/NEJMp2015897) [DOI] [PubMed] [Google Scholar]
- 27.Kucirka LM, Lauer SA, Laeyendecker O, Boon D, Lessler J. 2020. Variation in false-negative rate of reverse transcriptase polymerase chain reaction—based SARS-CoV-2 tests by time since exposure. Ann. Intern. Med. 173, 262-267. ( 10.7326/M20-1495) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 28.Daon Y, Thompson RN, Obolski U. 2020. Estimating covid-19 outbreak risk through air travel. J. Travel Med. 27, taaa093. ( 10.1093/jtm/taaa093) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 29.Thompson RN, Lovell-Read FA, Obolski U. 2020. Time from symptom onset to hospitalisation of coronavirus disease 2019 (COVID-19) cases: implications for the proportion of transmissions from infectors with few symptoms. J. Clin. Med. 9, 1297. ( 10.3390/jcm9051297) [DOI] [PMC free article] [PubMed] [Google Scholar]
- 30.Mina MJ, Andersen KG. 2021. COVID-19 testing: one size does not fit all. Science 371, 126-127. ( 10.1126/science.abe9187) [DOI] [PubMed] [Google Scholar]
- 31.Daon Y, Huppert A, Obolsik U. 2021. DOPE: D-Optimal Pooling Experimental design with application to SARS-CoV-2 screening. (http://arxiv.org/abs/2103.03706)
- 32.Daon Y. 2021. yairdaon/errorates: Release for archiving.
Associated Data
This section collects any data citations, data availability statements, or supplementary materials included in this article.


