Skip to main content
PLOS Pathogens logoLink to PLOS Pathogens
. 2026 Apr 9;22(4):e1014066. doi: 10.1371/journal.ppat.1014066

Cell-free RNA reveals host and microbial correlates of broadly neutralizing antibody development against HIV

Mark Kowarsky 1, Mercedes Dalman 2,3, Mira N Moufarrej 4, Jennifer Okamoto 5, Yike Xie 2,3, Norma F Neff 5, Salim S Abdool Karim 6,7, Nigel Garrett 6,8,9, Penny L Moore 6,10,11,12,*, Joan Camunas-Soler 2,3,13,*, Stephen R Quake 4,5,*
Editor: Vijayakumar Velu14
PMCID: PMC13065077  PMID: 41955174

Abstract

A small number of people living with HIV (PLWH) develop broadly neutralizing antibodies (bNAbs) targeting multiple HIV strains. Although several viral and immune factors contribute to bNAb development, the genetic and environmental factors driving this response remain largely unknown. We performed combined cell-free DNA (cfDNA) and cell-free RNA (cfRNA) sequencing in 42 plasma samples from a longitudinal cohort of 14 PLWH (7 who develop bNAbs and 7 matched controls). This approach enabled us to non-invasively monitor the host transcriptome, viral genetic variation, and microbiome composition during HIV infection, and to identify molecular correlates of bNAb development. We find that development of bNAbs is associated with a transcriptomic signature of early immune activation characterized by elevated levels of MHC class I antigen presentation genes. This signature is independent of viral load or CD4 count and declines over time. In addition to host features, we recovered sufficient viral reads to reconstruct HIV consensus sequences, supporting the utility of cfRNA for viral genotyping. Finally, we also identified an enrichment of several microbial taxa in bNAb producers and increased levels of GB virus C (GBV-C), a non-pathogenic lymphotropic virus. Our findings suggest a distinct early immune activation profile in PLWH who develop bNAbs. More broadly, we show that combined cfDNA/cfRNA sequencing can reveal relationships between a protective immunogenic response to HIV infection, the host immune system, and microbiome, highlighting its potential for biomarker discovery in future vaccine and therapeutic studies.

Author summary

A subset of people living with HIV develop broadly neutralizing antibodies (bNAbs), a phenomenon that remains poorly understood but relevant to vaccine development. We applied combined cell-free RNA and DNA sequencing to 42 longitudinal plasma samples (7 bNAb producers, 7 controls) to monitor the host transcriptome, viral variation, and microbiome composition during bNAb development. Early in infection, bNAb producers showed an MHC class I-linked immune activation signature that declined over time, alongside changes in commensal microbial sequences and increased levels of GB virus C in cfRNA. These findings support the use of cfDNA/cfRNA sequencing for biomarker discovery in HIV research.

Introduction

HIV has infected over 36 million people worldwide, reaching epidemic proportions in Eastern and Southern Africa with an estimated adult prevalence close to 7% [1]. Antiretroviral therapy (ART) has greatly improved outcomes, allowing HIV to be a manageable condition throughout life. However, global access to ART is difficult to achieve and also requires lifelong daily adherence to a pill taking regime. Moreover, despite effective treatment, transmission of HIV remains common in vulnerable populations, and people living with HIV (PLWH) are at increased risk of aging-related and chronic diseases [2].

Although virtually all PLWH develop strain-specific autologous responses [3–6], few develop broadly neutralizing antibodies (bNAbs), which can neutralize multiple strains of the virus [7], and when administered passively can prevent acquisition of susceptible HIV strains and suppress viral replication in PLWH [8]. There is great interest in understanding the drivers of bNAbs, as this could inform the development of an effective vaccine. Efforts to design an effective HIV vaccine have increasingly focused on strategies to induce bNAbs by guiding B-cell affinity maturation, with studies demonstrating promising immunization approaches through the rational design of immunogens [9–11]. Despite these advances, which specific cellular processes give rise to bNAbs during infection remains incompletely described. So far, there is evidence that viral and host genetic factors might play a role [12–16], and that a combination of high viral titres and reduced CD4 counts early in infection is associated with the development of bNAbs in some individuals [17,18]. Additionally, changes in frequency of certain immune subpopulations, such as T follicular helper cells (Tfh) might also influence affinity maturation towards bNAb production [19–23]. Moreover, there is some evidence that bNAbs might arise from pre-existing memory B cells initially stimulated by bacterial antigens -likely from the gut microbiome- which later cross-react with the HIV glycan shield and develop neutralizing capabilities [24,25]. Although microbial antigen cross-reactivity has also been shown to lead to non-protective immune responses in HIV vaccine trials [26], a better understanding of the interplay between the host immune response, microbiome and a protective bNAb response could aid in vaccine development.

Sequencing of plasma cell-free DNA (cfDNA) has become an established tool in clinical medicine, with applications ranging from non-invasive prenatal testing and oncology to organ transplant monitoring [27,28]. In contrast, plasma cell-free RNA (cfRNA) provides an emerging complementary approach to measure dynamic changes in the host transcriptome [29]. Because transcripts released into circulation can originate from many different cell types and tissues [30], cfRNA enables non-invasive monitoring of systemic physiological and immune states [31,32]. To date, cfRNA sequencing has been used in diverse biomedical settings, including the monitoring of placental and pregnancy complications for maternal-fetal health [33–36], to follow bone marrow reconstitution after stem cell transplantation [37], and to characterize host immune responses to tuberculosis infection [38]. However, integrative analyses that capture host, viral, and microbial nucleic acids simultaneously remain uncommon, and cfRNA sequencing approaches have been rarely applied to monitor retroviral infections such as HIV. Here, we conducted a pilot study to investigate whether combined cfDNA and cfRNA sequencing of plasma could reveal molecular features associated with the development of bNAbs in a longitudinal cohort of PLWH.

Results

Study design, sequencing overview and viral analysis

We analyzed 42 plasma samples from 14 PLWH from the CAPRISA studies, of whom 7 developed bNAbs and 7 did not. For each participant, plasma was collected at approximately 6 months, 1 year, and 3 years post-HIV acquisition and prior to ART (Fig 1A). At the 6-month time point, both groups had comparable CD4 counts and viral loads (S1 Fig). The extent of bNAb development was assessed by measuring plasma neutralization breadth against a panel of 18 Env-pseudotyped HIV viruses, summarized as percentage breadth, a continuous measure of bNAb development (Fig 1B and S1 Table and Methods).

Fig 1. Study design and development of neutralization breadth.

Fig 1

A) Longitudinal trajectories of neutralization breadth in bNAb producers (red) and controls (blue). Time zero corresponds to the estimated HIV infection point, with breadth defined as the ability of plasma antibodies to neutralize a cross-clade panel of 18 env-pseudotyped viruses (S2 Table). B) Heatmap of serum neutralization titers (IC50) against the 18-virus panel, grouped by individual, with time points sorted from first (top) to last (bottom). Colors represent log10-transformed IC50 neutralization titers, with higher values (yellow) indicating stronger neutralization potency and black indicating values below the limit of detection.

We performed combined cfDNA and cfRNA sequencing in each of the 42 samples (Methods). After quality control and filtering, we obtained enough reads to sample the host transcriptome (80% of total reads for cfRNA) and microbiome (0.2% and 20% reads of cfDNA and cfRNA respectively) (S2 Fig). Sequencing depth before and after sample QC processing (cleaning) is similar for both cfDNA and cfRNA, with comparable metrics between bNAb producers and non-producers (S3 Fig). Similarly, cell-type deconvolution of the host transcriptome showed comparable contributions of the most abundant cell types (e.g., red blood cells, immune cells, platelets) across both groups (S4 Fig). We detected HIV reads in the cfRNA assay for all samples, with coverage sufficient to genotype the majority strain in 38 out of 42 samples (S5A and S5B Fig). For high coverage samples, the nucleotide identity of HIV clustered within individuals based on previous single-genome amplified envelope sequences (Fig 2A) except for one mislabeled sample that was excluded from further analyses (Methods).

Fig 2. HIV consensus sequences from cfRNA.

Fig 2

A) Genomic similarity matrix of consensus HIV sequences obtained from cfRNA, showing clustering within individuals. B) Maximum-likelihood phylogeny of cfRNA-derived HIV consensus sequences (colored labels) with previously published envelope sequences (bold, CAPRISA). Shaded boxes highlight individual-level clustering. With the exception of one sample, cfRNA consensus sequences consistently cluster with the corresponding envelope sequences.

We validated the accuracy of cfRNA-derived consensus sequences by comparing them to previously published envelope sequences from the same CAPRISA participants (S2 Table), observing consistent clustering and strain concordance across time points and sequencing methods (Figs 2B and S5C). By comparing differences in consensus sequences within individuals (S5D Fig), we found that most mutations occur in the env gene, which encodes the surface protein responsible for host cell binding. The Env protein is the sole target of bNAbs and is known to be the most rapidly evolving part of the HIV genome [39–41]. Using cfRNA we estimated that the dominant viral strain in each individual accumulates approximately 10 mutations per year (S5E Fig), in line with other estimates from virion or PBMC-derived studies [42,43]. Altogether, these results show that cfRNA is a quantitative tool to perform genotype analysis of viral pathogens and to track strain-specific mutational dynamics over time.

Analysis of read lengths and mapping start and end positions indicates that HIV-derived reads in plasma cfRNA (cfHIV) predominantly consist of short RNA fragments (200–500 bases) and do not show systematic positional biases along the HIV genome (S6 Fig). This is consistent with cfHIV being largely derived from fragmented cell-free viral RNA rather than libraries dominated by intact full-length viral genomes. In addition, to assess the quality of HIV genome recovery from cfRNA, we performed a subsampling analysis showing that approximately 50,000 cfHIV reads are sufficient to recover >90% of the genome at ≥50X depth (S7 Fig). In this cohort, this level of coverage was achieved for most samples (83%).

Host transcriptome correlates of cfHIV and clinical parameters

To characterize host responses to HIV infection, we measured the correlation between the host transcriptome and the total abundance of cfHIV, as well as with clinical parameters such as viral load and CD4 counts. Spearman correlation of cfHIV, viral load and CD4 counts to the human transcriptome revealed several genes associated with these measurements (Figs 3A and S8A and S8B). Analysis of genes significantly correlated to cfHIV revealed 19 genes typically associated with host immune responses to viral infection (FDR < 0.05), whereas only 1 for each of viral load and CD4 count, respectively (Fig 3B). CD4 count levels showed an association with IL4, a cytokine inducing differentiation of naive T helper cells. Although the relation between cytokine profile and HIV progression is controversial, increased secretion of IL4 in relation to IFN-gamma is a known hallmark of switching from Th1 to Th2 lymphokine secretion in PLWH [44]. Our observation of IL4 being increased in low CD4 samples is compatible with an increase of IL-4 levels upon disease progression and T helper cell decrease. Top correlates of cfHIV include: CCL5, also known as RANTES, an HIV suppression factor released by CD8+ T cells, TFRC (CD71), a marker of T-cell activation, and CCL2, a chemokine that attracts immune cells to sites of infection (Figs 3C and S8C). These genes do not show a differential behaviour between bNAb producers and controls, indicating that they represent shared responses to HIV infection. This shows that cfRNA sequencing captures shared host transcriptional responses associated with peripheral viral burden during HIV infection.

Fig 3. Correlation between cfHIV levels and host transcriptome.

Fig 3

A) Number of genes significantly correlated with cfHIV counts, viral load (VL) and CD4 counts (Spearman’s rank correlation, see Methods). B) Subset of genes from panel A previously described as involved in host immune responses to viral infection. C) Example correlations between cfHIV counts and three HIV-associated genes. Spearman correlation coefficients indicated. No partitioning is observed between bNAb producers (red) and controls (blue).

Development of bNAbs is associated with a transcriptional phenotype of immune activation

We next compared the transcriptome of bNAb producers and non-producers. We performed differential gene abundance analysis finding 256 genes elevated in bNAbs producers (FDR < 0.05, fold-change>1, Fig 4A, Methods). This variation could not be attributed to environmental factors such as site of collection, that showed little to no correlation to these transcriptomic signatures. Remarkably, 52% (n = 131) of elevated genes in bNAb producers are immune-related (primarily associated with adaptive immunity), compared with only 14% (n = 39) of the 281 decreased genes. Among the elevated genes, 25 of them are associated with production of RANTES (CCL5), a chemokine involved in homing and migration of effector and memory T cells. Although CCL5 itself is not enriched in bNAb producers in our dataset, elevated RANTES, along with other inflammatory markers, have previously been linked to the development of bNAbs in spontaneous controllers of HIV [45].

Fig 4. Gene level differences in cfRNA between bNAb producers and controls.

Fig 4

A) Volcano plot of differentially abundant genes. Genes with a fold change > 2 and FDR < 0.05 are highlighted (red = increased, blue = decreased in bNAb producers). Labelled points are genes involved in host immune responses to viral infection. B) Violin plots of the distribution of elevated genes in bNAb producers, stratified by group and neutralization breadth (number of viral strains neutralized), used here as a proxy for progression along bNAb development. C) Genes in the MHC class I antigen presentation pathway. Most show negative correlations with breadth, and some are also differentially abundant between bNAb producers and controls. Correlations with breadth were assessed using Spearman’s rank correlation (Spearman’s ρ indicated next to each gene, see Methods). * correlated with breadth at FDR < 0.05, † differentially abundant between bNAb producers and controls at FDR < 0.05, ‡ correlated with cfHIV at FDR < 0.05.

A pathway analysis confirmed these results and further revealed an enrichment of molecular functions related to an immunologic response in bNAb producers (S9A Fig). Terms such as receptor activity, IgG and peptide antigen binding, cytokine signaling and NF-kB signaling were enriched in bNAb positive samples. Analyses of the genes decreased in the bNAb producers showed lower enrichment scores and did not point to the existence of other host immunoregulatory effects (S9B Fig). Overall, host cfRNA sequencing provides evidence for an increased immune activation in participants who developed bNAbs, despite having similar CD4 counts and viral loads to non-bNAb producers.

To investigate this immune activation in more detail, we performed pseudotime analysis on the two groups, using plasma neutralization breadth to define the temporal progression (Fig 4B, Methods). Pseudotime is a trajectory-inference approach that orders samples by relative biological state rather than chronological time, allowing comparison across individuals at different stages of bNAb development. We found that the immunologic gene signature is elevated only early in breadth development in bNAb producers (when only a limited number of viral strains are neutralized) and later converges to levels comparable to non-bNAb producers as neutralization breadth increases. This aligns with previous studies showing that factors such as CD4 levels inversely correlate with bNAb development only early in infection [4]. However, in our study, both groups had similar CD4 counts at the first time point. Focusing on the genes significantly correlated with breadth in both groups (S9C and S9D Fig, FDR < 0.05) revealed almost the entirety of the MHC class I antigen presentation pathway was anticorrelated with the development of breadth (i.e., pseudotime) and sometimes differentially altered or correlated with the level of cfHIV measured (Fig 4C). Hence, MHC class I antigen genes are highly elevated early in infection in participants who go on to develop bNAbs, and their abundance decrease as breadth develops, becoming comparable to that of participants who do not develop bNAbs. MHC class I antigen presentation is present in almost all cells of the body, presents antigens to be killed by cytotoxic T cells (CD8) but is not generally associated with antigen presenting cells or the production of antibodies by B cells [4]. Taken together, these results indicate that early immune activation associated with MHC class I antigen presentation, independent of viral load or CD4 count, is a relevant factor in the development of bNAbs.

Microbiome associated with bNAbs and the development of breadth

Development of bNAbs has been hypothesized to be triggered by antigen cross-reactivity between HIV and bacterial components present in the gut microbiome [24,25,46,47]. To test whether a differentiated microbiota could be observed in participants who developed bNAbs, we mapped all non-human reads (both cfDNA and cfRNA) to a database of bacterial and viral genomes. After discarding reads that mapped to potential reagent contaminants (S10 Fig, Methods) [48], we identified 249 taxa in cfDNA and 4366 taxa in cfRNA that could be confidently attributed to the participants microbiome. We combined data from all participants to assemble a phylogenetic tree of the human microbiome during HIV infection as measured from circulating nucleic acids (S11 Fig). This accounted for ~70% of non-human reads in both cfDNA and cfRNA, and spanned several areas of the tree of life including archaea, bacteria and both DNA and RNA viruses. Among other features, we identified the presence of several DNA and RNA viruses such as Anelloviruses, Flaviviruses, Herpesviruses indicative of concomitant viral infections for several participants. In cfRNA, we found that ribosomal reads encompass 37% of the nonhost microbiome reads, and we also identified a moderate fraction of reads (~20%) that could not be attributed to specific species and mapped to uncultured bacteria.

We then compared the microbiome of both groups (bNAb producers/controls) finding 19 genera that were differentially abundant between both groups (FDR < 0.05, |log2FC| > 1, Fig 5A). Among genera enriched in bNAb producers, we identified several taxa that include opportunistic pathogens and are commonly associated with mucosal surfaces and the gastrointestinal tract, including Enterobacter, Serratia and Leclercia (S12 Fig). By performing a temporal analysis, we found that the detection of these pathogens does not differ across timepoints and does not precede development of breadth (S13 Fig). Overall, and despite the relatively small sample size of this study, our data shows the existence of differences in the microbiota compositions in participants who develop bNAbs from those who do not.

Fig 5. Microbial differences and GBV-C enrichment in bNAb producers.

Fig 5

A) Volcano plot of differentially abundant microbial genera between bNAb producers and non-producers, colored by phylum (FDR < 0.05, Benjamini-Hochberg). Positive log2 fold change indicates enrichment in bNAb producers. B) Scatterplots showing cfGBV-C abundance in samples partitioned by time and study group (left, bNAbs producers, right, controls). The dashed line indicates the threshold used to define elevated cfGBV-C levels. Using a logistic mixed-effects model with participant as a random effect, bNAb producers showed higher odds of elevated GBV-C levels (odds ratio [OR]: 6.4, 95% CI: 2.3-17.9). In some participants, GBV-C read counts exceeded those of HIV up to 100-fold.

Additionally, metagenomic analysis of cfRNA revealed the presence of GB virus C (GBV-C, previously known as Hepatitis G) in several participants (Fig 5B). GBV-C is a flavivirus that infects lymphocytes, has no known pathology, and has been associated with improved survival in people with HIV infection [49,50]. Given the phylogenetic relationship between GBV-C and hepatitis C virus (HCV), we also examined reads mapping to HCV. However, we did not find evidence of HCV signal, whereas reads mapping to GBV-C showed broad coverage across the viral genome (S13 Fig). In our study, all bNAb producers had non-zero levels of GBV-C (p < 0.05, one-sided Fisher exact test) and showed higher odds of elevated GBV-C levels (Fig 5B). Differential gene abundance analysis on samples with high levels of GBV-C (defined as >1/2000 nonhost reads) revealed 42 host genes that are also elevated in these participants. Among top candidates we identified several genes of the MHC class II antigen presentation complex (HLA-DQA1, HLA-DQB1, HLA-DRB1), an HIV suppressive factor involved in inflammatory signaling (CCL3), several proinflammatory cytokines (IL6, EBI3), and immunoglobulin receptors (KIR3DL2). Notably, increased expression of inhibitory KIRs, including KIR3DL2, has been associated with bNAb development, suggesting a potential link between natural killer cell maturation and the regulation of antibody production [51]. These findings suggest that high levels of GBV-C might be associated with an increased activity of antigen presenting cells.

Discussion

In this work, we demonstrate simultaneous monitoring of viral infection and host response through sequencing of circulating nucleic acids (DNA and RNA) in plasma samples of PLWH prior to ART. We show that cfRNA can be used to characterize circulating HIV viral strains and, at the same time, to profile the host transcriptome, allowing characterization of immune activity during HIV infection. Analysis of fragment length and genome coverage suggests that HIV-derived cfRNA likely reflects fragmented viral RNA circulating in plasma rather than being dominated by intact full-length viral genomes. Subsampling analyses further shows that near-complete genome coverages (>90%) at moderate depths (≥50X) can be recovered from cfRNA, supporting the feasibility of viral genome analyses.

As cfRNA contains transcripts released from diverse tissues and cell types across the body [30], it offers a systemic view of immune processes that may be otherwise difficult to capture. We find that overall cfHIV levels—defined as HIV-derived fragments in cfRNA—correlate with abundance of host genes involved in immune responses to viral infections, including CCL5, TFRC, and CCL2, consistent with responses to peripheral viral burden. For instance, CCL5 encodes RANTES, a chemokine with HIV-suppressive activity that can block CCR5-mediated viral entry, while TFRC (CD71) and CCL2 (MCP-1) are involved in pathways commonly perturbed during infection [52–54]. However, when comparing participants who develop bNAbs with those who did not, we observed a distinct and orthogonal transcriptomic signature suggestive of an enhanced immune activation early in infection that was specific to bNAb producers. This response is independent of viral load and CD4 counts, and is enriched in MHC class I antigen presentation genes, RANTES-associated pathways (though not CCL5 itself), and other immune activation pathways.

Previous studies have highlighted the central role of lymph nodes in shaping humoral responses during HIV infection, yet their direct study remains challenging [23,55,56]. Cell-free RNA may provide a non-invasive window into immune activity occurring in tissue compartments, including lymphoid tissues where Tfh-B cell interactions drive affinity maturation. Hence, it is plausible that the immune-related cfRNA transcripts we report here reflects immune activity within lymph nodes, consistent with previous evidence that cfRNA can capture tissue-resident hematopoietic activity such as bone marrow progenitor transcripts [37]. However, we emphasize that cfRNA represents an integrated signal from multiple sources, and we cannot exclude the possibility that increased cfRNA levels are due to contributions of circulating immune cells or other tissues.

Analysis of non-human reads further allowed us to assess the microbiome and virome, revealing a reduced set of taxa whose presence or absence is associated with the development of bNAbs. Microbial sequences detected in plasma cfRNA/cfDNA most likely reflect leakage of microbial products into circulation due to barrier dysfunction and dysbiosis, a process well described in chronic HIV infection [57–60]. While gut microbiota has been speculated to increase antigen cross-reactivity, our data cannot establish a direct role in bNAb development. Surprisingly, we found that participants who developed bNAbs consistently had higher levels of GBV-C. GBV-C is a non-pathogenic lymphotropic flavivirus that has been associated with improved outcomes in PLWH [49], with some studies suggesting this may reflect an indirect rather than a causal association [50]. Mechanistically, co-infection with GBV-C has been reported to inhibit HIV replication and modulate cytokine release, co-receptor expression, and T-cell homeostasis [50,61,62]. In our study, GBV-C abundance correlated with increased abundance of proinflammatory cytokines and adaptative immune genes, raising the possibility that co-infection influences the host immune milieu in a way that could facilitate bNAb development. While causality cannot be inferred from our dataset, understanding such modulatory viral co-infections early after HIV acquisition could inform strategies to mimic beneficial immune environments in the context of HIV vaccine design.

Our study represents a small pilot investigation (42 samples, 14 PLWH) in a well-characterized cohort of young women from South Africa [63], followed from prior to HIV acquisition into acute infection and before initiation of ART. Biobanked samples from such cohorts are an invaluable resource for studies of factors associated with the development of breadth, as earlier ARV treatment regimens are now standard of care. Although this provides a relatively homogeneous and well-controlled population for an initial discovery study, the small sample size limits generalizability and statistical power. For example, associations between viral load or CD4 counts and bNAb induction reported in other studies may not be detectable here due to the limit size of this pilot cohort. Expanded studies within the CAPRISA cohort or in other settings, and with finer temporal resolution, will be important to validate the findings reported here and their relationship to bNAb development. Despite these limitations, this work illustrates the potential of cfRNA/cfDNA sequencing as a non-invasive approach to interrogate host, viral, and microbial processes simultaneously from a single plasma sample. This approach could be of use to support biomarker discovery in vaccine studies and help identify biological contexts that favor the induction of bNAbs.

Materials and methods

Ethics statement

The CAPRISA 002 Acute Infection study was reviewed and approved by the research ethics committees of the University of KwaZulu-Natal (E013/04), the University of Cape Town (025/2004), and the University of the Witwatersrand (MM040202). All participants provided written informed consent for screening, enrolment and specimen storage.

Participants

The study utilized biobanked plasma specimen of PLWH with known neutralization status enrolled in the CAPRISA cohorts. The CAPRISA 002 Acute Infection cohort was established in 2004 and has been following women from the early/acute stage of infection [63]. We included participants from two CAPRISA clinical sites, Ethekwini in Durban (site 1) and Vulindlela (site 2), and accounted for potential site-specific effects in all downstream analyses.

Neutralization measurements

Neutralization breadth was measured using a standardized pseudovirus assay that included 18 Env-pseudotyped HIV viruses (S1 Table) [64,65]. Briefly, HIV pseudoviruses were produced by co-transfecting HEK293T cells with each HIV-1 Env plasmid and a pSG3 backbone. Serial dilutions of heat-inactivated plasma were preincubated with each pseudovirus, followed by the addition of TZM-bl cells. After 48h, infection was quantified by luciferase read-out, and titers reported as the reciprocal dilution where 50% of the virus is neutralized (IC50). Breadth was defined as the percentage of pseudoviruses neutralized at an IC50 ≥ 1:45.

Participants were classified as bNAb producers if they achieved ≥20% breadth within year of infection or ≥40% breadth against the 18-pseudovirus panel at any three years post-infection.

Isolation of cfDNA and cfRNA and library preparation

Blood samples were collected in sodium citrate tubes and plasma extracted according to the standard clinical centrifugation protocol of the HIV/AIDS Network Coordination, and stored at -80 C. Each plasma sample was thawed once and split into two aliquots for cfDNA/cfRNA extraction (0.4-1 mL per aliquot). To minimize confounding effects from hemolysis and platelet-derived RNA, plasma samples were visually inspected prior to processing for signs of hemolysis (e.g., pink discoloration), which was not visible for any sample. Immediately before nucleic acid extraction, plasma samples were subjected to a final centrifugation step (1 min at 16,000 rpm) to remove any residual debris. cfDNA was extracted using the Qiagen Circulating Nucleic Acid kit. cfRNA was extracted using the Plasma/Serum Circulating RNA and Exosomal Purification kit (Norgen, cat 42800), followed by DNA digestion using Baseline-ZERO DNase (Epicentre) and then cleaned up using the RNA Clean and Concentrator-5 kit (Zymo) [32,66]. cfDNA/cfRNA libraries were prepared using the automated Ovation SP Ultralow Library Systems (Mondrian ST) and the SMARTer Stranded Total RNA Pico Input Mammalian kit (Clontech) respectively [32,66]. A negative control (1 mL PBS, Gibco) was processed in parallel for every batch to control for potential contaminants [67].

Sequencing methods

Paired end Illumina sequencing (2 × 75 bp) was performed on the Nextseq 500 for cfDNA and on the Novaseq 6000 for cfRNA. This resulted in a median of 69 million (cfDNA) and 109 million (cfRNA) reads per sample (S3 Fig).

Sequencing analysis

Sequencing reads from cfRNA and cfDNA were processed using a shared initial pipeline. First, low quality bases and adapter sequences were trimmed by Trimmomatic v0.38 [68], and overlapping read pairs were merged using FLASH v1.2.11 [69]. Merged reads were then aligned to the UniVec database using bowtie2 to remove reads mapping to common vectors, linkers and control sequences [70]. These processing QC steps removed only a fraction of reads in cfRNA and cfDNA libraries from most plasma samples (S2 and S3 Figs). Higher cleaning rates were observed in cfDNA negative controls (~50%) and in three cfRNA plasma samples (~25%), that may reflect lower sample input or quality. Overall, no differences in QC processing were observed between bNAb producers and non-producers.

The remaining reads were first aligned to the human genome (GRCh38, Ensembl) using bowtie2, retaining unmapped reads, which were subsequently passed through a second homology filtering by BLASTN against the NCBI nt database restricted to taxid 9606, to obtain non-human reads for microbial and viral analysis. Taxonomic classification of non-human reads was performed by BLASTN (NCBI BLAST+ v2.8) against the NCBI nt database, using curated taxid lists limited to viral, bacterial and fungal references [66,67]. BLAST hits were filtered by requiring ≥80% nucleotide identity and ≥80% per-read query coverage, also de-duplicating counts to the best-supported accession to avoid counting the same read twice across hits. Subsequently, per-accession counts and coverage profiles were generated and aggregated at taxonomic ranks for downstream analysis.

For cfRNA, reads were additionally aligned to the human genome using STAR v.25.4a [71] and duplicate-filtered using Picard v2.18.2,and gene abundance was quantified using HTSeq [72]. Downstream analysis of gene abundance data is detailed in the following section.

For HIV-specific analysis, non-human cfRNA reads were re-aligned to the HIV-1 group M consensus reference (CON-C.fa) using bowtie2. Consensus sequences were generated with samtools mpileup and bcftools consensus [73], followed by multiple sequence alignment using MAFFT v7 [74]. A maximum-likelihood phylogenetic tree was constructed using IQ-TREE [75], and visualized with Bio.Phylo [76], with metadata-based label coloring and shaded backgrounds highlighting predefined sample clusters. Previously published HIV reference sequences from the same CAPRISA participants were obtained from the Los Alamos National Laboratory (LANL) HIV Sequence Database (S2 Table). Sample identity was primarily validated by assessing HIV sequence consistency across all the three time-points per participant (Figs 2 and S3). As an orthogonal check, we performed HLA-B typing on cfRNA using seq2HLA (v2.3, 4-digit resolution) [77], which confirmed exclusion of one sample (CAP200, time point 1), where a mismatch was detected in both the HIV sequence analysis and HLA-B typing.

To control for potential environmental contamination in both cfDNA and cfRNA data, six negative controls were used to create a database of background environmental taxa that were blacklisted in downstream analysis. This filtering preserved 72% of microbial reads in cfRNA and 17% in cfDNA. In addition, non-template control barcodes (n = 8) were spiked into the cfRNA libraries to estimate sample cross-contamination, which was found to be < 0.5% and compatible with index-hopping during sequencing. This threshold was used to define the detection limit for each individual microbial taxon.

Differential gene abundance analysis

Differential Gene Abundance Analysis was performed using edgeR with a generalized linear model and quasi-likelihood F-tests. Site of sample collection was included as a covariate in the model to account for potential confounding effects. All p-values were corrected for multiple hypothesis testing using the Benjamini-Hochberg method, and results are reported as false discovery rate (FDR). For downstream analysis, we selected genes with FDR < 0.05 and absolute fold change > 1. Associations between gene abundance and clinical metadata (e.g., viral load, CD4 + counts) were measured using nonparametric Spearman correlation unless otherwise stated (SciPy package, Python). Differential abundance of microbial taxa was also assessed in edgeR using generalized linear models, with the same thresholds and covariate adjustments. Overall, no significant site-associated differences were observed in either the transcriptomic or microbial enrichment analyses.

Pseudotime analysis

Pseudotime analysis was used as an alternative approach to investigate factors associated with the development of neutralization breadth, given the limited number of longitudinal samples and heterogeneity in bNAb development across individuals. Samples were ordered according to the number of pseudoviruses neutralized by each plasma sample, which was used as a proxy for progression along the bNAb development continuum. This trajectory-inference approach allows us to account for differences in the pace of bNAb development across individuals. Spearman’s rank correlation analyses were then performed against this pseudotime variable to evaluate changes both between bNAb producers and non-producers and along the continuous trajectory of breadth progression.

GBV-C analysis

GBV-C reads were normalized to the total number of non-human reads per sample and scaled by a factor of 1,000 (GBV-C reads per 1,000 non-human reads), analogous to the normalization used for cfHIV. Differences in GBV-C abundance between bNAb producers and non-producers were assessed using a logistic mixed-effects model (statsmodels package, Python), where we model GBV-C abundance as a binary outcome and included participant as a random effect to account for repeated sampling across time points.

Supporting information

S1 Fig. Clinical information for blood samples.

Longitudinal trajectories for CD4 (orange) and viral load (purple). Diamonds indicate timepoints sequenced in this study. Grey bars indicate values for CD4 count < 200.

(TIF)

ppat.1014066.s001.tif (4.1MB, tif)
S2 Fig. Sequencing quality control metrics per sample.

A) Number of reads (2x75 bp) sequenced in each sample and negative controls. Two cfDNA samples failed extraction/library preparation. B) Percentage of reads cleaned in QC steps (low quality or that align to the UniVec Core database). Percentages are generally very low, except in cfDNA negative controls, which had a high proportion of primer/library-derived sequences. C) Percentage of reads mapping to the human genome. D) Total number of non-human reads, typically more than a million for the cfRNA samples and around 10 thousand for the cfDNA samples. Samples are ordered by time within each participant.

(TIF)

ppat.1014066.s002.tif (8.7MB, tif)
S3 Fig. Sequencing quality control metrics summarized by study group.

A) Total sequencing depth per sample (raw read counts). B) Read counts retained after QC cleaning steps (see Methods) C) Percentage of reads removed during cleaning D) Percentage of human-mapped reads after cleaning. Metrics are shown for cfRNA (left) and cfDNA (right) libraries across bNAb producers, non-producers, and negative controls. Data are shown as mean ± SEM.

(TIF)

ppat.1014066.s003.tif (6.7MB, tif)
S4 Fig. Cell-type deconvolution analysis of cfRNA.

(a) Relative cell-type contributions to the cfRNA composition for each sample. Samples are grouped by participant and study group (bNAb producer or control), and less abundant cell types are aggregated into ‘Other’. (b) Box plots showing the relative abundance of each cell type in bNAb producers and controls. Statistical significance was assessed using a Mann–Whitney U test, with p-values adjusted for multiple hypothesis testing using Benjamini–Hochberg.

(TIF)

ppat.1014066.s004.tif (24.7MB, tif)
S5 Fig. HIV genotypes, coverage and mutations.

A) Heatmap of coverage across the genome in each sample. The four low abundance samples are marked with black squares on the left hand panel. B) Total number of read fragments aligning to HIV in each sample. C) Phylogenetic tree of HIV consensus genotypes for each sample obtained in this study together with 200 other genotypes of HIV-1 group M, subtype C from South Africa. D) Locations of SNVs local to each participant across the three time points. Most cluster within the env gene. E) Mutation rate observed in the samples using differences in the consensus genotypes within participants.

(TIF)

ppat.1014066.s005.tif (42MB, tif)
S6 Fig. Fragment length and coverage of cfHIV reads.

A) Distribution of paired-end genomic spans for reads mapping to the HIV genome, showing a predominance of short fragments (200–600 bp). B) Mean paired-end genomic span as a function of genomic start position, showing no systematic trends that could indicate priming from long RNA templates. (c) Genomic coverage profile across the HIV genome, shown alongside the reference organization of the HIV genome.

(TIF)

ppat.1014066.s006.tif (26.3MB, tif)
S7 Fig. Subsampling analysis of HIV genome coverage from cfRNA.

A) Fraction of the HIV genome covered above increasing minimum depth thresholds (1X-50X) at different subsampling levels of sample CAP261. B) Number of HIV-mapped reads versus mean coverage depth across the HIV genome, showing a linear relationship. C) Genome coverage achieved for each sample in the cohort at ≥10X (left) and ≥50X (right) depth thresholds.

(TIF)

ppat.1014066.s007.tif (21.7MB, tif)
S8 Fig. Correlations of genes with different disease measurements.

A) Distributions of Spearman correlations of all genes with either cfHIV counts (top, blue), CD4 counts (middle, orange), or viral load (bottom, purple). Gray bands indicate regions of significant correlation as determined by the Spearman test on the data. The small ticks (rugplot) at the bottom of each plot indicate known HIV genes. B) Tables of the number of genes strongly correlated/anticorrelated between genes and measurements (as in A). C) Scatter plots of cfHIV abundance vs gene expression (z-score) for the 14 significantly correlated known HIV-associated genes. Besides HLA-A, none of these show strong differences between the two study groups.

(TIF)

S9 Fig. Enriched gene pathways in bNAb producers.

A) Top gene ontologies (molecular function, biological process) and pathways enriched in genes that are elevated in bNAb producers. Numbers indicate the number of selected genes compared with the total number assigned to that category. B) Similar to A, but for the genes with decreased abundance in bNAb producers. Note that scores are much lower for decreased genes than for elevated genes and no clear enriched groups of genes are present in decreased genes. C) Distribution of genes correlated with breadth in bNAb producers (upper, red) and controls (lower, blue). D) Tables of the number of genes meeting correlation thresholds in (C).

(TIF)

ppat.1014066.s009.tif (10.3MB, tif)
S10 Fig. Background microbial sequences in negative controls.

Volcano plot showing differential abundance of microbial genera between plasma samples and negative controls (NC) (FDR < 0.05, Benjamini–Hochberg). Negative log2 fold-change values indicate genera enriched in NC, which were used to define background levels. Several NC-enriched taxa showed large effect sizes (|log2FC| > 4, FDR < 10-8), primarily comprising environmental taxa previously reported in DNA extraction kits and reagents in low-biomass sequencing.

(TIF)

ppat.1014066.s010.tif (4.1MB, tif)
S11 Fig. Phylogenetic tree of microbial sequences from circulating nucleic acids in HIV infection (all participants).

A) Phylogenetic tree of cfDNA-derived microbiome reads pooled across all participants (PLWH who do and do not develop bNAbs). The main branches represent archaea, bacteria and viruses (from left to right); representative viruses including Torque Teno Viruses (TTV), Epstein-Barr virus (EBV) and cytomegalovirus (CMV) are highlighted. B) Phylogenetic tree of cfRNA-derived microbiome reads pooled across all participants. The main branches represent archaea, bacteria and viruses (from left to right). Within the viral branch, HIV and GB virus C are the main contributors to the virome.

(TIF)

ppat.1014066.s011.tif (9.9MB, tif)
S12 Fig. Subset of microbial genera enriched in bNAb producers.

Boxplots showing normalized abundances of selected microbial genera across negative controls (NC), bNAb non-producers, and bNAb producers. Values are shown as log2-transformed abundances relative to the median abundance in negative controls.

(TIF)

ppat.1014066.s012.tif (3.8MB, tif)
S13 Fig. Pseudo-temporal dynamics in differentially abundant genera between bNAb producers and non-producers.

Scatter plots of the abundance (z-score) vs breadth production (%) for the 22 genera found to be differentially abundant between case (bNAbs) and control (no bNAbs). No apparent time differences found.

(TIF)

ppat.1014066.s013.tif (14.2MB, tif)
S14 Fig. Coverage profiles of HCV and GBV-C genomes across samples.

A) Read coverage across the HCV genome. B) Read coverage across the GBV-C genome. Coverage for individual samples is shown in red (HCV) and blue (GBV-C); mean coverage is shown in black.

(TIFF)

ppat.1014066.s014.tiff (9.9MB, tiff)
S1 Table. Neutralization data for the panel of 18 Env-pseudotyped HIV viruses.

(XLSX)

ppat.1014066.s015.xlsx (17.1KB, xlsx)
S2 Table. Participant codes and corresponding accession numbers in the HIV database (https://www.hiv.lanl.gov/content/index) for the reference FASTA sequences included in this study.

(XLSX)

ppat.1014066.s016.xlsx (9.4KB, xlsx)

Acknowledgments

Finally, we are indebted to the study participants for their generous support of scientific research.

Data Availability

The datasets generated and analyzed in the study are available in the NCBI Gene Expression Omnibus (GEO) and Sequence Read Archive (SRA) with accession number GSE313190. The code to reproduce the analyses in this study is publicly available at https://github.com/CamunasLab/cfHIV.

Funding Statement

This work was supported by the funding from the Chan Zuckerberg Biohub to S.R.Q, the Bill & Melinda Gates Foundation (OPP1113682) and the Global Health Vaccine Accelerator Platform (GH-VAP-SI-ID-14) to S.R.Q and P.L.M. P.L.M. is supported by South African Medical Research Council (MRC) SHIP program and the South African Research Chairs Initiative of the Department of Science and Technology and the NRF (Grant No 98341). J.C.-S. is supported by the Knut and Alice Wallenberg Foundation (Wallenberg Molecular Medicine Fellow, the Swedish Research Council (2021-05109) and the Erling Perssons Stiftelse (Swedish Foundations’ Starting Grant). M.N.M. was supported by the Stanford Bio-X Bowes Fellowship. The funders had no role in study design, data collection and analysis, decision to publish, or preparation of the manuscript.

References

  • 1.UNAIDS Report 2017. Ending AIDS. Progress towards the 90-90-90 targets. (Global Aids Update 2017). Joint United Nations Programme on HIV/AIDS (UNAIDS); 2017. [Google Scholar]
  • 2.Deeks SG, Overbaugh J, Phillips A, Buchbinder S. HIV infection. Nat Rev Dis Primers. 2015;1:15035. doi: 10.1038/nrdp.2015.35 [DOI] [PubMed] [Google Scholar]
  • 3.Moore PL, Gray ES, Choge IA, Ranchobe N, Mlisana K, Abdool Karim SS, et al. The c3-v4 region is a major target of autologous neutralizing antibodies in human immunodeficiency virus type 1 subtype C infection. J Virol. 2008;82(4):1860–9. doi: 10.1128/JVI.02187-07 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 4.Gray ES, Moore PL, Choge IA, Decker JM, Bibollet-Ruche F, Li H, et al. Neutralizing antibody responses in acute human immunodeficiency virus type 1 subtype C infection. J Virol. 2007;81(12):6187–96. doi: 10.1128/JVI.00239-07 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 5.Li B, Decker JM, Johnson RW, Bibollet-Ruche F, Wei X, Mulenga J, et al. Evidence for potent autologous neutralizing antibody titers and compact envelopes in early infection with subtype C human immunodeficiency virus type 1. J Virol. 2006;80(11):5211–8. doi: 10.1128/JVI.00201-06 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 6.Richman DD, Wrin T, Little SJ, Petropoulos CJ. Rapid evolution of the neutralizing antibody response to HIV type 1 infection. Proc Natl Acad Sci U S A. 2003;100(7):4144–9. doi: 10.1073/pnas.0630530100 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 7.Landais E, Huang X, Havenar-Daughton C, Murrell B, Price MA, Wickramasinghe L, et al. Broadly neutralizing antibody responses in a large longitudinal sub-Saharan HIV primary infection cohort. PLoS Pathog. 2016;12(1):e1005369. doi: 10.1371/journal.ppat.1005369 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 8.Corey L, Gilbert PB, Juraska M, Montefiori DC, Morris L, Karuna ST, et al. Two randomized trials of neutralizing antibodies to prevent HIV-1 acquisition. N Engl J Med. 2021;384:1003–14. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 9.Caniels TG, Prabhakaran M, Ozorowski G, MacPhee KJ, Wu W, van der Straten K, et al. Precise targeting of HIV broadly neutralizing antibody precursors in humans. Science. 2025;389(6759):eadv5572. doi: 10.1126/science.adv5572 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 10.Willis JR, Prabhakaran M, Muthui M, Naidoo A, Sincomb T, Wu W, et al. Vaccination with mRNA-encoded nanoparticles drives early maturation of HIV bnAb precursors in humans. Science. 2025;389(6759):eadr8382. doi: 10.1126/science.adr8382 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 11.Leggat DJ, Cohen KW, Willis JR, Fulp WJ, deCamp AC, Kalyuzhniy O, et al. Vaccination induces HIV broadly neutralizing antibody precursors in humans. Science. 2022;378(6623):eadd6502. doi: 10.1126/science.add6502 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 12.Rusert P, Kouyos RD, Kadelka C, Ebner H, Schanz M, The Swiss HIV Cohort Study. Determinants of HIV-1 broadly neutralizing antibody induction. Nat Med. 2016;22:1260–7. [DOI] [PubMed] [Google Scholar]
  • 13.Moody MA, Pedroza-Pacheco I, Vandergrift NA, Chui C, Lloyd KE, Parks R, et al. Immune perturbations in HIV-1-infected individuals who make broadly neutralizing antibodies. Sci Immunol. 2016;1(1):aag0851. doi: 10.1126/sciimmunol.aag0851 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 14.Haynes BF, Shaw GM, Korber B, Kelsoe G, Sodroski J, Hahn BH, et al. HIV-host interactions: implications for vaccine design. Cell Host Microbe. 2016;19(3):292–303. doi: 10.1016/j.chom.2016.02.002 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 15.Roskin KM, Jackson KJL, Lee J-Y, Hoh RA, Joshi SA, Hwang K-K, et al. Aberrant B cell repertoire selection associated with HIV neutralizing antibody breadth. Nat Immunol. 2020;21(2):199–209. doi: 10.1038/s41590-019-0581-0 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 16.Landais E, Moore PL. Development of broadly neutralizing antibodies in HIV-1 infected elite neutralizers. Retrovirology. 2018;15(1):61. doi: 10.1186/s12977-018-0443-0 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 17.Euler Z, van Gils MJ, Bunnik EM, Phung P, Schweighardt B, Wrin T, et al. Cross-reactive neutralizing humoral immunity does not protect from HIV type 1 disease progression. J Infect Dis. 2010;201(7):1045–53. doi: 10.1086/651144 [DOI] [PubMed] [Google Scholar]
  • 18.Gray ES, Madiga MC, Hermanus T, Moore PL, Wibmer CK, Tumba NL, et al. The neutralization breadth of HIV-1 develops incrementally over four years and is associated with CD4+ T cell decline and high viral load during acute infection. J Virol. 2011;85(10):4828–40. doi: 10.1128/JVI.00198-11 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 19.Cottrell CA, Hu X, Lee JH, Skog P, Luo S, Flynn CT, et al. Heterologous prime-boost vaccination drives early maturation of HIV broadly neutralizing antibody precursors in humanized mice. Sci Transl Med. 2024;16(748):eadn0223. doi: 10.1126/scitranslmed.adn0223 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 20.Cohen K, Altfeld M, Alter G, Stamatatos L. Early preservation of CXCR5+ PD-1+ helper T cells and B cell activation predict the breadth of neutralizing antibody responses in chronic HIV-1 infection. J Virol. 2014;88(22):13310–21. doi: 10.1128/JVI.02186-14 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 21.Locci M, Havenar-Daughton C, Landais E, Wu J, Kroenke MA, Arlehamn CL. Human circulating PD-1 CXCR3−CXCR5 memory tfh cells are highly functional and correlate with broadly neutralizing HIV antibody responses. Immunity. 2013;39:758–69. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 22.Perreau M, Savoye A-L, De Crignis E, Corpataux J-M, Cubas R, Haddad EK, et al. Follicular helper T cells serve as the major CD4 T cell compartment for HIV-1 infection, replication, and production. J Exp Med. 2013;210(1):143–56. doi: 10.1084/jem.20121932 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 23.Crotty S. T follicular helper cell biology: a decade of discovery and diseases. Immunity. 2019;50:1132–48. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 24.Liao H-X, Chen X, Munshaw S, Zhang R, Marshall DJ, Vandergrift N, et al. Initial antibodies binding to HIV-1 gp41 in acutely infected subjects are polyreactive and highly mutated. J Exp Med. 2011;208(11):2237–49. doi: 10.1084/jem.20110363 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 25.Trama AM, Moody MA, Alam SM, Jaeger FH, Lockwood B, Parks R, et al. HIV-1 envelope gp41 antibodies can originate from terminal ileum B cells that share cross-reactivity with commensal bacteria. Cell Host Microbe. 2014;16(2):215–26. doi: 10.1016/j.chom.2014.07.003 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 26.Williams WB, Liao H-X, Moody MA, Kepler TB, Alam SM, Gao F, et al. HIV-1 VACCINES. Diversion of HIV-1 vaccine-induced immunity by gp41-microbiota cross-reactive antibodies. Science. 2015;349(6249):aab1253. doi: 10.1126/science.aab1253 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 27.Loy C, Ahmann L, De Vlaminck I, Gu W. Liquid biopsy based on cell-free DNA and RNA. Annu Rev Biomed Eng. 2024;26(1):169–95. doi: 10.1146/annurev-bioeng-110222-111259 [DOI] [PubMed] [Google Scholar]
  • 28.Moufarrej MN, Bianchi DW, Shaw GM, Stevenson DK, Quake SR. Noninvasive prenatal testing using circulating DNA and RNA: advances, challenges, and possibilities. Annu Rev Biomed Data Sci. 2023;6:397–418. doi: 10.1146/annurev-biodatasci-020722-094144 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 29.Koh W, Pan W, Gawad C, Fan HC, Kerchner GA, Wyss-Coray T, et al. Noninvasive in vivo monitoring of tissue-specific global gene expression in humans. Proc Natl Acad Sci U S A. 2014;111(20):7361–6. doi: 10.1073/pnas.1405528111 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 30.Vorperian SK, Moufarrej MN, Tabula Sapiens Consortium, Quake SR. Cell types of origin of the cell-free transcriptome. Nat Biotechnol. 2022;40(6):855–61. doi: 10.1038/s41587-021-01188-9 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 31.Loy CJ, Servellita V, Sotomayor-Gonzalez A, Bliss A, Lenz JS, Belcher E, et al. Plasma cell-free RNA signatures of inflammatory syndromes in children. Proc Natl Acad Sci U S A. 2024;121(37):e2403897121. doi: 10.1073/pnas.2403897121 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 32.Pan W, Ngo TTM, Camunas-Soler J, Song C-X, Kowarsky M, Blumenfeld YJ, et al. Simultaneously monitoring immune response and microbial infections during pregnancy through plasma cfRNA sequencing. Clin Chem. 2017;63(11):1695–704. doi: 10.1373/clinchem.2017.273888 [DOI] [PubMed] [Google Scholar]
  • 33.Camunas-Soler J, Gee EPS, Reddy M, Mi JD, Thao M, Brundage T, et al. Predictive RNA profiles for early and very early spontaneous preterm birth. Am J Obstet Gynecol. 2022;227(1):72.e1-72.e16. doi: 10.1016/j.ajog.2022.04.002 [DOI] [PubMed] [Google Scholar]
  • 34.Ngo TTM, Moufarrej MN, Rasmussen M-LH, Camunas-Soler J, Pan W, Okamoto J, et al. Noninvasive blood tests for fetal development predict gestational age and preterm delivery. Science. 2018;360(6393):1133–6. doi: 10.1126/science.aar3819 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 35.Rasmussen M, Reddy M, Nolan R, Camunas-Soler J, Khodursky A, Scheller NM, et al. RNA profiles reveal signatures of future health and disease in pregnancy. Nature. 2022;601(7893):422–7. doi: 10.1038/s41586-021-04249-w [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 36.Moufarrej MN, Vorperian SK, Wong RJ, Campos AA, Quaintance CC, Sit RV, et al. Early prediction of preeclampsia in pregnancy with cell-free RNA. Nature. 2022;602(7898):689–94. doi: 10.1038/s41586-022-04410-z [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 37.Ibarra A, Zhuang J, Zhao Y, Salathia NS, Huang V, Acosta AD, et al. Non-invasive characterization of human bone marrow stimulation and reconstitution by cell-free messenger RNA sequencing. Nat Commun. 2020;11(1):400. doi: 10.1038/s41467-019-14253-4 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 38.Chang A, Loy CJ, Eweis-LaBolle D, Lenz JS, Steadman A, Andgrama A, et al. Circulating cell-free RNA in blood as a host response biomarker for detection of tuberculosis. Nat Commun. 2024;15(1):4949. doi: 10.1038/s41467-024-49245-6 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 39.Burton DR, Stanfield RL, Wilson IA. Antibody vs. HIV in a clash of evolutionary titans. Proc Natl Acad Sci U S A. 2005;102(42):14943–8. doi: 10.1073/pnas.0505126102 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 40.Haddox HK, Dingens AS, Bloom JD. Experimental estimation of the effects of all amino-acid mutations to HIV’s envelope protein on viral replication in cell culture. PLoS Pathog. 2016;12(12):e1006114. doi: 10.1371/journal.ppat.1006114 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 41.Burton DR, Mascola JR. Antibody responses to envelope glycoproteins in HIV-1 infection. Nat Immunol. 2015;16(6):571–6. doi: 10.1038/ni.3158 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 42.Cuevas JM, Geller R, Garijo R, López-Aldeguer J, Sanjuán R. Extremely high mutation rate of HIV-1 in vivo. PLoS Biol. 2015;13(9):e1002251. doi: 10.1371/journal.pbio.1002251 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 43.Zanini F, Puller V, Brodin J, Albert J, Neher RA. In vivo mutation rates and the landscape of fitness costs of HIV-1. Virus Evol. 2017;3. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 44.Meyaard L, Hovenkamp E, Keet IP, Hooibrink B, de Jong IH, Otto SA, et al. Single cell analysis of IL-4 and IFN-gamma production by T cells from HIV-infected individuals: decreased IFN-gamma in the presence of preserved IL-4 production. J Immunol. 1996;157(6):2712–8. doi: 10.4049/jimmunol.157.6.2712 [DOI] [PubMed] [Google Scholar]
  • 45.Dugast A-S, Arnold K, Lofano G, Moore S, Hoffner M, Simek M, et al. Virus-driven inflammation is associated with the development of bNAbs in spontaneous controllers of HIV. Clin Infect Dis. 2017;64(8):1098–104. doi: 10.1093/cid/cix057 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 46.Mascola JR, Haynes BF. HIV-1 neutralizing antibodies: understanding nature’s pathways. Immunol Rev. 2013;254(1):225–44. doi: 10.1111/imr.12075 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 47.Moir S, Fauci AS. B-cell responses to HIV infection. Immunol Rev. 2017;275(1):33–48. doi: 10.1111/imr.12502 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 48.Salter SJ, Cox MJ, Turek EM, Calus ST, Cookson WO, Moffatt MF, et al. Reagent and laboratory contamination can critically impact sequence-based microbiome analyses. BMC Biol. 2014;12:87. doi: 10.1186/s12915-014-0087-z [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 49.Xiang J, Wünschmann S, Diekema DJ, Klinzman D, Patrick KD, George SL, et al. Effect of coinfection with GB virus C on survival among patients with HIV infection. N Engl J Med. 2001;345(10):707–14. doi: 10.1056/NEJMoa003364 [DOI] [PubMed] [Google Scholar]
  • 50.Schwarze-Zander C, Blackard JT, Rockstroh JK. Role of GB virus C in modulating HIV disease. Expert Rev Anti Infect Ther. 2012;10(5):563–72. doi: 10.1586/eri.12.37 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 51.Bradley T, Peppa D, Pedroza-Pacheco I, Li D, Cain DW, Henao R, et al. RAB11FIP5 expression and altered natural killer cell function are associated with induction of HIV broadly neutralizing antibody responses. Cell. 2018;175(2):387-399.e17. doi: 10.1016/j.cell.2018.08.064 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 52.Mengozzi M, De Filippi C, Transidico P, Biswas P, Cota M, Ghezzi S, et al. Human immunodeficiency virus replication induces monocyte chemotactic protein-1 in human macrophages and U937 promonocytic cells. Blood. 1999;93(6):1851–7. doi: 10.1182/blood.v93.6.1851.406k12_1851_1857 [DOI] [PubMed] [Google Scholar]
  • 53.Packard TA, Schwarzer R, Herzig E, Rao D, Luo X, Egedal JH, et al. CCL2: a chemokine potentially promoting early seeding of the latent HIV reservoir. mBio. 2022;13(5):e0189122. doi: 10.1128/mbio.01891-22 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 54.Madrid R, Janvier K, Hitchin D, Day J, Coleman S, Noviello C, et al. Nef-induced alteration of the early/recycling endosomal compartment correlates with enhancement of HIV-1 infectivity. J Biol Chem. 2005;280(6):5032–44. doi: 10.1074/jbc.M401202200 [DOI] [PubMed] [Google Scholar]
  • 55.Cirelli KM, Carnathan DG, Nogal B, Martin JT, Rodriguez OL, Upadhyay AA, et al. Slow delivery immunization enhances HIV neutralizing antibody and germinal center responses via modulation of immunodominance. Cell. 2019;177(5):1153-1171.e28. doi: 10.1016/j.cell.2019.04.012 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 56.Lindqvist M, van Lunzen J, Soghoian DZ, Kuhl BD, Ranasinghe S, Kranias G, et al. Expansion of HIV-specific T follicular helper cells in chronic HIV infection. J Clin Invest. 2012;122(9):3271–80. doi: 10.1172/JCI64314 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 57.Alzahrani J, Hussain T, Simar D, Palchaudhuri R, Abdel-Mohsen M, Crowe SM, et al. Inflammatory and immunometabolic consequences of gut dysfunction in HIV: parallels with IBD and implications for reservoir persistence and non-AIDS comorbidities. EBioMedicine. 2019;46:522–31. doi: 10.1016/j.ebiom.2019.07.027 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 58.Das Adhikari U, Froehle LM, Pipkin AN, Baharlou H, Linder AH, Shah P, et al. Immunometabolic defects of CD8+ T cells disrupt gut barrier integrity in people with HIV. Cell. 2025;188(20):5666-5679.e19. doi: 10.1016/j.cell.2025.08.024 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 59.Brenchley JM, Price DA, Schacker TW, Asher TE, Silvestri G, Rao S, et al. Microbial translocation is a cause of systemic immune activation in chronic HIV infection. Nat Med. 2006;12(12):1365–71. doi: 10.1038/nm1511 [DOI] [PubMed] [Google Scholar]
  • 60.Nazli A, Chan O, Dobson-Belaire WN, Ouellet M, Tremblay MJ, Gray-Owen SD, et al. Exposure to HIV-1 directly impairs mucosal epithelial barrier integrity allowing microbial translocation. PLoS Pathog. 2010;6(4):e1000852. doi: 10.1371/journal.ppat.1000852 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 61.Stapleton JT. Human pegivirus type 1: a common human virus that is beneficial in immune-mediated disease? Front Immunol. 2022;13:887760. doi: 10.3389/fimmu.2022.887760 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 62.Acchioni C, Sandini S, Acchioni M, Sgarbanti M. Co-infections and superinfections between HIV-1 and other human viruses at the cellular level. Pathogens. 2024;13:349. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 63.van Loggerenberg F, Mlisana K, Williamson C, Auld SC, Morris L, Gray CM, et al. Establishing a cohort at high risk of HIV infection in South Africa: challenges and experiences of the CAPRISA 002 acute infection study. PLoS One. 2008;3(4):e1954. doi: 10.1371/journal.pone.0001954 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 64.Montefiori DC. Evaluating neutralizing antibodies against HIV, SIV, and SHIV in luciferase reporter gene assays. In: CP in immunology, vol. 64; 2005. p. 12.11.1-12.11.17. [DOI] [PubMed] [Google Scholar]
  • 65.Moyo-Gwete T, Scheepers C, Makhado Z, Kgagudi P, Mzindle NB, Ziki R, et al. Enhanced neutralization potency of an identical HIV neutralizing antibody expressed as different isotypes is achieved through genetically distinct mechanisms. Sci Rep. 2022;12(1):16473. doi: 10.1038/s41598-022-20141-7 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 66.De Vlaminck I, Khush KK, Strehl C, Kohli B, Luikart H, Neff NF, et al. Temporal response of the human virome to immunosuppression and antiviral therapy. Cell. 2013;155(5):1178–87. doi: 10.1016/j.cell.2013.10.034 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 67.Kowarsky M, Camunas-Soler J, Kertesz M, De Vlaminck I, Koh W, Pan W, et al. Numerous uncharacterized and highly divergent microbes which colonize humans are revealed by circulating cell-free DNA. Proc Natl Acad Sci U S A. 2017;114(36):9623–8. doi: 10.1073/pnas.1707009114 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 68.Bolger AM, Lohse M, Usadel B. Trimmomatic: a flexible trimmer for Illumina sequence data. Bioinformatics. 2014;30(15):2114–20. doi: 10.1093/bioinformatics/btu170 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 69.Magoč T, Salzberg SL. FLASH: fast length adjustment of short reads to improve genome assemblies. Bioinformatics. 2011;27(21):2957–63. doi: 10.1093/bioinformatics/btr507 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 70.Langmead B, Salzberg SL. Fast gapped-read alignment with Bowtie 2. Nat Methods. 2012;9(4):357–9. doi: 10.1038/nmeth.1923 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 71.Dobin A, Davis CA, Schlesinger F, Drenkow J, Zaleski C, Jha S, et al. STAR: ultrafast universal RNA-seq aligner. Bioinformatics. 2013;29(1):15–21. doi: 10.1093/bioinformatics/bts635 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 72.Putri GH, Anders S, Pyl PT, Pimanda JE, Zanini F. Analysing high-throughput sequencing data in Python with HTSeq 2.0. Bioinformatics. 2022;38(10):2943–5. doi: 10.1093/bioinformatics/btac166 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 73.Danecek P, Bonfield JK, Liddle J, Marshall J, Ohan V, Pollard MO, et al. Twelve years of SAMtools and BCFtools. Gigascience. 2021;10(2):giab008. doi: 10.1093/gigascience/giab008 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 74.Katoh K, Standley DM. MAFFT multiple sequence alignment software version 7: improvements in performance and usability. Mol Biol Evol. 2013;30(4):772–80. doi: 10.1093/molbev/mst010 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 75.Minh BQ, Schmidt HA, Chernomor O, Schrempf D, Woodhams MD, von Haeseler A, et al. Corrigendum to: IQ-TREE 2: new models and efficient methods for phylogenetic inference in the genomic era. Mol Biol Evol. 2020;37(8):2461. doi: 10.1093/molbev/msaa131 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 76.Talevich E, Invergo BM, Cock PJA, Chapman BA. Bio.Phylo: a unified toolkit for processing, analyzing and visualizing phylogenetic trees in Biopython. BMC Bioinform. 2012;13:209. doi: 10.1186/1471-2105-13-209 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 77.Boegel S, Löwer M, Schäfer M, Bukur T, de Graaf J, Boisguérin V, et al. HLA typing from RNA-Seq sequence reads. Genome Med. 2012;4(12):102. doi: 10.1186/gm403 [DOI] [PMC free article] [PubMed] [Google Scholar]

Decision Letter 0

Vijayakumar Velu

10 Dec 2025

Cell-free RNA reveals host and microbial correlates of broadly neutralizing antibody development against HIV

PLOS Pathogens

Dear Dr. Camunas Soler,

Thank you for submitting your manuscript to PLOS Pathogens. After careful consideration, we feel that it has merit but does not fully meet PLOS Pathogens's publication criteria as it currently stands. Therefore, we invite you to submit a revised version of the manuscript that addresses the points raised during the review process.

Please submit your revised manuscript by Feb 08 2026 11:59PM. If you will need more time than this to complete your revisions, please reply to this message or contact the journal office at plospathogens@plos.org. When you're ready to submit your revision, log on to https://www.editorialmanager.com/ppathogens/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.

Please include the following items when submitting your revised manuscript:

* A rebuttal letter that responds to each point raised by the editor and reviewer(s). You should upload this letter as a separate file labeled 'Response to Reviewers'. This file does not need to include responses to any formatting updates and technical items listed in the 'Journal Requirements' section below.

* A marked-up copy of your manuscript that highlights changes made to the original version. You should upload this as a separate file labeled 'Revised Manuscript with Track Changes'.

* An unmarked version of your revised paper without tracked changes. You should upload this as a separate file labeled 'Manuscript'.

If you would like to make changes to your financial disclosure, competing interests statement, or data availability statement, please make these updates within the submission form at the time of resubmission. Guidelines for resubmitting your figure files are available below the reviewer comments at the end of this letter.

We look forward to receiving your revised manuscript.

Kind regards,

Vijayakumar Velu, ph.D.

Academic Editor

PLOS Pathogens

Susan Ross

Section Editor

PLOS Pathogens

Editor-in-Chief

PLOS Pathogens

orcid.org/0000-0003-2946-9497

Editor-in-Chief

PLOS Pathogens

orcid.org/0000-0002-7699-2064

Journal Requirements:

1) We ask that a manuscript source file is provided at Revision. Please upload your manuscript file as a .doc, .docx, .rtf or .tex. If you are providing a .tex file, please upload it under the item type u2018LaTeX Source Fileu2019 and leave your .pdf version as the item type u2018Manuscriptu2019.

2) Please upload all main figures as separate Figure files in .tif or .eps format. For more information about how to convert and format your figure files please see our guidelines:

https://journals.plos.org/plospathogens/s/figures

3) We have noticed that you have uploaded Supporting Information files, but you have not included a list of legends. Please add a full list of legends for your Supporting Information files after the references list.

4) In the online submission form, you indicated that "The datasets generated and analyzed in the study are available in the NCBI Gene Expression Omnibus (GEO) and Sequence Read Archive (SRA) and can be accessed upon request". All PLOS journals now require all data underlying the findings described in their manuscript to be freely available to other researchers, either

1. In a public repository

2. Within the manuscript itself

3. Uploaded as supplementary information.

This policy applies to all data except where public deposition would breach compliance with the protocol approved by your research ethics board. If your data cannot be made publicly available for ethical or legal reasons (e.g., public availability would compromise patient privacy), please explain your reasons by return email and your exemption request will be escalated to the editor for approval. Your exemption request will be handled independently and will not hold up the peer review process, but will need to be resolved should your manuscript be accepted for publication. One of the Editorial team will then be in touch if there are any issues.

Reviewers' Comments:

Reviewer's Responses to Questions

Part I - Summary

Please use this section to discuss strengths/weaknesses of study, novelty/significance, general execution and scholarship.

Reviewer #1: This is a pilot study of the use of cell free RNA and DNA measurements from a plasma repository of the extraordinary CAPRISA cohort. This is an admirable attempt to bring new technology to an important problem, that of defining the mechanisms of bnAb induction in PLWH. The use of statistics throughout is not consistent for each set of data.

Reviewer #2: Kowarsky et al. sampled cell-free nucleic acids longitudinally from HIV cases before antiviral therapy with the goal of finding differences between people who (go on to) develop broadly neutralising antibodies and people who do not. The hope is to find actionable environmental factors that can be tinkered with for vaccine design. Overall, the find multiple associations between HIV, host immune, and commensal or co-infecting microbial sectors. Perhaps the most interesting discovery is that during the first year after (presumed) infection, patients that will later develop broad neutralisation already show a stronger MHC-I genetic signature in their cell-free nucleic acids. They also find that bNAb producers tend to have higher relative abundance of GBV-C reads, potentially implicating an “immune priming” mechanism for early response to HIV.

In general, the study appears to have been conducted thoroughly and the cohort is really interesting for vaccine design purposes. Although only one main experimental method was used throughout, the computational analyses are diverse and informative. My impression is that, although already a useful reference as is, this manuscript could improve significantly by providing a few, relatively straightforward additions to clarify both technical and biomedical aspects of the findings. See below for specific comments.

**********

Part II – Major Issues: Key Experiments Required for Acceptance

Please use this section to detail the key new experiments or modifications of existing experiments that should be absolutely required to validate study conclusions.

Generally, there should be no more than 3 such required experiments or major modifications for a "Major Revision" recommendation. If more than 3 experiments are necessary to validate the study conclusions, then you are encouraged to recommend "Reject".

Reviewer #1: 1. Success in cfRNA, cfDNA interpretation is control for RBC hemolysis and for release of platelet RNA. How were platelet and RBC contamination controlled for in this study? what percentate of RNA was from platelet or RBCs? In a paper by the senior author’s group (PMID: 35132263) 60% of cfRNA was platelet + RBC derived.

2. That this study did not find an association between high viral load and CD4 counts with bnAb induction compared to other studies may only reflect the small sample size in this study. The authors did note that this was a pilot study and conclusions are limited.

3. In Figure 1A the difference in those with bnAbs and those with fewer bnAbs is not well demarcated. In 1A the lowest breadth % of the bnAb group and the highest breadth % of the low neutralizing group overlap at 3 years of follow up. That is unlike other such studies, as here, there is no clean break between the groups.

4. In Figure 1B, it is unclear what the boxes mean, and what log 10 affinity means in the context of pseudovirus neutralization assays. Does affinity correlate with EC50 or EC80?

5. Figure 3A, it is unclear what test was used for the p values shown.

6. Figure 4, correlations are presented but no tests are indicated how they were obtained.

7. In Figure 5 there does not appear to be statistically significant differences in GBV-C levels in the two groups. No statistical methods noted to analyze the data.

8. In Figure S2 there appear to be as many sequences in the water control as in the samples. Is this usual? How are relevant cf genetic sequences identified with such high background of contamination?

9. Also in Figure 2 why were so many contaminating sequences “cleaned” from the bnAb group but not in the control group?

10. In Figure S3, was the RNA outside of cells and outside of virions or does this sequencing simply represent virion-associated RNA? Were the samples pelleted by ultracentrifugation to remove intact virions?

11. Figure S6 is hard to understand. Are these data from only the PLWH who made high bnAbs? Or both groups?

12. Figure S7. What are pseudo-temporal dynamics? Definition would help the reader.

13. Several uses of the word “significantly” (as in line 193) with no reference to statistical test or giving a P value.

14. Similarly the word “correlated” or “anti-correlated” was used (as in lines 209-210) with no correlation coefficient or statistical test described.

Reviewer #2: • Data should be on GEO with identifier, and code should be available e.g. on GitHub.

• Although the authors demonstrate that they can reconstruct HIV genomes using cell-free nucleic acids, it is unclear to me how many reads they need to achieve what kind of assembly quality. Some downsampling to get a sense of how many cfRNA reads you would need to assemble an HIV genome to what degree/quality would really help future researchers tailor their experimental design.

• On line 170, the authors state that CCL5 etc. are responses to HIV infection. Because these genes have very broad immune functions, the authors should explore the immunology literature to clarify whether these genes are responses not just shared between HIV bNAb groups, but also with patients infected by other viruses or even bacteria. The authors should comment on this issue based on a literature search and potentially cross-examine third-party data from previous publications to check the specificity of this signature.

• The authors mention GBV-C but do not elaborate whether other flaviruses were also found. HCV-HIV coinfection is not that uncommon and would cause a quite different immunological response, therefore the authors should clarify whether they specifically see Hepatitis C reads in their samples.

• Did the pseudovirus neutralisation assay in Figure 1B for non-bNAb producers match the strain that was found in their cfRNA reads? In other words, can their serum at least neutralise narrowly their own strain?

**********

Part III – Minor Issues: Editorial and Data Presentation Modifications

Please use this section for editorial suggestions as well as relatively minor modifications of existing data that would enhance clarity.

Reviewer #1: minor comments:

15. The participants in the CAPRISA cohort study are referred to as either “donors” or “patients” in the paper (line 235). Perhaps the generic term “participants” might be more consistent.

16. Vague phrases throughout. Line 43, “activation that involves MHC class I antigen presentation”; line 62 “changes in commensal microbiota and increased GB virus C co-infection” (increased levels or increased number of participants?); GB virus C is mentioned several times in the paper before it is defined in line 251. It should be defined early on in the paper.

Reviewer #2: The first sentence of the abstract is not strictly true: a narrow response is ok to prevent viral infections if the strain matches. Hence the yellow fever vaccine. Rephrasing is needed.

Not sure CCL5 etc. are “covariates of HIV infection progression” as much as they are proxies of high viral titers in the periphery. The two things are not quite the same as progression is clinically defined based on CD4 counts rather than viral load.

The authors dutifully cite many of the software packages used for the analyses but miss others. For instance, they presumably used HTSeq 2.0 but fail to cite it. They should amend this inconsistency.

The discussion section about lymph nodes is highly speculative and should be tempered to avoid giving the impression that this manuscript provides hard evidence that specific signature found herein can be attributed unequivocally to lymph node processes.

In addition to the above, I have a specific suggestion that, albeit not 100% necessary for the manuscript, would elevate the conclusions of the paper and potentially shed light on new biology:

• The correlation between higher MHC-I production transcripts and the potential to later develop bNAbs is very interesting. MHC-I presentation is usually seen as an immune pioneering process that then activates cytotoxic T cells (CTLs), thereby contributing to the early stages of the immune activation cascade. However, some viral proteins (e.g. HIV Nef https://www.nature.com/articles/nri2575) can downregulate MHC-I presentation as an escape mechanism from CTLs. These data suggests that perhaps in bNAb producers, early inhibition of MHC-I by viral factors is less effective, buying enough time for the immune system to take the bNAb route. If so, anti-Nef drugs or a Nef-mutant viral strain would be an interesting direction for vaccine efforts. The authors could test the hypothesis that patients with stronger MHC-I production early on (before returning to baseline) also have relatively fewer Nef reads (compared to the total of HIV reads) and/or mutations in Nef that might decrease its effectiveness (i.e. nonsynonymous mutations in relatively conserved residues of the protein). Potentially, the authors could also take the Nef sequences from samples with high vs low MHC-I production early on and predict structures using AlphaFold 3 (online server) to test the hypothesis that these specific Nef strains might differ structurally from the ones derived from non bNAbs producers.

**********

PLOS authors have the option to publish the peer review history of their article (what does this mean? ). If published, this will include your full peer review and any attached files.

If you choose “no”, your identity will remain anonymous but your review may still be made public.

Do you want your identity to be public for this peer review? For information about this choice, including consent withdrawal, please see our Privacy Policy .

Reviewer #1: No

Reviewer #2: Yes: Fabio Zanini

[NOTE: If reviewer comments were submitted as an attachment file, they will be attached to this email and accessible via the submission site. Please log into your account, locate the manuscript record, and check for the action link "View Attachments". If this link does not appear, there are no attachment files.]

Figure resubmission:

Reproducibility:

?>

Decision Letter 1

Vijayakumar Velu

5 Mar 2026

Dear Dr. Camunas Soler,

We are pleased to inform you that your manuscript 'Cell-free RNA reveals host and microbial correlates of broadly neutralizing antibody development against HIV' has been provisionally accepted for publication in PLOS Pathogens.

Before your manuscript can be formally accepted you will need to complete some formatting changes, which you will receive in a follow up email. A member of our team will be in touch with a set of requests.

Please note that your manuscript will not be scheduled for publication until you have made the required changes, so a swift response is appreciated.

IMPORTANT: The editorial review process is now complete. PLOS will only permit corrections to spelling, formatting or significant scientific errors from this point onwards. Requests for major changes, or any which affect the scientific understanding of your work, will cause delays to the publication date of your manuscript.

Should you, your institution's press office or the journal office choose to press release your paper, you will automatically be opted out of early publication. We ask that you notify us now if you or your institution is planning to press release the article. All press must be co-ordinated with PLOS.

Thank you again for supporting Open Access publishing; we are looking forward to publishing your work in PLOS Pathogens.

Best regards,

Vijayakumar Velu, ph.D.

Academic Editor

PLOS Pathogens

Susan Ross

Section Editor

PLOS Pathogens

Sumita Bhaduri-McIntosh

Editor-in-Chief

PLOS Pathogens

orcid.org/0000-0003-2946-9497

Michael Malim

Editor-in-Chief

PLOS Pathogens

orcid.org/0000-0002-7699-2064

***********************************************************

Reviewer Comments (if any, and for reference):

Acceptance letter

Vijayakumar Velu

Dear Dr. Camunas Soler,

We are delighted to inform you that your manuscript, "Cell-free RNA reveals host and microbial correlates of broadly neutralizing antibody development against HIV," has been formally accepted for publication in PLOS Pathogens.

We have now passed your article onto the PLOS Production Department who will complete the rest of the pre-publication process. All authors will receive a confirmation email upon publication.

The corresponding author will soon be receiving a typeset proof for review, to ensure errors have not been introduced during production. Please review the PDF proof of your manuscript carefully, as this is the last chance to correct any scientific or type-setting errors. Please note that major changes, or those which affect the scientific understanding of the work, will likely cause delays to the publication date of your manuscript. Note: Proofs for Front Matter articles (Pearls, Reviews, Opinions, etc...) are generated on a different schedule and may not be made available as quickly.

Soon after your final files are uploaded, the early version of your manuscript, if you opted to have an early version of your article, will be published online. The date of the early version will be your article's publication date. The final article will be published to the same URL, and all versions of the paper will be accessible to readers.

For Research Articles, you will receive an invoice from PLOS for your publication fee after your manuscript has reached the completed accept phase. If you receive an email requesting payment before acceptance or for any other service, this may be a phishing scheme. Learn how to identify phishing emails and protect your accounts at https://explore.plos.org/phishing.

Thank you again for supporting open-access publishing; we are looking forward to publishing your work in PLOS Pathogens.

Best regards,

Sumita Bhaduri-McIntosh

Editor-in-Chief

PLOS Pathogens

orcid.org/0000-0003-2946-9497

Michael Malim

Editor-in-Chief

PLOS Pathogens

orcid.org/0000-0002-7699-2064

Associated Data

    This section collects any data citations, data availability statements, or supplementary materials included in this article.

    Supplementary Materials

    S1 Fig. Clinical information for blood samples.

    Longitudinal trajectories for CD4 (orange) and viral load (purple). Diamonds indicate timepoints sequenced in this study. Grey bars indicate values for CD4 count < 200.

    (TIF)

    ppat.1014066.s001.tif (4.1MB, tif)
    S2 Fig. Sequencing quality control metrics per sample.

    A) Number of reads (2x75 bp) sequenced in each sample and negative controls. Two cfDNA samples failed extraction/library preparation. B) Percentage of reads cleaned in QC steps (low quality or that align to the UniVec Core database). Percentages are generally very low, except in cfDNA negative controls, which had a high proportion of primer/library-derived sequences. C) Percentage of reads mapping to the human genome. D) Total number of non-human reads, typically more than a million for the cfRNA samples and around 10 thousand for the cfDNA samples. Samples are ordered by time within each participant.

    (TIF)

    ppat.1014066.s002.tif (8.7MB, tif)
    S3 Fig. Sequencing quality control metrics summarized by study group.

    A) Total sequencing depth per sample (raw read counts). B) Read counts retained after QC cleaning steps (see Methods) C) Percentage of reads removed during cleaning D) Percentage of human-mapped reads after cleaning. Metrics are shown for cfRNA (left) and cfDNA (right) libraries across bNAb producers, non-producers, and negative controls. Data are shown as mean ± SEM.

    (TIF)

    ppat.1014066.s003.tif (6.7MB, tif)
    S4 Fig. Cell-type deconvolution analysis of cfRNA.

    (a) Relative cell-type contributions to the cfRNA composition for each sample. Samples are grouped by participant and study group (bNAb producer or control), and less abundant cell types are aggregated into ‘Other’. (b) Box plots showing the relative abundance of each cell type in bNAb producers and controls. Statistical significance was assessed using a Mann–Whitney U test, with p-values adjusted for multiple hypothesis testing using Benjamini–Hochberg.

    (TIF)

    ppat.1014066.s004.tif (24.7MB, tif)
    S5 Fig. HIV genotypes, coverage and mutations.

    A) Heatmap of coverage across the genome in each sample. The four low abundance samples are marked with black squares on the left hand panel. B) Total number of read fragments aligning to HIV in each sample. C) Phylogenetic tree of HIV consensus genotypes for each sample obtained in this study together with 200 other genotypes of HIV-1 group M, subtype C from South Africa. D) Locations of SNVs local to each participant across the three time points. Most cluster within the env gene. E) Mutation rate observed in the samples using differences in the consensus genotypes within participants.

    (TIF)

    ppat.1014066.s005.tif (42MB, tif)
    S6 Fig. Fragment length and coverage of cfHIV reads.

    A) Distribution of paired-end genomic spans for reads mapping to the HIV genome, showing a predominance of short fragments (200–600 bp). B) Mean paired-end genomic span as a function of genomic start position, showing no systematic trends that could indicate priming from long RNA templates. (c) Genomic coverage profile across the HIV genome, shown alongside the reference organization of the HIV genome.

    (TIF)

    ppat.1014066.s006.tif (26.3MB, tif)
    S7 Fig. Subsampling analysis of HIV genome coverage from cfRNA.

    A) Fraction of the HIV genome covered above increasing minimum depth thresholds (1X-50X) at different subsampling levels of sample CAP261. B) Number of HIV-mapped reads versus mean coverage depth across the HIV genome, showing a linear relationship. C) Genome coverage achieved for each sample in the cohort at ≥10X (left) and ≥50X (right) depth thresholds.

    (TIF)

    ppat.1014066.s007.tif (21.7MB, tif)
    S8 Fig. Correlations of genes with different disease measurements.

    A) Distributions of Spearman correlations of all genes with either cfHIV counts (top, blue), CD4 counts (middle, orange), or viral load (bottom, purple). Gray bands indicate regions of significant correlation as determined by the Spearman test on the data. The small ticks (rugplot) at the bottom of each plot indicate known HIV genes. B) Tables of the number of genes strongly correlated/anticorrelated between genes and measurements (as in A). C) Scatter plots of cfHIV abundance vs gene expression (z-score) for the 14 significantly correlated known HIV-associated genes. Besides HLA-A, none of these show strong differences between the two study groups.

    (TIF)

    S9 Fig. Enriched gene pathways in bNAb producers.

    A) Top gene ontologies (molecular function, biological process) and pathways enriched in genes that are elevated in bNAb producers. Numbers indicate the number of selected genes compared with the total number assigned to that category. B) Similar to A, but for the genes with decreased abundance in bNAb producers. Note that scores are much lower for decreased genes than for elevated genes and no clear enriched groups of genes are present in decreased genes. C) Distribution of genes correlated with breadth in bNAb producers (upper, red) and controls (lower, blue). D) Tables of the number of genes meeting correlation thresholds in (C).

    (TIF)

    ppat.1014066.s009.tif (10.3MB, tif)
    S10 Fig. Background microbial sequences in negative controls.

    Volcano plot showing differential abundance of microbial genera between plasma samples and negative controls (NC) (FDR < 0.05, Benjamini–Hochberg). Negative log2 fold-change values indicate genera enriched in NC, which were used to define background levels. Several NC-enriched taxa showed large effect sizes (|log2FC| > 4, FDR < 10-8), primarily comprising environmental taxa previously reported in DNA extraction kits and reagents in low-biomass sequencing.

    (TIF)

    ppat.1014066.s010.tif (4.1MB, tif)
    S11 Fig. Phylogenetic tree of microbial sequences from circulating nucleic acids in HIV infection (all participants).

    A) Phylogenetic tree of cfDNA-derived microbiome reads pooled across all participants (PLWH who do and do not develop bNAbs). The main branches represent archaea, bacteria and viruses (from left to right); representative viruses including Torque Teno Viruses (TTV), Epstein-Barr virus (EBV) and cytomegalovirus (CMV) are highlighted. B) Phylogenetic tree of cfRNA-derived microbiome reads pooled across all participants. The main branches represent archaea, bacteria and viruses (from left to right). Within the viral branch, HIV and GB virus C are the main contributors to the virome.

    (TIF)

    ppat.1014066.s011.tif (9.9MB, tif)
    S12 Fig. Subset of microbial genera enriched in bNAb producers.

    Boxplots showing normalized abundances of selected microbial genera across negative controls (NC), bNAb non-producers, and bNAb producers. Values are shown as log2-transformed abundances relative to the median abundance in negative controls.

    (TIF)

    ppat.1014066.s012.tif (3.8MB, tif)
    S13 Fig. Pseudo-temporal dynamics in differentially abundant genera between bNAb producers and non-producers.

    Scatter plots of the abundance (z-score) vs breadth production (%) for the 22 genera found to be differentially abundant between case (bNAbs) and control (no bNAbs). No apparent time differences found.

    (TIF)

    ppat.1014066.s013.tif (14.2MB, tif)
    S14 Fig. Coverage profiles of HCV and GBV-C genomes across samples.

    A) Read coverage across the HCV genome. B) Read coverage across the GBV-C genome. Coverage for individual samples is shown in red (HCV) and blue (GBV-C); mean coverage is shown in black.

    (TIFF)

    ppat.1014066.s014.tiff (9.9MB, tiff)
    S1 Table. Neutralization data for the panel of 18 Env-pseudotyped HIV viruses.

    (XLSX)

    ppat.1014066.s015.xlsx (17.1KB, xlsx)
    S2 Table. Participant codes and corresponding accession numbers in the HIV database (https://www.hiv.lanl.gov/content/index) for the reference FASTA sequences included in this study.

    (XLSX)

    ppat.1014066.s016.xlsx (9.4KB, xlsx)
    Attachment

    Submitted filename: Response_File.pdf

    ppat.1014066.s018.pdf (2.5MB, pdf)

    Data Availability Statement

    The datasets generated and analyzed in the study are available in the NCBI Gene Expression Omnibus (GEO) and Sequence Read Archive (SRA) with accession number GSE313190. The code to reproduce the analyses in this study is publicly available at https://github.com/CamunasLab/cfHIV.


    Articles from PLOS Pathogens are provided here courtesy of PLOS

    RESOURCES