ABSTRACT
Liquid biopsies are transforming oncology, enabling earlier diagnosis, dynamic treatment guidance, and personalized precision medicine, yet current approaches focusing mainly on circulating host cell-free DNA (cfDNA) neglect crucial information within co-existing microbial cell-free DNA (mcfDNA). This review argues for the combined potential of simultaneously analyzing host and microbial signals from samples like blood, specifically focusing on circulating tumor DNA (ctDNA) as the key host component. While ctDNA analysis is already used to guide treatment decisions, the detection of mcfDNA—although present in smaller amounts compared to total cfDNA—offers a distinct and complementary opportunity to identify disease-causing microbes and investigate the host-associated microbiome in the context of cancer. Leveraging machine learning strategies is essential to integrate these multi-view data sets and realize their full potential for enhancing liquid biopsy applications, particularly in early cancer detection.
INTRODUCTION
UNLOCKING MICROBIAL SIGNATURES FOR DIAGNOSTIC INNOVATION
The human microbiome is commonly perceived as our second genome (1) due to its genetic diversity, versatility, importance for immune modulation, and human metabolism (2). Human evolution is taking place in a microbial world (3). Microbes educate our immune system and occupy niches that could be exploited by pathogens (4). Microbes accompany human beings for their entire life and are omnipresent during disease, treatment, and until death. Hence, microbial traces or signatures can be powerful biomarkers to assess the medical conditions of the human body, and even predict the course of disease and success of a therapy (5–10). Our understanding of the human microbiome was mainly driven by investigating stool samples and applying next-generation sequencing (NGS). Using stool samples has the advantage of providing high microbial biomass and low human background noise for downstream sequencing analysis. Nevertheless, it remains questionable whether stool samples are representative of the entire gastrointestinal tract and its complex dynamics (11). Generally, diagnostics in the medical field often rely on sample types other than human stool alone—for instance, blood, urine, or biopsy samples. It is therefore not surprising that these alternative sample types have already been used for microbiome profiling and microbiome-informed biomarker searches for potential diagnostic applications (12). Still, the differentiation between correlation and causality—a common challenge in microbiome research—is evident across alternative sample types. For instance, while viruses like papillomavirus for cervical cancer (13) or bacterial pathogens like Helicobacter pylori (Cag-A) for stomach cancer (14) and Fusobacterium nucleatum (adhesin A, autotransporter protein 2) for colon cancer (15) have been widely accepted as causative agents of cancer, other signals of the microbiome are still part of a lively debate. Just recently, Gihawi et al. (16) questioned the cancer-type-specific microbial signatures reported by Poore and colleagues (12) and attributed many observations and results to bioinformatic problems, including flawed data, problematic machine learning (ML), normalization artifacts, and database contamination. Meanwhile, Sepich-Poore et al. retracted their original manuscript from 2020 but reaffirmed the robustness of cancer-associated microbiome signals using updated bioinformatic workflows that included batch correction, improved host read depletion with a newer human genome reference, cross-method consistency analyses, and additional validation steps (17). Irrespective of one’s angle on this ongoing debate, the scientific debate emphasizes a common challenge for microbiome research in low-biomass environments: the need for rigorous controls, specific adaptation, and standardized bioinformatic workflows, and the difficulty of distinguishing correlations from causal relationships. Hence, in addition to these manifold technical hurdles, exploiting genetic traces of the microbiome through whole-genome sequencing (WGS) remains a promising target, especially in a joint multi-view approach for modern cancer diagnostics (Box 1) .
Box 1. Glossary.
Feature
A specific, measurable characteristic of the data that serves as input for a machine learning model. For example, a feature might represent the cfDNA fragment length or the abundance of a specific microbe.
Modality
A distinct category or source of data. Therefore, a multimodal analysis integrates different data types, such as cfDNA fragmentomics and microbial composition data, to form a more comprehensive view.
Autoencoder
An unsupervised neural network that learns to efficiently encode data into a compressed format and then decode it back to its original form. This process forces it to identify the most critical patterns in the data.
Penalized logistic regression
A classification method that adds a regularization term to the loss function to prevent overfitting by shrinking model coefficients, commonly using L1 (lasso) or L2 (ridge) penalties.
Multidimensional scaling (MDS)
A statistical technique that visualizes the similarity or dissimilarity between data points by placing them in a low-dimensional space while preserving their pairwise distances as accurately as possible.
Support vector machine (SVM)
A supervised learning algorithm that finds the optimal hyperplane to separate data into distinct classes by maximizing the margin between them.
Meta-learner
A higher-level model in machine learning that learns how to combine or optimize the performance of multiple base models, often used in ensemble methods or meta-learning frameworks.
Classifier
A machine learning model that assigns input data to predefined categories or classes based on learned patterns from labeled training data.
Regressor
A machine learning model that predicts continuous numerical values based on input features by learning patterns from labeled training data.
Neural networks
Computational models inspired by the structure and function of biological neurons that learn patterns from data to perform tasks like classification, prediction, and pattern recognition.
Graph convolutional neural network (GCN)
A deep learning architecture that generalizes convolution operations to graph-structured data, enabling the extraction of node features by aggregating information from neighboring nodes.
In this minireview, we aim to address the potential of liquid biopsy (LB) to simultaneously extract tumor-specific alterations and microbial signatures to enhance cancer diagnostics. We will cover how these signals can be extracted using a single assay (i.e., WGS), the methodological limitations and challenges, and strategies for integrating these signals using ML to harness the potential of a multi-view approach .
THE ROLE OF THE MICROBIOME IN CANCER
The impact of the microbiome within our body was long overlooked. However, today, it is a widely discussed topic that extends its influence to various aspects such as aging, physical health, and overall well-being. Influenced by factors like sex, age, lifestyle, and diet, the trillions of microbes in the human body—including bacteria, archaea, viruses, and eukaryotes like fungi and protozoans—play a crucial role in maintaining our health through a symbiotic relationship with the human body (18–20). Most of the human microbiome’s biomass is located in the gastrointestinal tract, especially in the colon, i.e., the gut microbiome (21). Changes in the composition of these microbes have been linked to numerous diseases, including depression, diabetes, inflammatory bowel disease, and autoimmune disorders (22).
In oncology, studies over the past decade have begun to uncover intricate links between microbes and cancer (23–27). More recently, the gut microbiome has drawn increasing attention in the era of precision oncology, as gut dysbiosis, a condition characterized by an imbalance in the types of microbes in the gut, contributes substantially to gastrointestinal cancers by causing, modulating, and affecting the progression and therapy response (28). In addition to the importance of the gut microbiome in cancer, intra-tumoral microbial profiles may provide valuable insights into cancerous diseases. It has been reported that microbial profiles not only differ between primary tumors and non-malignant tissues but also among different tumor entities, each exhibiting unique intra-tumoral microbial fingerprints (12, 29). Nejman et al. (29) demonstrated that the bacterial composition varies across seven cancer types, including breast, lung, ovary, pancreas, melanoma, bone, and brain tumors. Moreover, the research by Narunsky-Haziza and colleagues (30) has revealed that despite low levels of fungal DNA across many major human cancers, there are cancer-type-specific differences in the composition of fungal ecologies.
Recent data show that differences in gut and intra-tumoral microbe populations may lead to cancer progression and discrepancies in the efficacy of cancer therapies (28, 31–33), including chemotherapy, radiotherapy, immunotherapy, and surgery, between patients. For example, patients harboring specific gut microbe taxa were associated with better responses to immunotherapy treatment across different tumor types (28). In addition to the gut microbiome, altered intra-tumoral microbiota are associated with resistance to cancer treatment (31). Battaglia et al. (34) highlighted that lung cancer patients whose metastases contain Fusobacterium tend to show resistance to immune checkpoint blockade treatment compared to patients without this bacterium. Furthermore, intra-tumoral bacteria can influence tumor biology through multiple, conserved mechanisms across cancer types. For instance, certain bacterial genotoxins, such as colibactin produced by Escherichia coli and cytolethal distending toxin from Campylobacter spp., can directly induce DNA double-strand breaks and genomic instability, thereby promoting mutagenesis and tumor evolution (35). Bacterial metabolites and cell wall components can also alter chromatin accessibility, modulate histone modifications, and reshape DNA methylation patterns, ultimately reprogramming oncogenic signaling and stress-response pathways (35, 36). Moreover, tumor-associated microbes can influence local and systemic immunity by affecting antigen presentation, cytokine production, and immune cell infiltration, creating an immunosuppressive microenvironment that fosters tumor progression and therapeutic resistance (36).
The growing recognition that microbial ecosystems significantly influence cancer development, progression, and therapeutic response underscores the critical need for diagnostic tools that can capture this complex biological interplay alongside the tumor’s own characteristics. Effectively monitoring cancer diseases, including potential microbial influences, requires moving beyond traditional methods. This clinical need has driven the rapid development and adoption of innovative approaches, prominently featuring LB.
THE STRENGTH OF LB IN CANCER DIAGNOSTICS
LB represents a major advance, offering a minimally invasive window into tumor biology that overcomes many drawbacks of conventional techniques. Clinicians require a precise diagnosis that includes details about the tumor’s biology and genetics to make confident cancer treatment decisions. Molecular fingerprints of the tumor are essential for monitoring how the cancer responds to therapy and tracking its progression over time. However, conventional tissue biopsies are invasive and provide only a limited, time-specific, and spatially restricted snapshot of the tumor. Moreover, certain tumor locations are particularly challenging to access. In addition, radiological imaging has its own drawbacks, including high costs, patient exposure to radiation, and limited molecular information.
In contrast, LB has emerged as a powerful, minimally invasive strategy for detecting, tracking, and treating cancer (37). A key focus—though not the only one—is the analysis of cell-free DNA (cfDNA), which mainly derives from the hematopoietic system, with smaller fractions from solid tissues (38). First reported by Mandel and Metais (39), cfDNA has been extensively studied across various medical fields, particularly in oncology. In cancer patients, tumor cells can release DNA fragments into the bloodstream, adding circulating tumor DNA (ctDNA) to the cfDNA pool (38). The ability to analyze these fragments via a simple blood draw has transformed ctDNA into a promising alternative to conventional biopsy and a widely used prognostic and predictive biomarker (see recent reviews [37, 40, 41]). LB unlocks a broad range of clinical possibilities for both advanced and early-stage cancer settings.
One of the central applications of LB for advanced-stage tumors, now implemented into clinical practice, is to guide treatment decisions by detecting actionable genetic alterations and resistance mutations (42). Since ctDNA is thought to correlate with tumor burden, quantitative changes in ctDNA levels over time have proven to be a valuable indicator of therapeutic efficacy (42–44). The non-invasive nature of ctDNA analysis enables more frequent assessments than tissue biopsies and imaging techniques, further underscoring its potential for tracking treatment response. More recently, the application of LB has expanded to earlier disease stages, where ctDNA-based approaches aim to detect minimal residual disease after curative treatments, enabling early relapse detection and informing adjuvant therapy decisions to adjust treatment intensity (45–48). Notably, intensive research is underway to apply ctDNA approaches for non-invasive cancer screening (49–52).
CIRCULATING BIOMARKERS IN BODY FLUIDS
LB is an umbrella term describing the analysis of circulating biomarkers found in peripheral blood and other body fluids, such as cerebrospinal fluid, pleural fluid, peritoneal fluid, saliva, and urine (53). Traditionally, LB has focused on detecting biomarkers of human self-origin—molecules released by the patient’s own cells, whether healthy or diseased. These include cfDNA, ctDNA, circulating tumor cells, extracellular vesicles, proteins, and messenger and non-coding RNA (41) (Fig. 1A). Each biomarker, individually and in combination, provides valuable insights into physiological processes and disease states.
Fig 1.
(A) Schematic representation of potential circulating biomarkers from LB samples. (B) Single-analyte approaches examine either ctDNA or microbial cell-free DNA (mcfDNA) individually. The multi-view strategy integrates these distinct layers of information for a more comprehensive picture. For instance, a multi-view analysis might integrate genetic and nongenetic features of ctDNA alongside the total microbial diversity and specific pathogens from mcfDNA. Created in BioRender (E. Heitzer, 2025).
Beyond these self-derived biomarkers, LB is also applied to detect human biomarkers from different origins, such as fetal cfDNA in maternal circulation for prenatal screening (54) or donor-derived cfDNA in transplant patients for graft monitoring (55).
More recently, it has become evident that not only human cells release nucleic acids into the bloodstream (Fig. 1). A small and variable but biologically significant fraction of cfDNA—approximately <0.1% to 1%—originates from non-human sources (56–60), particularly from microbes colonizing or invading the host. Most hypotheses about the origin of this microbial DNA in the bloodstream are based on direct shedding from tumor sites (tumor necrosis), barrier translocation (leaky gut), microbial extracellular vesicles carrying DNA, transient bacteremia, host immune cell phagocytosis, and artifacts such as technical contamination (false positives) (56). Studies such as those by Tong et al. (61) showed that most of this mcfDNA comes from bacteria, while archaeal, fungal, and viral sequences are also detectable, with viral cfDNA offering particular diagnostic and prognostic utility in virally driven cancers (62). Although mcfDNA represents only a minor fraction of total cfDNA, its presence offers a unique opportunity to detect pathogenic microbes and explore the host-associated microbiome’s biodiversity.
In essence, biomarkers derived from the host (e.g., ctDNA) and those from the associated microbiome (e.g., mcfDNA) represent distinct but complementary layers of biological information accessible through LB. Examining them in isolation, i.e., single-analyte approach, provides only a partial view. However, integrating these signals offers the potential for a far more comprehensive understanding of tumor biology, laying the foundation for enhanced diagnostics, treatment, and early detection strategies. This concept aligns with the growing shift toward a multi-view or multimodal LB, which seeks to combine multiple layers of information for a more complete picture of the disease (Fig. 1B). A thorough analysis of microbial cfDNA—capturing the diversity of bacteria, archaea, fungi, and viruses—and its integration with an expanded ctDNA profile that includes both genetic (single-nucleotide variants and copy number alterations) and nongenetic features (methylomics, fragmentomics, and nucleosomics), holds the potential to boost LB sensitivity significantly.
EXPLORING ctDNA AND MICROBIAL FINGERPRINTS USING WGS
For the analysis of cfDNA, including both ctDNA and mcfDNA, two main methodological approaches are commonly used: PCR-based and NGS-based methods (63). PCR-based assays are highly sensitive but limited to detecting predefined targets. In contrast, NGS-based methods are more versatile and can be designed as targeted or untargeted approaches. Targeted NGS focuses on specific (pathogen) genes or regions. A notable example is the study by Chan et al. (64), where targeted sequencing of the entire Epstein-Barr virus (EBV) genome in plasma enabled more comprehensive and unbiased detection of EBV DNA, providing better risk stratification for nasopharyngeal carcinoma patients compared to conventional PCR approaches.
In contrast, untargeted methods such as WGS enable comprehensive profiling of nucleic acids within a sample, rather than focusing solely on specific regions. Although WGS shows promise for mcfDNA analysis, its implementation is challenged by the typically low abundance of microbial DNA in LB samples. This limitation reduces sensitivity and requires high sequencing depth, thereby increasing costs. Recent studies have explored enrichment strategies prior to sequencing since mcfDNA fragments tend to be shorter in size than host-derived cfDNA. One such approach achieved a remarkable 204-fold enrichment of the mcfDNA fraction by selectively removing longer host cfDNA fragments and utilizing single-stranded library methods to capture shorter fragments. Nevertheless, this enrichment technique was reported to increase background noise at the genus level, potentially limiting sensitivity for pathogen detection (65).
However, the unique strength of untargeted WGS lies in its intrinsic capacity to simultaneously profile both host-derived cfDNA (including ctDNA) and non-human microbial DNA within a single LB sample. Most LB workflows are designed primarily for human nucleic acid analysis. Consequently, microbial sequences, although co-sequenced, are often removed during bioinformatic processing, wasting valuable microbial biomarker information (66). A study by Cheng and colleagues (67) highlighted the value of integrating both host and microbial information: by sequencing host urine cfDNA, they could detect tissue damage caused by microbial infections. Such a comprehensive approach expands the spectrum of potential biomarkers and provides deeper insights into tumor biology and its interplay with associated microbial fingerprints. Harnessing host and microbial signals from WGS could unlock new diagnostic and prognostic biomarkers, advancing LB toward a more integrated and insightful cancer assessment tool.
UNIFICATION OF ctDNA AND MICROBIAL FINGERPRINT ANALYSES USING DATA FUSION APPROACHES
To realize the potential of combining cfDNA/ctDNA and microbial fingerprint analyses, robust data integration (data fusion) strategies are essential. The sheer volume and complexity of data generated, particularly from WGS, necessitate ML methods that can integrate multimodal inputs—i.e., different data types, such as host cfDNA/ctDNA and mcfDNA signals. Although dedicated host–microbe fusion is still an emerging area, established ML frameworks provide a strong starting point for investigation (68, 69). Approaches to integrating these distinct biomarker types typically include early, late, and intermediate fusion strategies (Fig. 2).
Fig 2.
Combining different modalities using early, late, and intermediate fusion. Each fusion approach has its advantages. While early fusion provides possible orthogonal information, late fusion techniques become experts in their own modality (e.g., genetic and nongenetic alterations of ctDNA and mcfDNA). Intermediate fusion serves as a flexible bridge between early and late fusion, combining features or representations in a stepwise manner. Variants include (A) fusion through a joint embedding integrated with an additional modality and (B) sequential fusion of modality-specific embeddings. Created in BioRender (E. Heitzer, 2025).
Early fusion
Early fusion is the simplest and most widely used data fusion strategy: it combines features from all modalities (see Box 1) into a single feature matrix. This enables the model to learn directly which combinations of features are most informative (69, 70). Early fusion performs best when the modalities contribute a comparable number of features, and these features are measured on similar numerical scales. To prevent one modality from dominating the model when feature numbers differ, dimensionality-reduction methods are often applied to reduce redundancy and balance the contribution of each modality (68). Early fusion is particularly appropriate when most features already carry high intrinsic predictive value, meaning they are informative on their own. For example, Medina et al. (71) developed an early detection model for ovarian cancer by linking cfDNA fragment length profiles with protein biomarkers using penalized logistic regression (see Box 1). Because the protein markers already have a high intrinsic predictive power, the cfDNA fragment length data (originally >400 bins) were first summarized as a short-to-long fragment ratio (<151 bp/>150 bp). Principal component analysis was then used to condense these ratios into two components further. These high-value cfDNA features were subsequently combined with the informative protein markers.
Additional feature reduction approaches, such as multidimensional scaling (65) or autoencoders (72), are also frequently employed to compress high-dimensional data into a smaller set of informative features (see Box 1).
The key principle of early fusion is that features from different modalities provide complementary information that the model can exploit directly. If the modalities require separate processing to become informative, then intermediate or late fusion strategies may be more suitable (69).
Late fusion
Late fusion trains a separate model for each modality and then combines their predictions into a single output. Each modality thus has its own dedicated classifier or regressor (see Box 1), and the final prediction is obtained by combining these individual outputs—either by simple averaging or by using a meta-learner, i.e., a simple model that combines the predictions. In this setup, each model serves as an expert for its modality, contributing an independent assessment to the final decision. For example, Kim et al. (70) trained dedicated classifiers for cfDNA methylation, copy number, and fragmentation size to predict both cancer detection and tissue of origin (TOO). By averaging the support vector machine (SVM) classifier outputs across these three modalities (see Box 1), they improved sensitivity for cancer detection and TOO prediction. Similarly, pancreatic cancer detection was enhanced using a late fusion strategy in which fragment sizes were distilled into the DELFI score (52), and fragment end patterns served as surrogates for methylation signatures. A simple logistic regression meta-learner then combined both scores to generate the final prediction (73).
According to Ramachandram and Taylor (68), neither early nor late fusion is inherently superior; the optimal approach depends on the specific problem context. This observation was further supported by the SPOT-MAS model for multi-cancer detection and TOO prediction (74). In this study, both fusion techniques were evaluated using features such as 450 target methylation regions, genome-wide methylation, fragment size, copy number alterations, and fragment end motifs. The late fusion approach, which integrated modality-specific predictions via a meta-learner (see Box 1), notably improved discrimination between healthy and cancer samples. Conversely, the early fusion approach, implemented via a graph convolutional neural network (see Box 1), achieved superior accuracy for TOO prediction (74). These findings suggest that while late fusion excels at distinguishing healthy from cancer cases, it may fail to fully capture cross-modal relationships that are critical for more complex predictions, such as TOO classification.
Intermediate fusion
Intermediate fusion aims to overcome the limitations of both early and late fusion. Early fusion may combine features too soon, before each modality’s features are individually informative, while late fusion may overlook relationships between modalities by merging only final predictions. Intermediate fusion provides a stepwise strategy that allows gradual integration of information across modalities (68). It can be implemented as a slow fusion process, in which information is progressively integrated across multiple stages (75).
Intermediate fusion is a flexible framework that can be adapted to the degree of complementarity or overlap between modalities. When modalities contribute distinct yet complementary signals, they can be combined into a shared embedding—a compact representation that unifies the informative signals across modalities. This embedding can then be fused with additional data types or fed into a downstream classifier for prediction (Fig. 2A). For instance, genetic and nongenetic features of ctDNA capture similar underlying tumor-specific signals but reveal distinct patterns. Combining them into a shared embedding (e.g., via an autoencoder) helps the model learn connections between these related modalities before integrating more distinct but complementary information, such as mcfDNA, for prediction.
Alternatively, each modality can be distilled independently—using autoencoders, neural networks, or other feature-extraction models—to generate modality-specific embeddings (see Box 1). These embeddings are then combined stepwise into a joint representation that integrates cross-modal information for downstream prediction (Fig. 2B). While not yet applied in the LB field, this approach has been successfully used to diagnose Alzheimer’s disease by integrating magnetic resonance imaging and positron emission tomography data using a stacked autoencoder. This joint representation was then fed into an SVM for classification (76). Together, these strategies make intermediate fusion a flexible bridge between early and late fusion, preserving the level of complementarity or overlap across modalities.
Data integration strategies—early, intermediate, and late fusion—are key for combining ctDNA with microbial fingerprints. Because each method has distinct strengths, the choice of fusion strategy should be guided by the data characteristics and the underlying biological question.
LIMITATIONS, DEVELOPMENTS, AND FUTURE OUTLOOK
While LB has the potential to extract microbial cfDNA, WGS analysis faces critical hurdles, including low sensitivity, specificity, and reproducibility. Foremost among these is the lack of standardized, validated methods for robustly characterizing the low-abundance mcfDNA component while accounting for the complex, dynamic, and personalized nature of the host-microbiome system. Specific technical challenges include mitigating pervasive contamination through rigorous controls and filtering and overcoming low microbial yields relative to host DNA using validated enrichment or depletion strategies. In addition, accurate bioinformatic classification must be ensured despite short reads and database limitations, requiring the use of stringent algorithms and curated reference resources. Furthermore, validating the significance of combined ctDNA-mcfDNA signatures and disentangling these signals from noise remains a significant research challenge.
Key developments addressing these issues include optimized laboratory protocols for co-analysis and potentially leveraging alternative sequencing technologies, such as real-time long-read sequencing, to improve microbial classification. Future priorities must therefore include (i) establishing standardized, validated end-to-end protocols and pipelines for the parallel analysis of ctDNA and mcfDNA; (ii) developing and rigorously evaluating novel, reliable, and clinically relevant biomarkers based on combined host and microbial signatures, demonstrating their reproducibility, predictive power, and utility compared to single-analyte approaches; (iii) conducting further studies—potentially including longitudinal, interventional, and relevant model system approaches—to elucidate causal links between specific microbial signatures and host cancer biology; and (iv) finally, unlocking the full potential of combined ctDNA–mcfDNA analysis will depend on advanced ML-based data fusion strategies and tailored bioinformatic approaches to capture all relevant molecular and microbial insights.
Supplementary Material
ACKNOWLEDGMENTS
During the preparation of this work, the authors used ChatGPT for language editing. The authors carefully reviewed and revised the content and take full responsibility for the content of the publication.
Contributor Information
Alexander Mahnert, Email: alexander.mahnert@medunigraz.at.
Pedro H. Oliveira, Génomique Métabolique, Genoscope, Evry, France
SUPPLEMENTAL MATERIAL
The following material is available online at https://doi.org/10.1128/msystems.00039-24.
An accounting of the reviewer comments and feedback.
ASM does not own the copyrights to Supplemental Material that may be linked to, or accessed through, an article. The authors have granted ASM a non-exclusive, world-wide license to publish the Supplemental Material files. Please contact the corresponding author directly for reuse.
REFERENCES
- 1. Grice EA, Segre JA. 2012. The human microbiome: our second genome. Annu Rev Genomics Hum Genet 13:151–170. doi: 10.1146/annurev-genom-090711-163814 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 2. Tamburini S, Shen N, Wu HC, Clemente JC. 2016. The microbiome in early life: implications for health outcomes. Nat Med 22:713–722. doi: 10.1038/nm.4142 [DOI] [PubMed] [Google Scholar]
- 3. Gilbert JA, Blaser MJ, Caporaso JG, Jansson JK, Lynch SV, Knight R. 2018. Current understanding of the human microbiome. Nat Med 24:392–400. doi: 10.1038/nm.4517 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 4. Lloyd-Price J, Abu-Ali G, Huttenhower C. 2016. The healthy human microbiome. Genome Med 8:51. doi: 10.1186/s13073-016-0307-y [DOI] [PMC free article] [PubMed] [Google Scholar]
- 5. Owokotomo OE, Sengupta R, Shkedy Z. 2023. Development of microbiome biomarkers in intervention studies. J Appl Microbiol 134:lxac052. doi: 10.1093/jambio/lxac052 [DOI] [PubMed] [Google Scholar]
- 6. Tackmann J, Arora N, Schmidt TSB, Rodrigues JFM, von Mering C. 2018. Ecologically informed microbial biomarkers and accurate classification of mixed and unmixed samples in an extensive cross-study of human body sites. Microbiome 6:192. doi: 10.1186/s40168-018-0565-6 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 7. Manor O, Dai CL, Kornilov SA, Smith B, Price ND, Lovejoy JC, Gibbons SM, Magis AT. 2020. Health and disease markers correlate with gut microbiome composition across thousands of people. Nat Commun 11:5206. doi: 10.1038/s41467-020-18871-1 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 8. Pribyl AL, Hugenholtz P, Cooper MA. 2025. A decade of advances in human gut microbiome-derived biotherapeutics. Nat Microbiol 10:301–312. doi: 10.1038/s41564-024-01896-3 [DOI] [PubMed] [Google Scholar]
- 9. Baas FS, Brusselaers N, Nagtegaal ID, Engstrand L, Boleij A. 2024. Navigating beyond associations: opportunities to establish causal relationships between the gut microbiome and colorectal carcinogenesis. Cell Host Microbe 32:1235–1247. doi: 10.1016/j.chom.2024.07.008 [DOI] [PubMed] [Google Scholar]
- 10. Zitvogel L, Derosa L, Routy B, Loibl S, Heinzerling L, de Vries IJM, Engstrand L, Segata N, Kroemer G, ONCOBIOME Network . 2025. Impact of the ONCOBIOME network in cancer microbiome research. Nat Med 31:1085–1098. doi: 10.1038/s41591-025-03608-8 [DOI] [PubMed] [Google Scholar]
- 11. Donaldson GP, Lee SM, Mazmanian SK. 2016. Gut biogeography of the bacterial microbiota. Nat Rev Microbiol 14:20–32. doi: 10.1038/nrmicro3552 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 12. Poore GD, Kopylova E, Zhu Q, Carpenter C, Fraraccio S, Wandro S, Kosciolek T, Janssen S, Metcalf J, Song SJ, Kanbar J, Miller-Montgomery S, Heaton R, Mckay R, Patel SP, Swafford AD, Knight R. 2020. Microbiome analyses of blood and tissues suggest cancer diagnostic approach. Nature 579:567–574. doi: 10.1038/s41586-020-2095-1 [DOI] [PMC free article] [PubMed] [Google Scholar] [Retracted]
- 13. Bosch FX, Lorincz A, Munoz N, Meijer CJLM, Shah KV. 2002. The causal relation between human papillomavirus and cervical cancer. J Clin Pathol 55:244–265. doi: 10.1136/jcp.55.4.244 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 14. Blaser MJ, Perez-Perez GI, Kleanthous H, Cover TL, Peek RM, Chyou PH, Stemmermann GN, Nomura A. 1995. Infection with Helicobacter pylori strains possessing cagA is associated with an increased risk of developing adenocarcinoma of the stomach. Cancer Res 55:2111–2115. [PubMed] [Google Scholar]
- 15. Zepeda-Rivera M, Minot SS, Bouzek H, Wu H, Blanco-Míguez A, Manghi P, Jones DS, LaCourse KD, Wu Y, McMahon EF, Park SN, Lim YK, Kempchinsky AG, Willis AD, Cotton SL, Yost SC, Sicinska E, Kook JK, Dewhirst FE, Segata N, Bullman S, Johnston CD. 2024. A distinct Fusobacterium nucleatum clade dominates the colorectal cancer niche. Nature 628:424–432. doi: 10.1038/s41586-024-07182-w [DOI] [PMC free article] [PubMed] [Google Scholar]
- 16. Gihawi A, Ge Y, Lu J, Puiu D, Xu A, Cooper CS, Brewer DS, Pertea M, Salzberg SL. 2023. Major data analysis errors invalidate cancer microbiome findings. mBio 14:e01607-23 doi: 10.1128/mbio.01607-23 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 17. Sepich-Poore GD, McDonald D, Kopylova E, Guccione C, Zhu Q, Austin G, Carpenter C, Fraraccio S, Wandro S, Kosciolek T, Janssen S, Metcalf JL, Song SJ, Kanbar J, Miller-Montgomery S, Heaton R, Mckay R, Patel SP, Swafford AD, Korem T, Knight R. 2024. Robustness of cancer microbiome signals over a broad range of methodological variation. Oncogene 43:1127–1148. doi: 10.1038/s41388-024-02974-w [DOI] [PMC free article] [PubMed] [Google Scholar]
- 18. de la Cuesta-Zuluaga J, Kelley ST, Chen Y, Escobar JS, Mueller NT, Ley RE, McDonald D, Huang S, Swafford AD, Knight R, Thackray VG. 2019. Age- and sex-dependent patterns of gut microbial diversity in human adults. mSystems 4:e00261-19. doi: 10.1128/mSystems.00261-19 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 19. Bashir H, Bawa A, Kumar R. 2022. Human microbiome: implication of age and external factors, p 1–26. In Human microbiome. Springer Nature Singapore, Singapore. [Google Scholar]
- 20. Mohammadzadeh R, Mahnert A, Zurabischvili T, Wink L, Kumpitsch C, Habisch H, Sprengel J, Filek K, Mertelj P, Pernitsch D, Hingerl K, Gorkiewicz G, Diener C, Loy A, Kolb D, Trautwein C, Madl T, Moissl-Eichinger C. 2025. Methanobrevibacter smithii associates with colorectal cancer through trophic control of the cancer bacteriome. bioRxiv. doi: 10.1101/2025.06.12.659283 [DOI] [PMC free article] [PubMed]
- 21. Ruan W, Engevik MA, Spinler JK, Versalovic J. 2020. Healthy human gastrointestinal microbiome: composition and function after a decade of exploration. Dig Dis Sci 65:695–705. doi: 10.1007/s10620-020-06118-4 [DOI] [PubMed] [Google Scholar]
- 22. Clemente JC, Ursell LK, Parfrey LW, Knight R. 2012. The impact of the gut microbiota on human health: an integrative view. Cell 148:1258–1270. doi: 10.1016/j.cell.2012.01.035 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 23. Mundo L, Del Porro L, Granai M, Siciliano MC, Mancini V, Santi R, Marcar L, Vrzalikova K, Vergoni F, Di Stefano G, Schiavoni G, Segreto G, Onyango N, Nyagol JA, Amato T, Bellan C, Anagnostopoulos I, Falini B, Leoncini L, Tiacci E, Lazzi S. 2020. Frequent traces of EBV infection in Hodgkin and non-Hodgkin lymphomas classified as EBV-negative by routine methods: expanding the landscape of EBV-related lymphomas. Mod Pathol 33:2407–2421. doi: 10.1038/s41379-020-0575-3 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 24. Su ZY, Siak PY, Leong CO, Cheah SC. 2023. The role of Epstein-Barr virus in nasopharyngeal carcinoma. Front Microbiol 14:1116143. doi: 10.3389/fmicb.2023.1116143 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 25. Woodman CBJ, Collins SI, Young LS. 2007. The natural history of cervical HPV infection: unresolved issues. Nat Rev Cancer 7:11–22. doi: 10.1038/nrc2050 [DOI] [PubMed] [Google Scholar]
- 26. Wroblewski LE, Peek RM, Wilson KT. 2010. Helicobacter pylori and gastric cancer: factors that modulate disease risk. Clin Microbiol Rev 23:713–739. doi: 10.1128/CMR.00011-10 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 27. Levrero M, Zucman-Rossi J. 2016. Mechanisms of HBV-induced hepatocellular carcinoma. J Hepatol 64:S84–S101. doi: 10.1016/j.jhep.2016.02.021 [DOI] [PubMed] [Google Scholar]
- 28. Routy B, Le Chatelier E, Derosa L, Duong CPM, Alou MT, Daillère R, Fluckiger A, Messaoudene M, Rauber C, Roberti MP, et al. 2018. Gut microbiome influences efficacy of PD-1-based immunotherapy against epithelial tumors. Science 359:91–97. doi: 10.1126/science.aan3706 [DOI] [PubMed] [Google Scholar]
- 29. Nejman D, Livyatan I, Fuks G, Gavert N, Zwang Y, Geller LT, Rotter-Maskowitz A, Weiser R, Mallel G, Gigi E, et al. 2020. The human tumor microbiome is composed of tumor type-specific intracellular bacteria. Science 368:973–980. doi: 10.1126/science.aay9189 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 30. Narunsky-Haziza L, Sepich-Poore GD, Livyatan I, Asraf O, Martino C, Nejman D, Gavert N, Stajich JE, Amit G, González A, et al. 2022. Pan-cancer analyses reveal cancer-type-specific fungal ecologies and bacteriome interactions. Cell 185:3789–3806. doi: 10.1016/j.cell.2022.09.005 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 31. Geller LT, Barzily-Rokni M, Danino T, Jonas OH, Shental N, Nejman D, Gavert N, Zwang Y, Cooper ZA, Shee K, et al. 2017. Potential role of intratumor bacteria in mediating tumor resistance to the chemotherapeutic drug gemcitabine. Science 357:1156–1160. doi: 10.1126/science.aah5043 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 32. Gopalakrishnan V, Spencer CN, Nezi L, Reuben A, Andrews MC, Karpinets TV, Prieto PA, Vicente D, Hoffman K, Wei SC, et al. 2018. Gut microbiome modulates response to anti-PD-1 immunotherapy in melanoma patients. Science 359:97–103. doi: 10.1126/science.aan4236 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 33. Matson V, Fessler J, Bao R, Chongsuwat T, Zha Y, Alegre M-L, Luke JJ, Gajewski TF. 2018. The commensal microbiome is associated with anti-PD-1 efficacy in metastatic melanoma patients. Science 359:104–108. doi: 10.1126/science.aao3290 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 34. Battaglia TW, Mimpen IL, Traets JJH, van Hoeck A, Zeverijn LJ, Geurts BS, de Wit GF, Noë M, Hofland I, Vos JL, Cornelissen S, Alkemade M, Broeks A, Zuur CL, Cuppen E, Wessels L, van de Haar J, Voest E. 2024. A pan-cancer analysis of the microbiome in metastatic cancer. Cell 187:2324–2335. doi: 10.1016/j.cell.2024.03.021 [DOI] [PubMed] [Google Scholar]
- 35. El Tekle G, Garrett WS. 2023. Bacteria in cancer initiation, promotion and progression. Nat Rev Cancer 23:600–618. doi: 10.1038/s41568-023-00594-2 [DOI] [PubMed] [Google Scholar]
- 36. Yang L, Li A, Wang Y, Zhang Y. 2023. Intratumoral microbiota: roles in cancer initiation, development and therapeutic efficacy. Signal Transduct Target Ther 8:35. doi: 10.1038/s41392-022-01304-4 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 37. Heitzer E, Haque IS, Roberts CES, Speicher MR. 2019. Current and future perspectives of liquid biopsies in genomics-driven oncology. Nat Rev Genet 20:71–88. doi: 10.1038/s41576-018-0071-5 [DOI] [PubMed] [Google Scholar]
- 38. Moser T, Kühberger S, Lazzeri I, Vlachos G, Heitzer E. 2023. Bridging biological cfDNA features and machine learning approaches. Trends Genet 39:285–307. doi: 10.1016/j.tig.2023.01.004 [DOI] [PubMed] [Google Scholar]
- 39. Mandel P, Metais P. 1948. Les acides nucléiques du plasma sanguin chez l'homme [Nuclear Acids In Human Blood Plasma]. C R Seances Soc Biol Fil 142:241–243. [PubMed] [Google Scholar]
- 40. Wan JCM, Massie C, Garcia-Corbacho J, Mouliere F, Brenton JD, Caldas C, Pacey S, Baird R, Rosenfeld N. 2017. Liquid biopsies come of age: towards implementation of circulating tumour DNA. Nat Rev Cancer 17:223–238. doi: 10.1038/nrc.2017.7 [DOI] [PubMed] [Google Scholar]
- 41. Alix-Panabières C, Pantel K. 2025. Advances in liquid biopsy: from exploration to practical application. Cancer Cell 43:161–165. doi: 10.1016/j.ccell.2024.11.009 [DOI] [PubMed] [Google Scholar]
- 42. Weber S, van der Leest P, Donker HC, Schlange T, Timens W, Tamminga M, Hasenleithner SO, Graf R, Moser T, Spiegl B, Yaspo M-L, Terstappen L, Sidorenkov G, Hiltermann Tj, Speicher MR, Schuuring E, Heitzer E, Groen HJM. 2021. Dynamic changes of circulating tumor DNA predict clinical outcome in patients with advanced non-small-cell lung cancer treated with immune checkpoint inhibitors. JCO Precis Oncol 5:1540–1553. doi: 10.1200/PO.21.00182 [DOI] [PubMed] [Google Scholar]
- 43. Zhou Q, Gampenrieder SP, Frantal S, Rinnerthaler G, Singer CF, Egle D, Pfeiler G, Bartsch R, Wette V, Pichler A, Petru E, Dubsky PC, Bago-Horvath Z, Fesl C, Rudas M, Ståhlberg A, Graf R, Weber S, Dandachi N, Filipits M, Gnant M, Balic M, Heitzer E. 2022. Persistence of ctDNA in patients with breast cancer during neoadjuvant treatment is a significant predictor of poor tumor response. Clin Cancer Res 28:697–707. doi: 10.1158/1078-0432.CCR-21-3231 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 44. Moser T, Waldispuehl-Geigl J, Belic J, Weber S, Zhou Q, Hasenleithner SO, Graf R, Terzic JA, Posch F, Sill H, Lax S, Kashofer K, Hoefler G, Schoellnast H, Heitzer E, Geigl JB, Bauernhofer T, Speicher MR. 2020. On-treatment measurements of circulating tumor DNA during FOLFOX therapy in patients with colorectal cancer. npj Precis Oncol 4:30. doi: 10.1038/s41698-020-00134-3 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 45. Tie J, Wang Y, Tomasetti C, Li L, Springer S, Kinde I, Silliman N, Tacey M, Wong H-L, Christie M, et al. 2016. Circulating tumor DNA analysis detects minimal residual disease and predicts recurrence in patients with stage II colon cancer. Sci Transl Med 8. doi: 10.1126/scitranslmed.aaf6219 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 46. Abbosh C, Birkbak NJ, Wilson GA, Jamal-Hanjani M, Constantin T, Salari R, Le Quesne J, Moore DA, Veeriah S, Rosenthal R, et al. 2017. Phylogenetic ctDNA analysis depicts early-stage lung cancer evolution. Nature 545:446–451. doi: 10.1038/nature22364 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 47. Reinert T, Henriksen TV, Christensen E, Sharma S, Salari R, Sethi H, Knudsen M, Nordentoft I, Wu H-T, Tin AS, et al. 2019. Analysis of plasma cell-free DNA by ultradeep sequencing in patients with stages I to III colorectal cancer. JAMA Oncol 5:1124–1131. doi: 10.1001/jamaoncol.2019.0528 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 48. McDonald BR, Contente-Cuomo T, Sammut S-J, Odenheimer-Bergman A, Ernst B, Perdigones N, Chin S-F, Farooq M, Mejia R, Cronin PA, Anderson KS, Kosiorek HE, Northfelt DW, McCullough AE, Patel BK, Weitzel JN, Slavin TP, Caldas C, Pockaj BA, Murtaza M. 2019. Personalized circulating tumor DNA analysis to detect residual disease after neoadjuvant therapy in breast cancer. Sci Transl Med 11. doi: 10.1126/scitranslmed.aax7392 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 49. Cohen JD, Li L, Wang Y, Thoburn C, Afsari B, Danilova L, Douville C, Javed AA, Wong F, Mattox A, et al. 2018. Detection and localization of surgically resectable cancers with a multi-analyte blood test. Science 359:926–930. doi: 10.1126/science.aar3247 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 50. Liu MC, Oxnard GR, Klein EA, Swanton C, Seiden MV, Liu MC, Oxnard GR, Klein EA, Smith D, Richards D, et al. 2020. Sensitive and specific multi-cancer detection and localization using methylation signatures in cell-free DNA. Ann Oncol 31:745–759. doi: 10.1016/j.annonc.2020.02.011 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 51. Lennon AM, Buchanan AH, Kinde I, Warren A, Honushefsky A, Cohain AT, Ledbetter DH, Sanfilippo F, Sheridan K, Rosica D, et al. 2020. Feasibility of blood testing combined with PET-CT to screen for cancer and guide intervention. Science 369:eabb9601. doi: 10.1126/science.abb9601 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 52. Cristiano S, Leal A, Phallen J, Fiksel J, Adleff V, Bruhm DC, Jensen SØ, Medina JE, Hruban C, White JR, et al. 2019. Genome-wide cell-free DNA fragmentation in patients with cancer. Nature 570:385–389. doi: 10.1038/s41586-019-1272-6 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 53. Tivey A, Church M, Rothwell D, Dive C, Cook N. 2022. Circulating tumour DNA - looking beyond the blood. Nat Rev Clin Oncol 19:600–612. doi: 10.1038/s41571-022-00660-y [DOI] [PMC free article] [PubMed] [Google Scholar]
- 54. Dennis Lo YM, Corbetta N, Chamberlain PF, Rai V, Sargent IL, G Redman CW, Wainscoat JS. 1997. Presence of fetal DNA in maternal plasma and serum. Lancet 350:485–487. doi: 10.1016/S0140-6736(97)02174-0 [DOI] [PubMed] [Google Scholar]
- 55. Bloom RD, Bromberg JS, Poggio ED, Bunnapradist S, Langone AJ, Sood P, Matas AJ, Mehta S, Mannon RB, Sharfuddin A, Fischbach B, Narayanan M, Jordan SC, Cohen D, Weir MR, Hiller D, Prasad P, Woodward RN, Grskovic M, Sninsky JJ, Yee JP, Brennan DC, Circulating Donor-Derived Cell-Free DNA in Blood for Diagnosing Active Rejection in Kidney Transplant Recipients (DART) Study Investigators . 2017. Cell-Free DNA and active rejection in kidney allografts. J Am Soc Nephrol 28:2221–2232. doi: 10.1681/ASN.2016091034 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 56. Zhang Z, Ren W, Wang Y, Zhai T, Huang J. 2025. The origin of circulating microbial DNA in the blood: where does it come from? Ann Med 57:2560605. doi: 10.1080/07853890.2025.2560605 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 57. Zhai T, Ren W, Ji X, Wang Y, Chen H, Jin Y, Liang Q, Zhang N, Huang J. 2024. Distinct compositions and functions of circulating microbial DNA in the peripheral blood compared to fecal microbial DNA in healthy individuals. mSystems 9:e00008-24. doi: 10.1128/msystems.00008-24 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 58. Zhao Y, Zhang W, Zhang X. 2024. Application of metagenomic next-generation sequencing in the diagnosis of infectious diseases. Front Cell Infect Microbiol 14. doi: 10.3389/fcimb.2024.1458316 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 59. Huang YF, Chen YJ, Fan TC, Chang NC, Chen YJ, Midha MK, Chen TH, Yang HH, Wang YT, Yu AL, Chiu KP. 2018. Analysis of microbial sequences in plasma cell-free DNA for early-onset breast cancer patients and healthy females. BMC Med Genomics 11:16. doi: 10.1186/s12920-018-0329-y [DOI] [PMC free article] [PubMed] [Google Scholar]
- 60. Kowarsky M, Camunas-Soler J, Kertesz M, De Vlaminck I, Koh W, Pan W, Martin L, Neff NF, Okamoto J, Wong RJ, Kharbanda S, El-Sayed Y, Blumenfeld Y, Stevenson DK, Shaw GM, Wolfe ND, Quake SR. 2017. Numerous uncharacterized and highly divergent microbes which colonize humans are revealed by circulating cell-free DNA. Proc Natl Acad Sci USA 114:9623–9628. doi: 10.1073/pnas.1707009114 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 61. Tong X, Yu X, Du Y, Su F, Liu Y, Li H, Liu Y, Mu K, Liu Q, Li H, Zhu J, Xu H, Xiao F, Li Y. 2022. Peripheral blood microbiome analysis via noninvasive prenatal testing reveals the complexity of circulating microbial cell-free DNA. Microbiol Spectr 10:e00414-22. doi: 10.1128/spectrum.00414-22 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 62. Amponsah RD, Gyabaah EA, Amidu M, Cheng Z, Talukdar FR. 2025. Harnessing viral footprints in circulating free DNA (cfDNA) for early cancer detection: a focus on liquid-biopsy-based screening. Int J Cancer. doi: 10.1002/ijc.70121 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 63. Hu Y, Zhao Y, Zhang Y, Chen W, Zhang H, Jin X. 2025. Cell-free DNA: a promising biomarker in infectious diseases. Trends Microbiol 33:421–433. doi: 10.1016/j.tim.2024.06.005 [DOI] [PubMed] [Google Scholar]
- 64. Chan DCT, Lam WKJ, Hui EP, Ma BBY, Chan CML, Lee VCT, Cheng SH, Gai W, Jiang P, Wong KCW, Mo F, Zee B, King AD, Le QT, Chan ATC, Chan KCA, Lo YMD. 2022. Improved risk stratification of nasopharyngeal cancer by targeted sequencing of Epstein-Barr virus DNA in post-treatment plasma. Ann Oncol 33:794–803. doi: 10.1016/j.annonc.2022.04.068 [DOI] [PubMed] [Google Scholar]
- 65. Dominguez EG, McDonald BR, Zhang H, Stephens MD, Dietmann EC, Nedden M, Byington N, Thompson S, Junak M, Pepperell CS, Kisat MT. 2025. Enrichment of microbial DNA in plasma to improve pathogen detection in sepsis: a pilot study. Clin Chem 71:567–576. doi: 10.1093/clinchem/hvaf011 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 66. Kataria R, Shoaie S, Grigoriadis A, Wan JCM. 2023. Leveraging circulating microbial DNA for early cancer detection. Trends Cancer 9:879–882. doi: 10.1016/j.trecan.2023.08.001 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 67. Cheng AP, Burnham P, Lee JR, Cheng MP, Suthanthiran M, Dadhania D, De Vlaminck I. 2019. A cell-free DNA metagenomic sequencing assay that integrates the host injury response to infection. Proc Natl Acad Sci USA 116:18738–18744. doi: 10.1073/pnas.1906320116 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 68. Ramachandram D, Taylor GW. 2017. Deep multimodal learning: a survey on recent advances and trends. IEEE Signal Process Mag 34:96–108. doi: 10.1109/MSP.2017.2738401 [DOI] [Google Scholar]
- 69. Stahlschmidt SR, Ulfenborg B, Synnergren J. 2022. Multimodal deep learning for biomedical data fusion: a review. Brief Bioinform 23:bbab569. doi: 10.1093/bib/bbab569 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 70. Kim SY, Jeong S, Lee W, Jeon Y, Kim YJ, Park S, Lee D, Go D, Song SH, Lee S, Woo HG, Yoon JK, Park YS, Kim YT, Lee SH, Kim KH, Lim Y, Kim JS, Kim HP, Bang D, Kim TY. 2023. Cancer signature ensemble integrating cfDNA methylation, copy number, and fragmentation facilitates multi-cancer early detection. Exp Mol Med 55:2445–2460. doi: 10.1038/s12276-023-01119-5 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 71. Medina JE, Annapragada AV, Lof P, Short S, Bartolomucci AL, Mathios D, Koul S, Niknafs N, Noë M, Foda ZH, et al. 2025. Early detection of ovarian cancer using cell-free DNA fragmentomes and protein biomarkers. Cancer Discov 15:105–118. doi: 10.1158/2159-8290.CD-24-0393 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 72. Hinton GE, Salakhutdinov RR. 2006. Reducing the dimensionality of data with neural networks. Science 313:504–507. doi: 10.1126/science.1127647 [DOI] [PubMed] [Google Scholar]
- 73. Noë M, Mathios D, Annapragada AV, Koul S, Foda ZH, Medina JE, Cristiano S, Cherry C, Bruhm DC, Niknafs N, Adleff V, Ferreira L, Easwaran H, Baylin S, Phallen J, Scharpf RB, Velculescu VE. 2024. DNA methylation and gene expression as determinants of genome-wide cell-free DNA fragmentation. Nat Commun 15:6690. doi: 10.1038/s41467-024-50850-8 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 74. Nguyen VTC, Nguyen TH, Doan NNT, Pham TMQ, Nguyen GTH, Nguyen TD, Tran TTT, Vo DL, Phan TH, Jasmine TX, et al. 2023. Multimodal analysis of methylomics and fragmentomics in plasma cell-free DNA for multi-cancer early detection and localization. Elife 12:RP89083. doi: 10.7554/eLife.89083 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 75. Karpathy A, Toderici G, Shetty S, Leung T, Sukthankar R, Fei-Fei L. 2014. Large-scale video classification with convolutional neural networks. 2014 IEEE Conference on Computer Vision and Pattern Recognition (CVPR); Columbus, OH, USA: , p 1725–1732IEEE. doi: 10.1109/CVPR.2014.223 [DOI] [Google Scholar]
- 76. Liu S, Liu S, Cai W, Che H, Pujol S, Kikinis R, Feng D, Fulham MJ. 2015. Multimodal neuroimaging feature learning for multiclass diagnosis of Alzheimer’s disease. IEEE Trans Biomed Eng 62:1132–1140. doi: 10.1109/TBME.2014.2372011 [DOI] [PMC free article] [PubMed] [Google Scholar]
Associated Data
This section collects any data citations, data availability statements, or supplementary materials included in this article.
Supplementary Materials
An accounting of the reviewer comments and feedback.


