Abstract
Introduction
Observational studies represent an alternative to estimate real-world causal effects in the absence of available randomised controlled trials (RCTs). Target trial emulation is a framework for the application of RCT design principles to emulate a hypothetical open-label RCT (the hypothetical target trial) using existing observational data as the primary data source as opposed to the prospective recruitment and measurement of randomised units. The aim of this systematic review is to investigate the practices of studies applying the target trial emulation framework to evaluate the effectiveness of interventions.
Methods and analysis
We will systematically search in Medline (via Ovid), Embase (via Ovid, entries from medRxiv are included), PsycINFO (via Ovid), SCOPUS, Web of Science, Cochrane Library, the ISRCTN registry and ClinicalTrials.gov for all study reports and protocols which used the trial emulation framework (without time restriction). We will extract information concerning study design, data source, analysis, results, interpretation and dissemination. Two reviewers will perform study selection, data extraction and quality assessment. Disagreements between reviewers will be resolved by a third reviewer. A narrative approach will be used to synthesise and report qualitative and quantitative data. Reporting of the review will be informed by Preferred Reporting Items for Systematic Review and Meta-Analysis guidance (PRISMA).
Ethics and dissemination
Ethical approval is not required as it is a protocol for a systematic review. Findings will be disseminated through peer-reviewed publications and conference presentations.
Keywords: statistics & research methods, oncology, registries, systematic review
STRENGTHS AND LIMITATIONS OF THIS STUDY.
This systematic review protocol details a comprehensive search strategy consisting of academic databases, Cochrane library and academic repositories and archives.
This systematic review protocol pre-specifies qualitative research methods to synthesise extracted text data.
This systematic review protocol follows the Preferred Reporting Items for Systematic Review and Meta-Analysis Protocols guidelines.
Risk of bias assessments will not be included in the planned systematic review.
Exclusion of non-English electronic databases may introduce language bias.
Introduction
Observational studies represent an alternative to estimate real-world causal effects in the absence of available randomised controlled trials (RCTs). Undoubtedly, RCTs are positioned toward the top of the evidence hierarchy in the evaluation of the efficacy and safety of treatments. However, reasons to conduct observational studies instead of RCTs may include better external validity, feasibility, ethical considerations, reduced associated costs and greater time efficiency.1 2 While the use of observational studies to attempt to estimate causal effects is not a new phenomenon—the practice of which dates back as early as Rubin’s work3 on counterfactual theory—the application of RCT design principles to overcome biases in and improve the quality of observational research has been a relatively new development.
Target trial emulation is a framework for the application of RCT design principles to emulate a hypothetical open-label RCT (the hypothetical target trial) using existing observational data as the primary data source as opposed to the prospective recruitment and measurement of randomised units.4 Because treatment assignment is neither blind nor random, valid causal treatment effects are estimated if the identification principles of causal inference are satisfied.4–8 These include the assumptions of exchangeability (participants in a study may be swapped between treatment and control groups without changing the average treatment effect), consistency (no multiple versions of the treatment options being investigated), positivity (every patient has a non-zero probability to receive every treatment option under investigation given their measured covariates) and non-interference (the treatment given to one patient does not affect the outcome of another patient).7 9–13 A table of acronyms, abbreviations and technical terms are defined in online supplemental appendix.
bmjopen-2022-070963supp001.pdf (120.9KB, pdf)
A study employing target trial emulation is characterised by an explicit description of the hypothetical target trial on several design components: the eligibility criteria, treatment strategy, assignment procedure, follow-up, primary outcome definition, the causal contrast of interest and a statistical analysis plan.4 Additionally, an accompanying target trial emulation specification is described and makes explicit how the hypothetical target trial is being emulated along the aforementioned design elements. Stating both the hypothetical target trial and the target trial emulation specifications has important benefits. It assists investigators in avoiding biases and pitfalls that are common to the analysis of observational data14 15 and permits the sceptical reader to judge if the design of the hypothetical target trial was sensible and if emulation was satisfactory; for example, the eligibility criteria may not be faithfully emulated owing to limitations with the observational data.
Target trial emulation studies can be credible sources of real-world evidence (RWE) to support regulatory decisions in the absence of RCTs, for instance, within the Food and Drug Administration’s Real-World Evidence Program16 and the European Medicines Agency’s OPerational, TechnIcal, and MethodologicAL framework (OPTIMAL).17 Furthermore, with the increasing availability of large-scale real-world data (RWD) especially those seen in research cohorts, registry, hospital and general practitioner databases, there is impetus for increased popularity of the application of target trial emulation.
To our knowledge, no study has reviewed the extant work applying the target trial emulation framework. We argue that a review of this nature is needed for several reasons that have implications for the continued growth of target trial emulation.
First, a systematic review (SR) that reports on the design, conduct and reporting of target trial emulation studies will help facilitate new research hoping to use the methodology. Investigators seeking to conduct their own target trial emulation study may draw guidance from what is currently published, and so an SR of the literature serves as a convenient repository of knowledge to new investigators who may be unfamiliar with terminology, choices of designs and analyses, and reporting practices. An additional benefit of the review to both new and experienced investigators is the increased awareness of available datasets that may be pertinent to their field of expertise.
Second, target trial emulation studies attempt to mimic RCTs and thus the Consolidated Standards of Reporting Trials (CONSORT) checklist18 may be a valid guideline for reporting standards. However, given that target trial emulation studies are necessarily observational studies, the Strengthening the Reporting of Observational Studies in Epidemiology (STROBE) guidelines also apply.19 Thus, it is essential to capture what (or whether) established reporting guideline(s) have been adopted within the existing trial emulation literature, so that future studies may follow suit. Furthermore, with the recent movement towards the language of the estimand framework in clinical trials,20 it may be useful to capture whether the target trial emulation literature has also displayed such a shift. It may also be that existing checklists are not sufficient to capture all the nuances that exist in target trial emulation. For example, explicit statements on quality assessment such as how the data have been evaluated and judged to be fit-for-purpose for trial emulation are not expected by existing guidelines. Thus, any available reports on data quality may be selective. Clear reporting of these steps is critical especially for regulatory purposes that may require relevance assessments demonstrating fit-for-purpose of the RWD.16 Additionally, pre-registration is not expected for observational studies but is expected for RCTs. At present, there is uncertainty over whether pre-registration should be mandatory, or at least encouraged, for target trial emulation studies.
Third, it may be useful to understand how published studies have discussed the biases and study limitations that preclude perfect emulation of the idealised hypothetical target trial. Recent findings from the Randomised, Controlled Trials Duplicated Using Prospective Longitudinal Insurance Claims: Applying Techniques of Epidemiology (RCT DUPLICATE) initiative compared results of RCTs with their emulation counterparts and found that they do not necessarily agree because perfect emulation may not be possible.21 This may be due to, but not limited to, differences in treatment adherence rates between RCTs and real-world practice, partial emulation of eligibility criteria due to database limitations, measurement error in the primary outcome and the challenges associated with emulating a placebo add-on arm. Although not every study applying the target trial emulation framework will be able to make comparisons with results from an actual RCT, it is still crucial for authors to postulate on the factors that may strengthen or qualify their interpretation of their results.
Research aims and objectives
The aims and objectives of this SR are to investigate the practices of studies applying the target trial emulation framework to evaluate the risk-benefit of interventions. We aim to examine multiple facets of the study process, specifically pre-specification, study design, data source, analysis, results, interpretation, and dissemination.
Objectives
To review how trial emulation studies have designed their target trial specification (eligibility criteria, treatment strategies, treatment assignment, outcomes, time zero and follow-up, causal contrasts, statistical analysis).
To review how the causal inference assumptions were justified.
To review pre-registration, reporting, and dissemination practices.
To review what study limitations and biases in target trial emulation are commonly discussed.
Methods
Database search
The following electronic databases will be searched: Medline (via Ovid), Embase (via Ovid; entries from medRxiv are included), PsycINFO (via Ovid), SCOPUS, Web of Science, Cochrane Library, the ISRCTN registry and ClinicalTrials.gov. Search terms for each database are available in online supplemental table 1. The search will be limited to articles published in English only and no publication date limits will be applied. We anticipate a duration of 1.5 years for this SR (from July 2023 until December 2024).
bmjopen-2022-070963supp002.pdf (72.9KB, pdf)
Study inclusion and exclusion criteria
Eligible studies must first meet the definition of a target trial emulation study based on Hernán and Robins:4
Presence of a hypothetical target trial specification which explicitly describes the trial design components of eligibility criteria, treatment strategy, assignment procedure, follow-up, primary outcome definition, the causal contrast of interest, and a statistical analysis plan; and
Presence of a target trial emulation specification which explicitly describes how the hypothetical target trial is being emulated. The specifications may be displayed in tabular format14 though this need not strictly be the case so long as the information can be retrieved from the text; and
The study must have used at least one observational dataset to conduct their trial emulation.
Additional inclusion criteria are as follows: (1) meets study definition of target trial emulation study; (2) peer-reviewed reports, protocols (published or archived), study registrations and conference presentations on original research studies that used/plan to use the trial emulation, from all fields of research; and (3) available in English language. Studies which sought to replicate RCT findings with RWD will also be included since the RCT protocol contains all the information needed in a hypothetical target trial specification and the target trial emulation specification may be inferred from the methods section.
Exclusion criteria are: (a) conference abstracts, book chapters, opinions or commentaries, purely qualitative papers, SRs and meta-analyses; (b) conceptual/methodological papers that do not use participant data; and (c) non-human subjects.
The reference sections of the excluded SRs, meta-analyses and methodological papers will be checked to identify eligible papers. Additional articles will also be sourced from studies that cited either Hernán and Robins4 or a closely associated paper published in the same year by Hernán et al.15
Selection of studies for inclusion in the review
TB and SH will read all titles/and or abstracts retrieved from the database search process and eliminate obviously irrelevant records. The full text of the remaining possibly relevant studies will then be obtained and read, with each record classified as clearly relevant (meets all inclusion criteria), clearly irrelevant or insufficient information to decide. For those cases with insufficient information, corresponding author(s) will be contacted for further information to make a final decision on relevance. In the instance that the full text of a possibly relevant study cannot be retrieved, the corresponding author(s) will be contacted and if no reply is received, the record will be excluded from the final review. Disagreements on eligibility of studies will be arbitrated by a third co-author. References will be managed on Rayyan.22
Data extraction and management
TB and SH will extract data from all included records according to a bespoke data capture form (DCF) developed in Excel and designed around the key research objectives. The DCF will be designed to help review co-authors elicit the required information from the study text with specific questions and prompts. Pre-defined responses will be available for all straightforward yes/no-type questions and certain multiple-choice questions. The multiple-choice questions with pre-defined responses will be those in which the investigators have prior knowledge on a list of possible responses; lists that are judged to be non-exhaustive will be supplemented by an ‘Other’ response. For questions without this prior knowledge, we will ask the co-author extracting the information to copy the appropriate text verbatim from the paper. The text data will then be subjected to a coding process (see Synthesis of qualitative data section). The extracted data will take the form of close-ended choices (eg, yes/no and other pre-defined responses), numerical data (eg, sample size) and qualitative data.
V.1.0 of the DCF will be developed by adapting information from existing reporting frameworks that are judged to be applicable to target trial emulation. These are :
CONSORT,18
STROBE (combined checklist)19
REporting of studies Conducted using Observational Routinely-collected health Data (RECORD),23
The estimand framework,20 24
Risk Of Bias In Non-randomised Studies—of Interventions (ROBINS-I) assessment tool25 and
The STRengthening Analytical Thinking for Observational Studies (STRATOS) publications on initial data analysis,26 missing data27 and causal inference.28
V.1.0 will then be piloted on 10 of the included papers and adjustments made where appropriate via discussion with additional authors (JW and DT). V.2.0 will then be used for all papers, including the 10 used during the piloting. Key study features to be extracted are outlined in online supplemental table 2; the table functions as a starting point for data collection, which is an evolving process as will be described by version controls of the DCF.
bmjopen-2022-070963supp003.pdf (114.7KB, pdf)
Synthesis of qualitative data
Qualitative data may constitute a mixture of brief commentary or discussion points as well as richer qualitative data describing processes and experiences. These data will be analysed thematically drawing on guidelines for synthesising qualitative research in SRs29 consisting of a three-stage process: (1) TB and SH will independently code each line of extracted text based on content and meaning; (2) initial codes will be grouped together according to similarities to form descriptive themes; and (3) analytic themes will be developed based on interpretation of these data in the context of the aims of this review. Each stage of this process will be undertaken by TB and SH independently followed by discussion to seek agreement of code and theme generation. Where agreement cannot be reached, an additional co-author will arbitrate.
Summarising and reporting of results
The current protocol adheres to the Preferred Reporting Items for Systematic reviews and Meta-Analyses–Protocol (PRISMA-P) reporting guidelines.30 Our SR will adhere to the PRISMA reporting guidelines.31 32
We will describe the frequency (%) of studies displaying each categorical descriptor of a study feature. For example, when describing types of bias discussed, we may report the frequency and percentage of studies with the following descriptors: ‘None stated’, ‘Unmeasured confounding’, ‘Selection bias’ and so on.
Risk of bias assessment
No risk of bias assessment will be carried out. It is not within the scope of our objectives to render a judgement on whether the hypothetical target trial was successfully emulated or judge if the study results are credible evidence to inform the effectiveness of a treatment. The evidence being synthesised will primarily be the characteristics of the studies, methods and reporting patterns, as opposed to their results.
Patients and public involvement
No patient involved.
Conclusion
This SR of target trial emulation studies will give future researchers a detailed overview of how target trial emulation studies have been conducted and reported. Understanding the current strengths and weaknesses of how current studies are being reported will help to improve the design of trial emulation studies and inform bespoke reporting guidelines.
Supplementary Material
Acknowledgments
We thank Marian Rixham from the Newcastle University Walton library for their expertise in optimising the systematic search strategy.
Footnotes
Contributors: Design of the protocol—TB and SH; draft of the manuscript—TB and SH; review and final approval of the manuscript—TB, SH, JW, DT, AB and MB.
Funding: TB is funded by the National Institute for Health and Care Research (NIHR) Applied Research Collaboration North East and North Cumbria (Grant/Award Number: NIHR200173). SH (Pre-Doctoral Fellowship, NIHR302746) and JW (Research Professorship, NIHR301614) are funded by the National Institute for Health Research (NIHR) for this research project.
Disclaimer: The views expressed in this publication are those of the author(s) and not necessarily those of the NIHR, National Health Service (NHS), or the UK Department of Health and Social Care.
Competing interests: None declared.
Patient and public involvement: Patients and/or the public were not involved in the design, or conduct, or reporting, or dissemination plans of this research.
Provenance and peer review: Not commissioned; externally peer reviewed.
Supplemental material: This content has been supplied by the author(s). It has not been vetted by BMJ Publishing Group Limited (BMJ) and may not have been peer-reviewed. Any opinions or recommendations discussed are solely those of the author(s) and are not endorsed by BMJ. BMJ disclaims all liability and responsibility arising from any reliance placed on the content. Where the content includes any translated material, BMJ does not warrant the accuracy and reliability of the translations (including but not limited to local regulations, clinical guidelines, terminology, drug names and drug dosages), and is not responsible for any error and/or omissions arising from translation and adaptation or otherwise.
Ethics statements
Patient consent for publication
Not applicable.
References
- 1. Kennedy-Martin T, Curtis S, Faries D, et al. A literature review on the Representativeness of randomized controlled trial samples and implications for the external validity of trial results. Trials 2015;16:495. 10.1186/s13063-015-1023-4 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 2. Maringe C, Benitez Majano S, Exarchakou A, et al. Reflection on modern methods: trial emulation in the presence of immortal-time bias. assessing the benefit of major surgery for elderly lung cancer patients using observational data. Int J Epidemiol 2020;49:1719–29. 10.1093/ije/dyaa057 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 3. Rubin DB. Estimating causal effects of treatments in randomized and Nonrandomized studies. J Educat Psychol 1974;66:688–701. 10.1037/h0037350 [DOI] [Google Scholar]
- 4. Hernán MA, Robins JM. Using big data to emulate a target trial when a randomized trial is not available. Am J Epidemiol 2016;183:758–64. 10.1093/aje/kwv254 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 5. García-Albéniz X, Hsu J, Hernán MA. The value of explicitly emulating a target trial when using real world evidence: an application to colorectal cancer screening. Eur J Epidemiol 2017;32:495–500. 10.1007/s10654-017-0287-2 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 6. Hernán MA, Alonso A, Logan R, et al. Observational studies analyzed like randomized experiments: an application to postmenopausal hormone therapy and coronary heart disease. Epidemiology 2008;19:766–79. 10.1097/EDE.0b013e3181875e61 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 7. Widome R, Wahlstrom KL, Laska MN, et al. The start study: an evaluation to study the impact of a natural experiment in high school start times on adolescent weight and related behaviors. Obs Stud 2020;6:66–86. [PMC free article] [PubMed] [Google Scholar]
- 8. Lodi S, Phillips A, Lundgren J, et al. Effect estimates in randomized trials and observational studies: comparing apples with apples. Am J Epidemiol 2019;188:1569–77. 10.1093/aje/kwz100 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 9. Igelström E, Craig P, Lewsey J, et al. Causal inference and effect estimation using observational data. J Epidemiol Community Health 2022;76:960–6. 10.1136/jech-2022-219267 [DOI] [Google Scholar]
- 10. Oakes JM. n.d. Effect identification in comparative effectiveness research. EGEMs;1:4. 10.13063/2327-9214.1004 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 11. Cole SR, Frangakis CE. The consistency statement in causal inference: a definition or an assumption. Epidemiology 2009;20:3–5. 10.1097/EDE.0b013e31818ef366 [DOI] [PubMed] [Google Scholar]
- 12. Zhu Y, Hubbard RA, Chubak J, et al. Core concepts in Pharmacoepidemiology: violations of the positivity assumption in the causal analysis of observational data: consequences and statistical approaches. Pharmacoepidemiol Drug Saf 2021;30:1471–85. 10.1002/pds.5338 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 13. Hernán MA. Does water kill? A call for less casual causal inferences. Ann Epidemiol 2016;26:674–80. 10.1016/j.annepidem.2016.08.016 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 14. Dickerman BA, García-Albéniz X, Logan RW, et al. Avoidable flaws in observational analyses: an application to Statins and cancer. Nat Med 2019;25:1601–6. 10.1038/s41591-019-0597-x [DOI] [PMC free article] [PubMed] [Google Scholar]
- 15. Hernán MA, Sauer BC, Hernández-Díaz S, et al. Specifying a target trial prevents immortal time bias and other self-inflicted injuries in observational analyses. J Clin Epidemiol 2016;79:70–5. 10.1016/j.jclinepi.2016.04.014 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 16. US Food and Drug Administration . Framework for FDA’s real-world evidence program. 2018. [Google Scholar]
- 17. Cave A, Kurz X, Arlett P. Real-world data for regulatory decision making: challenges and possible solutions for Europe. Clin Pharmacol Ther 2019;106:36–9. 10.1002/cpt.1426 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 18. Schulz KF, Altman DG, Moher D, et al. Statement: updated guidelines for reporting parallel group randomized trials. Ann Intern Med 2010;152:726–32. 10.7326/0003-4819-152-11-201006010-00232 [DOI] [PubMed] [Google Scholar]
- 19. von Elm E, Altman DG, Egger M, et al. The strengthening the reporting of observational studies in epidemiology (STROBE) statement: guidelines for reporting observational studies. Ann Intern Med 2007;147:573–7. 10.7326/0003-4819-147-8-200710160-00010 [DOI] [PubMed] [Google Scholar]
- 20. Ratitch B, Bell J, Mallinckrodt C, et al. Choosing Estimands in clinical trials: putting the ICH E9(R1) into practice. Ther Innov Regul Sci 2020;54:324–41. 10.1007/s43441-019-00061-x [DOI] [PubMed] [Google Scholar]
- 21. Franklin JM, Patorno E, Desai RJ, et al. Emulating randomized clinical trials with Nonrandomized real-world evidence studies: first results from the RCT DUPLICATE initiative. Circulation 2021;143:1002–13. 10.1161/CIRCULATIONAHA.120.051718 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 22. Ouzzani M, Hammady H, Fedorowicz Z, et al. Rayyan-a web and mobile App for systematic reviews. Syst Rev 2016;5:210. 10.1186/s13643-016-0384-4 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 23. Benchimol EI, Smeeth L, Guttmann A, et al. The reporting of studies conducted using observational routinely-collected health data (RECORD) statement. PLOS Med 2015;12:e1001885. 10.1371/journal.pmed.1001885 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 24. Lipkovich I, Ratitch B, Mallinckrodt CH. Causal inference and Estimands in clinical trials. Stat Biopharmaceut Res 2020;12:54–67. 10.1080/19466315.2019.1697739 [DOI] [Google Scholar]
- 25. Sterne JA, Hernán MA, Reeves BC, et al. ROBINS-I: a tool for assessing risk of bias in non-randomised studies of interventions. BMJ 2016:i4919. 10.1136/bmj.i4919 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 26. Huebner M, le Cessie S, Schmidt CO, et al. A contemporary conceptual framework for initial data analysis. Observational Studies 2018;4:171–92. 10.1353/obs.2018.0014 [DOI] [Google Scholar]
- 27. Lee KJ, Tilling KM, Cornish RP, et al. Framework for the treatment and reporting of missing data in observational studies: the treatment and reporting of missing data in observational studies framework. J Clin Epidemiol 2021;134:79–88. 10.1016/j.jclinepi.2021.01.008 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 28. Goetghebeur E, le Cessie S, De Stavola B, et al. Formulating causal questions and principled statistical answers. Stat Med 2020;39:4922–48. 10.1002/sim.8741 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 29. Thomas J, Harden A. Methods for the thematic synthesis of qualitative research in systematic reviews. BMC Med Res Methodol 2008;8:45. 10.1186/1471-2288-8-45 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 30. Moher D, Shamseer L, Clarke M, et al. Preferred reporting items for systematic review and meta-analysis protocols (PRISMA-P) 2015 statement. Syst Rev 2015;4:1. 10.1186/2046-4053-4-1 [DOI] [PMC free article] [PubMed] [Google Scholar]
- 31. Moher D, Liberati A, Tetzlaff J, et al. Preferred reporting items for systematic reviews and meta-analyses: the PRISMA statement. Int J Surg 2010;8:336–41. 10.1016/j.ijsu.2010.02.007 [DOI] [PubMed] [Google Scholar]
- 32. Page MJ, McKenzie JE, Bossuyt PM, et al. The PRISMA 2020 statement: an updated guideline for reporting systematic reviews. Int J Surg 2021;88:S1743-9191(21)00040-6. 10.1016/j.ijsu.2021.105906 [DOI] [PubMed] [Google Scholar]
Associated Data
This section collects any data citations, data availability statements, or supplementary materials included in this article.
Supplementary Materials
bmjopen-2022-070963supp001.pdf (120.9KB, pdf)
bmjopen-2022-070963supp002.pdf (72.9KB, pdf)
bmjopen-2022-070963supp003.pdf (114.7KB, pdf)
