PT - JOURNAL ARTICLE AU - Shirlee Wohl AU - John R. Giles AU - Justin Lessler TI - Sample Size Calculation for Phylogenetic Case Linkage AID - 10.1101/2020.07.10.20150920 DP - 2020 Jan 01 TA - medRxiv PG - 2020.07.10.20150920 4099 - http://medrxiv.org/content/early/2020/07/11/2020.07.10.20150920.short 4100 - http://medrxiv.org/content/early/2020/07/11/2020.07.10.20150920.full AB - Sample size calculations are an essential component of the design and evaluation of scientific studies. However, there is a lack of clear guidance for determining the sample size needed for phylogenetic studies, which are becoming an essential part of studying pathogen transmission. We introduce a statistical framework for determining the number of true infector-infectee transmission pairs identified by a phylogenetic study, given the size and population coverage of that study. We then show how characteristics of the criteria used to determine linkage and aspects of the study design can influence our ability to correctly identify transmission links, in sometimes counterintuitive ways. We test the overall approach using outbreak simulations and provide guidance for calculating the sensitivity and specificity of the linkage criteria, the key inputs to our approach. The framework is freely available as the R package phylosamp, and is broadly applicable to designing and evaluating a wide array of pathogen phylogenetic studies.Competing Interest StatementThe authors have declared no competing interest.Funding StatementFunding was provided by Bill and Melinda Gates Foundation OPP1195157 (S.W. and J.L.).Author DeclarationsI confirm all relevant ethical guidelines have been followed, and any necessary IRB and/or ethics committee approvals have been obtained.YesThe details of the IRB/oversight body that provided approval or exemption for the research described are given below:Not applicable: no clinical or health data was used in this study (simulations only).All necessary patient/participant consent has been obtained and the appropriate institutional forms have been archived.YesI understand that all clinical trials and any other prospective interventional studies must be registered with an ICMJE-approved registry, such as ClinicalTrials.gov. I confirm that any such study reported in the manuscript has been registered and the trial registration ID is provided (note: if posting a prospective study registered retrospectively, please provide a statement in the trial ID field explaining why the study was not registered in advance).YesI have followed all appropriate research reporting guidelines and uploaded the relevant EQUATOR Network research reporting checklist(s) and other pertinent material as supplementary files, if applicable.YesAll simulations and code used as a part of this manuscript are publicly available on github. https://github.com/HopkinsIDD/phylosamplesize https://github.com/HopkinsIDD/phylosamp