Makadia, RupaShoaibi, AzzaRao, Gowtham A.Ostropolets, AnnaRijnbeek, PeterVoss, Erica A.Duarte Salles, Talita, 1985-Ramírez Anguita, Juan ManuelMayer, Miguel Ángel, 1960-Maljkovic, FilipDenaxas, SpirosNyberg, FredrikPapez, VaclavSena, Anthony G.Alshammari, Thamir M.Lai, Lana Y. H.Haynes, KevinSuchard, Marc A.Hripcsak, GeorgeRyan, Patrick B.2024-10-222024-10-222023Makadia R, Shoaibi A, Rao GA, Ostropolets A, Rijnbeek PR, Voss EA, et al. Evaluating the impact of alternative phenotype definitions on incidence rates across a global data network. JAMIA Open. 2023 Nov 21;6(4):ooad096. DOI: 10.1093/jamiaopen/ooad0962574-2531http://hdl.handle.net/10230/68283Objective: Developing accurate phenotype definitions is critical in obtaining reliable and reproducible background rates in safety research. This study aims to illustrate the differences in background incidence rates by comparing definitions for a given outcome. Materials and methods: We used 16 data sources to systematically generate and evaluate outcomes for 13 adverse events and their overall background rates. We examined the effect of different modifications (inpatient setting, standardization of code set, and code set changes) to the computable phenotype on background incidence rates. Results: Rate ratios (RRs) of the incidence rates from each computable phenotype definition varied across outcomes, with inpatient restriction showing the highest variation from 1 to 11.93. Standardization of code set RRs ranges from 1 to 1.64, and code set changes range from 1 to 2.52. Discussion: The modification that has the highest impact is requiring inpatient place of service, leading to at least a 2-fold higher incidence rate in the base definition. Standardization showed almost no change when using source code variations. The strength of the effect in the inpatient restriction is highly dependent on the outcome. Changing definitions from broad to narrow showed the most variability by age/gender/database across phenotypes and less than a 2-fold increase in rate compared to the base definition. Conclusion: Characterization of outcomes across a network of databases yields insights into sensitivity and specificity trade-offs when definitions are altered. Outcomes should be thoroughly evaluated prior to use for background rates for their plausibility for use across a global network.application/pdfeng© The Author(s) 2023. Published by Oxford University Press on behalf of the American Medical Informatics Association. This is an Open Access article distributed under the terms of the Creative Commons Attribution License (https://creativecommons.org/licenses/by/4.0/), which permits unrestricted reuse, distribution, and reproduction in any medium, provided the original work is properly cited.Evaluating the impact of alternative phenotype definitions on incidence rates across a global data networkinfo:eu-repo/semantics/articlehttp://dx.doi.org/10.1093/jamiaopen/ooad096AlgorithmsElectronic health recordIncidence studyPhenotypeinfo:eu-repo/semantics/openAccess