Bick AG, Metcalf GA, Mayo KR, Lichtenstein L, Rura S, Carroll RJ, Musick A, Linder JE, Jordan IK, Nagar SD, Sharma S, Meller R, Basford M, Boerwinkle E, Cicek MS, Doheny KF, Eichler EE, Gabriel S, Gibbs RA, Glazer D, Harris PA, Jarvik GP, Philippakis A, Rehm HL, Roden DM, Thibodeau SN, Topper S, Blegen AL, Wirkus SJ, Wagner VA, Meyer JG, Cicek MS, Muzny DM, Venner E, Mawhinney MZ, Griffith SML, Hsu E, Ling H, Adams MK, Walker K, Hu J, Doddapaneni H, Kovar CL, Murugan M, Dugan S, Khan Z, Boerwinkle E, Lennon NJ, Austin-Tse C, Banks E, Gatzen M, Gupta N, Henricks E, Larsson K, McDonough S, Harrison SM, Kachulis C, Lebo MS, Neben CL, Steeves M, Zhou AY, Smith JD, Frazar CD, Davis CP, Patterson KE, Wheeler MM, McGee S, Lockwood CM, Shirts BH, Pritchard CC, Murray ML, Vasta V, Leistritz D, Richardson MA, Buchan JG, Radhakrishnan A, Krumm N, Ehmen BW, Schwartz S, Aster MMT, Cibulskis K, Haessly A, Asch R, Cremer A, Degatano K, Shergill A, Gauthier LD, Lee SK, Hatcher A, Grant GB, Brandt GR, Covarrubias M, Banks E, Able A, Green AE, Carroll RJ, Zhang J, Condon HR, Wang Y, Dillon MK, Albach CH, Baalawi W, Choi SH, Wang X, Rosenthal EA, Ramirez AH, Lim S, Nambiar S, Ozenberger B, Wise AL, Lunt C, Ginsburg GS, Denny JC. Genomic data in the All of Us Research Program.
Nature 2024;
627:340-346. [PMID:
38374255 PMCID:
PMC10937371 DOI:
10.1038/s41586-023-06957-x]
[Citation(s) in RCA: 0] [Impact Index Per Article: 0] [Reference Citation Analysis] [Abstract] [Key Words] [MESH Headings] [Grants] [Track Full Text] [Journal Information] [Subscribe] [Scholar Register] [Received: 07/22/2022] [Accepted: 12/08/2023] [Indexed: 02/21/2024]
Abstract
Comprehensively mapping the genetic basis of human disease across diverse individuals is a long-standing goal for the field of human genetics1-4. The All of Us Research Program is a longitudinal cohort study aiming to enrol a diverse group of at least one million individuals across the USA to accelerate biomedical research and improve human health5,6. Here we describe the programme's genomics data release of 245,388 clinical-grade genome sequences. This resource is unique in its diversity as 77% of participants are from communities that are historically under-represented in biomedical research and 46% are individuals from under-represented racial and ethnic minorities. All of Us identified more than 1 billion genetic variants, including more than 275 million previously unreported genetic variants, more than 3.9 million of which had coding consequences. Leveraging linkage between genomic data and the longitudinal electronic health record, we evaluated 3,724 genetic variants associated with 117 diseases and found high replication rates across both participants of European ancestry and participants of African ancestry. Summary-level data are publicly available, and individual-level data can be accessed by researchers through the All of Us Researcher Workbench using a unique data passport model with a median time from initial researcher registration to data access of 29 hours. We anticipate that this diverse dataset will advance the promise of genomic medicine for all.
Collapse