Herrera-Luis, E., Benke, K., Volk, H., Ladd-Acosta, C. & Wojcik, G. L. Gene–environment interactions in human health. Nat. Rev. Genet. 25, 768–784 (2024).
Boye, C., Nirmalan, S., Ranjbaran, A. & Luca, F. Genotype × environment interactions in gene regulation and complex traits. Nat. Genet. 56, 1057–1068 (2024).
Motsinger-Reif, A. A. et al. Gene–environment interactions within a precision environmental health framework. Cell Genom. 4, 100591 (2024).
Virolainen, S. J., VonHandorf, A., Viel, K. C. M. F., Weirauch, M. T. & Kottyan, L. C. Gene–environment interactions and their impact on human health. Genes Immunity 24, 1–11 (2022).
Bi, W. et al. A fast and accurate method for genome-wide scale phenome-wide G × E analysis and its application to UK Biobank. Am. J. Hum. Genet. 105, 1182–1192 (2019).
Zhong, W., Chhibber, A., Luo, L., Mehrotra, D. V. & Shen, J. A fast and powerful linear mixed model approach for genotype–environment interaction tests in large-scale GWAS. Brief. Bioinform. 24, bbac547 (2023).
Ma, Y., Zhao, Y., Zhang, J. F. & Bi, W. Efficient and accurate framework for genome-wide gene–environment interaction analysis in large-scale biobanks. Nat. Commun. 16, 3064 (2025).
Miao, J., Wu, Y. & Lu, Q. Statistical methods for gene–environment interaction analysis. WIREs Comput. Stat. 16, e1635 (2024).
Laird, N. M. & Ware, J. H. Random-effects models for longitudinal data. Biometrics 38, 963–974 (1982).
Bates, D., Mächler, M., Bolker, B. & Walker, S. Fitting linear mixed-effects models using lme4. J. Stat. Softw. 67, 1–48 (2015).
Warrington, N. M. et al. A genome-wide association study of body mass index across early life and childhood. Int. J. Epidemiol. 44, 700–712 (2015).
Ko, S. et al. GWAS of longitudinal trajectories at biobank scale. Am. J. Hum. Genet. 109, 433–445 (2022).
Xu, H. et al. SPA(GRM): effectively controlling for sample relatedness in large-scale genome-wide association studies of longitudinal traits. Nat. Commun. 16, 1413 (2025).
Sikorska, K. et al. Fast linear mixed model computations for genome-wide association studies with longitudinal data. Stat. Med. 32, 165–180 (2013).
Meirelles, O. D. et al. SHAVE: shrinkage estimator measured for multiple visits increases power in GWAS of quantitative traits. Eur. J. Hum. Genet. 21, 673–679 (2013).
Robinson-Cohen, C. et al. Genome-wide association study of CKD progression. J. Am. Soc. Nephrol. 34, 1547–1559 (2023).
Venkatesh, S. S. et al. Characterising the genetic architecture of changes in adiposity during adulthood using electronic health records. Nat. Commun. 15, 5801 (2024).
Savic, R. M. & Karlsson, M. O. Importance of shrinkage in empirical bayes estimates for diagnostics: problems and solutions. AAPS J. 11, 558–569 (2009).
Yuan, M. et al. A quick and accurate method for the estimation of covariate effects based on empirical Bayes estimates in mixed-effects modeling: Correction of bias due to shrinkage. Stat. Methods Med. Res. 28, 3568–3578 (2019).
Wiegrebe, S. et al. Analyzing longitudinal trait trajectories using GWAS identifies genetic variants for kidney function decline. Nat. Commun. 15, 10061 (2024).
Sikorska, K., Lesaffre, E., Groenen, P. J. F., Rivadeneira, F. & Eilers, P. H. C. Genome-wide analysis of large-scale longitudinal outcomes using penalization—GALLOP algorithm. Sci. Rep. 8, 6815 (2018).
Yuan, M. et al. SCEBE: an efficient and scalable algorithm for genome-wide association studies on longitudinal outcomes with mixed-effects modeling. Brief. Bioinform. 22, bbaa130 (2021).
Jiang, D., Mbatchou, J. & McPeek, M. S. Retrospective association analysis of binary traits: overcoming some limitations of the additive polygenic model. Hum. Hered. 80, 187–195 (2015).
Jakobsdottir, J. & McPeek, M. S. MASTOR: mixed-model association mapping of quantitative traits in samples with related individuals. Am. J. Hum. Genet. 92, 652–666 (2013).
Kuonen, D. Miscellanea. Saddlepoint approximations for distributions of quadratic forms in normal variables. Biometrika 86, 929–935 (1999).
Dey, R., Schmidt, E. M., Abecasis, G. R. & Lee, S. A fast and accurate algorithm to test for binary phenotypes and its application to PheWAS. Am. J. Hum. Genet. 101, 37–49 (2017).
Yang, J. et al. Genomic inflation factors under polygenic inheritance. Eur. J. Hum. Genet. 19, 807–812 (2011).
Tchetgen Tchetgen, E. Robust discovery of genetic associations incorporating gene–environment interaction and independence. Epidemiology 22, 262–272 (2011).
Liu, Y., Li, P., Song, L., Yu, K. & Qin, J. Retrospective versus prospective score tests for genetic association with case-control data. Biometrics 77, 102–112 (2021).
Winkler, T. W. et al. Genetic-by-age interaction analyses on complex traits in UK Biobank and their potential to identify effects on longitudinal trait change. Genome Biol. 25, 300 (2024).
McCarthy, M. I. et al. Genetic determinants of long-term changes in blood lipid concentrations: 10-year follow-up of the GLACIER Study. PLoS Genet. 10, e1004388 (2014).
Kemper, K. E. et al. Genetic influence on within-person longitudinal change in anthropometric traits in the UK Biobank. Nat. Commun. 15, 3776 (2024).
Weiler, E. S., Szabo, T. G., Garcia-Carpio, I. & Villunger, A. PIDD1 in cell cycle control, sterile inflammation and cell death. Biochem. Soc. Trans. 50, 813–824 (2022).
Zhu, X., Zhu, L., Wang, H., Cooper, R. S. & Chakravarti, A. Genome-wide pleiotropy analysis identifies novel blood pressure variants and improves its polygenic risk scores. Genet. Epidemiol. 46, 105–121 (2022).
Consortium, G. T. The Genotype-Tissue Expression (GTEx) project. Nat. Genet. 45, 580–585 (2013).
Shirts, B. H., Hasstedt, S. J., Hopkins, P. N. & Hunt, S. C. Evaluation of the gene–age interactions in HDL cholesterol, LDL cholesterol, and triglyceride levels: the impact of the SORT1 polymorphism on LDL cholesterol levels is age dependent. Atherosclerosis 217, 139–141 (2011).
Gorski, M. et al. Genetic loci and prioritization of genes for kidney function decline derived from a meta-analysis of 62 longitudinal genome-wide association studies. Kidney Int. 102, 624–639 (2022).
Young, A. I., Wauthier, F. & Donnelly, P. Multiple novel gene-by-environment interactions modify the effect of FTO variants on body mass index. Nat. Commun. 7, 12724 (2016).
Dang, L. C. et al. FTO affects food cravings and interacts with age to influence age-related decline in food cravings. Physiol. Behav. 192, 188–193 (2018).
He, D. et al. FTO gene variant and risk of hypertension: a meta-analysis of 57,464 hypertensive cases and 41,256 controls. Metabolism 63, 633–639 (2014).
Signer, R. et al. BMI interacts with the genome to regulate gene expression globally, with emphasis in the brain and gut. Preprint at medRxiv https://doi.org/10.1101/2024.11.26.24317923 (2024).
Sveinbjornsson, G. et al. Multiomics study of nonalcoholic fatty liver disease. Nat. Genet. 54, 1652–1663 (2022).
Xi, B. et al. The common SNP (rs9939609) in the FTO gene modifies the association between obesity and high blood pressure in Chinese children. Mol. Biol. Rep. 40, 773–778 (2013).
Westerman, K. E. & Sofer, T. Many roads to a gene–environment interaction. Am. J. Hum. Genet. 111, 626–635 (2024).
Johnson, S. M. et al. PNPLA3 is a triglyceride lipase that mobilizes polyunsaturated fatty acids to facilitate hepatic secretion of large-sized very low-density lipoprotein. Nat. Commun. 15, 4847 (2024).
Sookoian, S. & Pirola, C. J. Genetics in non-alcoholic fatty liver disease: the role of risk alleles through the lens of immune response. Clin. Mol. Hepatol. 29, S184–S195 (2023).
Stojkovic, I. A. et al. The PNPLA3 Ile148Met interacts with overweight and dietary intakes on fasting triglyceride levels. Genes Nutr. 9, 388 (2014).
Stender, S. et al. Adiposity amplifies the genetic risk of fatty liver disease conferred by multiple loci. Nat. Genet. 49, 842–847 (2017).
Del Bosque-Plata, L., Martinez-Martinez, E., Espinoza-Camacho, M. A. & Gragnoli, C. The role of TCF7L2 in type 2 diabetes. Diabetes 70, 1220–1228 (2021).
Li, L. et al. Interaction analysis of gene variants of TCF7L2 and body mass index and waist circumference on type 2 diabetes. Clin. Nutr. 39, 192–197 (2020).
Dudbridge, F. & Fletcher, O. Gene–environment dependence creates spurious gene–environment interaction. Am. J. Hum. Genet. 95, 301–307 (2014).
Wu, Y. et al. Genome-wide association study of medication-use and associated disease in the UK Biobank. Nat. Commun. 10, 1891 (2019).
Kiiskinen, T. et al. Genetic predictors of lifelong medication-use patterns in cardiometabolic diseases. Nat. Med. 29, 209–218 (2023).
Sadowski, M. et al. Characterizing the genetic architecture of drug response using gene-context interaction methods. Cell Genom. 4, 100722 (2024).
Svishcheva, G. R., Axenovich, T. I., Belonogova, N. M., van Duijn, C. M. & Aulchenko, Y. S. Rapid variance components-based method for whole-genome association analysis. Nat. Genet. 44, 1166–1170 (2012).
Loh, P. R. et al. Efficient Bayesian mixed-model analysis increases association power in large cohorts. Nat. Genet. 47, 284–290 (2015).
Zhou, W. et al. Efficiently controlling for case-control imbalance and sample relatedness in large-scale genetic association studies. Nat. Genet. 50, 1335–1341 (2018).
Thornton, T. et al. Estimating kinship in admixed populations. Am. J. Hum. Genet. 91, 122–138 (2012).
Chow, C. & Liu, C. Approximating discrete probability distributions with dependence trees. IEEE Trans. Inform. Theory 14, 462–467 (1968).
Barndorff-Nielsen, O. E. Approximate interval probabilities. J. R. Stat. Soc. Series B 52, 485–496 (1990).
Jiang, L. et al. A resource-efficient tool for mixed model association analysis of large-scale data. Nat. Genet. 51, 1749–1755 (2019).
Gorski, M. et al. Bias-corrected serum creatinine from UK Biobank electronic medical records generates an important data resource for kidney function trajectories. Sci. Rep. 15, 3540 (2025).
Inker, L. A. et al. New creatinine- and cystatin C-based equations to estimate GFR without race. N. Engl. J. Med. 385, 1737–1749 (2021).
Sadler, M. C. et al. Leveraging large-scale biobank EHRs to enhance pharmacogenetics of cardiometabolic disease medications. Nat. Commun. 16, 2913 (2025).
Loh, P. R. et al. Reference-based phasing using the Haplotype Reference Consortium panel. Nat. Genet. 48, 1443–1448 (2016).
Wang, K., Li, M. & Hakonarson, H. ANNOVAR: functional annotation of genetic variants from high-throughput sequencing data. Nucleic Acids Res. 38, e164 (2010).
Mutie, P. M. et al. Investigating the causal relationships between excess adiposity and cardiometabolic health in men and women. Diabetologia 66, 321–335 (2023).
Xu, H. & Bi, W. Leveraging longitudinal data from electronic health records to boost statistical power for gene-environment interaction analysis. Zenodo https://doi.org/10.5281/zenodo.16345203 (2025).
Xu, H. HeXuPKU/SAGELD: SAGELD analyses. Zenodo https://doi.org/10.5281/zenodo.19556109 (2026).