Individual dynamic prediction of clinical endpoint from large dimensional longitudinal biomarker history: a landmark approach

被引:8
作者
Devaux, Anthony [1 ]
Genuer, Robin [1 ,2 ]
Peres, Karine [1 ]
Proust-Lima, Cecile [1 ]
机构
[1] Univ Bordeaux, INSERM, U1219, BPH, Bordeaux, France
[2] INRIA Bordeaux Sud Ouest, Talence, France
关键词
Individual prediction; Landmark; Longitudinal data; Survival data; Machine learning methods; TIME-TO-EVENT; REGRESSION; MODELS;
D O I
10.1186/s12874-022-01660-3
中图分类号
R19 [保健组织与事业(卫生事业管理)];
学科分类号
摘要
Background The individual data collected throughout patient follow-up constitute crucial information for assessing the risk of a clinical event, and eventually for adapting a therapeutic strategy. Joint models and landmark models have been proposed to compute individual dynamic predictions from repeated measures to one or two markers. However, they hardly extend to the case where the patient history includes much more repeated markers. Our objective was thus to propose a solution for the dynamic prediction of a health event that may exploit repeated measures of a possibly large number of markers. Methods We combined a landmark approach extended to endogenous markers history with machine learning methods adapted to survival data. Each marker trajectory is modeled using the information collected up to the landmark time, and summary variables that best capture the individual trajectories are derived. These summaries and additional covariates are then included in different prediction methods adapted to survival data, namely regularized regressions and random survival forests, to predict the event from the landmark time. We also show how predictive tools can be combined into a superlearner. The performances are evaluated by cross-validation using estimators of Brier Score and the area under the Receiver Operating Characteristic curve adapted to censored data. Results We demonstrate in a simulation study the benefits of machine learning survival methods over standard survival models, especially in the case of numerous and/or nonlinear relationships between the predictors and the event. We then applied the methodology in two prediction contexts: a clinical context with the prediction of death in primary biliary cholangitis, and a public health context with age-specific prediction of death in the general elderly population. Conclusions Our methodology, implemented in R, enables the prediction of an event using the entire longitudinal patient history, even when the number of repeated markers is large. Although introduced with mixed models for the repeated markers and methods for a single right censored time-to-event, the technique can be used with any other appropriate modeling technique for the markers and can be easily extended to competing risks setting.
引用
收藏
页数:14
相关论文
共 41 条
  • [1] Albert PS, 2010, BIOMETRICS, V66, P983, DOI [10.1111/j.1541-0420.2009.01324.x, 10.1111/j.1541-0420.2009.01324_1.x]
  • [2] Deviance residuals-based sparse PLS and sparse kernel PLS regression for censored data
    Bastien, Philippe
    Bertrand, Frederic
    Meyer, Nicolas
    Maumy-Bertrand, Myriam
    [J]. BIOINFORMATICS, 2015, 31 (03) : 397 - 404
  • [3] Quantifying and Comparing Dynamic Predictive Accuracy of Joint Models for Longitudinal Marker and Time-to-Event in Presence of Censoring and Competing Risks
    Blanche, Paul
    Proust-Lima, Cecile
    Loubere, Lucie
    Berr, Claudine
    Dartigues, Jean-Francois
    Jacqmin-Gadda, Helene
    [J]. BIOMETRICS, 2015, 71 (01) : 102 - 113
  • [4] Estimating and comparing time-dependent areas under receiver operating characteristic curves for censored event times with competing risks
    Blanche, Paul
    Dartigues, Jean-Francois
    Jacqmin-Gadda, Helene
    [J]. STATISTICS IN MEDICINE, 2013, 32 (30) : 5381 - 5397
  • [5] Breiman L., 2001, Machine Learning, V45, P5
  • [6] Sparse partial least squares regression for simultaneous dimension reduction and variable selection
    Chun, Hyonho
    Keles, Suenduez
    [J]. JOURNAL OF THE ROYAL STATISTICAL SOCIETY SERIES B-STATISTICAL METHODOLOGY, 2010, 72 : 3 - 25
  • [7] Individual dynamic predictions using landmarking and joint modelling: Validation of estimators and robustness assessment
    Ferrer, Loic
    Putter, Hein
    Proust-Lima, Cecile
    [J]. STATISTICAL METHODS IN MEDICAL RESEARCH, 2019, 28 (12) : 3649 - 3666
  • [8] L1 Penalized Estimation in the Cox Proportional Hazards Model
    Goeman, Jelle J.
    [J]. BIOMETRICAL JOURNAL, 2010, 52 (01) : 70 - 84
  • [9] Moving beyond regression techniques in cardiovascular risk prediction: applying machine learning to address analytic challenges
    Goldstein, Benjamin A.
    Navar, Ann Marie
    Carter, Rickey E.
    [J]. EUROPEAN HEART JOURNAL, 2017, 38 (23) : 1805 - 1814
  • [10] Super Learner for Survival Data Prediction
    Golmakani, Marzieh K.
    Polley, Eric C.
    [J]. INTERNATIONAL JOURNAL OF BIOSTATISTICS, 2020, 16 (02)