Detecting irregularities in randomized controlled trials using machine learning

被引：0

作者：

Nelson, Walter ^{[1
]}

Petch, Jeremy ^{[1
,2
,3
]}

Ranisau, Jonathan

Zhao, Robin

Balasubramanian, Kumar

Bangdiwala, Shrikant, I ^{[2
]}

机构：

[1] Hamilton Hlth Sci, Ctr Data Sci & Digital Hlth, Hamilton, ON, Canada

[2] Populat Hlth Res Inst, 20 Copeland Ave, Hamilton L8L 2X2, ON, Canada

[3] McMaster Univ, Dept Med, Hamilton, ON, Canada

来源：

CLINICAL TRIALS | 2024年

关键词：

Central statistical monitoring; machine learning; artificial intelligence; outlier detection; quality assurance; data quality; randomized controlled trials; DABIGATRAN;

D O I：

10.1177/17407745241297947

中图分类号：

R-3 [医学研究方法]; R3 [基础医学];

学科分类号：

1001 ;

摘要：

Background: Over the course of a clinical trial, irregularities may arise in the data. Trialists implement human-intensive, expensive central statistical monitoring procedures to identify and correct these irregularities before the results of the trial are analyzed and disseminated. Machine learning algorithms have shown promise for identifying center-level irregularities in multi-center clinical trials with minimal human intervention. We aimed to characterize the form-level data irregularities in several historical clinical trials and evaluate the ability of a machine learning-based outlier detection algorithm to identify them.Methods: Data irregularities previously identified by humans in historical clinical trials were ascertained by comparing preliminary snapshots of the trial databases to the final, locked databases. We measured the ability of a machine learning based outlier detection algorithm to identify form-level irregularities using concordance (area under the receiver operator characteristic), positive predictive value (precision), and sensitivity (recall).Results: We examined preliminary snapshots of seven historical clinical trials which randomized a total of 77,001 participants. We extracted a total of 1,267,484 completed entries from 358 case report forms containing irregularities from all snapshots across all trials, containing a total of 24,850 form-wide irregularities (median per-form form-level irregularity rate: 1.81%). Our proposed machine learning algorithm detects form-level irregularities with a median concordance of 0.74 (interquartile range = 0.57-0.89), slightly exceeding the performance of a previously proposed machine learning approach with a median area under the receiver operator characteristic of 0.73 (interquartile range = 0.54-0.88).Conclusion: Data irregularities in historical clinical trials were ascertained by comparing preliminary snapshots of the trial database to the final database. These irregularities can be categorized according to their scope. Irregularities can be successfully detected by a machine learning algorithm as early or earlier than a human can, without human intervention. Such an approach may complement existing techniques for central statistical monitoring in large multi-center randomized controlled trials and possibly improve the efficiency of costly data verification processes.

引用

页码：178 / 187

页数：10

共 50 条

[21] Machine learning analysis plans for randomised controlled trials: detecting treatment effect heterogeneity with strict control of type I error
James A. Watson
Chris C. Holmes
Trials, 21
[22] Detecting Fake News Using Machine Learning Algorithms
Bharath, G.
Manikanta, K. J.
Prakash, G. Bhanu
Sumathi, R.
Chinnasamy, P.
2021 INTERNATIONAL CONFERENCE ON COMPUTER COMMUNICATION AND INFORMATICS (ICCCI), 2021,
[23] Detecting BGP Anomalies Using Machine Learning Techniques
Ding, Qingye
Li, Zhida
Batta, Prerna
Trajkovic, Ljiljana
2016 IEEE INTERNATIONAL CONFERENCE ON SYSTEMS, MAN, AND CYBERNETICS (SMC), 2016, : 3352 - 3355
[24] Detecting Suspicious Texts Using Machine Learning Techniques
Sharif, Omar
Hoque, Mohammed Moshiul
Kayes, A. S. M.
Nowrozy, Raza
Sarker, Iqbal H.
APPLIED SCIENCES-BASEL, 2020, 10 (18):
[25] Detecting Under-Resolved Flow Physics Using Supervised Machine Learning
Hedayat, Amirpasha
Ollivier-Gooch, Carl
AIAA JOURNAL, 2023, 61 (09) : 3958 - 3975
[26] RANDOMIZED, CONTROLLED TRIALS USING THE METRO FIRM SYSTEM
CEBUL, RD
MEDICAL CARE, 1991, 29 (07) : JS9 - JS18
[27] Machine learning analysis plans for randomised controlled trials: detecting treatment effect heterogeneity with strict control of type I error
Watson, James A.
Holmes, Chris C.
TRIALS, 2020, 21 (01)
[28] An Explainable Machine Learning-Based Phenomapping Strategy for Adaptive Predictive Enrichment in Randomized Controlled Trials of Preventive Interventions
Oikonomou, Evangelos K.
Thangaraj, Phyllis M.
Bhatt, Deepak L.
Ross, Joseph S.
Young, Lawrence H.
Krumholz, Harlan M.
Suchard, Marc A.
Khera, Rohan
CIRCULATION, 2023, 148
[29] Machine-learning approaches to predict individualized treatment effect using a randomized controlled trial
Hamaya, Rikuta
Hara, Konan
Manson, JoAnn E.
Rimm, Eric B.
Sacks, Frank M.
Xue, Qiaochu
Qi, Lu
Cook, Nancy R.
EUROPEAN JOURNAL OF EPIDEMIOLOGY, 2025, : 151 - 166
[30] Comparing machine and human reviewers to evaluate the risk of bias in randomized controlled trials
Armijo-Olivo, Susan
Craig, Rodger
Campbell, Sandy
RESEARCH SYNTHESIS METHODS, 2020, 11 (03) : 484 - 493

← 1 2 3 4 5 →