Finding and understanding pedal misapplication crashes using a deep learning natural language model

被引:4
|
作者
Bareiss, Max [1 ]
Smith, Colin [1 ]
Gabler, Hampton C. [1 ]
机构
[1] Virginia Tech, Dept Biomed Engn, Blacksburg, VA USA
关键词
Pedal misapplication; NMVCCS; deep learning; BERT; NLP;
D O I
10.1080/15389588.2021.1982616
中图分类号
R1 [预防医学、卫生学];
学科分类号
1004 ; 120402 ;
摘要
Objective The objective of this study was to develop a system which used the BERT natural language understanding model to identify pedal misapplication (PM) crashes from their crash narratives and validate the accuracy of the system. Methods The training dataset used for this study was 11 cases from the NMVCCS study and 952 cases from the North Carolina state crash database. Cases for this study were selected from their respective full datasets using a keyword search algorithm containing terms indicative of a pedal-related mistake. A BERT language model was used to classify each case narrative as either no pedal misapplication, PM by vehicle 1, PM by vehicle 2, or PM by vehicle 3. After training, the language model was used to determine the incidence of pedal misapplication in a test dataset of 8,668 North Carolina and NMVCCS cases and these results were compared to a manual review of the dataset. After manual review, 2,969 cases were pedal misapplications. Results The model's AUC ROC performance at detecting PM was quantified on the entire testing dataset to evaluate the power of the system to generalize to case narratives unseen at training time. The AUC ROC value was 0.9835, indicating strong generalization to all crash narratives. By choosing the optimal threshold using the ROC curve, the system correctly identified PM in 95.7% of crash narratives. When pedal misapplication was correctly identified, the correct vehicle was identified in 95.9% of cases. A total of 3,062 pedal misapplications were identified. The model labeled cases 353 times faster than a researcher. Conclusions The strong performance of the model suggests that the automated interpretation of case narratives can be used for future research studies without any manual review. This would save time and enable the use of datasets where manual review would be infeasible. The automated extraction of information from crash narratives using deep learning natural language models has not been demonstrated previously in the literature, to the best of the authors' knowledge. This technique can be applied to large, infrequently used datasets of crash narratives and extended to extract useful vehicle, occupant, or environment information to make these datasets amenable to traditional statistical analyses.
引用
收藏
页码:S169 / S172
页数:4
相关论文
共 50 条
  • [41] Two-stream video-based deep learning model for crashes and near-crashes
    Shi, Liang
    Guo, Feng
    TRANSPORTATION RESEARCH PART C-EMERGING TECHNOLOGIES, 2024, 166
  • [42] A Survey of the Usages of Deep Learning for Natural Language Processing
    Otter, Daniel W.
    Medina, Julian R.
    Kalita, Jugal K.
    IEEE TRANSACTIONS ON NEURAL NETWORKS AND LEARNING SYSTEMS, 2021, 32 (02) : 604 - 624
  • [43] Inflectional Review of Deep Learning on Natural Language Processing
    Fahad, S. K. Ahammad
    Yahya, Abdulsamad Ebrahim
    2018 INTERNATIONAL CONFERENCE ON SMART COMPUTING AND ELECTRONIC ENTERPRISE (ICSCEE), 2018,
  • [44] Are Deep Learning Approaches Suitable for Natural Language Processing?
    Alshahrani, S.
    Kapetanios, E.
    NATURAL LANGUAGE PROCESSING AND INFORMATION SYSTEMS, NLDB 2016, 2016, 9612 : 343 - 349
  • [45] A Deep Learning Architecture for Psychometric Natural Language Processing
    Ahmad, Faizan
    Abbasi, Ahmed
    Li, Jingjing
    Dobolyi, David G.
    Netemeyer, Richard G.
    Clifford, Gari D.
    Chen, Hsinchun
    ACM TRANSACTIONS ON INFORMATION SYSTEMS, 2020, 38 (01)
  • [46] Using deep learning and natural language processing models to detect child physical abuse
    Shahi, Niti
    Shahi, Ashwani K.
    Phillips, Ryan
    Shirek, Gabrielle
    Lindberg, Daniel M.
    Moulton, Steven L.
    JOURNAL OF PEDIATRIC SURGERY, 2021, 56 (12) : 2326 - 2332
  • [47] An approach to detect offence in Memes using Natural Language Processing(NLP) and Deep learning
    Giri, Roushan Kumar
    Gupta, Subhash Chandra
    Gupta, Umesh Kumar
    2021 INTERNATIONAL CONFERENCE ON COMPUTER COMMUNICATION AND INFORMATICS (ICCCI), 2021,
  • [48] ENHANCING NATURAL LANGUAGE PROCESSING USING WALRUS OPTIMIZER WITH SELF-ATTENTION DEEP LEARNING MODEL IN APPLIED LINGUISTICS
    Hassan, Abdulkhaleq Q. A.
    Alshammari, Alya
    Zaqaibeh, Belal
    Alzaidi, Muhammad Swaileh a.
    Allafi, Randa
    Alazwari, Sana
    Aljabri, Jawhara
    Nouri, Amal M.
    FRACTALS-COMPLEX GEOMETRY PATTERNS AND SCALING IN NATURE AND SOCIETY, 2024,
  • [49] Fake News Detection Using Feature Extraction, Natural Language Processing, Curriculum Learning, and Deep Learning
    Madani, Mirmorsal
    Motameni, Homayun
    Roshani, Reza
    INTERNATIONAL JOURNAL OF INFORMATION TECHNOLOGY & DECISION MAKING, 2024, 23 (03) : 1063 - 1098
  • [50] Understanding Human Language: Can NLP and Deep Learning Help?
    Manning, Christopher
    SIGIR'16: PROCEEDINGS OF THE 39TH INTERNATIONAL ACM SIGIR CONFERENCE ON RESEARCH AND DEVELOPMENT IN INFORMATION RETRIEVAL, 2016, : 1 - 1