Physics-informed graph neural networks for predicting cetane number with systematic data quality analysis

被引:20
作者
Kim, Yeonjoon [1 ]
Cho, Jaeyoung [2 ]
Naser, Nimal [2 ]
Kumar, Sabari [1 ]
Jeong, Keunhong [1 ]
McCormick, Robert L. [2 ]
John, Peter C. St. [2 ]
Kim, Seonah [1 ,2 ]
机构
[1] Colorado State Univ, Ft Collins, CO 80523 USA
[2] Natl Renewable Energy Lab, 15013 Denver W Pkwy, Golden, CO 80401 USA
关键词
Cetane number; Machine learning; Graph neural network; Data curation; BIOFUEL; COMBUSTION; GASOLINE;
D O I
10.1016/j.proci.2022.09.059
中图分类号
O414.1 [热力学];
学科分类号
摘要
Designing alternative fuels for advanced compression ignition engines necessitates a predictive model for cetane number (CN). In this study, the physics-informed graph neural networks are introduced for a reliable CN prediction by considering molecular features pertinent to the physical properties of molecules that affect CN. The reliability of measured data is another key factor to consider for improving the predictive model. Various experimental instruments for measuring CN exist, including standard and non-standard methods. In this regard, a systematic data quality analysis was carried out for the total 630 CNs collected from literature and new measurements in this study using Advanced Fuel Ignition Delay Analyzer (AFIDA). The results from this data curation process were reflected in the model by imposing lower sample weights on the data coming from less reliable measurement techniques. This approach effectively maximized the prediction accuracy while incorporating data from all available sources. Using the sample weights decreased the mean absolute error (MAE) up to 0.8 CN units. The accuracy was also improved by introducing the CN-related physical properties (the number of hydrogen bond donors and acceptors); the test set MAE is 5.74 and 7.01 for the model with and without such properties, respectively. Investigating molecular structural effects on CN was also carried out to gain chemical insights into factors used to design new fuel candidates. The dimensionality reduction analysis of feature vectors showed a clear clustering in terms of functional groups and CN and the structural effect derived from the model was consistent with the physicochemical insights. This physics-informed model and data curation would be helpful for accurate CN prediction and inform rational fuel design.& COPY; 2022 The Combustion Institute. Published by Elsevier Inc. All rights reserved.
引用
收藏
页码:4969 / 4978
页数:10
相关论文
共 33 条
  • [1] Abadi M, 2016, PROCEEDINGS OF OSDI'16: 12TH USENIX SYMPOSIUM ON OPERATING SYSTEMS DESIGN AND IMPLEMENTATION, P265
  • [2] [Anonymous], 2018, ASTM D8183-18
  • [3] [Anonymous], 2018, ASTM D613-18ae1
  • [4] [Anonymous], 2018, D689018 ASTM INT
  • [5] [Anonymous], 2020, Standard Specification for Wrought Titanium6Aluminum4Vanadium Alloy for Surgical Implant Applications (UNS R56400)
  • [6] HEATS OF VAPORIZATION OF HYDROGEN-BONDED SUBSTANCES
    BONDI, A
    SIMKIN, DJ
    [J]. AICHE JOURNAL, 1957, 3 (04) : 473 - 479
  • [7] Cho J., SUSTAINABLE ENERGY F
  • [8] Cho K., 2009, SAE TECHNICAL PAPER
  • [9] Gaspar D. J., 2021, TOP 13 BLENDSTOCKS D
  • [10] Gilmer J, 2017, PR MACH LEARN RES, V70