Scholay

学术搜索 · AI 审稿 · LaTeX 协作

Effect of Molecular Descriptors on the Development of Machine Learning Models for the Prediction of Yield Sooting Index

作者:Quan‐De Wang, Lan Du, Yao Qian, Jinhu Liang, Bi-Yao Wang, Ping Zeng, Zuxi Xia · 发表于:ACS Omega · 年份:2025 · DOI:10.1021/acsomega.5c08720 · 被引用次数:2 · 研究领域:Advanced Combustion Engine Technologies、Combustion and flame dynamics、Thermochemical Biomass Conversion Processes

High Resolution Image Download MS PowerPoint Slide Sooting propensity is a critical property to estimate the combustion efficiency and pollution emissions of a fuel and also to discover the next generation of cleaner and more efficient fuels. Yield sooting index (YSI) is an important metric to characterize the sooting propensity; however, it is inefficient to measure this experimentally. Thus, the development of machine learning (ML)-based predictive models exists as an important instrument to predict the YSI for fuel design. Herein, this work compares the accuracies and interpretability of four ML models to predict the YSI based on different kinds of descriptors. It is demonstrated that the developed best ML models using different kinds of descriptors are different. The multilayer perceptron (MLP) regressor neural network (NN), gradient boosting (GB), and random forest (RF) models are the best models for the PaDEL, mordred, and quantum mechanical (QM) descriptors, respectively. The NN model is suitable for the combination of QM descriptors with full PaDEL and mordred descriptors, while the RF model is better for the combination of QM descriptors with PaDEL and mordred descriptors after the permutation feature importance (PFI) filtering procedure. The usage of QM descriptors can slightly improve the deep-learning-based ML model performance. The developed ML models can all predict the YSI with high accuracy, i.e., the coefficient of determination ( R 2 ) is close to 1.0, and t...