返回
Interpretable machine learning-based text classification method for construction quality defect reports
DOI:10.1016/j.jobe.2024.109330.png)
摘要
En 中文
Efficient identification and remediation of construction defects are critical for ensuring the quality and success of engineering projects. However, the complexity of construction environments poses challenges to this endeavor. Current research predominantly relies on statistical and causal analyses of defect detection reports, yet these methods are time-consuming and errorprone due to the unstructured nature of such reports. To address this, machine learning techniques have been applied to classify defect texts rapidly and accurately. However, existing studies primarily focus on model performance enhancement, neglecting interpretability and the effect of imbalanced data. This study introduces RF-SMOTE, an oversampling technique based on Random Forest (RF), to address the limitations of traditional methods like SMOTE. Comparative analyses demonstrated the efficacy of RF-SMOTE in mitigating imbalanced data effects. Further, the application of SHAP-based interpretability methods in construction management decision -making was explored, filling gaps in existing research. Contributions include providing interpretable machine learning solutions, discussing the effect of imbalanced data, and proposing SHAP-based application scenarios.
Keyword:
Machine learning
Construction defects
Text classification
SHAP
SMOTE
期刊
IF:
7.4
论文数:
1.7W
被引数:
6.6W
机构
引用论文
Text mining-based construction site accident classification using hybrid supervised machine learning基于文本挖掘的混合监督机器学习施工现场事故分类
Occlusion Aware Facial Expression Recognition Using CNN With Attention Mechanism基于注意力机制的CNN遮挡感知面部表情识别
Construction accident narrative classification: An evaluation of text mining techniques建筑事故叙事分类: 文本挖掘技术的评估
Arsenic removal from aqueous solutions by adsorption using novel MIL-53(Fe) as a highly efficient adsorbent使用新型MIL-53(Fe) 作为高效吸附剂通过吸附从水溶液中去除砷
RSC Advances
IF0
Multiclass and binary SVM classification: Implications for training and classification users多类和二进制SVM分类: 对培训和分类用户的影响

