arrow
返回

Efficient, interpretable and automated feature engineering for bank data

delete2025-04-01
delete0
PRE
AI
A
Atilla Karaahmetoğlu
M
Mehmet Yıldız
E
Erdem Ünal
U
Uğur Aydın
K
Koran, Murat
A
Akgun, Barin *
DOI:10.1016/j.bdr.2025.100524delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Banks rely on expert-generated features and simple models to have high performance and interpretability at the same time. Interpretability is needed for internal assessment and regulatory compliance for specific problems such as risk assessment and both expert generated features and simple models satisfy this need. However, feature generation by experts is a time-consuming process and susceptible to bias. In addition, features need to be generated fairly often due to the dynamic nature of bank data, and in case of significant changes or new data sources, expertise might take a while to build up. Complex models, such as deep neural networks, may be able to remedy this. However, interpretability/explainability approaches for complex models are not satisfactory from the banks' point of view. In addition, such models do not always work well with tabular data which is abundant in banking applications. This paper introduces an automated feature synthesis pipeline that creates informative and domain-interpretable features which iconsumes significantly less time than brute-force methods. We create novel feature synthesis steps, define elimination rules to rule out uninterpretable features, and combine performance-based feature selection methods to pick desirable ones to build our models. Our results on two different datasets show that the features generated with our pipeline; (1) perform on par or better than features generated by existing methods, (2) are obtained faster, and (3) are domain-interpretable.
Keyword:
Automated feature engineering
Feature selection
Banking data

期刊

Big Data Research 封面图
Big Data Research
IF:
4.2
论文数:
416
被引数:
1.1K

机构

K
koc university
学者数:
5.7K
论文数: 4.5K
被引数: 48
引用论文

引用论文

暂无论文信息