arrow
返回

Causal keyword driven reliable text classification with large language model feedback

delete2025-03-01
delete0
PRE
AI
R
Rui Song
Y
Yingji Li
M
Mingjie Tian
H
Hanwen Wang
F
Fausto Giunchiglia
J
Jian Li *
DOI:10.1016/j.ipm.2024.103964delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Recent studies show Pre-trained Language Models (PLMs) tend to shortcut learning, reducing effectiveness with Out-Of-Distribution (OOD) samples, prompting research on the impact of shortcuts and robust causal features by interpretable methods for text classification. However, current approaches encounter two primary challenges. Firstly, black-box interpretable methods often yield incorrect causal keywords. Secondly, existing methods do not differentiate between shortcuts and causal keywords, often employing a unified approach to deal with them. To address the first challenge, we propose a framework that incorporates Large Language Model's feedback into the process of identifying shortcuts and causal keywords. Specifically, we transform causal feature extraction into a word-level binary labeling task with the aid of ChatGPT. For the second challenge, we introduce a multi-grained shortcut mitigation framework. This framework includes two auxiliary tasks aimed at addressing shortcuts and causal features separately: shortcut reconstruction and counterfactual contrastive learning. These tasks enhance PLMs at both the token and sample granularity levels, respectively. Experimental results show that the proposed method achieves an average performance improvement of more than 1% under the premise of four different language model as the backbones for sentiment classification and toxicity detection tasks on 8 datasets compared with the most recent baseline methods.
Keyword:
LLM feedback
Contrastive learning
Reliable text classification

期刊

I
Information Processing and Management
IF:
6.9
论文数:
5.2K
被引数:
1.4W

机构

J
Jilin University
学者数:
8.7W
论文数: 5.6W
被引数: 8.9K
引用论文

引用论文

Molecular Forms of Serum Insulin-Like Growth Factor (IGF)-Binding Proteins in Man: Relationships with Growth Hormone and IGFs and Physiological Significance*
err1989-12-01
err0
PREAI
errSYLVIE HARDOUIN; MICHELINE GOURMELEN; PATRICIA NOGUIEZ; DANIELLE SEURIN; MONIREH ROGHANI; YVES LE BOUC; GUILHERME POVOA; THOMAS J. MERIMEE; PAUL HOSSENLOPP; MICHEL BINOUX
err分享
err收藏
Direct Relay Pathways from Lemniscal Auditory Thalamus to Secondary Auditory Field in Mice
err2018-10-01
err0
errOAAI
errShinpei Ohga; Hiroaki Tsukano; Masao Horie; Hiroki Terashima; Nana Nishio; Yamato Kubota; Kuniyuki Takahashi; Ryuichi Hishida; Hirohide Takebayashi; Katsuei Shibuki
err分享
err收藏
Shortcut learning in deep neural networks深度神经网络中的捷径学习
err2020-11-10
err489
PREAI
errGeirhos, Robert; Jacobsen, Joern-Henrik; Michaelis, Claudio; Zemel, Richard; Brendel, Wieland; Bethge, Matthias; Wichmann, Felix A.
err分享
err收藏
INSERT-seq enables high resolution mapping of genomically integrated DNA using nanopore sequencing
err
IF0
err2022-05-25
err0
errOAAI
errDimitrije Ivančić; Júlia Mir-Pedrol; Jessica Jaraba-Wallace; Núria Rafel; Avencia Sanchez-Mejias; Marc Güell
err分享
err收藏
err分享
err收藏
学者 查看更多内容