arrow
Return

Reasoning-Grounded Natural Language Explanations for Language Models

delete2026-01-01
delete0
delete
OA
AI
V
Vojtěch Cahlík *
R
Rodrigo Alves
P
Pavel Kordík
DOI:10.1007/978-3-032-08327-2_1delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
We propose a large language model explainability technique for obtaining faithful natural language explanations by grounding the explanations in a reasoning process. When converted to a sequence of tokens, the outputs of the reasoning process can become part of the model context and later be decoded to natural language as the model produces either the final answer or the explanation. To improve the faithfulness of the explanations, we propose to use a joint predict-explain approach, in which the answers and explanations are inferred directly from the reasoning sequence, without the explanations being dependent on the answers and vice versa. We demonstrate the plausibility of the proposed technique by achieving a high alignment between answers and explanations in several problem domains, observing that language models often simply copy the partial decisions from the reasoning sequence into the final answers or explanations. Furthermore, we show that the proposed use of reasoning can also improve the quality of the answers.
Keywords:
Explainable AI
Large Language Models
Natural Language Explanations
Reasoning
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

E
EXPLAINABLE ARTIFICIAL INTELLIGENCE, XAI 2025, PT III
IF:
0
Papers:
20
Citations:
0

Organization

C
czech technical university prague
Scholars:
6.5K
Papers: 5.3K
Citations: 3