1
Return

Aligning Large Language Models With Human Feedback: Mathematical foundations and algorithm design [Special Issue on the Mathematics of Deep Learning]

delete2026-06-12
delete0
PRE
AI
S
Siliang Zeng
L
Luca Viano
C
Chenliang Li
J
Jiaxiang Li
M
Markus Wulfmeier
S
Stefano Ermon
A
Alfredo Garcia
M
Mingyi Hong
DOI:10.1109/MSP.2026.3666824delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
This article provides an introduction to the mathematical foundations and algorithmic frameworks used to align large language models (LLMs) with human intentions, preferences, and values. We discuss standard alignment techniques, such as supervised fine-tuning (SFT), reinforcement learning with human feedback (RLHF), and direct preference optimization (DPO). We also explore the theoretical underpinnings of learning from human preferences, drawing connections to inverse reinforcement learning (IRL), and discrete choice models. We present state-of-the-art algorithms in a tutorial style, discuss their advantages and limitations, and offer insights into practical implementation. Our exposition is intended to serve as a comprehensive resource for researchers and practitioners, providing both a foundational understanding of alignment methodologies and a framework for developing more robust and scalable alignment techniques.
Keywords:
Learning (artificial intelligence)
Large language models
Mathematical models
Algorithm design and analysis
Feedback
Human factors
Supervised learning
Reinforcement learning
User preference

Journal

IEEE Signal Processing Magazine cover
IEEE Signal Processing Magazine
IF:
9.6
Papers:
1.1W
Citations:
1.7W

Organization

B
bytedance seed
Scholars:
7
Papers: 3
Citations: 0
G
google deepmind
Scholars:
220
Papers: 59
Citations: 44
T
texas a&m university
Scholars:
2.4K
Papers: 951
Citations: 0
S
stanford university
Scholars:
9.2K
Papers: 3.6K
Citations: 0
M
Meta
Scholars:
145
Papers: 37
Citations: 14
E
ecole polytechnique federale de lausanne
Scholars:
796
Papers: 389
Citations: 0
U
university of minnesota
Scholars:
3.2K
Papers: 1.5K
Citations: 0
Cited Papers

Cited Papers

Citing Papers

Citing Papers