arrow
Return

Random Direct Preference Optimization for Radiography Report Generation

delete2026-01-01
delete0
PRE
AI
S
Samokhin, Valentin
S
Shirokikh, Boris *
G
Goncharov, Mikhail
U
Umerenkov, Dmitriy
B
Bobrin, Maksim
O
Oseledets, Ivan
R
Rydylov, Dmit
B
Belyaev, Mikhail
DOI:10.1007/978-3-032-07845-2_17delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Radiography Report Generation (RRG) has gained significant attention in medical image analysis as a promising tool for alleviating the growing workload of radiologists. However, despite numerous advancements, existing methods have yet to achieve the quality required for deployment in real-world clinical settings. Meanwhile, large Visual Language Models (VLMs) have demonstrated remarkable progress in the general domain by adopting training strategies originally designed for Large Language Models (LLMs), such as alignment techniques. In this paper, we introduce a model-agnostic framework to enhance RRG accuracy using Direct Preference Optimization (DPO). Our approach leverages random contrastive sampling to construct training pairs, eliminating the need for reward models or human preference annotations. Experiments on supplementing three state-of-the-art models with our Random DPO show that our method improves clinical performance metrics by up to 5%, without requiring any additional training data.
Keywords:
Radiography
Report Generation
Visual Language Models
Alignment
Direct Preference Optimization

Journal

F
FOUNDATION MODELS FOR GENERAL MEDICAL AI, MEDAGI 2025
IF:
0
Papers:
16
Citations:
0

Organization

A
airi - artificial intelligence research institute
Scholars:
32
Papers: 10
Citations: 0