arrow
Return

Context-aware Korean sentence ordering using large language models

delete2026-04-01
delete0
PRE
AI
P
Park, Sangchun
L
Lee, Hyewon
S
Seo, Jungmin
Y
Yeo, Seo-yeon
K
Kim, Daeho
K
Kwak, Il-Youp *
DOI:10.5351/KJAS.2026.39.2.221delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Sentence order critically affects textual coherence and comprehension, yet real-world data often exhibit disrupted ordering. This study investigates Korean context-aware sentence ordering with Large Language Models, comparing three approaches-Pairwise, Sequence, and Global-through fine-tuning of pretrained models such as KLUE-BERT, KoELECTRA, KLUE-RoBERTa, and T5. Experiments were conducted on the DACON Context-Aware Sentence Ordering AI Competition dataset, comprising 7,350 training and 1,780 test samples. The Pair-wise approach effectively captured local sentence relations but failed to model global coherence. The Sequence approach provided an intuitive framework, yet its performance degraded with longer inputs due to overfitting. By contrast, the Global approach, formulated as a classification problem over all permutations, exhibited the most consistent and superior results. Notably, the KLUE-RoBERTa-based Global model achieved the highest score of 83.71% on the private leaderboard.
Keywords:
deep learning
sentence order prediction
large language model
Korean natural language processing

Journal

K
Korean Journal of Applied Statistics
IF:
0
Papers:
20
Citations:
0

Organization

C
Chung Ang University
Scholars:
1.3W
Papers: 1.4W
Citations: 133