arrow
Return

SimDiff: Depth pruning via similarity and difference

delete2026-08-10
delete0
PRE
AI
Y
Yuli Chen
S
Shuhao Zhang
F
Fanshen Meng
B
Bo Cheng *
J
Jiale Han
Q
Qiang Tong
X
Xiulei Liu *
DOI:10.1016/j.eswa.2026.133944delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
• SimDiff prunes LLMs by jointly measuring inter-layer similarity and difference. • Two metrics, MSSD and MASD, capture distinct and complementary layer behaviors. • SimDiff preserves >91% performance at 25% pruning and speeds up inference 1.49 × .
Keywords:
Depth pruning
Model compression
Large language models

Journal

Expert Systems with Applications cover
Expert Systems with Applications
IF:
7.5
Papers:
2.9W
Citations:
10.2W

Organization

H
hong kong university of science and technology
Scholars:
836
Papers: 471
Citations: 1
B
beijing university of posts and telecommunications
Scholars:
2.0K
Papers: 748
Citations: 0
B
beijing information science and technology university
Scholars:
394
Papers: 164
Citations: 0
researcher View more organizations