arrow
Return

Multi-View Single-Scan Visual State-Space Network for Efficient Image Super-Resolution

delete2026-07-24
delete0
PRE
AI
H
Hong Yang
W
Weihua Liu
X
Xianqiang Yang
DOI:10.1109/tip.2026.3714851delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Single image super-resolution (SISR) seeks to reconstruct high-resolution images from low-resolution inputs under the tight latency and memory budgets of edge devices. Recent visual state space models enable linear-time sequence modeling, but they often approximate two-dimensional dependencies through multiple directional scans, which increases computation and memory. We propose a Multi-View Single-Scan Visual State Space Network (MVSSN) for efficient image super-resolution. MVSSN contains three main components. A shuffled input stacking module (SISM) organizes replicated and shuffled input channels before a lightweight projection, where the learned projection filters help produce less redundant shallow responses. A multi-view single-scan block (MSSB) uses an alternating scan axis and invertible geometric transforms to change the serialization order observed by one selective scan in each block. Across stacked blocks, this provides a lightweight cross-layer approximation to multi-view context modeling. A multi-scale local feature block (MLFB) complements global aggregation with depthwise convolutions of complementary receptive fields and a compact MLP to restore local details. Experiments on standard SISR benchmarks show that MVSSN achieves competitive or better PSNR and SSIM with fewer than one million parameters and low FLOPs. Additional ablations, RealSR evaluations, and downstream detection and segmentation studies further examine its efficiency and practical behavior. We also discuss the limitation under unknown real degradation, where bicubic-trained models may still suffer from domain gaps.
Keywords:
Image super-resolution
visual state space models
lightweight neural network
geometric transform
local–global feature modeling

Journal

IEEE Transactions on Image Processing cover
IEEE Transactions on Image Processing
IF:
13.7
Papers:
1.0W
Citations:
8.4W

Organization

Y
yongjiang laboratory
Scholars:
333
Papers: 256
Citations: 2
H
Harbin Institute of Technology
Scholars:
1.4W
Papers: 4.5K
Citations: 8.5W