Efficient Visual State Space Model for Image Deblurring

2024-05-23CVPR 2025Code Available2· sign in to hype

Lingshun Kong, Jiangxin Dong, Ming-Hsuan Yang, Jinshan Pan

Code Available — Be the first to reproduce this paper.

Code

github.com/kkkls/evssm
OfficialIn papernone★ 137

Abstract

Convolutional neural networks (CNNs) and Vision Transformers (ViTs) have achieved excellent performance in image restoration. ViTs typically yield superior results in image restoration compared to CNNs due to their ability to capture long-range dependencies and input-dependent characteristics. However, the computational complexity of Transformer-based models grows quadratically with the image resolution, limiting their practical appeal in high-resolution image restoration tasks. In this paper, we propose a simple yet effective visual state space model (EVSSM) for image deblurring, leveraging the benefits of state space models (SSMs) to visual data. In contrast to existing methods that employ several fixed-direction scanning for feature extraction, which significantly increases the computational cost, we develop an efficient visual scan block that applies various geometric transformations before each SSM-based module, capturing useful non-local information and maintaining high efficiency. Extensive experimental results show that the proposed EVSSM performs favorably against state-of-the-art image deblurring methods on benchmark datasets and real-captured images.

Tasks

Deblurring Image Deblurring Image Restoration model State Space Models

Benchmark Results

Dataset	Model	Metric	Claimed	Verified	Status
GoPro	EVSSM	PSNR	34.5	—	Unverified
HIDE	EVSSM	PSNR	31.97	—	Unverified
RealBlur-J	EVSSM	PSNR	34.15	—	Unverified
RealBlur-R	EVSSM	PSNR	41.04	—	Unverified
Real-world Dataset	EVSSM	PSNR	48.78	—	Unverified

Efficient Visual State Space Model for Image Deblurring

Code

Abstract

Tasks

Benchmark Results

Reproductions