Cross-Scale KNN Image Transformer for Image Restoration

Hunsang Lee,Hyesong Choi,Dongbo Min,Kwanghoon Sohn

doi:10.1109/access.2023.3242556

Hunsang Lee, Hyesong Choi + Show 2 more

Open Access

https://doi.org/10.1109/access.2023.3242556

Copy DOI

Abstract

Numerous image restoration approaches have been proposed based on attention mechanism, achieving superior performance to convolutional neural networks (CNNs) based counterparts. However, they do not leverage the attention model in a form fully suited to the image restoration tasks. In this paper, we propose an image restoration network with a novel attention mechanism, called cross-scale <italic xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">k</i> -NN image Transformer (CS-KiT), that effectively considers several factors such as locality, non-locality, and cross-scale aggregation, which are essential to the image restoration. To achieve locality and non-locality, the CS-KiT builds <italic xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">k</i> -nearest neighbor relation of local patches and aggregates similar patches through a local attention. To induce cross-scale aggregation, we ensure that each local patch embraces different scale information with scale-aware patch embedding (SPE) which predicts an input patch scale through a combination of multi-scale convolution branches. We show the effectiveness of the CS-KiT with experimental results, outperforming state-of-the-art restoration approaches on image denoising, deblurring, and deraining benchmarks.

Full Text