Activating More Pixels in Image Super-Resolution Transformer

Chen, Xiangyu; Wang, Xintao; Zhou, Jiantao; Dong, Chao

Electrical Engineering and Systems Science > Image and Video Processing

arXiv:2205.04437v1 (eess)

[Submitted on 9 May 2022 (this version), latest version 19 Mar 2023 (v3)]

Title:Activating More Pixels in Image Super-Resolution Transformer

Authors:Xiangyu Chen, Xintao Wang, Jiantao Zhou, Chao Dong

View PDF

Abstract:Transformer-based methods have shown impressive performance in low-level vision tasks, such as image super-resolution. However, we find that these networks can only utilize a limited spatial range of input information through attribution analysis. This implies that the potential of Transformer is still not fully exploited in existing networks. In order to activate more input pixels for reconstruction, we propose a novel Hybrid Attention Transformer (HAT). It combines channel attention and self-attention schemes, thus making use of their complementary advantages. Moreover, to better aggregate the cross-window information, we introduce an overlapping cross-attention module to enhance the interaction between neighboring window features. In the training stage, we additionally propose a same-task pre-training strategy to bring further improvement. Extensive experiments show the effectiveness of the proposed modules, and the overall method significantly outperforms the state-of-the-art methods by more than 1dB. Codes and models will be available at this https URL.

Subjects:	Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2205.04437 [eess.IV]
	(or arXiv:2205.04437v1 [eess.IV] for this version)
	https://doi.org/10.48550/arXiv.2205.04437

Submission history

From: Xiangyu Chen [view email]
[v1] Mon, 9 May 2022 17:36:58 UTC (10,813 KB)
[v2] Mon, 16 May 2022 09:13:36 UTC (10,813 KB)
[v3] Sun, 19 Mar 2023 01:25:49 UTC (13,712 KB)

Electrical Engineering and Systems Science > Image and Video Processing

Title:Activating More Pixels in Image Super-Resolution Transformer

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Electrical Engineering and Systems Science > Image and Video Processing

Title:Activating More Pixels in Image Super-Resolution Transformer

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators