Skip to main navigation Skip to search Skip to main content

Progressive local-to-global vision transformer for occluded face hallucination

  • Huan Wang
  • , Jianning Chi
  • , Chengdong Wu
  • , Xiaosheng Yu
  • , Hao Wu
  • Northeastern University China
  • The University of Sydney

Research output: Contribution to journalArticlepeer-review

3 Citations (Scopus)

Abstract

Hallucinating a photo-realistic high-resolution (HR) face image from an occluded low-resolution (LR) face image is beneficial for a series of face-related applications. However, previous efforts focused on either super-resolving HR face images from non-occluded LR counterparts or inpainting occluded HR faces. It is necessary to address all these challenges jointly for real-world face images in unconstrained environment. In this paper, we develop a novel Local-to-Global Face Hallucination Transformer (LGFH-Transformer), which simultaneously handles the occluded LR face image super-resolution (SR) and inpainting in a unified framework. Specifically, the LGFH-Transformer is built on self-attention modules which excel at modeling long-range information between image patch sequences. Meanwhile, we introduce a mask-guided convolution and gated mechanism into the building modules (i.e., multi-head attention and feed-forward network) of each Transformer block, which can bring in the complimentary strength of convolution operation to emphasize on the spatially local context. Moreover, equipped with the delicate designed local-to-global feature reasoning mechanism in the phase of encoder, we exploit facial geometry priors (i.e., facial parsing maps) as the semantic guidance during the hallucination process in the phase of decoder to reconstruct more realistic facial details. Extensive experiments demonstrate the effectiveness and advancement of LGFH-Transformer.
Original languageEnglish
Pages (from-to)8219–8240
Number of pages22
JournalMultimedia Tools and Applications
Volume83
Issue number3
DOIs
Publication statusPublished - Jan 2024
Externally publishedYes

Keywords

  • Image processing
  • Super-resolution
  • Face inpainting
  • Vision transfomer

Fingerprint

Dive into the research topics of 'Progressive local-to-global vision transformer for occluded face hallucination'. Together they form a unique fingerprint.

Cite this