cs.CVJul 10, 2026

Simon-SR: Spatially Adaptive Modulation and Visual Prompt Adaptation for Text-Reinforced Super-Resolution

Authors: Haotong ChengYuxuan LiZijie CuiRongling TanChenyuan Wang

Organizations: College of Electronic Science and Engineering, Jilin University, Changchun China

Abstract

Single Image Super-Resolution (SISR) reconstructs high-quality images from low-resolution inputs. While recent multi-modal methods improve perceptual quality, they remain sensitive to erroneous priors and require expensive annotations. To address these issues, we propose Simon-SR, a multi-modal SISR framework leveraging learnable prompts for efficient semantic mining and robust text-image fusion. Our approach combines Contrastive Prompt Learning with Prompt-Guided Spatially Adaptive Refinement to enhance multi-modal alignment. Experiments demonstrate that Simon-SR surpasses state-of-the-art methods, achieving maximum improvements of 0.50 dB in PSNR, 0.0133 in SSIM, and 0.0695 in LPIPS. Code will be released.

Explore similar work

CardsList
  1. GramSR: Visual Feature Conditioning for Diffusion-Based Super-Resolution

    Apr 28, 2026Fabio D'Oronzio, Federico Putamorsi, Leonardo Zini +2Low-Rank Adaptation