Lune

CVPR2025Top-tier venue

The Power of Context: How Multimodality Improves Image Super-Resolution

Kangfu Mei, Hossein Talebi, Mojtaba Ardakani, Vishal M. Patel, Peyman Milanfar, Mauricio Delbracio

2025Year
8Top-tier citations

Abstract

Inputs Outputs Reference A close-up of a male lion with a dark mane, light tan face, and pink tongue sticking out . . . LR LR (Zoomed) Caption PASD SeeSR MMSR (Ours) HR Depth Segmentation Edge PASD (Zoomed) SeeSR (Zoomed) MMSR (Zoomed) HR (Zoomed) Figure 1. Our Multimodal Super-Resolution (MMSR) method leverages the rich context of multimodal guidance, including image captions, depth maps, semantic segmentation maps, and edges inferred from LR. MMSR surpasses state-of-the-art methods by producing more realistic results and suppressing artifacts that, while plausible, are inconsistent with the information present in the LR input.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 0d180834-0a9a-4259-8d73-4b5f4ffe4936

Cited by top-tier papers8

Ask how each one uses it

Builds on40

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines