Rectification Reimagined: A Unified Mamba Model for Image Correction and Rectangling with Prompts

Linwei Qiu; Gongzhe Li; Xiaozhe Zhang; Qinlin Sun; Fengying Xie

Rectification Reimagined: A Unified Mamba Model for Image Correction and Rectangling with Prompts

Linwei Qiu, Gongzhe Li, Xiaozhe Zhang, Qinlin Sun, Fengying Xie

TL;DR

<3-5 sentence high-level summary> UniRect reframes image correction and rectangling as a single distortion-rectification problem, introducing a general distortion model that unifies portrait, wide-angle, stitched, and rotation distortions. It provides a two-module architecture: a Deformation Module based on Residual Progressive Thin-Plate Spline (RP-TPS) and a Restoration Module built with Residual Mamba Blocks (RMBs), augmented by a Sparse Mixture-of-Experts (SMoEs) to enable four tasks in one model. The framework uses task-guiding prompts and specialized losses to constrain deformation and restoration, achieving state-of-the-art performance across four tasks and demonstrating strong cross-task generalization and real-world applicability. While computationally intensive, UniRect offers a scalable path toward unified, edge-friendly rectification pipelines on mobile devices.</paper_summary>

Abstract

Image correction and rectangling are valuable tasks in practical photography systems such as smartphones. Recent remarkable advancements in deep learning have undeniably brought about substantial performance improvements in these fields. Nevertheless, existing methods mainly rely on task-specific architectures. This significantly restricts their generalization ability and effective application across a wide range of different tasks. In this paper, we introduce the Unified Rectification Framework (UniRect), a comprehensive approach that addresses these practical tasks from a consistent distortion rectification perspective. Our approach incorporates various task-specific inverse problems into a general distortion model by simulating different types of lenses. To handle diverse distortions, UniRect adopts one task-agnostic rectification framework with a dual-component structure: a {Deformation Module}, which utilizes a novel Residual Progressive Thin-Plate Spline (RP-TPS) model to address complex geometric deformations, and a subsequent Restoration Module, which employs Residual Mamba Blocks (RMBs) to counteract the degradation caused by the deformation process and enhance the fidelity of the output image. Moreover, a Sparse Mixture-of-Experts (SMoEs) structure is designed to circumvent heavy task competition in multi-task learning due to varying distortions. Extensive experiments demonstrate that our models have achieved state-of-the-art performance compared with other up-to-date methods.

Rectification Reimagined: A Unified Mamba Model for Image Correction and Rectangling with Prompts

TL;DR

Abstract

Rectification Reimagined: A Unified Mamba Model for Image Correction and Rectangling with Prompts

TL;DR

Abstract

Paper Structure

Table of Contents

Figures (16)