vix.ing · top · new · best · stats · spec

TDM: Temporally-Consistent Diffusion Model for All-in-One Real-World Video Restoration

2025/01/04 by Yizhou Li, Zihua Liu, Li, Yizhou +5 · 1 citation
Computer Science · Medicine · #Advanced Image Processing Techniques #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Image and Signal Denoising Methods #Medical Imaging Techniques and Applications

paper · pdf · doi:10.48550/arxiv.2501.02269

openalex publication_date 2025/01/04 · openalex created_date 2025/01/08 · openalex updated_date 2026/07/28

Abstract

In this paper, we propose the first diffusion-based all-in-one video restoration method that utilizes the power of a pre-trained Stable Diffusion and a fine-tuned ControlNet. Our method can restore various types of video degradation with a single unified model, overcoming the limitation of standard methods that require specific models for each restoration task. Our contributions include an efficient training strategy with Task Prompt Guidance (TPG) for diverse restoration tasks, an inference strategy that combines Denoising Diffusion Implicit Models~(DDIM) inversion with a novel Sliding Window Cross-Frame Attention (SW-CFA) mechanism for enhanced content preservation and temporal consistency, and a scalable pipeline that makes our method all-in-one to adapt to different video restoration tasks. Through extensive experiments on five video restoration tasks, we demonstrate the superiority of our method in generalization capability to real-world videos and temporal consistency preservation over existing state-of-the-art methods. Our method advances the video restoration task by providing a unified solution that enhances video quality across multiple applications.

Cited by

Related