Perceptual Kalman Filters: Online State Estimation under a Perfect Perceptual-Quality Constraint
Dror Freirich, Tomer Michaeli, Ron Meir
摘要
Many practical settings call for the reconstruction of temporal signals from corrupted or missing data. Classic examples include decoding, tracking, signal enhancement and denoising. Since the reconstructed signals are ultimately viewed by humans, it is desirable to achieve reconstructions that are pleasing to human perception. Mathematically, perfect perceptual-quality is achieved when the distribution of restored signals is the same as that of natural signals, a requirement which has been heavily researched in static estimation settings (i.e. when a whole signal is processed at once). Here, we study the problem of optimal causal filtering under a perfect perceptual-quality constraint, which is a task of fundamentally different nature. Specifically, we analyze a Gaussian Markov signal observed through a linear noisy transformation. In the absence of perceptual constraints, the Kalman filter is known to be optimal in the MSE sense for this setting. Here, we show that adding the perfect perceptual quality constraint (i.e. the requirement of temporal consistency), introduces a fundamental dilemma whereby the filter may have to "knowingly" ignore new information revealed by the observations in order to conform to its past decisions. This often comes at the cost of a significant increase in the MSE (beyond that encountered in static settings). Our analysis goes beyond the classic innovation process of the Kalman filter, and introduces the novel concept of an unutilized information process. Using this tool, we present a recursive formula for perceptual filters, and demonstrate the qualitative effects of perfect perceptual-quality estimation on a video reconstruction problem.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper4
- Video Diffusion ModelsJonathan Ho, Tim Salimans, Alexey A. Gritsenko, William Chan 等NeurIPS 2022 · 被引用 2,948 次
- Make-A-Video: Text-to-Video Generation without Text-Video DataUriel Singer, Adam Polyak, Thomas Hayes, Xi Yin 等ICLR 2023 · 被引用 313 次
- Learning temporal coherence via self-supervision for GAN-based video generationMengyu Chu, You Xie, Jonas Mayer, Laura Leal-Taixé 等SIGGRAPH 2020 · 被引用 198 次
- A Theory of the Distortion-Perception Tradeoff in Wasserstein SpaceDror Freirich, Tomer Michaeli, Ron MeirNeurIPS 2021 · 被引用 77 次
相关 Paper
- On the choice of Perception Loss Function for Learned Video CompressionSadaf Salehkalaibar, Buu Phan, Jun Chen, Wei Yu 等NeurIPS 2023 · 被引用 23 次
- On Perceptual Lossy Compression: The Cost of Perceptual Reconstruction and An Optimal Training FrameworkZeyu Yan, Fei Wen, Rendong Ying, Chao Ma 等ICML 2021 · 被引用 48 次
- Posterior-Mean Rectified Flow: Towards Minimum MSE Photo-Realistic Image RestorationGuy Ohayon, Tomer Michaeli, Michael EladICLR 2025
- Reasons for the Superiority of Stochastic Estimators over Deterministic Ones: Robustness, Consistency and Perceptual QualityGuy Ohayon, Theo Joseph Adrai, Michael Elad, Tomer MichaeliICML 2023 · 被引用 18 次
- Deep Optimal Transport: A Practical Algorithm for Photo-realistic Image RestorationTheo Adrai, Guy Ohayon, Michael Elad, Tomer MichaeliNeurIPS 2023 · 被引用 22 次
