Perceptual Kalman Filters: Online State Estimation under a Perfect Perceptual-Quality Constraint
Dror Freirich, Tomer Michaeli, Ron Meir
Abstract
Many practical settings call for the reconstruction of temporal signals from corrupted or missing data. Classic examples include decoding, tracking, signal enhancement and denoising. Since the reconstructed signals are ultimately viewed by humans, it is desirable to achieve reconstructions that are pleasing to human perception. Mathematically, perfect perceptual-quality is achieved when the distribution of restored signals is the same as that of natural signals, a requirement which has been heavily researched in static estimation settings (i.e. when a whole signal is processed at once). Here, we study the problem of optimal causal filtering under a perfect perceptual-quality constraint, which is a task of fundamentally different nature. Specifically, we analyze a Gaussian Markov signal observed through a linear noisy transformation. In the absence of perceptual constraints, the Kalman filter is known to be optimal in the MSE sense for this setting. Here, we show that adding the perfect perceptual quality constraint (i.e. the requirement of temporal consistency), introduces a fundamental dilemma whereby the filter may have to "knowingly" ignore new information revealed by the observations in order to conform to its past decisions. This often comes at the cost of a significant increase in the MSE (beyond that encountered in static settings). Our analysis goes beyond the classic innovation process of the Kalman filter, and introduces the novel concept of an unutilized information process. Using this tool, we present a recursive formula for perceptual filters, and demonstrate the qualitative effects of perfect perceptual-quality estimation on a video reconstruction problem.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c5cbcab2-0637-420a-8931-97047253dcefBuilds on4
- Video Diffusion ModelsJonathan Ho, Tim Salimans, Alexey A. Gritsenko, William Chan et al.NeurIPS 2022 · 2,948 citations
- Make-A-Video: Text-to-Video Generation without Text-Video DataUriel Singer, Adam Polyak, Thomas Hayes, Xi Yin et al.ICLR 2023 · 313 citations
- Learning temporal coherence via self-supervision for GAN-based video generationMengyu Chu, You Xie, Jonas Mayer, Laura Leal-Taixé et al.SIGGRAPH 2020 · 198 citations
- A Theory of the Distortion-Perception Tradeoff in Wasserstein SpaceDror Freirich, Tomer Michaeli, Ron MeirNeurIPS 2021 · 77 citations
Related papers
- On the choice of Perception Loss Function for Learned Video CompressionSadaf Salehkalaibar, Buu Phan, Jun Chen, Wei Yu et al.NeurIPS 2023 · 23 citations
- On Perceptual Lossy Compression: The Cost of Perceptual Reconstruction and An Optimal Training FrameworkZeyu Yan, Fei Wen, Rendong Ying, Chao Ma et al.ICML 2021 · 48 citations
- Posterior-Mean Rectified Flow: Towards Minimum MSE Photo-Realistic Image RestorationGuy Ohayon, Tomer Michaeli, Michael EladICLR 2025
- Reasons for the Superiority of Stochastic Estimators over Deterministic Ones: Robustness, Consistency and Perceptual QualityGuy Ohayon, Theo Joseph Adrai, Michael Elad, Tomer MichaeliICML 2023 · 18 citations
- Deep Optimal Transport: A Practical Algorithm for Photo-realistic Image RestorationTheo Adrai, Guy Ohayon, Michael Elad, Tomer MichaeliNeurIPS 2023 · 22 citations
