Spatially Informed Autoencoders for Interpretable Visual Representation Learning
Dominik Sturm, Hiba Bensalem, Ivo F. Sbalzarini
Abstract
We introduce spatially informed variational autoencoders (SI-VAE) as self-supervised deep-learning models that use stochastic point processes to predict spatial organization patterns from images. Existing approaches to learning visual representations based on variational autoencoders (VAE) struggle to capture spatial correlations between objects or events, focusing instead on pixel intensities. We address this limitation by incorporating a point-process likelihood, derived from the Papangelou conditional intensity, as a self-supervision target. This results in a hybrid model that learns statistically interpretable representations of spatial localization patterns and enables zero-shot conditional simulation directly from images. Experiments with synthetic images show that SI-VAE improve the classification accuracy of attractive, repulsive, and uncorrelated point patterns from 48% (VAE) to over 80% in the worst case and 90% in the best case, while generalizing to unseen data. We apply SI-VAE to a real-world microscopy data set, demonstrating its use for studying the spatial organization of proteins in human cells and for using the representations in downstream statistical analysis.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on7
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Learning Energy-Based Models by Diffusion Recovery LikelihoodRuiqi Gao, Yang Song, Ben Poole, Ying Nian Wu et al.ICLR 2021 · 144 citations
- Is Score Matching Suitable for Estimating Point Processes?Haoqun Cao, Zizhuo Meng, Tianjun Ke, Feng ZhouNeurIPS 2024 · 7 citations
- CELLE-2: Translating Proteins to Pictures and Back with a Bidirectional Text-to-Image TransformerEmaad Khwaja, Yun Song, Aaron Agarunov, Bo HuangNeurIPS 2023 · 7 citations
- Unlocking Point Processes through Point Set DiffusionDavid Lüdke, Enric Rabasseda Raventós, Marcel Kollovieh, Stephan GünnemannICLR 2025
Related papers
- Representation Uncertainty in Self-Supervised Learning as Variational InferenceHiroki Nakamura, Masashi Okada, Tadahiro TaniguchiICCV 2023 · 27 citations
- Autoencoding Conditional Neural Processes for Representation LearningVictor Prokhorov, Ivan Titov, N. SiddharthICML 2024
- Structure by Architecture: Structured Representations without RegularizationFelix Leeb, Giulia Lanzillotta, Yashas Annadani, Michel Besserve et al.ICLR 2023 · 1 citation
- A Variational Autoencoder for Neural Temporal Point Processes with Dynamic Latent GraphsSikun Yang, Hongyuan ZhaAAAI 2024 · 7 citations
- Self-Guided Masked AutoencoderJeongwoo Shin, Inseo Lee, Junho Lee, Joonseok LeeNeurIPS 2024 · 18 citations
