OSKAR: Omnimodal Self-supervised Knowledge Abstraction and Representation
Mohamed Abdelfattah, Kaouther Messaoud, Alexandre Alahi
2025Year
1Citations
1Top-tier citations
Abstract
We present OSKAR, the first multimodal foundation model based on bootstrapped latent feature prediction. Unlike generative or contrastive methods, it avoids mem-orizing unnecessary details ( e
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c05b3620-e9a2-4188-b4bb-9d0197dc34a0Cited by top-tier papers1
Ask how each one uses itBuilds on56
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
Related papers
- A Generalization Theory for Zero-Shot PredictionRonak Mehta, Zaïd HarchaouiICML 2025
- VFMF: Dense Forecasting by Generating Foundation Model FeaturesGabrijel Boduljak, Yushi Lan, Christian Rupprecht, Andrea VedaldiICML 2026
- Conditional Generative Modeling via Learning the Latent SpaceSameera Ramasinghe, Kanchana Nisal Ranasinghe, Salman H. Khan, Nick Barnes et al.ICLR 2021 · 10 citations
- SleepMaMi: A Universal Sleep Foundation Model for Integrating Macro- and Micro-structuresKeondo Park, Younghoon Na, Yourim Choi, Hyunwoo Ryu et al.ICML 2026
- 3D-Spatial Multimodal MemoryXueyan Zou, Yuchen Song, Ri-Zhao Qiu, Xuanbin Peng et al.ICLR 2025
