Multi-modal Transfer Learning between Biological Foundation Models
Juan Jose Garau-Luis, Patrick Bordes, Liam Gonzalez, Masa Roller, Bernardo P. de Almeida, Christopher Blum, Lorenz Hexemer, Stefan Laurent, Maren Lang, Thomas Pierrot, Guillaume Richard
Abstract
Biological sequences encode fundamental instructions for the building blocks of life, in the form of DNA, RNA, and proteins. Modeling these sequences is key to understand disease mechanisms and is an active research area in computational biology. Recently, Large Language Models have shown great promise in solving certain biological tasks but current approaches are limited to a single sequence modality (DNA, RNA, or protein). Key problems in genomics intrinsically involve multiple modalities, but it remains unclear how to adapt general-purpose sequence models to those cases. In this work we propose a multi-modal model that connects DNA, RNA, and proteins by leveraging information from different pre-trained modality-specific encoders. We demonstrate its capabilities by applying it to the largely unsolved problem of predicting how multiple RNA transcript isoforms originate from the same gene (i.e. same DNA sequence) and map to different transcription expression levels across various human tissues. We show that our model, dubbed IsoFormer, is able to accurately predict differential transcript expression, outperforming existing methods and leveraging the use of multiple modalities. Our framework also achieves efficient transfer knowledge from the encoders pre-training as well as in between modalities. We open-source our model, paving the way for new multi-modal gene expression approaches.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6a014bae-d4c7-4df4-9632-c90f86741383Cited by top-tier papers5
- MergeDNA: Context-Aware Genome Modeling with Dynamic Tokenization Through Token MergingSiyuan Li, Kai Yu, Anna Wang, Zicheng Liu et al.AAAI 2026 · 2 citations
- PyTDC: A multimodal machine learning training, evaluation, and inference platform for biomedical foundation modelsAlejandro Velez-Arce, Marinka ZitnikICML 2025
- CDBridge: A Cross-omics Post-training Bridge Strategy for Context-aware Biological ModelingChang Yu, Siyuan Li, Zicheng Liu, Jingbo Zhou et al.ICLR 2026
- Causal Representation Learning from Multimodal Biomedical ObservationsYuewen Sun, Lingjing Kong, Guangyi Chen, Loka Li et al.ICLR 2025
- Bimodal masked language modeling for bulk RNA-seq and DNA methylation representation learningMaxence Gélard, Hakim Benkirane, Thomas Pierrot, Guillaume Richard et al.ICML 2026
Builds on16
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 11,349 citations
- BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language ModelsJunnan Li, Dongxu Li, Silvio Savarese, Steven C. H. HoiICML 2023 · 7,873 citations
- Flamingo: a Visual Language Model for Few-Shot LearningJean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech et al.NeurIPS 2022 · 6,707 citations
Related papers
- BioReason: Incentivizing Multimodal Biological Reasoning within a DNA-LLM ModelAdibvafa Fallahpour, Andrew Magnuson, Purav Gupta, Shihao Ma et al.NeurIPS 2025 · 53 citations
- CoPRA: Bridging Cross-domain Pretrained Sequence Models with Complex Structures for Protein-RNA Binding Affinity PredictionRong Han, Xiaohong Liu, Tong Pan, Jing Xu et al.AAAI 2025 · 7 citations
- SPATIA: Multimodal Generation and Prediction of Spatial Cell PhenotypesZhenglun Kong, Mufan Qiu, John Boesen, xiang lin et al.ICML 2026 · 1 citation
- HyperST: Hierarchical Hyperbolic Learning for Spatial Transcriptomics PredictionChen Zhang, Yilu An, Ying Chen, Hao Li et al.CVPR 2026
- BioToken and BioFM – Biologically-Informed Tokenization Enables Accurate and Efficient Genomic Foundation ModelsAleksandr Medvedev, Karthik Viswanathan, Praveenkumar Kanithi, Kirill Vishniakov et al.ICML 2026
