Lune

ICML2026Top-tier venue

Causes and Consequences of Representational Similarity in Machine Learning Models

Zeyu Michael Li, Hung Anh Vu, Damilola Awofisayo, Emily Wenger

2026Year
1Citations

Abstract

Numerous works have noted similarities in how machine learning models represent the world, even across modalities. Although much effort has been devoted to uncovering properties and metrics on which these models align, surprisingly little work has explored causes of this similarity. To advance this line of inquiry, this work explores how two factors—dataset overlap and task overlap—influence downstream model similarity. We evaluate the effects of both factors through experiments across model sizes and modalities, from small classifiers to large language models. We find that dataset and task overlap are positively associated with higher representational similarity across many of our settings, with clear evidence in vision/language classification and weaker trends in language generation experiments. Finally, we consider downstream consequences of representational similarity, showing that greater similarity is associated with increased vulnerability to transferable adversarial attacks in vision models.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext f0c92a82-9f03-4bd0-9428-47706db2793d

Builds on18

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines