Lune

CHI2025Top-tier venue

Abstraction Alignment: Comparing Model-Learned and Human-Encoded Conceptual Relationships

Angie W. Boggust, Hyemin Bang, Hendrik Strobelt, Arvind Satyanarayan

2025Year
4Citations
2Top-tier citations

Abstract

A model's confidence distribution is a reflection of its underlying knowledge.

Human abstractions represent the concepts and relationships we expect models to learn.

Abstraction alignment measures how much of a model's uncertainty can be explained by the human abstractions. SUBGRAPH PREFERENCE: Confidence in different regions of the abstraction. ABSTRACTION MATCH: Uncertainty reduced by a level of abstraction. CONCEPT CO-CONFUSION: Concepts the model regularly confuses.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext eaa63be1-5bf5-482d-a645-1f0c5b64e7a0

Cited by top-tier papers2

Ask how each one uses it

Builds on31

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines