Lune

ICML2026顶会

The Geometry of Sequential Learning: Lie-Bracket Prediction of Transfer Order

John Sweeney

2026年份
1顶会引用

摘要

Sequential learning is order-dependent: from Pile-style next-token domain adaptation to instruction-SFT and DPO, NN candidate sources induce N!N! possible curricula. We show that the local order effect is governed by a computable geometric quantity, the Lie-bracket commutator of gradient update fields, yielding a pairwise score for whether A→BA \to B or B→AB \to A is better for a target domain. The pairwise bracket primitive also defines a Lie-Bracket Tournament: with a shared θ0\theta_0 target-gradient reference, Hessian symmetry gives Borda/row-sum scores from one Hessian-vector product per source, O(N)O(N) dot products, and an O(Nlog⁡N)O(N\log N) sort, without materializing the O(N2)O(N^2) edge matrix. Empirically, the planner reaches 98.1%/98.9% pairwise accuracy at k=1k=1 for instruction-SFT/DPO, remains at 73.1%/72.2% at k=20k=20, and preserves the original pretraining-domain evidence with 82.4–92.0% accuracy across four LLMs and 91.1% on diffusion. At curriculum scale, it recovers the best of all 3!3! schedules in 87.5% of trials, ranks 85 Stack programming-language source domains for a Python target in the 99th sampled percentile, and reaches the 99.0–99.6th sampled percentile on 56 MMLU subjects, sharply above the reported descending gradient-norm baseline. These results reframe sequential learning as a geometric tournament problem: commutators provide both local pairwise order information and a scalable primitive for many-domain schedules.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

lune papers fulltext e1fa1bcd-9103-48ce-a244-15569f348c83

引用它的顶会 Paper1

问问它们各自怎么用它

它引用的顶会 Paper12

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖