Can Transformers Do Enumerative Geometry?
Baran Hashemi, Roderic Guigo Corominas, Alessandro Giacchetto
摘要
We introduce a Transformer-based approach to computational enumerative geometry, specifically targeting the computation of ψ-class intersection numbers on the moduli space of curves. Traditional methods for calculating these numbers suffer from factorial computational complexity, making them impractical to use. By reformulating the problem as a continuous optimization task, we compute intersection numbers across a wide value range from 10 -45 to 10 45 . To capture the recursive nature inherent in these intersection numbers, we propose the Dynamic Range Activator (DRA) 1 , a new activation function that enhances the Transformer's ability to model recursive patterns and handle severe heteroscedasticity. Given the precision required to compute these invariants, we quantify the uncertainty in the predictions using Conformal Prediction with a dynamic sliding window, adaptive to partitions of equivalent numbers of marked points. To the best of our knowledge, there has been no prior work on modeling recursive functions with such a high-variance and factorial growth. Beyond simply computing intersection numbers, we explore the enumerative "world-model" of Transformers. Our interpretability analysis reveals that the network is implicitly modeling the Virasoro constraints in a purely data-driven manner. Moreover, through abductive hypothesis testing, probing, and causal inference, we uncover evidence of an emergent internal representation of the the large-genus asymptotic of ψ-class intersection numbers. These findings suggest that the network internalizes the parameters of the asymptotic closed-form and the polynomiality phenomenon of ψ-class intersection numbers in a non-linear manner.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- From Euler to AI: Unifying Formulas for Mathematical ConstantsTomer Raz, Michael Shalyt, Elyasheev Leibtag, Rotem Kalisch 等NeurIPS 2025 · 被引用 4 次
- Polynomial, trigonometric, and tropical activationsIsmail Khalfaoui Hassani, Stefan KesselheimICLR 2026 · 被引用 1 次
它引用的顶会 Paper19
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Locating and Editing Factual Associations in GPTKevin Meng, David Bau, Alex Andonian, Yonatan BelinkovNeurIPS 2022 · 被引用 3,415 次
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun 等ICML 2021 · 被引用 2,942 次
- Principal Neighbourhood Aggregation for Graph NetsGabriele Corso, Luca Cavalleri, Dominique Beaini, Pietro Liò 等NeurIPS 2020 · 被引用 914 次
- Vision Transformers Need RegistersTimothée Darcet, Maxime Oquab, Julien Mairal, Piotr BojanowskiICLR 2024 · 被引用 769 次
相关 Paper
- Approximation Error Upper and Lower Bounds for Hölder Class with TransformersXin He, Yuling Jiao, Xiliang Lu, Jerry YangICML 2026
- Understanding In-Context Learning on Structured Manifolds: Bridging Attention to Kernel MethodsZhaiming Shen, Alexander Hsu, Rongjie Lai, Wenjing LiaoICLR 2026 · 被引用 13 次
- From Noise to Narrative: Tracing the Origins of Hallucinations in TransformersPraneet Suresh, Jack Stanley, Sonia Joseph, Luca Scimeca 等NeurIPS 2025 · 被引用 6 次
- MechaFormer: Sequence Learning for Kinematic Mechanism Design AutomationDiana Bolanos, Mohammadmehdi Ataei, Pradeep Kumar JayaramanAAAI 2026 · 被引用 1 次
- Why are Sensitive Functions Hard for Transformers?Michael Hahn, Mark RofinACL 2024 · 被引用 3 次
