Continuous Thought Machines
Luke Darlow, Ciaran Regan, Sebastian Risi, Jeffrey Seely, Llion Jones
摘要
Biological brains demonstrate complex neural activity, where neural dynamics are critical to how brains process information. Most artificial neural networks ignore the complexity of individual neurons. We challenge that paradigm. By incorporating neuron-level processing and synchronization, we reintroduce neural timing as a foundational element. We present the Continuous Thought Machine (CTM), a model designed to leverage neural dynamics as its core representation. The CTM has two innovations: (1) neuron-level temporal processing, where each neuron uses unique weight parameters to process incoming histories; and (2) neural synchronization as a latent representation. The CTM aims to strike a balance between neuron abstractions and biological realism. It operates at a level of abstraction that effectively captures essential temporal dynamics while remaining computationally tractable. We demonstrate the CTM's performance and versatility across a range of tasks, including solving 2D mazes, ImageNet-1K classification, parity computation, and more. Beyond displaying rich internal representations and offering a natural avenue for interpretation owing to its internal process, the CTM is able to perform tasks that require complex sequential reasoning. The CTM can also leverage adaptive compute, where it can stop earlier for simpler tasks, or keep computing when faced with more challenging instances. The goal of this work is to share the CTM and its associated innovations, rather than pushing for new state-of-the-art results. To that end, we believe the CTM represents a significant step toward developing more biologically plausible and powerful artificial intelligence systems. We provide an accompanying interactive online demonstration at https://pub.sakana.ai/ctm/ and an extended technical report at https://pub.sakana.ai/ctm/paper .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Selection, Reflection and Self-Refinement: Revisit Reasoning Tasks via a Causal LensYunlong Deng, Boyang Sun, Yan Li, Zeyu Tang 等ICLR 2026 · 被引用 2 次
- C-Voting: Confidence-Based Test-Time Voting without Explicit Energy FunctionsKenji Kubo, Shunsuke Kamiya, Masanori Koyama, Kohei Hayashi 等ICLR 2026 · 被引用 1 次
- Thinking in Scales: Accelerating Gigapixel Pathology Image Analysis via Adaptive Continuous ReasoningJiusong Ge, Yingkang Zhan, Wenjie Zhao, Di Zhang 等ICML 2026
它引用的顶会 Paper14
- Perceiver: General Perception with Iterative AttentionAndrew Jaegle, Felix Gimeno, Andy Brock, Oriol Vinyals 等ICML 2021 · 被引用 1,399 次
- Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth ApproachJonas Geiping, Sean McLeish, Neel Jain, John Kirchenbauer 等NeurIPS 2025 · 被引用 431 次
- Liquid Time-constant NetworksRamin M. Hasani, Mathias Lechner, Alexander Amini, Daniela Rus 等AAAI 2021 · 被引用 399 次
- Human Uncertainty Makes Classification More RobustJoshua C. Peterson, Ruairidh M. Battleday, Thomas L. Griffiths, Olga RussakovskyICCV 2019 · 被引用 362 次
- Recurrent Independent MechanismsAnirudh Goyal, Alex Lamb, Jordan Hoffmann, Shagun Sodhani 等ICLR 2021 · 被引用 357 次
相关 Paper
- Thalamus: a brain-inspired algorithm for biologically-plausible continual learning and disentangled representationsAli HummosICLR 2023 · 被引用 9 次
- Latent Equilibrium: Arbitrarily fast computation with arbitrarily slow neuronsPaul Haider, Benjamin Ellenberger, Laura Kriener, Jakob Jordan 等NeurIPS 2021 · 被引用 32 次
- SyncBrain: Exploring Brain Functional Dynamics Through Neural Oscillatory SynchronizationJiaqi Ding, Tingting Dan, Zhixuan Zhou, Guorong WuAAAI 2026
- Tracking objects that change in appearance with phase synchronySabine Muzellec, Drew Linsley, Alekh Karkada Ashok, Ennio Mingolla 等ICLR 2025
- Recurrent neural network dynamical systems for biological visionWayne Soo, Aldo Battista, Puria Radmard, Xiao-Jing WangNeurIPS 2024 · 被引用 7 次
