TinyTrain: Resource-Aware Task-Adaptive Sparse Training of DNNs at the Data-Scarce Edge
Young D. Kwon, Rui Li, Stylianos I. Venieris, Jagmohan Chauhan, Nicholas Donald Lane, Cecilia Mascolo
摘要
On-device training is essential for user personalisation and privacy. With the pervasiveness of IoT devices and microcontroller units (MCUs), this task becomes more challenging due to the constrained memory and compute resources, and the limited availability of labelled user data. Nonetheless, prior works neglect the data scarcity issue, require excessively long training time (e.g. a few hours), or induce substantial accuracy loss (≥10%). In this paper, we propose TinyTrain, an on-device training approach that drastically reduces training time by selectively updating parts of the model and explicitly coping with data scarcity. TinyTrain introduces a task-adaptive sparse-update method that dynamically selects the layer/channel to update based on a multiobjective criterion that jointly captures user data, the memory, and the compute capabilities of the target device, leading to high accuracy on unseen tasks with reduced computation and memory footprint. TinyTrain outperforms vanilla finetuning of the entire network by 3.6-5.0% in accuracy, while reducing the backward-pass memory and computation cost by up to 1,098× and 7.68×, respectively. Targeting broadly used realworld edge devices, TinyTrain achieves 9.5× faster and 3.5× more energy-efficient training over status-quo approaches, and 2.23× smaller memory footprint than SOTA methods, while remaining within the 1 MB memory envelope of MCU-grade platforms. Code is available at https://github.com/theyoungkwon/TinyTrain
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Benchmarking Ultra-Low-Power μNPUsJosh Millar, Yushan Huang, Sarab S. Sethi, Hamed Haddadi 等MobiCom 2025 · 被引用 13 次
- DEX: Data Channel Extension for Efficient CNN Inference on Tiny AI AcceleratorsTaesik Gong, Fahim Kawsar, Chulhong MinNeurIPS 2024 · 被引用 8 次
- Delta: A Cloud-assisted Data Enrichment Framework for On-Device Continual LearningChen Gong, Zhenzhe Zheng, Fan Wu, Xiaofeng Jia 等MobiCom 2024 · 被引用 6 次
- HierarchicalPrune: Position-Aware Compression for Large-Scale Diffusion ModelsYoung D. Kwon, Rui Li, Sijia Li, Da Li 等AAAI 2026 · 被引用 5 次
- Study of Training Dynamics for Memory-Constrained Fine-TuningAël Quélennec, Nour Hezbri, Pavlo Mozharovskyi, Van-Tam Nguyen 等ICLR 2026 · 被引用 1 次
它引用的顶会 Paper12
- A Baseline for Few-Shot Image ClassificationGuneet Singh Dhillon, Pratik Chaudhari, Avinash Ravichandran, Stefano SoattoICLR 2020 · 被引用 640 次
- TTT++: When Does Self-Supervised Test-Time Training Fail or Thrive?Yuejiang Liu, Parth Kothari, Bastien van Delft, Baptiste Bellot-Gurlet 等NeurIPS 2021 · 被引用 469 次
- On-Device Training Under 256KB MemoryJi Lin, Ligeng Zhu, Wei-Ming Chen, Wei-Chen Wang 等NeurIPS 2022 · 被引用 345 次
- Pushing the Limits of Simple Pipelines for Few-Shot Learning: External Data and Fine-Tuning Make a DifferenceShell Xu Hu, Da Li, Jan Stühmer, Minyoung Kim 等CVPR 2022 · 被引用 161 次
- Dynamic Tensor RematerializationMarisa Kirisame, Steven Lyubomirsky, Altan Haan, Jennifer Brennan 等ICLR 2021 · 被引用 115 次
相关 Paper
- Enabling On-Tiny-Device Model Personalization via Gradient Condensing and Alternant Partial UpdateZhenge Jia, Yiyang Shi, Zeyu Bao, Zirui Wang 等DAC 2025
- TinyFoA: Memory Efficient Forward-Only Algorithm for On-Device LearningBaichuan Huang, Amir AminifarAAAI 2025 · 被引用 3 次
- TinyTTA: Efficient Test-time Adaptation via Early-exit Ensembles on Edge DevicesHong Jia, Young D. Kwon, Alessio Orsino, Ting Dang 等NeurIPS 2024 · 被引用 25 次
- MCUNet: Tiny Deep Learning on IoT DevicesJi Lin, Wei-Ming Chen, Yujun Lin, John Cohn 等NeurIPS 2020 · 被引用 827 次
- RepNet: Efficient On-Device Learning via Feature ReprogrammingLi Yang, Adnan Siraj Rakin, Deliang FanCVPR 2022 · 被引用 18 次
