TinyTrain: Resource-Aware Task-Adaptive Sparse Training of DNNs at the Data-Scarce Edge
Young D. Kwon, Rui Li, Stylianos I. Venieris, Jagmohan Chauhan, Nicholas Donald Lane, Cecilia Mascolo
Abstract
On-device training is essential for user personalisation and privacy. With the pervasiveness of IoT devices and microcontroller units (MCUs), this task becomes more challenging due to the constrained memory and compute resources, and the limited availability of labelled user data. Nonetheless, prior works neglect the data scarcity issue, require excessively long training time (e.g. a few hours), or induce substantial accuracy loss (≥10%). In this paper, we propose TinyTrain, an on-device training approach that drastically reduces training time by selectively updating parts of the model and explicitly coping with data scarcity. TinyTrain introduces a task-adaptive sparse-update method that dynamically selects the layer/channel to update based on a multiobjective criterion that jointly captures user data, the memory, and the compute capabilities of the target device, leading to high accuracy on unseen tasks with reduced computation and memory footprint. TinyTrain outperforms vanilla finetuning of the entire network by 3.6-5.0% in accuracy, while reducing the backward-pass memory and computation cost by up to 1,098× and 7.68×, respectively. Targeting broadly used realworld edge devices, TinyTrain achieves 9.5× faster and 3.5× more energy-efficient training over status-quo approaches, and 2.23× smaller memory footprint than SOTA methods, while remaining within the 1 MB memory envelope of MCU-grade platforms. Code is available at https://github.com/theyoungkwon/TinyTrain
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8db2d4ec-1aff-4e8f-be17-6dbaf2309697Cited by top-tier papers6
- Benchmarking Ultra-Low-Power μNPUsJosh Millar, Yushan Huang, Sarab S. Sethi, Hamed Haddadi et al.MobiCom 2025 · 13 citations
- DEX: Data Channel Extension for Efficient CNN Inference on Tiny AI AcceleratorsTaesik Gong, Fahim Kawsar, Chulhong MinNeurIPS 2024 · 8 citations
- Delta: A Cloud-assisted Data Enrichment Framework for On-Device Continual LearningChen Gong, Zhenzhe Zheng, Fan Wu, Xiaofeng Jia et al.MobiCom 2024 · 6 citations
- HierarchicalPrune: Position-Aware Compression for Large-Scale Diffusion ModelsYoung D. Kwon, Rui Li, Sijia Li, Da Li et al.AAAI 2026 · 5 citations
- Study of Training Dynamics for Memory-Constrained Fine-TuningAël Quélennec, Nour Hezbri, Pavlo Mozharovskyi, Van-Tam Nguyen et al.ICLR 2026 · 1 citation
Builds on12
- A Baseline for Few-Shot Image ClassificationGuneet Singh Dhillon, Pratik Chaudhari, Avinash Ravichandran, Stefano SoattoICLR 2020 · 640 citations
- TTT++: When Does Self-Supervised Test-Time Training Fail or Thrive?Yuejiang Liu, Parth Kothari, Bastien van Delft, Baptiste Bellot-Gurlet et al.NeurIPS 2021 · 469 citations
- On-Device Training Under 256KB MemoryJi Lin, Ligeng Zhu, Wei-Ming Chen, Wei-Chen Wang et al.NeurIPS 2022 · 345 citations
- Pushing the Limits of Simple Pipelines for Few-Shot Learning: External Data and Fine-Tuning Make a DifferenceShell Xu Hu, Da Li, Jan Stühmer, Minyoung Kim et al.CVPR 2022 · 161 citations
- Dynamic Tensor RematerializationMarisa Kirisame, Steven Lyubomirsky, Altan Haan, Jennifer Brennan et al.ICLR 2021 · 115 citations
Related papers
- Enabling On-Tiny-Device Model Personalization via Gradient Condensing and Alternant Partial UpdateZhenge Jia, Yiyang Shi, Zeyu Bao, Zirui Wang et al.DAC 2025
- TinyFoA: Memory Efficient Forward-Only Algorithm for On-Device LearningBaichuan Huang, Amir AminifarAAAI 2025 · 3 citations
- TinyTTA: Efficient Test-time Adaptation via Early-exit Ensembles on Edge DevicesHong Jia, Young D. Kwon, Alessio Orsino, Ting Dang et al.NeurIPS 2024 · 25 citations
- MCUNet: Tiny Deep Learning on IoT DevicesJi Lin, Wei-Ming Chen, Yujun Lin, John Cohn et al.NeurIPS 2020 · 827 citations
- RepNet: Efficient On-Device Learning via Feature ReprogrammingLi Yang, Adnan Siraj Rakin, Deliang FanCVPR 2022 · 18 citations
