Lune

MobiCom2025Top-tier venue

Benchmarking Ultra-Low-Power μNPUs

Josh Millar, Yushan Huang, Sarab S. Sethi, Hamed Haddadi, Anil Madhavapeddy

2025Year
13Citations
2Top-tier citations

Abstract

Efficient on-device neural network (NN) inference offers predictable latency, improved privacy and reliability, and lower operating costs for vendors than cloud-based inference. This has sparked recent development of microcontroller-scale NN accelerators, also known as neural processing units (𝜇NPUs), designed specifically for ultra-low-power applications.

We present the first comparative evaluation of a number of commercially-available 𝜇NPUs, including the first independent benchmarks for multiple platforms. To ensure fairness, we develop and open-source a model compilation pipeline supporting consistent benchmarking of quantized models across diverse microcontroller hardware. Our resulting analysis uncovers both expected performance trends as well as surprising disparities between hardware specifications and actual performance, including certain 𝜇NPUs exhibiting unexpected scaling behaviors with model complexity. This work provides a foundation for ongoing evaluation of 𝜇NPU platforms, alongside offering practical insights for both hardware and software developers in this rapidly evolving space.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext bd9beb0d-a878-45d8-a974-b713a4a0d09c

Cited by top-tier papers2

Ask how each one uses it

Builds on8

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines