InstaNAS: Instance-Aware Neural Architecture Search
An-Chieh Cheng, Chieh Hubert Lin, Da-Cheng Juan, Wei Wei, Min Sun
Abstract
Conventional Neural Architecture Search (NAS) aims at finding a single architecture that achieves the best performance, which usually optimizes task related learning objectives such as accuracy. However, a single architecture may not be representative enough for the whole dataset with high diversity and variety. Intuitively, electing domain-expert architectures that are proficient in domain-specific features can further benefit architecture related objectives such as latency. In this paper, we propose InstaNAS—an instance-aware NAS framework—that employs a controller trained to search for a “distribution of architectures” instead of a single architecture; This allows the model to use sophisticated architectures for the difficult samples, which usually comes with large architecture related cost, and shallow architectures for those easy samples. During the inference phase, the controller assigns each of the unseen input samples with a domain expert architecture that can achieve high accuracy with customized inference costs. Experiments within a search space inspired by MobileNetV2 show InstaNAS can achieve up to 48.8% latency reduction without compromising accuracy on a series of datasets against MobileNetV2.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers9
- FasterSeg: Searching for Faster Real-time Semantic SegmentationWuyang Chen, Xinyu Gong, Xianming Liu, Qian Zhang et al.ICLR 2020 · 206 citations
- Towards Fast Adaptation of Neural Architectures with Meta LearningDongze Lian, Yin Zheng, Yintao Xu, Yanxiong Lu et al.ICLR 2020 · 95 citations
- Mitigating Forgetting in Online Continual Learning via Instance-Aware ParameterizationHung-Jen Chen, An-Chieh Cheng, Da-Cheng Juan, Wei Wei et al.NeurIPS 2020 · 50 citations
- Instance-Aware Dynamic Neural Network QuantizationZhenhua Liu, Yunhe Wang, Kai Han, Siwei Ma et al.CVPR 2022 · 38 citations
- Cocktailer: Analyzing and Optimizing Dynamic Control Flow in Deep LearningChen Zhang, Lingxiao Ma, Jilong Xue, Yining Shi et al.OSDI 2023 · 28 citations
Related papers
- You only search once: on lightweight differentiable architecture search for resource-constrained embedded platformsXiangzhong Luo, Di Liu, Hao Kong, Shuo Huai et al.DAC 2022 · 13 citations
- Rapid Neural Architecture Search by Learning to Generate Graphs from DatasetsHayeon Lee, Eunyoung Hyung, Sung Ju HwangICLR 2021 · 57 citations
- M-NAS: Meta Neural Architecture SearchJiaxing Wang, Jiaxiang Wu, Haoli Bai, Jian ChengAAAI 2020 · 34 citations
- Optimizing Network Simulation: Enhancing Performance Prediction Accuracy via Neural Architecture SearchShaoChen He, Zirui Zhuang, Haifeng Sun, Xiaoyuan Fu et al.ICML 2026
- Improving One-Shot NAS by Suppressing the Posterior FadingXiang Li, Chen Lin, Chuming Li, Ming Sun et al.CVPR 2020
