MONAS: Efficient Zero-Shot Neural Architecture Search for MCUs
Fuente:
arXiv
Saved in:
| Main Authors: | Qiao, Ye, Xu, Haocheng, Zhang, Yifan, Huang, Sitao |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MicroNAS: Zero-Shot Neural Architecture Search for MCUs
by: Qiao, Ye, et al.
Published: (2024)
by: Qiao, Ye, et al.
Published: (2024)
TG-NAS: Generalizable Zero-Cost Proxies with Operator Description Embedding and Graph Learning for Efficient Neural Architecture Search
by: Qiao, Ye, et al.
Published: (2024)
by: Qiao, Ye, et al.
Published: (2024)
Rethinking RoPE Scaling in Quantized LLM: Theory, Outlier, and Channel-Band Analysis with Weight Rescaling
by: Qiao, Ye, et al.
Published: (2025)
by: Qiao, Ye, et al.
Published: (2025)
RSEND: Retinex-based Squeeze and Excitation Network with Dark Region Detection for Efficient Low Light Image Enhancement
by: Li, Jingcheng, et al.
Published: (2024)
by: Li, Jingcheng, et al.
Published: (2024)
TeLLMe: An Energy-Efficient Ternary LLM Accelerator for Prefilling and Decoding on Edge FPGAs
by: Qiao, Ye, et al.
Published: (2025)
by: Qiao, Ye, et al.
Published: (2025)
COBRA: Algorithm-Architecture Co-optimized Binary Transformer Accelerator for Edge Inference
by: Qiao, Ye, et al.
Published: (2025)
by: Qiao, Ye, et al.
Published: (2025)
TeLLMe v2: An Efficient End-to-End Ternary LLM Prefill and Decode Accelerator with Table-Lookup Matmul on Edge FPGAs
by: Qiao, Ye, et al.
Published: (2025)
by: Qiao, Ye, et al.
Published: (2025)
Q-ROAR: Outlier-Aware Rescaling for RoPE Position Interpolation in Quantized Long-Context LLMs
by: Qiao, Ye, et al.
Published: (2025)
by: Qiao, Ye, et al.
Published: (2025)
Zero-Shot Neural Architecture Search with Weighted Response Correlation
by: Jing, Kun, et al.
Published: (2025)
by: Jing, Kun, et al.
Published: (2025)
Zero-Shot Neural Architecture Search: Challenges, Solutions, and Opportunities
by: Li, Guihong, et al.
Published: (2023)
by: Li, Guihong, et al.
Published: (2023)
Jet-Nemotron: Efficient Language Model with Post Neural Architecture Search
by: Gu, Yuxian, et al.
Published: (2025)
by: Gu, Yuxian, et al.
Published: (2025)
Low-Energy On-Device Personalization for MCUs
by: Huang, Yushan, et al.
Published: (2024)
by: Huang, Yushan, et al.
Published: (2024)
Characterizing State Space Model and Hybrid Language Model Performance with Long Context
by: Mitra, Saptarshi, et al.
Published: (2025)
by: Mitra, Saptarshi, et al.
Published: (2025)
SA-GNAS: Seed Architecture Expansion for Efficient Large-scale Graph Neural Architecture Search
by: Zhu, Guanghui, et al.
Published: (2024)
by: Zhu, Guanghui, et al.
Published: (2024)
Zero-shot Quantum Neural Architecture Search
by: Dao, Tung, et al.
Published: (2026)
by: Dao, Tung, et al.
Published: (2026)
vMCU: Coordinated Memory Management and Kernel Optimization for DNN Inference on MCUs
by: Zheng, Size, et al.
Published: (2024)
by: Zheng, Size, et al.
Published: (2024)
UnIT: Scalable Unstructured Inference-Time Pruning for MAC-efficient Neural Inference on MCUs
by: Neth, Ashe, et al.
Published: (2025)
by: Neth, Ashe, et al.
Published: (2025)
FR-NAS: Forward-and-Reverse Graph Predictor for Efficient Neural Architecture Search
by: Zhang, Haoming, et al.
Published: (2024)
by: Zhang, Haoming, et al.
Published: (2024)
SleepNetZero: Zero-Burden Zero-Shot Reliable Sleep Staging With Neural Networks Based on Ballistocardiograms
by: Li, Shuzhen, et al.
Published: (2024)
by: Li, Shuzhen, et al.
Published: (2024)
FASQ: Flexible Accelerated Subspace Quantization for Calibration-Free LLM Compression
by: Qiao, Ye, et al.
Published: (2026)
by: Qiao, Ye, et al.
Published: (2026)
LLM-NAS: LLM-driven Hardware-Aware Neural Architecture Search
by: Zhu, Hengyi, et al.
Published: (2025)
by: Zhu, Hengyi, et al.
Published: (2025)
ZeroDiff: Solidified Visual-Semantic Correlation in Zero-Shot Learning
by: Ye, Zihan, et al.
Published: (2024)
by: Ye, Zihan, et al.
Published: (2024)
DANCE: Resource-Efficient Neural Architecture Search with Data-Aware and Continuous Adaptation
by: Wang, Maolin, et al.
Published: (2025)
by: Wang, Maolin, et al.
Published: (2025)
HGNAS: Hardware-Aware Graph Neural Architecture Search for Edge Devices
by: Zhou, Ao, et al.
Published: (2024)
by: Zhou, Ao, et al.
Published: (2024)
One-Embedding-Fits-All: Efficient Zero-Shot Time Series Forecasting by a Model Zoo
by: Shi, Hao-Nan, et al.
Published: (2025)
by: Shi, Hao-Nan, et al.
Published: (2025)
ZeroShotOpt: Towards Zero-Shot Pretrained Models for Efficient Black-Box Optimization
by: Meindl, Jamison, et al.
Published: (2025)
by: Meindl, Jamison, et al.
Published: (2025)
SeqFusion: Sequential Fusion of Pre-Trained Models for Zero-Shot Time-Series Forecasting
by: Huang, Ting-Ji, et al.
Published: (2025)
by: Huang, Ting-Ji, et al.
Published: (2025)
Auto-Agent-Distiller: Towards Efficient Deep Reinforcement Learning Agents via Neural Architecture Search
by: Fu, Yonggan, et al.
Published: (2020)
by: Fu, Yonggan, et al.
Published: (2020)
NAS-Bench-Graph: Benchmarking Graph Neural Architecture Search
by: Qin, Yijian, et al.
Published: (2022)
by: Qin, Yijian, et al.
Published: (2022)
InstMeter: An Instruction-Level Method to Predict Energy and Latency of DL Model Inference on MCUs
by: Liu, Hao, et al.
Published: (2026)
by: Liu, Hao, et al.
Published: (2026)
Zero-Shot and Efficient Clarification Need Prediction in Conversational Search
by: Lu, Lili, et al.
Published: (2025)
by: Lu, Lili, et al.
Published: (2025)
QNAS: A Neural Architecture Search Framework for Accurate and Efficient Quantum Neural Networks
by: Maleki, Kooshan, et al.
Published: (2026)
by: Maleki, Kooshan, et al.
Published: (2026)
Kernel-Level Energy-Efficient Neural Architecture Search for Tabular Dataset
by: La, Hoang-Loc, et al.
Published: (2025)
by: La, Hoang-Loc, et al.
Published: (2025)
Anytime Neural Architecture Search on Tabular Data
by: Xing, Naili, et al.
Published: (2024)
by: Xing, Naili, et al.
Published: (2024)
Optimizing the Deployment of Tiny Transformers on Low-Power MCUs
by: Jung, Victor J. B., et al.
Published: (2024)
by: Jung, Victor J. B., et al.
Published: (2024)
Personalized Federated Instruction Tuning via Neural Architecture Search
by: Zhang, Pengyu, et al.
Published: (2024)
by: Zhang, Pengyu, et al.
Published: (2024)
Optimized Spatial Architecture Mapping Flow for Transformer Accelerators
by: Xu, Haocheng, et al.
Published: (2024)
by: Xu, Haocheng, et al.
Published: (2024)
Graph Neural Architecture Search with GPT-4
by: Wang, Haishuai, et al.
Published: (2023)
by: Wang, Haishuai, et al.
Published: (2023)
Zero-Shot Robustification of Zero-Shot Models
by: Adila, Dyah, et al.
Published: (2023)
by: Adila, Dyah, et al.
Published: (2023)
Efficient Reinforcement Learning for Zero-Shot Coordination in Evolving Games
by: Hui, Bingyu, et al.
Published: (2025)
by: Hui, Bingyu, et al.
Published: (2025)
Similar Items
-
MicroNAS: Zero-Shot Neural Architecture Search for MCUs
by: Qiao, Ye, et al.
Published: (2024) -
TG-NAS: Generalizable Zero-Cost Proxies with Operator Description Embedding and Graph Learning for Efficient Neural Architecture Search
by: Qiao, Ye, et al.
Published: (2024) -
Rethinking RoPE Scaling in Quantized LLM: Theory, Outlier, and Channel-Band Analysis with Weight Rescaling
by: Qiao, Ye, et al.
Published: (2025) -
RSEND: Retinex-based Squeeze and Excitation Network with Dark Region Detection for Efficient Low Light Image Enhancement
by: Li, Jingcheng, et al.
Published: (2024) -
TeLLMe: An Energy-Efficient Ternary LLM Accelerator for Prefilling and Decoding on Edge FPGAs
by: Qiao, Ye, et al.
Published: (2025)