← Back to all articles
arXiv cs.AIAugust 17, 2026

Joint Optimization of Memory and Computing Frequency for Energy-Efficient DNN Inference

Excerpt

arXiv:2608.13863v1 Announce Type: new Abstract: Deep neural network (DNN) inference on mobile devices often incurs high latency and energy consumption due to limited computing and memory resources. To enable energy-efficient DNN inference, most existing studies focus on dynamic voltage and frequency scaling (DVFS) for adjusting the computing frequency, while the impact of memory frequency on the inference performance has been greatly overlooked. In this paper, we consider the impact of memory fr