Mobile Machine Learning Hardware at ARM: A Systems-on-Chip (SoC) Perspective

Machine learning is playing an increasingly significant role in emerging mobile application domains such as AR/VR, ADAS, etc. Accordingly, hardware architects have designed customized hardware for machine learning algorithms, especially neural networks, to improve compute efficiency. However, machine learning is typically just one processing stage in complex end-to-end applications, which involve multiple components in a mobile Systems-on-a-chip (SoC). Focusing on just ML accelerators loses bigger optimization opportunity at the system (SoC) level. This paper argues that hardware architects should expand the optimization scope to the entire SoC. We demonstrate one particular case-study in the domain of continuous computer vision where camera sensor, image signal processor (ISP), memory, and NN accelerator are synergistically co-designed to achieve optimal system-level efficiency.

Paper

References (17)

10Jetson TX2 Modulehttp://www.nvidia.com/object/ embedded-systems-dev-kits-modules
11Arm Compute Librarygithub.com/ARM-software/ ComputeLibrary
12Apple’s Neural Engine Infuses the iPhone with AI Smartswww

Scroll for more · 5 remaining

Similar papers

© 2026 NYSGPT2525 LLC