Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
97 tokens/sec
GPT-4o
53 tokens/sec
Gemini 2.5 Pro Pro
43 tokens/sec
o3 Pro
4 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

Latency and Throughput Characterization of Convolutional Neural Networks for Mobile Computer Vision (1803.09492v1)

Published 26 Mar 2018 in cs.CV

Abstract: We study performance characteristics of convolutional neural networks (CNN) for mobile computer vision systems. CNNs have proven to be a powerful and efficient approach to implement such systems. However, the system performance depends largely on the utilization of hardware accelerators, which are able to speed up the execution of the underlying mathematical operations tremendously through massive parallelism. Our contribution is performance characterization of multiple CNN-based models for object recognition and detection with several different hardware platforms and software frameworks, using both local (on-device) and remote (network-side server) computation. The measurements are conducted using real workloads and real processing platforms. On the platform side, we concentrate especially on TensorFlow and TensorRT. Our measurements include embedded processors found on mobile devices and high-performance processors that can be used on the network side of mobile systems. We show that there exists significant latency--throughput trade-offs but the behavior is very complex. We demonstrate and discuss several factors that affect the performance and yield this complex behavior.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (6)
  1. Jussi Hanhirova (2 papers)
  2. Teemu Kämäräinen (4 papers)
  3. Sipi Seppälä (1 paper)
  4. Matti Siekkinen (13 papers)
  5. Vesa Hirvisalo (3 papers)
  6. Antti Ylä-Jääski (9 papers)
Citations (85)

Summary

We haven't generated a summary for this paper yet.