Characterizing and Understanding End-to-End Multi-Modal Neural Networks on GPUs.

Xiaofeng Hou,Cheng Xu,Jiacheng Liu,Xuehan Tang,Lingyu Sun,Chao Li,Kwang-Ting Cheng
DOI: https://doi.org/10.1109/lca.2022.3215718
IF: 2.3
2022-01-01
IEEE Computer Architecture Letters
Abstract:Multi-modal neural networks have become increasingly pervasive in many machine learning application domains due to their superior accuracy by fusing various modalities. However, they present many unique characteristics such as multi-stage execution, frequent synchronization and high heterogeneity, which are not well understood in the system and architecture community. In this article, we first present and characterize a set of multi-modal neural network workloads of different sizes at inference stage. We then explore their important implications from system and architecture aspects. We hope that our work can help guide future software/hardware design and optimization for efficient inference of multi-modal DNN applications.
What problem does this paper attempt to address?