跳到主要导航 跳到搜索 跳到主要内容

HACompBench: Co-designed Multimodal DNN Compression Evaluation for Edge Devices

  • Zhengyu Gan*
  • , Haohua Du
  • , Chengquan Feng
  • , Haisheng Tan
  • *此作品的通讯作者
  • School of Computer Science

科研成果: 书/报告/会议事项章节会议稿件同行评审

摘要

Deployment of deep neural networks on edge devices faces challenges from heterogeneous hardware and multimodal tasks, where existing compression evaluation frameworks overlook hardware co-design, leading to suboptimal performance. To address this, we introduce HACompBench, a new hardware-aware framework that defines compression evaluation as a multi-objective optimization problem and combines hardware metrics such as quantization efficiency ξ and sparsity compatibility η with a dynamic scoring function J. We performed comprehensive experiments across four leading SoC platforms: Snapdragon 888, Snapdragon 765G, Kirin 970, and Jetson Nano P3450, and tested ten DNN models covering vision, text, and speech modalities using compression techniques such as quantization, pruning, and weight sharing, revealing hardware-induced performance gaps, such as quantization yields J=32.3% on Snapdragon 888 but J=68.0% on Jetson Nano P3450 due to INT8 emulation overhead. These results systematically highlight compression variations from differences in parallel processing capabilities. HACompBench innovates by linking hardware features, supporting multimodal tasks, and surpassing MLPerf Tiny’s single-modality focus and AIoTBench’s lack of co-design through embedded metrics ξ and η, while modality-specific corrections improve accuracy by up to 12.6%. It provides a unified and robust framework for edge deployment.

源语言英语
主期刊名Algorithms and Architectures for Parallel Processing - 25th International Conference, ICA3PP 2025, Proceedings
编辑Huazhong Liu, Shadi Ibrahim, Thomas Rauber
出版商Springer Science and Business Media Deutschland GmbH
478-496
页数19
ISBN(印刷版)9789819584048
DOI
出版状态已出版 - 2026
活动25th International Conference on Algorithms and Architectures for Parallel Processing, ICA3PP 2025 - Zhengzhou, 中国
期限: 30 10月 20252 11月 2025

出版系列

姓名Lecture Notes in Computer Science
16383 LNCS
ISSN(印刷版)0302-9743
ISSN(电子版)1611-3349

会议

会议25th International Conference on Algorithms and Architectures for Parallel Processing, ICA3PP 2025
国家/地区中国
Zhengzhou
时期30/10/252/11/25

学术指纹

探究 'HACompBench: Co-designed Multimodal DNN Compression Evaluation for Edge Devices' 的科研主题。它们共同构成独一无二的学术指纹。

引用此