paper-with-me

Papers

BENCHIP: Benchmarking Intelligence Processors

2017-10-23 · Jinhua Tao, Zidong Du, Qi Guo, Huiying Lan, Lei Zhang, Shengyuan Zhou, Lingjie Xu, Cong Liu, Haifeng Liu, Shan Tang, Allen Rush, Willian Chen, Shaoli Liu, Yunji Chen, Tianshi Chen

The increasing attention on deep learning has tremendously spurred the design of intelligence processing hardware. The variety of emerging intelligence processors requires standard benchmarks for fair comparison and system optimization (in both software and hardware). However, existing benchmarks are unsuitable for benchmarking intelligence processors due to their non-diversity and nonrepresentativeness. Also, the lack of a standard benchmarking methodology further exacerbates this problem. In this paper, we propose BENCHIP, a benchmark suite and benchmarking methodology for intelligence processors. The benchmark suite in BENCHIP consists of two sets of benchmarks: microbenchmarks and macrobenchmarks. The microbenchmarks consist of single-layer networks. They are mainly designed for bottleneck analysis and system optimization. The macrobenchmarks contain state-of-the-art industrial networks, so as to offer a realistic comparison of different platforms. We also propose a standard benchmarking methodology built upon an industrial software stack and evaluation metrics that comprehensively reflect the various characteristics of the evaluated intelligence processors. BENCHIP is utilized for evaluating various hardware platforms, including CPUs, GPUs, and accelerators. BENCHIP will be open-sourced soon.

📄 PDF Abstract BibTeX arXiv:1710.08315

Code (0)

등록된 구현이 없습니다.

Tasks

BenchmarkingDiversity

Similar Papers 제목 키워드 기반

Pareto Optimal Benchmarking of AI Models on ARM Cortex Processors for Sustainable Embedded Systems

2026-02-19 · Pranay Jain, Maximilian Kasper, Göran Köber, Oliver Amft 외 arxiv

This work presents a practical benchmarking framework for optimizing artificial intelligence (AI) models on ARM Cortex processors (M0+, M4, M7), focusing on energy efficiency, accuracy, and resource utilization in embedd…

Strategies for Optimizing End-to-End Artificial Intelligence Pipelines on Intel Xeon Processors

2022-11-01 · Meena Arunachalam, Vrushabh Sanghavi, Yi A Yao, Yi A Zhou 외

End-to-end (E2E) artificial intelligence (AI) pipelines are composed of several stages including data preprocessing, data ingestion, defining and training the model, hyperparameter optimization, deployment, inference, po…

Hyperparameter OptimizationRecommendation Systems

Benchmarking the Performance and Energy Efficiency of AI Accelerators for AI Training

2019-09-15 · Yuxin Wang, Qiang Wang, Shaohuai Shi, Xin He 외

Deep learning has become widely used in complex AI applications. Yet, training a deep neural network (DNNs) model requires a considerable amount of calculations, long running time, and much energy. Nowadays, many-core AI…

BenchmarkingCPUDeep LearningGPU

Brain Co-Processors: Using AI to Restore and Augment Brain Function

2020-12-06 · Rajesh P. N. Rao

Brain-computer interfaces (BCIs) use decoding algorithms to control prosthetic devices based on brain signals for restoration of lost function. Computer-brain interfaces (CBIs), on the other hand, use encoding algorithms…

Quantum latent distributions in deep generative models

2025-08-27 · Omar Bacarreza, Thorin Farnsworth, Alexander Makarovskiy, Hugo Wallner 외 arxiv

Many successful families of generative models leverage a low-dimensional latent distribution that is mapped to a data distribution. Though simple latent distributions are often used, the choice of distribution has a stro…