paper-with-me

Papers

Tango: A Deep Neural Network Benchmark Suite for Various Accelerators

2019-01-14 · Aajna Karki, Chethan Palangotu Keshava, Spoorthi Mysore Shivakumar, Joshua Skow, Goutam Madhukeshwar Hegde, Hyeran Jeon

Deep neural networks (DNNs) have been proving the effectiveness in various computing fields. To provide more efficient computing platforms for DNN applications, it is essential to have evaluation environments that include assorted benchmark workloads. Though a few DNN benchmark suites have been recently released, most of them require to install proprietary DNN libraries or resource-intensive DNN frameworks, which are hard to run on resource-limited mobile platforms or architecture simulators. To provide a more scalable evaluation environment, we propose a new DNN benchmark suite that can run on any platform that supports CUDA and OpenCL. The proposed benchmark suite includes the most widely used five convolution neural networks and two recurrent neural networks. We provide in-depth architectural statistics of these networks while running them on an architecture simulator, a server- and a mobile-GPU, and a mobile FPGA.

📄 PDF Abstract BibTeX arXiv:1901.04987

Code (0)

등록된 구현이 없습니다.

Tasks

GPU

Methods 이 논문이 사용한 방법론

Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…

Similar Papers 제목 키워드 기반

Performance and Power: Systematic Evaluation of AI Workloads on Accelerators with CARAML

2024-09-19 · Chelsea Maria John, Stepan Nassyr, Carolin Penke, Andreas Herten

The rapid advancement of machine learning (ML) technologies has driven the development of specialized hardware accelerators designed to facilitate more efficient model training. This paper introduces the CARAML benchmark…

Tango: Taming Visual Signals for Efficient Video Large Language Models

2026-04-10 · Shukang Yin, Sirui Zhao, Hanchao Wang, Baozhi Jia 외 arxiv

Token pruning has emerged as a mainstream approach for developing efficient Video Large Language Models (Video LLMs). This work revisits and advances the two predominant token-pruning paradigms: attention-based selection…

Mustango: Toward Controllable Text-to-Music Generation

2023-11-14 · Jan Melechovsky, Zixun Guo, Deepanway Ghosal, Navonil Majumder 외

The quality of the text-to-music models has reached new heights due to recent advancements in diffusion models. The controllability of various musical aspects, however, has barely been explored. In this paper, we propose…

Data AugmentationDenoisingInformation RetrievalMusic Generation+2

LLM-Inference-Bench: Inference Benchmarking of Large Language Models on AI Accelerators

2024-10-31 · Krishna Teja Chitty-Venkata, Siddhisanket Raskar, Bharat Kale, Farah Ferdaus 외

Large Language Models (LLMs) have propelled groundbreaking advancements across several domains and are commonly used for text generation applications. However, the computational demands of these complex models pose signi…

BenchmarkingText Generation

BENCHIP: Benchmarking Intelligence Processors

2017-10-23 · Jinhua Tao, Zidong Du, Qi Guo, Huiying Lan 외

The increasing attention on deep learning has tremendously spurred the design of intelligence processing hardware. The variety of emerging intelligence processors requires standard benchmarks for fair comparison and syst…

BenchmarkingDiversity