paper-with-me

Papers

MLCommons Chakra: Advancing Performance Benchmarking and Co-design using Standardized Execution Traces

2026-05-11 · Srinivas Sridharan, Theodor-Adrian Badea, Andy Balogh, Bradford M. Beckmann, Brian Coutinho, Louis Feng, Sheng Fu, Sanshan Gao, Mehryar Garakani, Taekyung Heo, David Kanter, Josh Ladd, Ziwei Li, Winston Liu, Changhai Man, Dan Mihailescu, Spandan More, Joongun Park, Ashwin Ramachandran, Vinay Ramakrishnaiah, Saeed Rashidi, Vijay Janapa Reddi, Puneet Sharma, Phio Tian, William Won, Hanjiang Wu, Huan Xu, Jinsun Yoo, Tushar Krishna arxiv

The fast pace of artificial intelligence~(AI) innovation demands an agile methodology for observation, reproduction and optimization of distributed machine learning~(ML) workload behavior in production AI systems and enables efficient software-hardware~(SW-HW) co-design for future systems. We present Chakra, an open and portable ecosystem for performance benchmarking and co-design. The core component of Chakra is an open and interoperable graph-based representation of distributed AI/ML workloads, called Chakra execution trace~(ET). These ETs represent key operations, such as compute, memory, and communication, data and control dependencies, timing, and resource constraints. Additionally, Chakra includes a complementary set of tools and capabilities to enable the collection, analysis, generation, and adoption of Chakra ETs by a broad range of simulators, emulators, and replay tools. We present analysis of Chakra ETs collected on production AI clusters and demonstrate value via real-world case studies. Chakra has been adopted by MLCommons and has active contributions and engagement across the industry, including but not limited to NVIDIA, AMD, Meta, Keysight, HPE, and Scala, to name a few.

📄 PDF Abstract BibTeX arXiv:2605.11333

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Chakra: Advancing Performance Benchmarking and Co-design using Standardized Execution Traces

2023-05-23 · Srinivas Sridharan, Taekyung Heo, Louis Feng, Zhaodong Wang 외

Benchmarking and co-design are essential for driving optimizations and innovation around ML models, ML software, and next-generation hardware. Full workload benchmarks, e.g. MLPerf, play an essential role in enabling fai…

Benchmarking

MLHarness: A Scalable Benchmarking System for MLCommons

2021-11-09 · Yen-Hsiang Chang, Jianhao Pu, Wen-mei Hwu, JinJun Xiong

With the society's growing adoption of machine learning (ML) and deep learning (DL) for various intelligent solutions, it becomes increasingly imperative to standardize a common set of measures for ML/DL models with larg…

Benchmarking

Improvements & Evaluations on the MLCommons CloudMask Benchmark

2024-03-07 · Varshitha Chennamsetti, Laiba Mehnaz, Dan Zhao, Banani Ghosh 외

In this paper, we report the performance benchmarking results of deep learning models on MLCommons' Science cloud-masking benchmark using a high-performance computing cluster at New York University (NYU): NYU Greene. MLC…

Benchmarking

An MLCommons Scientific Benchmarks Ontology

2025-11-06 · Ben Hawks, Gregor von Laszewski, Matthew D. Sinclair, Marco Colombo 외 arxiv

Scientific machine learning research spans diverse domains and data modalities, yet existing benchmark efforts remain siloed and lack standardization. This makes novel and transformative applications of machine learning …

MLPerf Automotive

2025-10-31 · Radoyeh Shojaei, Predrag Djurdjevic, Mostafa El-Khamy, James Goel 외 arxiv

We present MLPerf Automotive, the first standardized public benchmark for evaluating Machine Learning systems that are deployed for AI acceleration in automotive systems. Developed through a collaborative partnership bet…

2D Semantic Segmentation2D Object Detection3D Object Detection