paper-with-me

Papers

QuArch: A Question-Answering Dataset for AI Agents in Computer Architecture

2025-01-03 · Shvetank Prakash, Andrew Cheng, Jason Yik, Arya Tschand, Radhika Ghosal, Ikechukwu Uchendu, Jessica Quaye, Jeffrey Ma, Shreyas Grampurohit, Sofia Giannuzzi, Arnav Balyan, Fin Amin, Aadya Pipersenia, Yash Choudhary, Ankita Nayak, Amir Yazdanbakhsh, Vijay Janapa Reddi

We introduce QuArch, a dataset of 1500 human-validated question-answer pairs designed to evaluate and enhance language models' understanding of computer architecture. The dataset covers areas including processor design, memory systems, and performance optimization. Our analysis highlights a significant performance gap: the best closed-source model achieves 84% accuracy, while the top small open-source model reaches 72%. We observe notable struggles in memory systems, interconnection networks, and benchmarking. Fine-tuning with QuArch improves small model accuracy by up to 8%, establishing a foundation for advancing AI-driven computer architecture research. The dataset and leaderboard are at https://harvard-edge.github.io/QuArch/.

📄 PDF Abstract BibTeX arXiv:2501.01892

Code (0)

등록된 구현이 없습니다.

Tasks

BenchmarkingQuestion Answering

Similar Papers 제목 키워드 기반

QuArch: A Benchmark for Evaluating LLM Reasoning in Computer Architecture

2025-10-24 · Shvetank Prakash, Andrew Cheng, Arya Tschand, Mark Mazumder 외 arxiv

The field of computer architecture, which bridges high-level software abstractions and low-level hardware implementations, remains absent from current large language model (LLM) evaluations. To this end, we present QuArc…

QUARCH: A New Quasi-Affine Reconstruction Stratum From Vague Relative Camera Orientation Knowledge

2019-10-01 · ICCV 2019 10 · Devesh Adlakha, Adlane Habed, Fabio Morbidi, Cedric Demonceaux 외

We present a new quasi-affine reconstruction of a scene and its application to camera self-calibration. We refer to this reconstruction as QUARCH (QUasi-Affine Reconstruction with respect to Camera centers and the Hodogr…

Multi-Agent Embodied Question Answering in Interactive Environments

2020-08-01 · ECCV 2020 8 · Sinan Tan, Weilai Xiang, Huaping Liu, Di Guo 외

We investigate a new AI task --- Multi-Agent Interactive Question Answering --- where several agents explore the scene jointly in interactive environments to answer a question. To cooperate efficiently and answer accurat…

3D ReconstructionEmbodied Question AnsweringQuestion Answering

TA-Student VQA: Multi-Agents Training by Self-Questioning

2020-06-01 · CVPR 2020 6 · Peixi Xiong, Ying Wu

There are two main challenges in Visual Question Answering (VQA). The first one is that each model obtains its strengths and shortcomings when applied to several questions; what is more, the "ceiling effect" for specific…

DiversityQuestion AnsweringVisual Question AnsweringVisual Question Answering (VQA)

DBLP-QuAD: A Question Answering Dataset over the DBLP Scholarly Knowledge Graph

2023-03-23 · Debayan Banerjee, Sushil Awale, Ricardo Usbeck, Chris Biemann

In this work we create a question answering dataset over the DBLP scholarly knowledge graph (KG). DBLP is an on-line reference for bibliographic information on major computer science publications that indexes over 4.4 mi…

Question Answering