paper-with-me

홈 › Papers

Pluto: A Benchmark for Evaluating Efficiency of LLM-generated Hardware Code

2025-10-16 · Manar Abdelatty, Maryam Nouh, Jacob K. Rosenstein, Sherief Reda arxiv

Large Language Models (LLMs) are increasingly used to automate hardware design tasks, including the generation of Verilog code. While early benchmarks focus primarily on functional correctness, efficient hardware design demands additional optimization for synthesis metrics such as area, delay, and power. Existing benchmarks fall short in evaluating these aspects comprehensively: they often lack optimized baselines or testbenches for verification. To address these gaps, we present Pluto, a benchmark and evaluation framework designed to assess the efficiency of LLM-generated Verilog designs. Pluto presents a comprehensive evaluation set of 114 problems with self-checking testbenches and multiple Pareto-optimal reference implementations. Experimental results show that state-of-the-art LLMs can achieve high functional correctness, reaching 78.3\% at pass@1, but their synthesis efficiency still lags behind expert-crafted implementations, with area efficiency of 63.8\%, delay efficiency of 65.9\%, and power efficiency of 64.0\% at eff@1. This highlights the need for efficiency-aware evaluation frameworks such as Pluto to drive progress in hardware-focused LLM research.

📄 PDF Abstract BibTeX arXiv:2510.14756

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

PLUTO-4: Frontier Pathology Foundation Models

2025-11-04 · Harshith Padigela, Shima Nofallah, Atchuth Naveen Chilaparasetti, Ryun Han 외 arxiv

Foundation models trained on large-scale pathology image corpora have demonstrated strong transfer capabilities across diverse histopathology tasks. Building on this progress, we introduce PLUTO-4, our next generation of…

PLUTO: Pathology-Universal Transformer

2024-05-13 · Dinkar Juyal, Harshith Padigela, Chintan Shah, Daniel Shenker 외

Pathology is the study of microscopic inspection of tissue, and a pathology diagnosis is often the medical gold standard to diagnose disease. Pathology images provide a unique challenge for computer-vision-based analysis…

Instance SegmentationSemantic Segmentation

PLUTO: Penalized Unbiased Logistic Regression Trees

2014-11-25 · Wenwen Zhang, Wei-Yin Loh

We propose a new algorithm called PLUTO for building logistic regression trees to binary response data. PLUTO can capture the nonlinear and interaction patterns in messy data by recursively partitioning the sample space.…

regressionSelection biasVariable Selection

ECCO: Can We Improve Model-Generated Code Efficiency Without Sacrificing Functional Correctness?

2024-07-19 · Siddhant Waghjale, Vishruth Veerendranath, Zora Zhiruo Wang, Daniel Fried

Although large language models (LLMs) have been largely successful in generating functionally correct programs, conditioning models to produce efficient solutions while ensuring correctness remains a challenge. Further, …

BenchmarkingCode GenerationIn-Context Learning

PLUTO: Automated Solutions for Patent Translation

2012-04-01 · WS 2012 4 · John Tinsley, Alex Ceausu, ru, Jian Zhang
Domain AdaptationMachine TranslationTranslation