paper-with-me

홈 › Papers

EvolveGen: Algorithmic Level Hardware Model Checking Benchmark Generation through Reinforcement Learning

2026-02-26 · Guangyu Hu, Xiaofeng Zhou, Wei Zhang, Hongce Zhang arxiv

Progress in hardware model checking depends critically on high-quality benchmarks. However, the community faces a significant benchmark gap: existing suites are limited in number, often distributed only in representations such as BTOR2 without access to the originating register-transfer-level (RTL) designs, and biased toward extreme difficulty where instances are either trivial or intractable. These limitations hinder rigorous evaluation of new verification techniques and encourage overfitting of solver heuristics to a narrow set of problems. To address this, we introduce EvolveGen, a framework for generating hardware model checking benchmarks by combining reinforcement learning (RL) with high-level synthesis (HLS). Our approach operates at an algorithmic level of abstraction in which an RL agent learns to construct computation graphs. By compiling these graphs under different synthesis directives, we produce pairs of functionally equivalent but structurally distinct hardware designs, inducing challenging model checking instances. Solver runtime is used as the reward signal, enabling the agent to autonomously discover and generate small-but-hard instances that expose solver-specific weaknesses. Experiments show that EvolveGen efficiently creates a diverse benchmark set in standard formats (e.g., AIGER and BTOR2) and effectively reveals performance bottlenecks in state-of-the-art model checkers.

📄 PDF Abstract BibTeX arXiv:2602.22609

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Verifying Non-friendly Formal Verification Designs: Can We Start Earlier?

2024-10-24 · Bryan Olmos, Daniel Gerl, Aman Kumar, Djones Lettnin

The design of Systems on Chips (SoCs) is becoming more and more complex due to technological advancements. Missed bugs can cause drastic failures in safety-critical environments leading to the endangerment of lives. To o…

IC3-Evolve: Proof-/Witness-Gated Offline LLM-Driven Heuristic Evolution for IC3 Hardware Model Checking

2026-01-18 · Mingkai Miao, Guangyu Hu, Ziyi Yang, Hongce Zhang arxiv

IC3, also known as property-directed reachability (PDR), is a commonly-used algorithm for hardware safety model checking. It checks if a state transition system complies with a given safety property. IC3 either returns U…

Procedural Refinement by LLM-driven Algorithmic Debugging for ARC-AGI-2

2026-03-20 · Yu-Ning Qiu, Lin-Feng Zou, Jiong-Da Wang, Xue-Rong Yuan 외 arxiv

In high-complexity abstract reasoning, a system must infer a latent rule from a few examples or structured observations and apply it to unseen instances. LLMs can express such rules as programs, but ordinary conversation…

Efficiently Checking Actual Causality with SAT Solving

2019-04-30 · Amjad Ibrahim, Simon Rehwald, Alexander Pretschner

Recent formal approaches towards causality have made the concept ready for incorporation into the technical world. However, causality reasoning is computationally hard; and no general algorithmic approach exists that eff…

The Price of Progress: Price Performance and the Future of AI

2025-11-28 · Hans Gundlach, Jayson Lynch, Matthias Mertens, Neil Thompson arxiv

Language models have seen enormous progress on advanced benchmarks in recent years, but much of this progress has only been possible by using more costly models. Benchmarks may therefore present a warped picture of progr…