Benchmark Radar
AI BENCHMARK PROFILE

RTL-Bench

Industrial & EngineeringKnowledge & Reasoning

RTL-BenchLS evaluates LLMs on RTL design generation and reasoning, containing over 10,000 formally verified Verilog designs. Tasks include specification-to-RTL generation, round-trip reasoning, masked-content reasoning, and repository-issue reasoning. All tasks are verified via formal equivalence checking without manual testbenches.

Released
2026-06-08
Readiness
Paper only
Primary field
Industrial & Engineering

Why it matters

Existing RTL benchmarks are small and saturate with frontier models. RTL-BenchLS provides a large-scale, challenging benchmark with self-supervised tasks, enabling tracking of progress on complex hardware design reasoning and generation.

Motivation

LLM-based RTL generation and reasoning is a promising direction for hardware design automation.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.