SpecRead: A Benchmark for Measuring Whether Language Models Understand Hardware Specifications
This work presents SpecRead, a benchmark that isolates specification comprehension from generation ability, and is automatically scorable by deterministic checks, with gray-zone cases counted wrong under the conservative main scoring.
Fei-Lian Huang
· 0 citations