TicTacBench: Benchmarking Timing Closure Capabilities of Coding Agents
Recent advances in large language models (LLMs) have led to the emergence of coding agents capable of performing complex engineering tasks, including register-transfer level (RTL) design and optimization. Existing RTL benchmarks mainly evaluate functional correctness and performance, power, and area (PPA) of the generated RTL designs, leaving agents' ability for \emph{timing closure}…
We haven't written up this one. arXiv cs.AI has the full story — the link below goes straight to it.