Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T19:59:24.526196Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 26 inbound Pith citation observations for arXiv:2506.14074.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T19:59:24.526196Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:08:22.502422Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T15:09:55.523834Z
26 of 26 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5a7838e3-f16a-4832-9b13-90fe09896501 · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Cursor: The ai code editor, 2025
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation becc77b3-2f9e-4970-9033-2b67cf9f6757 · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Large Language Model for Verilog Generation with Code-Structure-Guided Reinforcement Learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdc6561f-6e08-4e3c-ac6a-a7224fda8cc2 · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Craft RTL : High-quality synthetic data generation for verilog code models with correct-by-construction non-textual representations and targeted code repair
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4f091f46-d4e1-4384-98cd-aa0f2890d5a8 · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification VerilogEval : Evaluating large language models for Verilog code generation, 2023
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 59e82c01-446d-4df1-9992-4b23ef37607a · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification VerilogCoder: Autonomous Verilog Coding Agents with Graph-based Planning and Abstract Syntax Tree (AST)-based Waveform Tracing Tool
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9236c40c-e7bd-4c83-842d-0cefb2f1798c · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Rtllm: An open-source benchmark for design rtl generation with large language model, 2023
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea25378d-a37b-4a9a-98e1-80ba1852acef · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification OpenLLM-RTL: Open Dataset and Benchmark for LLM-Aided Design RTL Generation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7326f1d-2bd8-4b07-8af0-ca65a504e6e0 · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Revisiting verilogeval: A year of improvements in large-language models for hardware code generation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2148fa78-28b0-49ea-9d80-1ddf8b5fc72b · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Rtl-repo: A benchmark for evaluating llms on large-scale rtl design projects, 2024
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 91b7475d-605b-48fa-88ea-a52c599c26dd · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification LLM4DV: Using Large Language Models for Hardware Test Stimuli Generation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7c348e9-3467-4a16-977a-d8ca47d0aea8 · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification SWE -bench: Can language models resolve real-world github issues? In The Twelfth International Conference on Learning Representations, 2024
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa957e42-0ff2-4692-bc52-0e672e2e646a · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Copilot Evaluation Harness: Evaluating LLM-Guided Software Programming
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad66fe9e-216c-4ec2-ab35-6428c64f3bf2 · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Cocotb: Coroutine-based cosimulation testbench for vhdl and systemverilog, 2025
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ea19c1ab-1376-47dc-8b02-40a897753ab7 · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Bleu: a method for automatic evaluation of machine translation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b7dd254-d76b-4adc-ae35-3da22695ee7b · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Icarus verilog, 2025
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d4e3ab9a-785f-43d6-8b08-ea640f9ba7ba · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Yosys open synthesis suite, 2025
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 96b2e548-3ed9-4b6e-a3a9-3d1f3912dec3 · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Verilator: Open-source systemverilog simulator, 2025
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d8356655-b1d4-49b9-87ab-4155e3cd5c20 · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Xcelium logic simulator, 2025
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fe8cbc2b-7f8e-4bbe-bcae-ab30a9f10cbb · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Claude 3.7 sonnet, 2025
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 824be2fd-d8ef-4049-86c5-89e563adc852 · outbound
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 73c66bb6-35f6-4703-a187-c15f7211bc08 · outbound
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6553af7d-8cfd-4f66-808e-a4ae49435b65 · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Openai o4-mini, 2025 b
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3f5bf0cf-6cf5-41f7-84d6-d8f274f2d92f · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Meta llama 3.1 405b, 2024 a
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 914e6c69-9e97-4b42-b37d-40ea6b424762 · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Meta llama 3.1 70b, 2024 b
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b77c1f11-36b5-403b-8c70-3bc98e952442 · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Unsupervised k-means clustering algorithm
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8865192a-cfcb-4134-8fb2-33ccc13adb85 · outbound
Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification Understanding how dimension reduction tools work: An empirical approach to deciphering t-sne, umap, trimap, and pacmap for data visualization
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4fefa38d-41cd-4acc-9549-bf39075cadd6 · inbound
Revolution or Hype? Seeking the Limits of Large Models in Hardware Design Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4554590-ea58-42ce-b493-aeecd25816e8 · inbound
HWE-Bench: Benchmarking LLM Agents on Real-World Hardware Bug Repair Tasks Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6f2649cf-7615-4134-8db7-818d406a1eb7 · inbound
Dr. RTL: Autonomous Agentic RTL Optimization through Tool-Grounded Self-Improvement Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 05a66429-c6ca-49f4-a6b7-c14695a4b5dc · inbound
Spec2Cov: An Agentic Framework for Code Coverage Closure of Digital Hardware Designs Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3d8b9413-7a79-4d95-a73a-12b631e29fea · inbound
Spec2Cov: An Agentic Framework for Code Coverage Closure of Digital Hardware Designs Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e9be5ce3-29dd-49ba-a076-528c46bea872 · inbound
Understanding Inference-Time Token Allocation and Coverage Limits in Agentic Hardware Verification Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 66379d30-45c0-42e4-963b-2152a4175cc1 · inbound
ChipCraftBrain: Validation-First RTL Generation via Multi-Agent Orchestration Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c4b63f7b-e5d4-4c85-a536-402b82b2bc90 · inbound
SafeTune: Mitigating Data Poisoning in LLM Fine-Tuning for RTL Code Generation Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e755e713-d726-41ca-9ac0-5d652deb0f22 · inbound
RuC: HDL-Agnostic Rule Completion Benchmark Generation Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 70ad44aa-d7cd-4b57-b824-2abbb68a642a · inbound
ChipMATE: Multi-Agent Training via Reinforcement Learning for Enhanced RTL Generation Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 98ed2132-5527-46c7-aa76-26bd1d1860bf · inbound
RTL-BenchMT: Dynamic Maintenance of RTL Generation Benchmark Through Agent-Assisted Analysis and Revision Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1eca5840-96c5-4243-80c5-a90980d25c3c · inbound
AssertLLM2: A Comprehensive LLM Benchmark for Assertion Generation from Design Specifications Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6731d922-2cd0-4f9f-a721-5779b433838a · inbound
CASS-RTL: Correctness-Aware Subspace Steering for RTL Generation with LLMs Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ad382723-37ab-41e8-a3aa-e1d3cccd64ac · inbound
RTL-BenchLS: A Large-Scale Benchmark for RTL Reasoning and Generation with Large Language Models Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2d10777a-2324-4932-a571-b82ca9fbc2ac · inbound
Structured Testbench Generation for LLM-Driven HDL Design and Verification-Oriented Data Curation Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7449a3fe-d2c1-4234-9a82-afca4826b75a · inbound
Interpretable and Verifiable Hardware Generation with LLM-Driven Stepwise Refinement Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 222552d4-52ea-4ec0-abce-269bc2e165ce · inbound
Interpretable and Verifiable Hardware Generation with LLM-Driven Stepwise Refinement Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90f9680b-adf9-41a8-8e44-a8343be03cc5 · inbound
CHIA: An open-source framework for principled, agentic AI-driven hardware/software co-design research Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d767d608-007f-4b11-897a-32f4e6770163 · inbound
CHIA: An open-source framework for principled, agentic AI-driven hardware/software co-design research Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 71f95b0d-fe5c-455d-bc95-6b9a60d62bd0 · inbound
CHIA: An open-source framework for principled, agentic AI-driven hardware/software co-design research Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08b2b0cb-1415-4b92-9a54-412c598df5aa · inbound
Agentic Hardware Design as Repository-Level Code Evolution Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c54ee891-db8a-4e77-9d1d-edc26fd2d2d5 · inbound
ChipVerilog: A Large-Scale OpenCores-Derived Benchmark for LLM-Based Verilog RTL Generation Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c960fe16-40d7-4547-8447-a1ef52a6030a · inbound
When LLMs Over-Answer: Measuring and Mitigating Quality Issues in LLM-Based Hardware Description Language Question Answering Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c791175-4f57-49b6-a500-c8c73ff05263 · inbound
Can AI Agents Really Complete RTL-to-GDS? Lessons from Benchmarking Tool-Interactive EDA Workflows Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60de5fe6-2dfe-447c-b64e-78344d342b27 · inbound
FinHardBench: Can LLMs Generate Latency-Aware Hardware for Financial Computing? Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c31426c2-2a29-4db9-a2d4-bdfdb26e00af · inbound
GateTruth: Auditing the Rigor of RTL Design Benchmarks via Mutation Testing Comprehensive Verilog Design Problems: A Next-Generation Benchmark Dataset for Evaluating Large Language Models and Agents on RTL Design and Verification
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.