Pith. sign in

Paper Citation Record · LEDGER

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback

As of 22 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 2 inbound Pith citation observations for arXiv:2504.15804.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.15804 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:22:02.415434Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:20:54.929448Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T13:21:01.812135Z

Reference resolution

40 of 40 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved33
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c58f6ad4-dbcf-4bb4-828f-3b4cec844548 · outbound

This paper cites Concrete Problems in AI Safety.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Concrete Problems in AI Safety

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.253536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.253536Z digest=sha256:722e0f7c10b5db3d263837530fe349fa5937e57cca48c24967a4413124cf32bb

Observation d4756dad-289d-4630-abb4-508c25fde563 · outbound

This paper cites Claude 3 sonnet.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Claude 3 sonnet

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:22:03.138511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T11:22:02.258883Z digest=sha256:ed7c213ad2c2636d471e9533fb16e06f7b9e95e9df72afd2305176463990c6fc

Observation e1fcaac3-dbdf-45cf-bf69-c52d3a7fc120 · outbound

This paper cites Qwen Technical Report.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Qwen Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.263484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.263484Z digest=sha256:5eb1e93f86e5fdf635124948d4e1a49bbc1df263b8ba399996372e7f4a0df24d

Observation a71ee65b-7249-4a81-9999-bba745e645c1 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.267999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.267999Z digest=sha256:d80ad8c28bc999021bf1caf189a6b0024a3568cbb2f132f2f489abd7be659710

Observation d5dfe428-ed4b-49bf-bc20-c46f5cf8c082 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Evaluating Large Language Models Trained on Code

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.272188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.272188Z digest=sha256:0590ceb6f01676c5fe23ddd8362c5beed3c7225c6776005e562165fa1bcedeb8

Observation d6ba3cc4-d250-4185-8ac9-fb223539d838 · outbound

This paper cites OriGen:Enhancing RTL Code Generation with Code-to-Code Augmentation and Self-Reflection.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback OriGen:Enhancing RTL Code Generation with Code-to-Code Augmentation and Self-Reflection

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.276316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.276316Z digest=sha256:647bfa30d5694b9d2c80a9558ec8307a6db0f6b248255bc96bfaa31936d9416e

Observation e71a4b5c-9854-4a26-ac25-edb133ed392e · outbound

This paper cites an unresolved cited work.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:22:03.125067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T11:22:02.281137Z digest=sha256:5de2d63a8d1285282f7482591689b3b1ce6b543dc4158dc59d4f5d20aaedf9db

Observation ed6db6bd-d72d-4928-aca6-b8e6967cb4b5 · outbound

This paper cites The Llama 3 Herd of Models.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback The Llama 3 Herd of Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.285108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.285108Z digest=sha256:27b0bee58fcc9829049c732799d946524e0d7ae4f861b127e059b9a045b393de

Observation fdee0e80-56a3-41eb-868f-1286e364e14f · outbound

This paper cites DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.289182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.289182Z digest=sha256:6fc01d6f06bccc1bead73612b7859fbd0656b28e326be337f695e439ee8ebc71

Observation e49969f1-011c-4315-b1dc-f71802ef4032 · outbound

This paper cites an unresolved cited work.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:22:03.112002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T11:22:02.293483Z digest=sha256:1635eb8c256191ad4890cc20ea78fb11ab5cc5d409f1b600b562e7b65d8e0a54

Observation e7e84c63-5439-4290-8ea7-40cae1ca4571 · outbound

This paper cites Qwen2.5-Coder Technical Report.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Qwen2.5-Coder Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.297511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.297511Z digest=sha256:e8276200f6078dbbfc3c4ee7eafd0d3fa388fc8dcad5059ff2547e9f4080acdf

Observation fb3e77a7-fc51-4dca-99c7-a777517ceadf · outbound

This paper cites GPT-4o System Card.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback GPT-4o System Card

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.301876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.301876Z digest=sha256:5944f96cdfaa507dafc12e979886c10ecd06c67aba808a879c74522964b3db95

Observation 230ee76f-8083-45fa-8371-069eb9d25967 · outbound

This paper cites an unresolved cited work.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.305956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.305956Z digest=sha256:43df40e7fb805828b8186357f6cc7109a8ec46389d188c6f2b64b6c7d54d3f33

Observation d62b8600-16dc-4bc6-8aa8-6cf7468b1827 · outbound

This paper cites Decomposed Prompting: A Modular Approach for Solving Complex Tasks.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Decomposed Prompting: A Modular Approach for Solving Complex Tasks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.309754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.309754Z digest=sha256:1f980f720b9e2021c6b50f6287eadc173469711cf004ec7601cf5e79508104ec

Observation 59ff72a6-f134-47cf-aff2-fb3b8d1eafff · outbound

This paper cites an unresolved cited work.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:22:03.089908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T11:22:02.313898Z digest=sha256:1be33cd0954fd0ba7f0c02d164dff8adefe4dd590bdea9902427f26957430308

Observation 325eef15-5183-4a9c-bb2d-cea0fd84d919 · outbound

This paper cites DrAttack: Prompt Decomposition and Reconstruction Makes Powerful LLM Jailbreakers.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback DrAttack: Prompt Decomposition and Reconstruction Makes Powerful LLM Jailbreakers

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.318012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.318012Z digest=sha256:0e1539538fed0f35494201fc3b49a2b6abb2d391e3e990e167389da109ee3057

Observation 33b07571-1e13-4752-9dd0-70cbe4c81bd5 · outbound

This paper cites DeepSeek-V3 Technical Report.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback DeepSeek-V3 Technical Report

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.322270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.322270Z digest=sha256:333ce5fff4152ca6eb2ff738b62801bba06c5b7143a2de9a867fb98388933444

Observation c9d82cc4-7d88-4f1e-8ba3-297f19bf0d48 · outbound

This paper cites an unresolved cited work.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:22:03.076631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T11:22:02.326855Z digest=sha256:b9cf459215f5357064504fc5dfa81515218912eebe6039eace2828b581a22e68

Observation 3498f55a-3730-486a-a7b9-e5d268fb12e0 · outbound

This paper cites an unresolved cited work.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:22:03.063158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T11:22:02.331279Z digest=sha256:8022e1b923faeb2dad204cb917555280871dfbd8f92ff81ae7fc137c4fdc0c69

Observation c266b73f-8f04-4648-82b9-a2e743b2b97d · outbound

This paper cites OpenLLM-RTL: Open Dataset and Benchmark for LLM-Aided Design RTL Generation.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback OpenLLM-RTL: Open Dataset and Benchmark for LLM-Aided Design RTL Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.335336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.335336Z digest=sha256:8bbfa29a9b2a98234d87a5eff21900ede20f088977a4e0fa323f7c491d0b707a

Observation eeadc240-2756-493e-837c-110c894e72ef · outbound

This paper cites Statistical Rejection Sampling Improves Preference Optimization.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Statistical Rejection Sampling Improves Preference Optimization

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.339249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.339249Z digest=sha256:6d99d963856637fc3ff910f34c9beba3a93d2d69575bf5179a93894d070dd17e

Observation 3b9e470f-9cd1-4a3f-bbf4-4cb1a72ebf58 · outbound

This paper cites an unresolved cited work.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:22:03.049875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T11:22:02.343293Z digest=sha256:854e78afd7b6e331db63a9a2239c841b9c5d0bad7561512d5d104c0365d41461

Observation cc753b71-31f0-446e-9b02-c06c2affe186 · outbound

This paper cites Nadimi, G.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Nadimi, G

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.347113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.347113Z digest=sha256:a3d7c3fa712abd3a29c3d2938edd293016a4b1e7514b96176990c5d341cdf6af

Observation 460238ed-3808-4954-8697-e455cdd33166 · outbound

This paper cites Papineni, S.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Papineni, S

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.350952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.350952Z digest=sha256:97c7b1c72a24050a68339c5d3e7cf7422989f783001223b3874e52bdde4576b1

Observation f29cebc1-d808-49e4-8ad4-3ab91bcc46f9 · outbound

This paper cites Pinckney, C.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Pinckney, C

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:22:03.027443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T11:22:02.354739Z digest=sha256:af235c131f38954cab7309cf6db5aa412f188b1c0a5246c59ef6ff8576be8a35

Observation 9b4fc2fe-3480-49ee-9d7e-5ea68a03e22f · outbound

This paper cites an unresolved cited work.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:22:03.014262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T11:22:02.358954Z digest=sha256:d96fa5a147b3f4185445ae6b7c8f5513cfb96e9d84de82eeca4bb10340992f43

Observation 8e1fd8ad-250d-4bd9-9182-8ba4b811d1d7 · outbound

This paper cites Rafailov, A.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Rafailov, A

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:22:03.000295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T11:22:02.363216Z digest=sha256:f93979a6284a4d92ab574beb8d78e318c35702b36db94cada8a173c08f80c3c9

Observation cdc91899-0ed5-4657-837b-c41cc67446b0 · outbound

This paper cites A Survey of Hallucination in Large Foundation Models.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback A Survey of Hallucination in Large Foundation Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.367007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.367007Z digest=sha256:7059fbb15412b9e8327c2dbc3ad335de4bbd164e0eccd6467e21b04e93fae1bb

Observation 92da724b-f9d9-4dee-9011-09cb9cb139f0 · outbound

This paper cites Code Llama: Open Foundation Models for Code.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Code Llama: Open Foundation Models for Code

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.370914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.370914Z digest=sha256:e88fef2878071aeeb7683b43de0ee595a860e7b0299e1ae4ff4b006824480770

Observation 70a3d8dc-ff93-4f07-8a09-68e0fba445ba · outbound

This paper cites Execution-based Code Generation using Deep Reinforcement Learning.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Execution-based Code Generation using Deep Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.374907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.374907Z digest=sha256:6f179fd47bf058b265f329804754d351184421315d7a2cea26ce46d2bccd3a85

Observation 136d9cd7-1522-4099-b7dd-4f996dfa38a6 · outbound

This paper cites Takamaeda-Yamazaki.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Takamaeda-Yamazaki

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:22:02.987133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T11:22:02.378974Z digest=sha256:50b8a7467acd10c057a9edf176caa24b23aa89f70ff031fb4701fff56d88bc61

Observation f0dfdebf-fdd7-4518-9ece-f380a690f8d1 · outbound

This paper cites Thakur, B.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Thakur, B

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:22:02.971907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T11:22:02.382783Z digest=sha256:81d9f270a4dfe1ced1906a5825bc8fac93ba246d0c82e69b9a2200c4b524cbf9

Observation 2565d436-a6bd-4be2-b903-a4817a4b8a59 · outbound

This paper cites Large Language Model for Verilog Generation with Code-Structure-Guided Reinforcement Learning.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Large Language Model for Verilog Generation with Code-Structure-Guided Reinforcement Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.386711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.386711Z digest=sha256:b9725afa3313de8d9b68e2d2e9b5f4e264d0fb27279dd6edde641b2e28b615b9

Observation 0454af7f-1994-4569-af2f-9f874e967bda · outbound

This paper cites an unresolved cited work.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:22:02.957940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T11:22:02.390684Z digest=sha256:0fb74c6218277c9c48136f92b628d6a6a27fc26dfdb95344ffee8774113b67b0

Observation 36229d6a-0948-4106-956e-f176b2d74a70 · outbound

This paper cites an unresolved cited work.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:22:02.944571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T11:22:02.394898Z digest=sha256:549ed155f27f5de2c684d3433f64f5c14cf1667657b166b6f17af59f65097380

Observation 87db4575-ac21-4e9f-836d-d2d8008a2065 · outbound

This paper cites Zehua, H.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Zehua, H

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:22:02.930830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T11:22:02.398794Z digest=sha256:314c816f9732b49c779485abd22d94ff3bc6bc9dcf57049236cf777aa2c2c652

Observation 9cc610f7-a633-4152-8530-f563083ae995 · outbound

This paper cites $\textbf{PLUM}$: Improving Code LMs with Execution-Guided On-Policy Preference Learning Driven By Synthetic Test Cases.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback $\textbf{PLUM}$: Improving Code LMs with Execution-Guided On-Policy Preference Learning Driven By Synthetic Test Cases

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.402508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.402508Z digest=sha256:ef6bebaa5f14743089a0427e78cc5618abe14bb6d00266c5205624a0256418ab

Observation 55925e03-fa47-4c72-9f3b-47261a545f79 · outbound

This paper cites Zhang, Y.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Zhang, Y

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:22:02.917039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T11:22:02.407199Z digest=sha256:430a90964ca741633e0ee7eb45385bc03ef7505bb3b48314f156ece2416dddaf

Observation b0bcd161-aed8-4d82-bb37-6f7587a6a80d · outbound

This paper cites CodeV: Empowering LLMs with HDL Generation through Multi-Level Summarization.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback CodeV: Empowering LLMs with HDL Generation through Multi-Level Summarization

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.411161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.411161Z digest=sha256:7858c32e0c13903d940eff519e7037bdb41f760e35bde295665258d1e311bda8

Observation 0a241b47-aaa7-4789-bbc8-2c9c6d117b35 · outbound

This paper cites Least-to-Most Prompting Enables Complex Reasoning in Large Language Models.

Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback Least-to-Most Prompting Enables Complex Reasoning in Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T11:22:02.415434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:22:02.415434Z digest=sha256:91bfa625008b0a2222d60a8e888c1674a6476a5d3b2e885f510f0aea33a3cfae

Pith citing papers

Observation 3ec5f2a1-0732-431b-80ce-7f6d64f9296d · inbound

iDSE: Navigating Design Space Exploration in High-Level Synthesis Using LLMs cites this paper.

iDSE: Navigating Design Space Exploration in High-Level Synthesis Using LLMs Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:21:01.949937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T13:20:54.929448Z digest=sha256:c049f9e30bd111d43982c11ba9d227ae2d6fcd05c9fe76c9eb617ebbb4ecd0ee

Observation ab754f08-a62c-47d1-9883-98313771a904 · inbound

A Progressive Approach to Synthesizable RTL Design Generation Using LLMs cites this paper.

A Progressive Approach to Synthesizable RTL Design Generation Using LLMs Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T15:14:11.555202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:14:11.555202Z digest=sha256:f271d3aef4218134efb6d426568d9ab117b3a90162e7aebc3ab482c2160df4a1