Pith. sign in

Paper Citation Record · LEDGER

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny

As of 10 August 2026, this Paper Citation Record lists 100 of 101 outbound references and 6 inbound Pith citation observations for arXiv:2507.16331.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.16331 v4

Coverage vector

measured 100 of 101 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:20:17.762309Z

measured 106 of 106 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T23:52:36.891080Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

100 of 101 outbound references displayed

  • verified exact3
  • verified fuzzy19
  • unresolved76
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 209662fa-e112-499d-9859-5b1b9b67a6a6 · outbound

This paper cites write newline.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:10.030154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:10.030154Z digest=sha256:a16fb801d22bc6b5be4452817a8a7752f2ca05be0a93cb2ec0785ba8c9d5ba0a

Observation 2c44f09e-db6c-45a8-811d-c2381f389bfd · outbound

This paper cites GPT-4 Technical Report.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny GPT-4 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:10.126663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:10.126663Z digest=sha256:44fb226e2d2884a7534056bdfc0f902cbf6c830f9a3bda61be8fe46a80c5ecd0

Observation 02566cd0-1fa5-4b59-82c4-35ebb3cdd1b0 · outbound

This paper cites Alphacode 2 technical report.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Alphacode 2 technical report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:10.230005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:10.230005Z digest=sha256:652d73b88b3d79bf6f7549dd8b3049778bd06ceb6b09a5c17095c648a3dafe18

Observation f9a21c30-57a9-43b6-8cad-fc0c63416986 · outbound

This paper cites System card: Claude opus 4 & claude sonnet 4.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny System card: Claude opus 4 & claude sonnet 4

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:10.313716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:10.313716Z digest=sha256:3bd65e233f872ea69d2d596bd62bc5225c0624f127fa14afc54ee6b731a19295

Observation 795fb0d1-7bd8-4456-a8c7-a4c6cd8553db · outbound

This paper cites Program Synthesis with Large Language Models.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Program Synthesis with Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:10.412132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:10.412132Z digest=sha256:15e6b5a7460595f194bdc948c327d42e27db9c507a564414f62abaeb0d31f095

Observation e0412022-7c59-4522-98d8-4a53737f451f · outbound

This paper cites Y., Collignon, N., Neo, C., Lee, I., Paren, A., Bibi, A., Trager, R., Fornasiere, D., Yan, J., Elazar, Y., and Bengio, Y.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Y., Collignon, N., Neo, C., Lee, I., Paren, A., Bibi, A., Trager, R., Fornasiere, D., Yan, J., Elazar, Y., and Bengio, Y

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:10.494851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:10.494851Z digest=sha256:e3e0ed4c87e7d535571691a2c0fa610983ba38f341d54851a5f65e829811c683

Observation ac5644be-2809-4111-9edf-db13e3301b34 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Evaluating Large Language Models Trained on Code

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:10.590600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:10.590600Z digest=sha256:494cbaba82c1382185cb31a6c3cc29fde6e0c3677c6f7567a7b166b7d0aa2a05

Observation 60a789b0-35ac-4a97-b03a-25a13e8eb4b8 · outbound

This paper cites Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:10.705035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:10.705035Z digest=sha256:f774bb4b8755ac9e3f529a5976e881477e12330b832b203b45574489d1a2d92f

Observation 5795bd41-e36f-4944-a6bc-19e2b558e834 · outbound

This paper cites Reasoning Models Don't Always Say What They Think.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Reasoning Models Don't Always Say What They Think

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:10.865469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:10.865469Z digest=sha256:cc31935c0018ba7e27384b594d9211ab61c243d9c5756a185fa88335022accca

Observation 90753e9c-6ab2-4c7c-a713-2cbb39979498 · outbound

This paper cites A., Nielsen-Garcia, C., Mir, S., Li, S., Orender, J., et al.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny A., Nielsen-Garcia, C., Mir, S., Li, S., Orender, J., et al

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:10.997966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:10.997966Z digest=sha256:bc72eee8c54d738d2d8cc81fc8947fe7e4cc593240f410152ddebb57b988f3d0

Observation 25498567-1427-4997-97e9-f22990e8e5a0 · outbound

This paper cites On the Measure of Intelligence.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny On the Measure of Intelligence

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:11.175293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:11.175293Z digest=sha256:aabe95427529742f2006e5de0916b0c041f42115f98789b3af66ebc30f8d2dfd

Observation 26e30546-1922-4ffc-8ca8-ef287dabe3bf · outbound

This paper cites V., Levine, S., and Ma, Y.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny V., Levine, S., and Ma, Y

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:11.307777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:11.307777Z digest=sha256:bcf04c435513e34aabc75d6846cb91bc58240b20198fefb2d297ee8775ab75f1

Observation d1e27b7a-75f7-4d4e-b5d7-77d471bc4b2b · outbound

This paper cites an unresolved cited work.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:11.457296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:11.457296Z digest=sha256:e6f7c54fc8f5fe7db1f3a931b1366ec06137272a05d475167fc59986fb803858

Observation 2be1b238-97bc-4567-b15c-b40184d03117 · outbound

This paper cites Towards formal verification of llm-generated code from natural language prompts, 2025.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Towards formal verification of llm-generated code from natural language prompts, 2025

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:11.623988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:11.623988Z digest=sha256:326016a6b92edaf38e394ed54aacf15dee04838c4a4601418b0b7cd6349f669e

Observation 4991e340-828a-47c7-89e5-c1ea9f56ce2c · outbound

This paper cites Towards Guaranteed Safe AI: A Framework for Ensuring Robust and Reliable AI Systems.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Towards Guaranteed Safe AI: A Framework for Ensuring Robust and Reliable AI Systems

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:11.745072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:11.745072Z digest=sha256:4d35565e5d730e7597a246c658e97e0f9bcbcb2e61a1ed96b232cd6385fb78b4

Observation e6b136cd-1e6d-4670-a8ac-6019bbc6c03e · outbound

This paper cites and Bj rner, N.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny and Bj rner, N

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:11.837567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:11.837567Z digest=sha256:d59b240ed738773d624b06ef9ca0cd10fb5423cc5dde29958c989583ea84663b

Observation 3ec856cf-374a-452e-a71e-5e62e3ef5587 · outbound

This paper cites The lean theorem prover (system description).

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny The lean theorem prover (system description)

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:11.938086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:11.938086Z digest=sha256:72a913fc7b284b8a53938101c276e68697bb3faae83b34a237954e4c1575be22

Observation d6411b08-26cb-4ce9-abd7-6283c52494b4 · outbound

This paper cites DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:11.998704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:11.998704Z digest=sha256:b21a803b2389684af64f74186ea00ca1a0329fcc78a2b2548cde98734ee90141

Observation b52ef171-fe8c-4273-8e73-1971028cba1b · outbound

This paper cites an unresolved cited work.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Unresolved cited work

Reference 19

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T15:20:19.326180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:12.089275Z digest=sha256:a2cad31a5a5bbaae5e1f3b0c31c04f650a3adc3b1d687b62f34e4fabc70b7e95

Observation 01c4464f-5ba6-4ec1-8495-06d50c2f1192 · outbound

This paper cites Generalization or memorization: Data contamination and trustworthy evaluation for large language models.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Generalization or memorization: Data contamination and trustworthy evaluation for large language models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:12.171055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:12.171055Z digest=sha256:5f3ebea67ccfa38340696432432c7482fc63203c1d8df3f33dc0848177a38d80

Observation 4fd3d24a-fa49-4e37-9c26-61b12410f04b · outbound

This paper cites and Mehta, R.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny and Mehta, R

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:12.270776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:12.270776Z digest=sha256:98575d7a42a4da95d64ff1a0614f77789f3f3e7b61c4fa5acb3795e4f1a69f20

Observation cec73c0b-d6f1-418d-9a79-2494f8f60d25 · outbound

This paper cites Gemini 2.5: Pushing the frontier with advanced reasoning, multimodality, long context, and next generation agentic capabilities.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Gemini 2.5: Pushing the frontier with advanced reasoning, multimodality, long context, and next generation agentic capabilities

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:12.349174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:12.349174Z digest=sha256:fc2e5199bc0bbf7bc25eb2794de03c5f47a329953eb4e82fd8b2a678f51c5cc3

Observation ea85e27b-028d-485e-9367-e23dd18282df · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:12.410202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:12.410202Z digest=sha256:a840aff41a8e95adcad9d22368637d39ecd056b6fb78c34b5e25b9500b6cc58e

Observation 64a92f93-6be1-4ee8-869d-ae8a5868953f · outbound

This paper cites Measuring and improving semantic diversity of dialogue generation.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Measuring and improving semantic diversity of dialogue generation

Reference 24

Resolution
verified exact
doi, observed 2026-08-06T15:20:17.829554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:12.504634Z digest=sha256:886e978bc4ae6206da2f10d9f1b5880e913fc9e83e1ef5016c558d9528d1657f

Observation 6b15fe36-310d-49e9-b19c-5aa8163d853c · outbound

This paper cites DynaCode: A Dynamic Complexity-Aware Code Benchmark for Evaluating Large Language Models in Code Generation.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny DynaCode: A Dynamic Complexity-Aware Code Benchmark for Evaluating Large Language Models in Code Generation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:12.595812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:12.595812Z digest=sha256:044ffc4d5ede0051059f17619f5d6998fe90c277954d9b30c69868a39fb5b6c3

Observation b2fdcece-c286-43c5-ae3f-083c6e84d34e · outbound

This paper cites Does math reasoning improve general llm capabilities? understanding transferability of llm reasoning, 2025.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Does math reasoning improve general llm capabilities? understanding transferability of llm reasoning, 2025

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:12.662973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:12.662973Z digest=sha256:e1bcf223538059dbfcfb93eb80c253e4f6e3181ea5f80eeba06d229e63870a13

Observation 30ba6c9b-bfd9-4a36-ad2f-bae3d0d4eb87 · outbound

This paper cites Qwen2.5-Coder Technical Report.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Qwen2.5-Coder Technical Report

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:12.732299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:12.732299Z digest=sha256:8a6aa4940dd228ddb22f7ccd943a2f2603941818cbfa2209243dbefbf15d913d

Observation 05843a20-46b4-47a9-95ab-6461b612471b · outbound

This paper cites Thinking beyond the anthropomorphic paradigm benefits LLM research.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Thinking beyond the anthropomorphic paradigm benefits LLM research

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:12.807230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:12.807230Z digest=sha256:9e697ad12677f4bdd148967b3ee347307374023af085845e3cb799fc58890438

Observation 53afef7e-0c5c-4a08-86ee-cc062f8b3c81 · outbound

This paper cites AI safety via debate, May 2018.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny AI safety via debate, May 2018

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:12.897446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:12.897446Z digest=sha256:17e4c609023bc2dfddd4ba746112a4eb15b9f361cceae2c5ad806b2c9655347b

Observation b7325784-33cd-4604-be45-4593b05e499b · outbound

This paper cites Verifast: A powerful, sound, predictable, fast verifier for c and java.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Verifast: A powerful, sound, predictable, fast verifier for c and java

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:12.967941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:12.967941Z digest=sha256:dde35481e2240cf316dab0cccfb2182ee77eefde3696e89bc0bdfbc942284f3c

Observation a749aac7-cb21-403c-88e8-a6161cebdcde · outbound

This paper cites Do we need to verify step by step? rethinking process supervision from a theoretical perspective, February 2025.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Do we need to verify step by step? rethinking process supervision from a theoretical perspective, February 2025

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:13.057170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:13.057170Z digest=sha256:ac6134173f20a04c819ebc849fc13140a34173c1ebc24eabef981c9e125b62cb

Observation 48dfe5ee-5ce7-48f6-818c-5b95c39e1955 · outbound

This paper cites Can large language models understand intermediate representations in compilers?, February 2025.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Can large language models understand intermediate representations in compilers?, February 2025

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:13.131215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:13.131215Z digest=sha256:9285058694d22eff8d6fca76a8d2f9f76b7eb15eecaffd6b4a06b82d71f17beb

Observation aba1d025-7b60-4c73-933a-57588e033ef2 · outbound

This paper cites sel4: Formal verification of an os kernel.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny sel4: Formal verification of an os kernel

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:13.225730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:13.225730Z digest=sha256:bcd10be1f49b3617c8b5ca1d4f2f5e36df0d6b1b18a3ae12c4ea505f7808c4bb

Observation 8a58bc26-c624-463f-be6e-c5bc3fe1f6ab · outbound

This paper cites Chain of thought monitorability: A new and fragile opportunity for AI safety, July 2025.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Chain of thought monitorability: A new and fragile opportunity for AI safety, July 2025

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:13.291719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:13.291719Z digest=sha256:3b4e6212e210b7911ac4813504eb4d8430a3210c21006b38e9b5394f9b319db8

Observation 99a75dab-f5f2-430e-973f-fbec1f26a294 · outbound

This paper cites Gradual disempowerment: Systemic existential risks from incremental AI development, January 2025.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Gradual disempowerment: Systemic existential risks from incremental AI development, January 2025

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:13.384938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:13.384938Z digest=sha256:b3af4408bfaae299d21fba9d4e4cd4a621330b53b49a0ca0f49ab6da681c2ee7

Observation dc700f65-d7bd-49ec-895f-0a9c1db2e5c5 · outbound

This paper cites LLM Post-Training: A Deep Dive into Reasoning Large Language Models.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny LLM Post-Training: A Deep Dive into Reasoning Large Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:13.473496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:13.473496Z digest=sha256:849b4428a2d1ca4e6d2032445d92617045d4e3cc918d495b8b9df8ace7fc224a

Observation ffd8c634-6d62-4e5f-aa76-105e55296bda · outbound

This paper cites Measuring Faithfulness in Chain-of-Thought Reasoning.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Measuring Faithfulness in Chain-of-Thought Reasoning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:13.565564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:13.565564Z digest=sha256:a0f763d1898ae45b1eec41397f2ef74e03708f47d52b5c2f537fb20fa0e2eb45

Observation e0b3a6a7-e2f6-475e-8f4c-d92da744b0b1 · outbound

This paper cites CodeRL: Mastering Code Generation through Pretrained Models and Deep Reinforcement Learning.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny CodeRL: Mastering Code Generation through Pretrained Models and Deep Reinforcement Learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:13.692122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:13.692122Z digest=sha256:f8f05f250392ffd2b9b8e8f6e982f13e6eb719c8672cc64ddad43440434f5e66

Observation 574e74cd-6a1e-41ad-b3c1-a22270773492 · outbound

This paper cites How Well do LLMs Compress Their Own Chain-of-Thought? A Token Complexity Approach.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny How Well do LLMs Compress Their Own Chain-of-Thought? A Token Complexity Approach

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:13.782725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:13.782725Z digest=sha256:ca6373e0b9b95e94baa1599afcab392b52c991c5457da9f130ca7f09842426fb

Observation 9683c76b-3eaf-4dfa-8559-71a99d5006ac · outbound

This paper cites an unresolved cited work.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:13.858775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:13.858775Z digest=sha256:205f9610c5aba6a5506945ad0aa7eb2c9f6a63e1679268c1947ec36b5322d016

Observation c368a4e0-0e10-4287-af15-2a3acd3ddee6 · outbound

This paper cites CodeI/O: Condensing Reasoning Patterns via Code Input-Output Prediction.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny CodeI/O: Condensing Reasoning Patterns via Code Input-Output Prediction

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:13.911084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:13.911084Z digest=sha256:a56ff531ba6b0248c9f164f83ad80babeca8ec23229fd1b071283f49257fd65e

Observation 87b53095-17e4-4de2-bed9-d9440d4b3123 · outbound

This paper cites AutoTriton: Automatic Triton Programming with Reinforcement Learning in LLMs.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny AutoTriton: Automatic Triton Programming with Reinforcement Learning in LLMs

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:14.003355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:14.003355Z digest=sha256:1ee478ea0821deaf7c232d42914f456e5656ee80afe61f28df6b6a74cd5485b5

Observation daf5c64a-e53e-4567-a80d-e07a591b51a5 · outbound

This paper cites Combining Induction and Transduction for Abstract Reasoning.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Combining Induction and Transduction for Abstract Reasoning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:14.124136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:14.124136Z digest=sha256:c6f9481c51ed08e5810018c81851aaa0bd2f8a9ec9e11773f1c97f1666e5dc64

Observation 6f838c0b-5ad9-4e72-8eac-b49769b0d00e · outbound

This paper cites Competition-level code generation with alphacode.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Competition-level code generation with alphacode

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:14.214324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:14.214324Z digest=sha256:00c1de6124297c3e7ef13731df7f790d126b633e653b039b003df3a10f5d2a4c

Observation f062fcf9-6ca4-48f0-9a4f-5a4f8255694b · outbound

This paper cites Dafny as Verification-Aware Intermediate Language for Code Generation.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Dafny as Verification-Aware Intermediate Language for Code Generation

Reference 45

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T15:20:18.998311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:14.285945Z digest=sha256:253b353727d75e0160b638985547ee9105347911fcefab4b91861e7bc7b32ad6

Observation e600c9f4-9753-41e5-87cb-0070d7103c13 · outbound

This paper cites Towards Solving More Challenging IMO Problems via Decoupled Reasoning and Proving.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Towards Solving More Challenging IMO Problems via Decoupled Reasoning and Proving

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:14.369135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:14.369135Z digest=sha256:0e258e44b9cf5814f2b60fd9a8509ccddd97e7baaf28d5e09af9355c80924ed6

Observation 2c89db8b-9dd5-4694-8a5d-0d049739b0ad · outbound

This paper cites Goedel-prover-v2: The strongest open-source theorem prover to date, 2025.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Goedel-prover-v2: The strongest open-source theorem prover to date, 2025

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:20.022233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:14.440664Z digest=sha256:cd77e6a741695e73623528bb5e0220b25592724095e8960496ea9ec9141247f5

Observation 44aa9e5a-8ce3-4ece-b7bd-f98cf55671ea · outbound

This paper cites Safe: Enhancing Mathematical Reasoning in Large Language Models via Retrospective Step-aware Formal Verification.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Safe: Enhancing Mathematical Reasoning in Large Language Models via Retrospective Step-aware Formal Verification

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:20:18.953748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:14.514001Z digest=sha256:e69d35569886932669c416e8e9804b71208e026bd68cc5dcd3859b6e48d7b22e

Observation 1cd65a90-dbba-4744-bbdb-cdc1931fba4c · outbound

This paper cites ProRL: Prolonged Reinforcement Learning Expands Reasoning Boundaries in Large Language Models.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny ProRL: Prolonged Reinforcement Learning Expands Reasoning Boundaries in Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:14.585900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:14.585900Z digest=sha256:06fcc0e96a3f926a14393dbcdec93a6ebc8ce8b1afa489b19901bf29a7ad63e0

Observation 6ede6bfb-c111-489e-b696-e7aef64b5e20 · outbound

This paper cites DafnyBench: A Benchmark for Formal Software Verification.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny DafnyBench: A Benchmark for Formal Software Verification

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:14.678411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:14.678411Z digest=sha256:e28a4392b3f3f10626b1b004735ca951a5e16c082432304503962f7ce0053025

Observation 89ced50b-bbb4-4dcf-a20a-0d56e00a13c6 · outbound

This paper cites StarCoder 2 and The Stack v2: The Next Generation.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny StarCoder 2 and The Stack v2: The Next Generation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:14.747793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:14.747793Z digest=sha256:dcb277ea3546e1c9029ead57f186382d891f64338ec7053cbcbe31f63b841385

Observation 5c16fc10-6966-4293-b7ba-25e84c98c600 · outbound

This paper cites Reasoning models can be effective without thinking, April 2025.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Reasoning models can be effective without thinking, April 2025

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:20.007260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:14.839523Z digest=sha256:39296ba7ad93679ac6fb9819e4b4c6b33be9f2811a879507eb52da1f5f93b2a8

Observation b920a734-a6e2-4c40-8dfb-4b2310026e8f · outbound

This paper cites Potemkin Understanding in Large Language Models.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Potemkin Understanding in Large Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:14.933677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:14.933677Z digest=sha256:0fed3fb07adaeeff809dcee60e450bf201c179f2d8d97378ef37cb2a2e3634c4

Observation 6f496522-0762-4ba9-8b11-eb318a0481da · outbound

This paper cites an unresolved cited work.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:20:19.991659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:15.027056Z digest=sha256:d67ae5fa2b8c83159e2d0da756cb04bce07eec70e362e6b98536aaf3a150f395

Observation ec17f54a-88df-4573-86ee-4ea54b1fe2e5 · outbound

This paper cites an unresolved cited work.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:20:19.977094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:15.101001Z digest=sha256:23b971fb5a0f3f3eac07d2e731df71e91025252e6b43987d79f6b93fc3284ef5

Observation e03c38af-a9f6-442a-aa21-5a3b8c9380c6 · outbound

This paper cites AlphaEvolve: A coding agent for scientific and algorithmic discovery.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny AlphaEvolve: A coding agent for scientific and algorithmic discovery

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:15.195563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:15.195563Z digest=sha256:d74d1f4aa5d25dc98f1ab637688cd2b522e5114f067ad86dfc6ac22cdb41c24f

Observation 5279d530-2ec7-4e4b-ba4d-976e1581188c · outbound

This paper cites Training language models to follow instructions with human feedback.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Training language models to follow instructions with human feedback

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:15.273173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:15.273173Z digest=sha256:9003d7d3da35ff244161da3a8683bc8736d12a927fff7deae01d81f0bb9303cd

Observation 8bd3c511-51d1-44d7-a3b0-746a80f39589 · outbound

This paper cites How to Get Your LLM to Generate Challenging Problems for Evaluation.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny How to Get Your LLM to Generate Challenging Problems for Evaluation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:15.359475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:15.359475Z digest=sha256:da9bdf908df308c5d1ffdb69f2665fc7822f4d44881920096f6456bbcf16c83b

Observation ba17a31f-ffb2-494e-9825-eb5d5ba96334 · outbound

This paper cites How does code pretraining affect language model task performance? Transactions on Machine Learning Research, 2025, 2025.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny How does code pretraining affect language model task performance? Transactions on Machine Learning Research, 2025, 2025

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:19.946878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:15.429571Z digest=sha256:5cac51ef526288ede4ec0323dae511a17752a84b8884c2771c18c4ada51c5b07

Observation 9e12d4bf-8609-4b5f-b343-b3ba71f9a748 · outbound

This paper cites dafny-annotator: AI-Assisted Verification of Dafny Programs.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny dafny-annotator: AI-Assisted Verification of Dafny Programs

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:15.513785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:15.513785Z digest=sha256:427233c34ea0173cfd0c827adcc7d671d04064216e7f15c27339cf0c5f482c19

Observation 86e6799f-2ae5-4ef2-87ce-33b062d9a77d · outbound

This paper cites Qodo-Embed-1: State-of-the-Art Code Embedding Models.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Qodo-Embed-1: State-of-the-Art Code Embedding Models

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:19.928972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:15.587429Z digest=sha256:fd24cfc2040ee5de9c0467192d27bd1a6160ce0f4ef2ebc37708cdd06969ee53

Observation e721aa21-6722-47e4-b27f-dc6af21dffa1 · outbound

This paper cites What is ansible?, 2025.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny What is ansible?, 2025

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:19.912819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:15.676924Z digest=sha256:2bf3e628194a7aba63f859da59b2dbe3e0c8f9a03c997c47fd5f1ad0f94f0426

Observation cf835e8b-db0b-444a-b7a6-b6874a99dd82 · outbound

This paper cites Evaluating the ability of gpt-4o to generate verifiable specifications in verifast.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Evaluating the ability of gpt-4o to generate verifiable specifications in verifast

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:19.894470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:15.752691Z digest=sha256:c926500b6455b01def391fcb6bbef1e6cd46b3e626865527a6b136d83d01880e

Observation c8a69c7a-4f56-40da-a2a7-54fe9c1d5311 · outbound

This paper cites Quantifying contamination in evaluating code generation capabilities of language models.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Quantifying contamination in evaluating code generation capabilities of language models

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:19.878126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:15.842211Z digest=sha256:b393ee07294331131aac0d1af9aee7cca1f8157ceb288c2876a8304edb7ca187

Observation 2c5f3e9e-cc6b-42fd-b790-d480cd83c094 · outbound

This paper cites R., Gnaneshwar, D., Locatelli, A., Kirk, R., Rockt \"a schel, T., Grefenstette, E., and Bartolo, M.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny R., Gnaneshwar, D., Locatelli, A., Kirk, R., Rockt \"a schel, T., Grefenstette, E., and Bartolo, M

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:19.861737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:15.858731Z digest=sha256:d27a488ba709ddaf6dea6bde5f80b9cd59c86522e4e66d962390d2e7709f8a3a

Observation 2258a791-bb9d-42a0-b019-fa488b484dcc · outbound

This paper cites Boundless Socratic Learning with Language Games.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Boundless Socratic Learning with Language Games

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:15.942232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:15.942232Z digest=sha256:efdc50d60af283f76ea1ced7e0825c13eef542623875efce37e929fa4abaa921

Observation b77dad33-2e0f-4bfe-a08c-b004e621c283 · outbound

This paper cites Autoregressive Large Language Models are Computationally Universal.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Autoregressive Large Language Models are Computationally Universal

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:16.110693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:16.110693Z digest=sha256:747fa6d9fa790eea583e65a1496f60e35c9d781e1884df3b82dab046a4fa524a

Observation c539490b-f77b-45b2-8265-0ddd3ebb8665 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:16.225181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:16.225181Z digest=sha256:ab4e412570225ff02c5ea90adfad261523b7f25a1ffc29f0ad2cbb88b798b489

Observation b4e2337f-c5da-4e47-9406-159cce9bb7c6 · outbound

This paper cites The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:16.310347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:16.310347Z digest=sha256:5989229c9ded298e1b8a6c369be9ba20c73d26a857680841bd369f727500c52a

Observation d73bd658-e949-4a95-955f-bfb56d68dad2 · outbound

This paper cites and Sutton, R.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny and Sutton, R

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:19.844506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:16.396651Z digest=sha256:1319590d220f6bd85d0cc46fbd3eb1ee0c600a4a3ac887a1080690b293f66b35

Observation efb2e494-847d-4463-b65a-92102c56969a · outbound

This paper cites an unresolved cited work.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:20:19.825059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:16.472998Z digest=sha256:296729eb97fe2f324d71863e34951afdf19b166b38f3d466cef3abc8c4e693b3

Observation fe984bbe-0df2-48a7-8a76-f304ee47b47e · outbound

This paper cites Beyond semantics: The unreasonable effectiveness of reasonless intermediate tokens, May 2025.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Beyond semantics: The unreasonable effectiveness of reasonless intermediate tokens, May 2025

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:19.805865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:16.595792Z digest=sha256:8f5aa289f139033377e4ad65e624d05b3061f4615d3fb27f8ecd9539658e0a6a

Observation 594c5252-7b8d-4089-9d1a-b3b203337122 · outbound

This paper cites Clover: Closed-loop verifiable code generation.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Clover: Closed-loop verifiable code generation

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:19.790933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:16.778032Z digest=sha256:196a18c61c8f87356651c8dd8b7100e2979ef3d7b6213b93ea3f8788a92cfb9c

Observation 32cb581c-3213-4fb9-b9df-af707d223f04 · outbound

This paper cites OMEGA: Can LLMs Reason Outside the Box in Math? Evaluating Exploratory, Compositional, and Transformative Generalization.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny OMEGA: Can LLMs Reason Outside the Box in Math? Evaluating Exploratory, Compositional, and Transformative Generalization

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:16.918871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:16.918871Z digest=sha256:0fcfe2a6d930efc2346b12fe724a73fbc3bf79ce551764031c8d5d9519e0dc0d

Observation 710d8701-93ce-4763-9efa-36aad7749b72 · outbound

This paper cites The bitter lesson.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny The bitter lesson

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:19.773635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:17.053344Z digest=sha256:26d827e71bf51030d75bdd067bdc110af21e12d31bff9c3765955a1a1e5a19c2

Observation b1e6a56f-affc-46cd-b542-2f20a7c1d8e7 · outbound

This paper cites S., Barto, A.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny S., Barto, A

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:19.755062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:17.219865Z digest=sha256:cd896be3eaa8e92f871e7b06103657b617d64c073bebb0553ebf640c3c51ad7b

Observation 644e4e30-d1c3-4bc3-a2dd-69b31d6bab0b · outbound

This paper cites S., McAllester, D., Singh, S., and Mansour, Y.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny S., McAllester, D., Singh, S., and Mansour, Y

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:17.337992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:17.337992Z digest=sha256:5846f66cd93b9cdda39d6796685ba6c7f188c2fc0b1b3e62fff466634899c468

Observation 2bb15353-3e51-4d5a-9a2f-abe251a4a5e0 · outbound

This paper cites K., Fu, S., and Sundaresan, N.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny K., Fu, S., and Sundaresan, N

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:19.722750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:17.451410Z digest=sha256:bc215362f16f8b421536aeeba53401f251c33735a97d6ae0513f7877480dd206

Observation d07c9b94-f6be-4bb6-b3da-07a2646d8ea1 · outbound

This paper cites A promising path towards autoformalization and general artificial intelligence.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny A promising path towards autoformalization and general artificial intelligence

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:19.705505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:17.562609Z digest=sha256:39d1bc6dbd019c5045644b6c1999dd4e4a6c75f107bd1f931db766d02c0d3d7a

Observation d086b80a-4169-474c-828d-c623390bb884 · outbound

This paper cites Worldcoder, a model-based llm agent: Building world models by writing code and interacting with the environment.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Worldcoder, a model-based llm agent: Building world models by writing code and interacting with the environment

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:19.686216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:17.616061Z digest=sha256:31e411cee4665437473dde81603e1b3f8d3a118ffddfff948faa48ef7a926bbe

Observation 503ceaa8-cfa8-4190-ba38-ea235016521b · outbound

This paper cites Clever: A curated benchmark for formally verified code generation.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Clever: A curated benchmark for formally verified code generation

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:17.647324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:17.647324Z digest=sha256:80adb3971ffd2fcbb12fec70a25a6f7b6c67a307763ccd4acdfe3f2e8b3b1bc6

Observation 55fda506-3489-4624-acac-f48830017611 · outbound

This paper cites an unresolved cited work.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Unresolved cited work

Reference 82

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:20:19.670359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:17.652416Z digest=sha256:809590e82159158ea1a1996b668f9533488fcefac68cde4fe96e206f0d430ddc

Observation 6f133273-40af-4c26-9b6c-a7c4f96aaf22 · outbound

This paper cites DICE: Detecting In-distribution Contamination in LLM's Fine-tuning Phase for Math Reasoning.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny DICE: Detecting In-distribution Contamination in LLM's Fine-tuning Phase for Math Reasoning

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:17.657163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:17.657163Z digest=sha256:d3f7d321c317021e3bce28d073749eae97893149298c304d1e97126a48cac02c

Observation 135ff164-bace-489b-8e95-01e296fc3110 · outbound

This paper cites Rethinking the Illusion of Thinking.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Rethinking the Illusion of Thinking

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:17.663386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:17.663386Z digest=sha256:a7220fed26199f502a34b96613bea4e200adeae6296e9ddc7ca3734107345141

Observation 3940a24d-4c04-407a-83cc-ab957e930820 · outbound

This paper cites Kimina-Prover Preview: Towards Large Formal Reasoning Models with Reinforcement Learning.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Kimina-Prover Preview: Towards Large Formal Reasoning Models with Reinforcement Learning

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:17.669259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:17.669259Z digest=sha256:dccb713acc0ac1f3c991b8965c778cd3fabbe720d319f403b5d88e82ef1b6b11

Observation 6710087f-8224-4e35-ad50-9b82666406af · outbound

This paper cites Reasoning or memorization? unreliable results of reinforcement learning due to data contamination.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Reasoning or memorization? unreliable results of reinforcement learning due to data contamination

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:17.675213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:17.675213Z digest=sha256:3a651dd7ebb9051d76d5087018ea63823af376e06d8837a2c8379653916ea54a

Observation bb3a35eb-8f44-4213-afc6-37a9832981d2 · outbound

This paper cites When More is Less: Understanding Chain-of-Thought Length in LLMs.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny When More is Less: Understanding Chain-of-Thought Length in LLMs

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:17.681074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:17.681074Z digest=sha256:c773b9ac4b91c717f5f39b9b9d338f5053249885b70312a441c0b5ef7594ce61

Observation bd8c166a-4143-41dc-9d89-038eee7ac710 · outbound

This paper cites LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:17.686438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:17.686438Z digest=sha256:693ac2c17f2f9b23be28d6f8d39012c64ed19d10cac260f4b4876df7196f8124

Observation 0312daef-72cb-4a38-9b52-3a44af887e8b · outbound

This paper cites Formal Mathematical Reasoning: A New Frontier in AI.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Formal Mathematical Reasoning: A New Frontier in AI

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:17.692754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:17.692754Z digest=sha256:93d9d60bcc4e5bb2d0ae09b29d863fc96f9d07114d5d53d08e63ad044010e349

Observation 02da9f95-122e-4711-96f7-034f9c0f2115 · outbound

This paper cites Verina: Benchmarking verifiable code generation.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Verina: Benchmarking verifiable code generation

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:17.698237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:17.698237Z digest=sha256:2ea4dd30df881cecde950fb382af8a0b7246c939cbb90f04b81519ca95ac8ff8

Observation 430c51eb-9fed-4caf-a9db-e088615483e0 · outbound

This paper cites FormalMATH : Benchmarking formal mathematical reasoning of large language models, May 2025.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny FormalMATH : Benchmarking formal mathematical reasoning of large language models, May 2025

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:19.652533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:17.703588Z digest=sha256:5257826eca7d2ac45b3e813a6059f92ca92fe6870c5d946e227b1d5c70b38d62

Observation e8b21127-c75a-417a-a23d-bebf0320684d · outbound

This paper cites Does reinforcement learning really incentivize reasoning capacity in LLMs beyond the base model?, April 2025.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Does reinforcement learning really incentivize reasoning capacity in LLMs beyond the base model?, April 2025

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:19.634829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:17.708408Z digest=sha256:443f6b1f835ec0c767cbd193dd2d25ed4dfe9fb8509a5296cf32ca0f1ccac35f

Observation 3699dcec-3bec-4155-88fa-a731f34d77b2 · outbound

This paper cites Absolute Zero: Reinforced Self-play Reasoning with Zero Data.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Absolute Zero: Reinforced Self-play Reasoning with Zero Data

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:17.714526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:17.714526Z digest=sha256:2e3d587db6a9681f422eb220e3281553faf2003835a9854b41d4766fc922252c

Observation e4f32607-c403-4172-afa8-285d69505bee · outbound

This paper cites MiniF2F: a cross-system benchmark for formal Olympiad-level mathematics.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny MiniF2F: a cross-system benchmark for formal Olympiad-level mathematics

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:17.720789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:17.720789Z digest=sha256:0c9c9980ff41df65f6969c641061fbe6638247e387a5ff9a33045e9709fb55f8

Observation 5bd1a8ea-c079-4975-9f12-b3db1ff86f95 · outbound

This paper cites What Makes Large Language Models Reason in (Multi-Turn) Code Generation?.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny What Makes Large Language Models Reason in (Multi-Turn) Code Generation?

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:17.726515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:17.726515Z digest=sha256:5c7d3316e7313980089e1f7fb2c4e7dcfde9d7c99ba3cba389d8bc114cb6a07e

Observation ce79385c-3a09-4593-93cb-ba2fe6fb31bd · outbound

This paper cites Reasoning by superposition: A theoretical perspective on chain of continuous thought.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Reasoning by superposition: A theoretical perspective on chain of continuous thought

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:20:19.619677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:17.731672Z digest=sha256:1e1e91f3becfbdeb5b7508fc42ef2269bfd2a2634d3cb9ff5e5861bdbe868033

Observation 4afedf7c-1975-4939-aed2-58f01ada670f · outbound

This paper cites @esa (Ref.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny @esa (Ref

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:17.737922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:17.737922Z digest=sha256:092feb847d237be6e10c66a21cbc0d2cee0e039dd77eac047c3c5d68aa01d78c

Observation 14cf0dc2-8a5a-408b-a8c5-64e0b2998ac7 · outbound

This paper cites an unresolved cited work.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny Unresolved cited work

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:17.743149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:17.743149Z digest=sha256:5d9f221226b889597117f55c2533247a21a6aaf4c09f288b28aac673cb738f78

Observation a4963fce-d01f-401b-b5a9-2f72c91700a3 · outbound

This paper cites best exploration.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny best exploration

Reference 99

Resolution
verified exact
raw_fallback, observed 2026-08-06T15:20:18.027458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T15:20:17.749558Z digest=sha256:6fd0aff81095c36c1708f2b5768129cb6c5120cf6e0a3ae8405fd930873d934f

Observation 8e5d1c21-5536-4f7b-a76b-20ce97f17e2f · outbound

This paper cites write newline.

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny write newline

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:17.762309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:20:17.762309Z digest=sha256:61ee9f40a2d91c8b95bf92e59a4d298f1710e2b732cd1f763f41836e0245b021

Pith citing papers

Observation 8c4c22c5-3eba-4d9f-b069-19726470732d · inbound

Differentiable Evolutionary Reinforcement Learning cites this paper.

Differentiable Evolutionary Reinforcement Learning Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-07T02:15:55.820429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T22:34:47.984401Z digest=sha256:90f5f018a7d0742978fe4c20226d867f0d8820e9815315ca290a7fa7f87e1cf8

Observation bbe7d2f0-f928-4619-8b21-8e28f7b5c114 · inbound

SpecRL: Reinforcement Learning with Test-Based Completeness Rewards for Formal Specification Synthesis cites this paper.

SpecRL: Reinforcement Learning with Test-Based Completeness Rewards for Formal Specification Synthesis Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-07-07T02:15:55.820429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T19:06:10.718673Z digest=sha256:b9139f4a34d0e5849d620e25bd5f8805a248125069f4c9704b9a0ff6d665e47e

Observation bc463633-a4ff-4710-a2b0-84d885712431 · inbound

Rethinking Agentic Reinforcement Learning In Large Language Models cites this paper.

Rethinking Agentic Reinforcement Learning In Large Language Models Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny

Reference 109

Resolution
verified exact
arxiv_id, observed 2026-07-07T02:15:55.820429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T06:30:09.945371Z digest=sha256:894eea4244bb58f00046e4eff7ae40feb0366a9ee546018ee665650b8bf0a721

Observation 7f4adce5-b04a-4521-8593-4de917f2e2e4 · inbound

Rethinking Agentic Reinforcement Learning In Large Language Models cites this paper.

Rethinking Agentic Reinforcement Learning In Large Language Models Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny

Reference 109

Resolution
verified exact
arxiv_id, observed 2026-07-07T02:15:55.820429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T03:12:19.414358Z digest=sha256:9cfd2f9a5d0c339cfabb5138d692ca5183242d8417bffc418244e02b81c62b1d

Observation 727b29f7-7158-48e6-8be7-2c89ade6de60 · inbound

Rethinking Agentic Reinforcement Learning In Large Language Models cites this paper.

Rethinking Agentic Reinforcement Learning In Large Language Models Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny

Reference 109

Resolution
verified exact
arxiv_id, observed 2026-07-07T02:15:55.820429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T16:58:41.558250Z digest=sha256:6ca798a1ce7319a34188f6ff39e889dc4c4cd11b2317ee15e1660f83e31f9e4a

Observation e04202e0-9fbc-44b0-8f39-e67bafc14397 · inbound

Automating Formal Verification with Reinforcement Learning and Recursive Inference cites this paper.

Automating Formal Verification with Reinforcement Learning and Recursive Inference Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny

Reference 104

Resolution
metadata mismatch
arxiv_id, observed 2026-07-07T02:15:55.820429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T23:52:36.891080Z digest=sha256:3e4d6ac292517de0ff721b2c7780c7e35e259576c1fe42385551c97179f10b6f