Pith. sign in

Paper Citation Record · LEDGER

DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 52 inbound Pith citation observations for arXiv:2408.08152.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.08152 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 52 of 52 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T15:50:22.003332Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation eb0ca9fb-c7d5-43a6-98d7-ad840d762eda · inbound

Sigma: Differential Rescaling of Query, Key and Value for Efficient Language Models cites this paper.

Sigma: Differential Rescaling of Query, Key and Value for Efficient Language Models DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T15:50:22.003332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:50:22.003332Z digest=sha256:029f4a7a9fa9a85406dcd4ffa11a57189823d6dc6eceecbcee07e878ef1a2d5b

Observation bed49623-bcc8-4193-8992-de02fa411719 · inbound

Can We Generate Images with CoT? Let's Verify and Reinforce Image Generation Step by Step cites this paper.

Can We Generate Images with CoT? Let's Verify and Reinforce Image Generation Step by Step DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T15:32:46.351972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:32:46.351972Z digest=sha256:3e685fe534eb0cf30f562c3f350b91c003260be2e63f82988ae0028bb4284813

Observation 27f6e7b4-c7cd-4c8a-99a2-6dc6050eb82b · inbound

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning cites this paper.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.141252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.141252Z digest=sha256:54f3adef05fb8a1b11933e7e9beb1a545176a032d6e52589e776130aff7c2468

Observation d999a1af-3b39-4a21-9340-8439d6fac3ef · inbound

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning cites this paper.

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 214

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T04:32:33.137238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-23T04:30:38.804702Z digest=sha256:caf21bd5cf7042d66d263a390493ea7743d3d37480412e2a377268a7c43cef0d

Observation 81cb6d15-f1e8-4ac8-aac2-b0bdd20150fc · inbound

Simplifying Formal Proof-Generating Models with ChatGPT and Basic Searching Techniques cites this paper.

Simplifying Formal Proof-Generating Models with ChatGPT and Basic Searching Techniques DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T05:12:15.759555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T05:12:15.759555Z digest=sha256:0c9edaa8adff0d03cbbf367c3b9b9ce78e1a8660958b02ee39c18ee3a5e6f965

Observation 2be1dd12-3a24-4ba3-bc55-51dda81dff4c · inbound

L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning cites this paper.

L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T00:19:22.290253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T00:19:22.140009Z digest=sha256:87dd322bc3c52a9cb9169cd95f4e26dcca01730b3d6c3bea4263d214a487920c

Observation fa8798e2-dfe8-4a2a-948a-3063f9ce1879 · inbound

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace cites this paper.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:33.854739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:33.854739Z digest=sha256:349e48dc481e72c99a0621719aa659467c7f4950bda68c006004ffa9d8924acf

Observation 0f8ba24a-76af-410f-8f9d-0bf41e04bb7b · inbound

Step-Wise Formal Verification for LLM-Based Mathematical Problem Solving cites this paper.

Step-Wise Formal Verification for LLM-Based Mathematical Problem Solving DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T13:48:59.281333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:48:59.281333Z digest=sha256:c1760280824cb8d1bc2e8d14038fbe87d67c7de66adc2413eea262842271dd83

Observation 4fd7d77d-3e81-4f63-83c4-4d4f65a1cad8 · inbound

Faithful and Robust LLM-Driven Theorem Proving for NLI Explanations cites this paper.

Faithful and Robust LLM-Driven Theorem Proving for NLI Explanations DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:20.322445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:20.322445Z digest=sha256:e7cb8013fc6a36816bed6b31e5d367eb193316bd6d99ca617793a560900cc35d

Observation 0f166eca-8d7c-4951-8f63-e223a9113c85 · inbound

Rewarding the Unlikely: Lifting GRPO Beyond Distribution Sharpening cites this paper.

Rewarding the Unlikely: Lifting GRPO Beyond Distribution Sharpening DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:31:02.571996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:31:02.571996Z digest=sha256:f5d8e128c06a4d9013e06d2749e6b9f1a28ffa65145c12bc99dff8c017ef4148

Observation 0fffaa7c-c3cf-4c90-a600-e3b2b07038c1 · inbound

EPiC: Towards Lossless Speedup for Reasoning Training through Edge-Preserving CoT Condensation cites this paper.

EPiC: Towards Lossless Speedup for Reasoning Training through Edge-Preserving CoT Condensation DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:49.660122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:52:49.660122Z digest=sha256:555c516247f53abe370ecf912e84484f1145ca76f2b67e87a9ac8ce6cc570733

Observation 31f0da6c-8af6-4be9-8cf0-71392c1708dc · inbound

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? cites this paper.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:27.508982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:27.508982Z digest=sha256:2765852d31226d6d3e8f05302571c7871b2db099fc8cb75c0252db59c94cae08

Observation 74fc557b-0362-4402-a155-a95122ec2aa9 · inbound

Mathesis: Towards Formal Theorem Proving from Natural Languages cites this paper.

Mathesis: Towards Formal Theorem Proving from Natural Languages DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:50:37.770014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:50:37.770014Z digest=sha256:dbdddf8bc11a621b8bae2af75f32ec292960f309916778f999eb9402dadd02c6

Observation 4b6a7b2b-0bf4-4c5a-b127-2de85421d2a2 · inbound

Reviving DSP for Advanced Theorem Proving in the Era of Reasoning Models cites this paper.

Reviving DSP for Advanced Theorem Proving in the Era of Reasoning Models DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:09.031291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:09.031291Z digest=sha256:e5f254d4cdd46fe0901a84bf2278308104cdf6f95e36af208e914f9c7a9f5885

Observation e25544d2-3ac1-47ec-a3b3-b0cbeb20bfbf · inbound

Reasoning in machine vision by learning fast and slow thinking cites this paper.

Reasoning in machine vision by learning fast and slow thinking DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T22:16:18.677828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:16:18.677828Z digest=sha256:fcb8ef2106ce1d419e0c4c5c1976677c38a06e9750ad290fb456b506f4f8841f

Observation b8f3c95e-e4b9-4d45-abd3-496250eac20c · inbound

Clarifying Before Reasoning: A Coq Prover with Structural Context cites this paper.

Clarifying Before Reasoning: A Coq Prover with Structural Context DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:27.641174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:31:27.641174Z digest=sha256:a424043cd28ce58b7818844eb5f60067aefc9d083fdb6fc37ed18c66da40f904

Observation 17991571-cd2a-444c-a667-59dc82f1846b · inbound

Bourbaki: Self-Generated and Goal-Conditioned MDPs for Theorem Proving cites this paper.

Bourbaki: Self-Generated and Goal-Conditioned MDPs for Theorem Proving DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T20:27:36.371313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:27:36.371313Z digest=sha256:af1c2345855cfb39c996927c89d9fdfa70e497c16e6175219e0666c9ef68eb59

Observation 11e8fc58-59b8-4021-aa85-c1d3d3da460f · inbound

Advocate for Complete Benchmarks for Formal Reasoning with Formal/Informal Statements and Formal/Informal Proofs cites this paper.

Advocate for Complete Benchmarks for Formal Reasoning with Formal/Informal Statements and Formal/Informal Proofs DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T19:45:12.454495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:45:12.454495Z digest=sha256:540b4a5484aac8903d5bc801ef7f9cf67c5e7814071ac465784cdb599219b081

Observation ae4628e3-3ea9-4fca-92cf-4ce7715b3919 · inbound

CriticLean: Critic-Guided Reinforcement Learning for Mathematical Formalization cites this paper.

CriticLean: Critic-Guided Reinforcement Learning for Mathematical Formalization DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T19:14:20.435556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:14:20.435556Z digest=sha256:c7e96ab5ecc1e60d0fd4d34fe903d0a47e4e42d9561c033316119fda5a2f4dfb

Observation a890d5cb-06cf-41ab-b7c4-9cf0e25be14e · inbound

Towards Solving More Challenging IMO Problems via Decoupled Reasoning and Proving cites this paper.

Towards Solving More Challenging IMO Problems via Decoupled Reasoning and Proving DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T19:30:50.388008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:30:50.388008Z digest=sha256:c2d0437afbb04b1cef89653d7636920a9187baa0eb84d4b8c5135d34dc46cb64

Observation b592f4b3-63e6-4fdf-8045-4b2ed7246c7a · inbound

Leanabell-Prover-V2: Verifier-integrated Reasoning for Formal Theorem Proving via Reinforcement Learning cites this paper.

Leanabell-Prover-V2: Verifier-integrated Reasoning for Formal Theorem Proving via Reinforcement Learning DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T18:18:46.956668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:18:46.956668Z digest=sha256:38ee1dc7feed29fd1d24e4046f38b65831d05cc102bb73f1cc430fd460640612

Observation 9c665c9f-b7e4-4dc8-8d47-e0efc2e95bb7 · inbound

It's Not That Simple. An Analysis of Simple Test-Time Scaling cites this paper.

It's Not That Simple. An Analysis of Simple Test-Time Scaling DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:06.494941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:06.494941Z digest=sha256:2b6f97233435f8eaf0e7ff8c93b15a966bdc62ef3707002d0ce55e965b73c364

Observation d0835ab6-7595-4193-8269-c7e79b982a38 · inbound

LeanTree: Accelerating White-Box Proof Search with Factorized States in Lean 4 cites this paper.

LeanTree: Accelerating White-Box Proof Search with Factorized States in Lean 4 DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T15:54:12.937920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:54:12.937920Z digest=sha256:0a2897a914b28eae30459546e9ccaedf8d3ec6bffd5140e05f794cbd80147efe

Observation 08aa81db-f101-429f-9659-fd3c8cfb03d9 · inbound

StepFun-Prover Preview: Let's Think and Verify Step by Step cites this paper.

StepFun-Prover Preview: Let's Think and Verify Step by Step DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:37.354517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:37.354517Z digest=sha256:f24348ce9344f2024dfca00e2116e20e2db086b4ada2ce670374ff3f483806c4

Observation e2641853-d88e-4123-92cc-f015eff25bff · inbound

FormaRL: Enhancing Autoformalization with no Labeled Data cites this paper.

FormaRL: Enhancing Autoformalization with no Labeled Data DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T16:11:23.256524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:11:23.256524Z digest=sha256:957eb03ed3f5560b9a6dfe38965ec6e4ce743f2a1b2926eb3fda91403e64ecb0

Observation 038c7cfd-b0ef-4822-afaf-c0087f287699 · inbound

Ax-Prover: A Deep Reasoning Agentic Framework for Theorem Proving in Mathematics and Quantum Physics cites this paper.

Ax-Prover: A Deep Reasoning Agentic Framework for Theorem Proving in Mathematics and Quantum Physics DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-25T07:46:42.257511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-25T07:46:30.936002Z digest=sha256:1163ca9a7883d1a75f42d5aaf0d3aa6f9228e4f59cfe66defbb08de5dc517458

Observation 58982c8d-fb8f-4747-8deb-1779904b6727 · inbound

Attention Illuminates LLM Reasoning: The Preplan-and-Anchor Rhythm Enables Fine-Grained Policy Optimization cites this paper.

Attention Illuminates LLM Reasoning: The Preplan-and-Anchor Rhythm Enables Fine-Grained Policy Optimization DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-04T09:48:10.345855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T09:48:10.345855Z digest=sha256:ca591113067d5a34267715af33951ef78a8096e4e148000a60aa24eafd6a6c9a

Observation 67f4891e-ace1-4cfe-a0e9-efa6fee65cb3 · inbound

The Search for Constrained Random Generators cites this paper.

The Search for Constrained Random Generators DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:15:21.810561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T22:14:38.898617Z digest=sha256:ec4f8b77e5a252c82268e10e515fe5b6695643cb7f9d6429b72bdf0cd2374616

Observation 383b457d-6862-40b6-be22-445f1b488469 · inbound

Geometry of Reason: Spectral Signatures of Valid Mathematical Reasoning cites this paper.

Geometry of Reason: Spectral Signatures of Valid Mathematical Reasoning DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T13:03:23.184840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:03:23.184840Z digest=sha256:497f2a7f99cc597fee1b7ceca3083df4e49bc02fb3b39a480df58b67b3c5d989

Observation f95f7968-3759-442c-9096-7548cdb8b70c · inbound

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula cites this paper.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:37.574813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:37.574813Z digest=sha256:a5663e7ce96222b5394ff59e08be57cd1e7e51abc80670ce25326a07c06b3353

Observation 5f181f44-d589-4c21-a3e6-33112643c945 · inbound

A Minimal Agent for Automated Theorem Proving cites this paper.

A Minimal Agent for Automated Theorem Proving DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-15T18:46:29.322796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T18:44:35.600033Z digest=sha256:4a9388c9aee5c705c48812d29a947d9076c5d9f39b0162d59b167dbf46044ced

Observation 5a286a8d-a65d-482a-8764-e050a284fc7b · inbound

How Your Credentials Are Leaked by LLM Agent Skills: An Empirical Study cites this paper.

How Your Credentials Are Leaked by LLM Agent Skills: An Empirical Study DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-13T13:34:41.548490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:34:41.548490Z digest=sha256:120d66220e0b89e1df7d71f19fea2944a1917aa607233706eb6592830ccee9a1

Observation eda0421a-7286-46b2-95a5-efc298ac2970 · inbound

Automatic Textbook Formalization cites this paper.

Automatic Textbook Formalization DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:08:12.507942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T20:08:10.087342Z digest=sha256:16205c17cf905f6bebbfcd9a2ab0b708e1e71803cc539aef66933d7a0cc47148

Observation 194727a8-828a-4c7a-bd68-1ff9457a8f08 · inbound

Understanding Performance Gap Between Parallel and Sequential Sampling in Large Reasoning Models cites this paper.

Understanding Performance Gap Between Parallel and Sequential Sampling in Large Reasoning Models DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:20:52.408031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T18:36:01.200412Z digest=sha256:d357536a924dc67c3a700c79968b385ec2324b754071661cd62ebef9f77556a7

Observation c4b7b6ed-a3e2-444e-b601-eb544e0328b1 · inbound

Re$^2$Math: Benchmarking Theorem Retrieval in Research-Level Mathematics cites this paper.

Re$^2$Math: Benchmarking Theorem Retrieval in Research-Level Mathematics DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:11:16.043910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T02:07:39.648269Z digest=sha256:6693f26b3faf614baf04e6539a8e0b4f0d7c4d78be2bdc9310e4beedc015fbe9

Observation 8ce3d95b-1300-4c04-aa35-7a1f6e34763f · inbound

CauSim: Scaling Causal Reasoning with Increasingly Complex Causal Simulators cites this paper.

CauSim: Scaling Causal Reasoning with Increasingly Complex Causal Simulators DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:31:29.658752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T02:33:04.549851Z digest=sha256:65faf74208972f22f0916d61da6ddc9ada18fce90539d09b768c364d190542fc

Observation a93e1740-4a2e-4408-ac60-f1675d37b284 · inbound

Rethinking Supervision Granularity: Segment-Level Learning for LLM-Based Theorem Proving cites this paper.

Rethinking Supervision Granularity: Segment-Level Learning for LLM-Based Theorem Proving DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:12:22.950755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T06:07:29.492413Z digest=sha256:23efc0c892d3892e33cb24854bfda3df788ca085d762fd0760042ab15f5d9024

Observation 8bb3386d-ace6-4532-9af3-c49f0536c33d · inbound

CAM-Bench: A Benchmark for Computational and Applied Mathematics in Lean cites this paper.

CAM-Bench: A Benchmark for Computational and Applied Mathematics in Lean DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-20T13:38:19.322697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T13:35:04.729506Z digest=sha256:d6bbb18f54cefedefa95498fb287a7eab95dc5eaf23b224add37e2752bea3b94

Observation a957e35b-fb21-421a-9241-e23bd4181a01 · inbound

OProver: A Unified Framework for Agentic Formal Theorem Proving cites this paper.

OProver: A Unified Framework for Agentic Formal Theorem Proving DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 151

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T14:48:23.540141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:43:46.517807Z digest=sha256:4a7f20e87e410ebf686ca1b24280d051b52935d8a226c4fb143c11dac93fb6ae

Observation 9e820d5b-0f1f-4a28-8f02-a2405fbbc2ea · inbound

Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search cites this paper.

Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T08:54:05.907784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T08:51:59.101930Z digest=sha256:7119c9a9677eef4fc4761618220175505a2ab5cca01770cd1032edb7e08f5b3c

Observation fd45d7f6-e905-4bf8-8478-ce9b3d49c872 · inbound

Less Effort, Shorter Proofs: Reinforcement Learning for Security Protocol Analysis in Tamarin cites this paper.

Less Effort, Shorter Proofs: Reinforcement Learning for Security Protocol Analysis in Tamarin DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:10:19.043597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-25T04:09:10.588205Z digest=sha256:a05aa08c409a787a0aed570bcc4a935fe5de0bc1da12efbda22ab58fbe67875f

Observation 2dce0e3e-20cf-4784-8b4a-de35f7be0386 · inbound

Automating Formal Verification with Reinforcement Learning and Recursive Inference cites this paper.

Automating Formal Verification with Reinforcement Learning and Recursive Inference DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T23:52:48.466037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T23:52:36.891080Z digest=sha256:7e698fe61710fc259e689662edb46c1c95b74af83b38be5763c30730c6dd9cd3

Observation fbfd1f37-2dde-4a79-b344-867f49a65649 · inbound

LeanMarathon: Toward Reliable AI Co-Mathematicians through Long-Horizon Lean Autoformalization cites this paper.

LeanMarathon: Toward Reliable AI Co-Mathematicians through Long-Horizon Lean Autoformalization DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-02T08:36:48.809434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T05:48:56.691155Z digest=sha256:926c47d68d53632f289355dab3c933b6d428afbe4f039496f0846c8982eb46bd

Observation 0bf5baf4-a935-494f-a80f-9041b3cbc47f · inbound

TheoremBench: Evaluating LLMs on Theorem Proving in Formal Mathematics cites this paper.

TheoremBench: Evaluating LLMs on Theorem Proving in Formal Mathematics DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:47:31.552153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T16:19:11.123994Z digest=sha256:0208ca74d780935ac24a2223eb6d894685eeda230fc7b6d6123c3c94b16c8605

Observation 67954b27-223b-453a-8fee-c3c728491771 · inbound

LAMP: Lean-based Agentic framework with MCP and Proof Repair cites this paper.

LAMP: Lean-based Agentic framework with MCP and Proof Repair DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-06-30T08:44:27.799923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T08:35:40.617232Z digest=sha256:645dcc30b6606ff9d93c695a70c99d5137f6db3e880932f3de7ee0152a503e3b

Observation 2fe0c78d-98bf-479f-abba-c9b3a7a6ae9e · inbound

Beyond the Library: An Agentic Framework for Autoformalizing Research Mathematics cites this paper.

Beyond the Library: An Agentic Framework for Autoformalizing Research Mathematics DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:05:40.783835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-01T05:54:51.200436Z digest=sha256:5174905c4afe5f8cabadd7efb62035322047ca3c3765a87a7a1dc89c1e8b76d2

Observation 59748c15-fb55-439b-ad3c-3c571e05cdc7 · inbound

Beyond the Library: An Agentic Framework for Autoformalizing Research Mathematics cites this paper.

Beyond the Library: An Agentic Framework for Autoformalizing Research Mathematics DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-03T22:39:01.224077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-03T22:34:08.241014Z digest=sha256:7dc823c07c236be9d1d4314e485ce2e94080c4430141b16572e8360d56fb9ce1

Observation 5ca7d177-3904-4ede-adaa-05bf3a79e060 · inbound

ShannonProver: Towards Automating Formal Cryptographic Proofs cites this paper.

ShannonProver: Towards Automating Formal Cryptographic Proofs DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T06:39:03.623110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:39:03.623110Z digest=sha256:ff66905e0cdf05a3faa2af2103d1468ac903f5f250fb2ac2aca5aec7cf1ede1a

Observation b64c7d61-f431-4082-ae65-36ca537c6fe7 · inbound

AutoCedar: An Agentic Framework for Verifier-Guided Access Control Policy Synthesis cites this paper.

AutoCedar: An Agentic Framework for Verifier-Guided Access Control Policy Synthesis DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:42.929719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:42.929719Z digest=sha256:53df870d7bd0739a685574859c8cd4d66b09e93b0697433eefd9b8f01167bc28

Observation 30e699b7-0798-4495-a085-c045ca5cf4ee · inbound

From Solvers to Research: Large Language Model-Driven Formal Mathematics at the Research Frontier cites this paper.

From Solvers to Research: Large Language Model-Driven Formal Mathematics at the Research Frontier DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 268

Resolution
verified exact
local_arxiv, observed 2026-07-10T18:17:33.834916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T18:16:31.176239Z digest=sha256:3db90deecba21d1da2deb245cfd8af2aa93d517ae6443d2e545fe0fd543fb436

Observation 94f34faf-6829-4be2-a21b-42e4cb001e21 · inbound

TreeThink: A Modular Tree Search Library for Mathematical Reasoning with LLMs cites this paper.

TreeThink: A Modular Tree Search Library for Mathematical Reasoning with LLMs DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 79

Resolution
unresolved
no resolver link, observed 2026-07-14T05:57:23.399019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T05:57:23.399019Z digest=sha256:3cc3347d036ba6d416abd41e991b116fdbdd51b98c045d87966729a531928e22

Observation fa0d14ba-0200-452d-b3c5-d89428de5e54 · inbound

Albilich: Steerable Proof-State Orchestration for LLM-Based Mathematical Research with CAS Integration cites this paper.

Albilich: Steerable Proof-State Orchestration for LLM-Based Mathematical Research with CAS Integration DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T03:02:28.586204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:02:28.586204Z digest=sha256:62c01a999b957ba42b1ebf324b99d3fa0d16a2a16fcfc01ce29d4925e9963095