Pith. sign in

Paper Citation Record · LEDGER

DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 61 inbound Pith citation observations for arXiv:2408.08152.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.08152 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 61 of 61 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:46:23.465371Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation dba36800-f15c-4513-b4b8-c1c47c3b0ced · inbound

A Lean Dataset for International Math Olympiad: Small Steps towards Writing Math Proofs for Hard Problems cites this paper.

A Lean Dataset for International Math Olympiad: Small Steps towards Writing Math Proofs for Hard Problems DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T10:53:18.396109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T10:53:18.396109Z digest=sha256:3752fe04995457fb02779040819cea7a89f0883bbdd53d0d0f242f99c1d9c7c8

Observation 0421d40f-4115-4b95-ad19-e706cf235518 · inbound

HunyuanProver: A Scalable Data Synthesis Framework and Guided Tree Search for Automated Theorem Proving cites this paper.

HunyuanProver: A Scalable Data Synthesis Framework and Guided Tree Search for Automated Theorem Proving DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T23:21:23.688419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T23:21:23.688419Z digest=sha256:89a213379fceab7c12b47b9552c2405643def6f2f402a9ee79642a68c1d76bd0

Observation eb0ca9fb-c7d5-43a6-98d7-ad840d762eda · inbound

Sigma: Differential Rescaling of Query, Key and Value for Efficient Language Models cites this paper.

Sigma: Differential Rescaling of Query, Key and Value for Efficient Language Models DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T15:50:22.003332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:50:22.003332Z digest=sha256:610dc0f8d9625bc3c7b1999db37249afdd46ca139f3975741c66b272949791a2

Observation bed49623-bcc8-4193-8992-de02fa411719 · inbound

Can We Generate Images with CoT? Let's Verify and Reinforce Image Generation Step by Step cites this paper.

Can We Generate Images with CoT? Let's Verify and Reinforce Image Generation Step by Step DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T15:32:46.351972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:32:46.351972Z digest=sha256:d52ea02e08f400a414c88aeb72a7e98de9b2e6b68c0d1ae564225268be6a2d28

Observation 27f6e7b4-c7cd-4c8a-99a2-6dc6050eb82b · inbound

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning cites this paper.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.141252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.141252Z digest=sha256:691ec45dec2aade1b1bb637b414e88013b6ddac653a0bb54576592fb6935cb51

Observation d999a1af-3b39-4a21-9340-8439d6fac3ef · inbound

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning cites this paper.

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 214

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T04:32:33.137238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-23T04:30:38.804702Z digest=sha256:430aa198ed42398b96ea18e2dafe022f7ce3f235fcecb218d9c8d5a8e4e81449

Observation 81cb6d15-f1e8-4ac8-aac2-b0bdd20150fc · inbound

Simplifying Formal Proof-Generating Models with ChatGPT and Basic Searching Techniques cites this paper.

Simplifying Formal Proof-Generating Models with ChatGPT and Basic Searching Techniques DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T05:12:15.759555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T05:12:15.759555Z digest=sha256:c5cd84db1ae73ee4f2b38e516548243dc3cd64ef615aa1472cb8862c8166a656

Observation 2be1dd12-3a24-4ba3-bc55-51dda81dff4c · inbound

L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning cites this paper.

L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T00:19:22.290253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T00:19:22.140009Z digest=sha256:68ce5cea0a38790dfbd26e873569cb2357a139f44af0ec83a5222bd82f978a11

Observation 009f7eb1-480d-4de6-91ec-639df56984e0 · inbound

FormalMATH: Benchmarking Formal Mathematical Reasoning of Large Language Models cites this paper.

FormalMATH: Benchmarking Formal Mathematical Reasoning of Large Language Models DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T00:46:23.465371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:46:23.465371Z digest=sha256:9388711b6178f7d6281de5ec13bd6e375e4e1f912e3ffdb5bd7a96c9d21ef7c5

Observation 284f4414-06e2-4d28-8895-f4b2f4677108 · inbound

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics cites this paper.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T00:04:38.434922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:04:38.434922Z digest=sha256:2bc90f2e183a13b556551d52cd992a8b8034c1f2b73239088bf064ceeacad631

Observation 117f3649-3aa4-44f5-af69-6dba9d95bbf8 · inbound

Beyond Theorem Proving: Formulation, Framework and Benchmark for Formal Problem-Solving cites this paper.

Beyond Theorem Proving: Formulation, Framework and Benchmark for Formal Problem-Solving DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-15T23:31:49.376325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:31:49.376325Z digest=sha256:14925d24032fd1f363c4aaa21ff9a00fbefff87a58aa26a3ec5209d80f436738

Observation ef2701fa-8b87-405e-9596-b2a6a2e94294 · inbound

CRPE: Expanding The Reasoning Capability of Large Language Model for Code Generation cites this paper.

CRPE: Expanding The Reasoning Capability of Large Language Model for Code Generation DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T21:22:19.139788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:22:19.139788Z digest=sha256:90f1cb6cafed9c64ee1306f0fc1c0910b65c3b36db2bbd338cb7af150b8012c9

Observation 1e3ebc5b-71de-417f-9811-808b51f6d5d3 · inbound

MPS-Prover: Advancing Stepwise Theorem Proving by Multi-Perspective Search and Data Curation cites this paper.

MPS-Prover: Advancing Stepwise Theorem Proving by Multi-Perspective Search and Data Curation DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:25.046861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:05:25.046861Z digest=sha256:69e0ad7bc1f1a6dc70b1f702565ae69fcb7560cc5a663c50948c3d563296ff4a

Observation 72228aee-673f-4a8f-a004-1c6cb8bd4f4a · inbound

LLM-based Automated Theorem Proving Hinges on Scalable Synthetic Data Generation cites this paper.

LLM-based Automated Theorem Proving Hinges on Scalable Synthetic Data Generation DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T20:50:26.135416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:50:26.135416Z digest=sha256:8aca4b255b51215417f9779f27f91cf641a7fe754980cc9db43e3f6447063d4c

Observation fa8798e2-dfe8-4a2a-948a-3063f9ce1879 · inbound

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace cites this paper.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:33.854739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:33.854739Z digest=sha256:37665748d1f3b9341e2bacadf51dfa0f044266c1ae90cd097468d4efd3f1f94a

Observation 0f8ba24a-76af-410f-8f9d-0bf41e04bb7b · inbound

Step-Wise Formal Verification for LLM-Based Mathematical Problem Solving cites this paper.

Step-Wise Formal Verification for LLM-Based Mathematical Problem Solving DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T13:48:59.281333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:48:59.281333Z digest=sha256:766278a309b684c63a9c537d52bc6dd92aa57f8ef37ba002fd6e34de902501b6

Observation 4fd7d77d-3e81-4f63-83c4-4d4f65a1cad8 · inbound

Faithful and Robust LLM-Driven Theorem Proving for NLI Explanations cites this paper.

Faithful and Robust LLM-Driven Theorem Proving for NLI Explanations DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:20.322445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:20.322445Z digest=sha256:3fdd0759ec4fec42a47e776e16eb1c0148dea868ea85d8636802cfabe2279f3a

Observation 0f166eca-8d7c-4951-8f63-e223a9113c85 · inbound

Rewarding the Unlikely: Lifting GRPO Beyond Distribution Sharpening cites this paper.

Rewarding the Unlikely: Lifting GRPO Beyond Distribution Sharpening DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:31:02.571996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:31:02.571996Z digest=sha256:5357003ef7b3aa97bd0fa7dbb09f6244b375c8a2391cf3c1f308cf1238ca8282

Observation 0fffaa7c-c3cf-4c90-a600-e3b2b07038c1 · inbound

EPiC: Towards Lossless Speedup for Reasoning Training through Edge-Preserving CoT Condensation cites this paper.

EPiC: Towards Lossless Speedup for Reasoning Training through Edge-Preserving CoT Condensation DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:49.660122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:52:49.660122Z digest=sha256:3c538b415021b37ae53e051a6b44e776af34388577408da76faa79bb49674080

Observation 31f0da6c-8af6-4be9-8cf0-71392c1708dc · inbound

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? cites this paper.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:27.508982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:27.508982Z digest=sha256:a9379bc7bdbc94761f726df2c5db4eaba1144fccf45e8cd45d0c7ab880deedfa

Observation 74fc557b-0362-4402-a155-a95122ec2aa9 · inbound

Mathesis: Towards Formal Theorem Proving from Natural Languages cites this paper.

Mathesis: Towards Formal Theorem Proving from Natural Languages DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:50:37.770014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:50:37.770014Z digest=sha256:5eca57aa22b50b6a59381281bfaacadebfdc0c102dd5fc8285b46411a6a2784e

Observation 4b6a7b2b-0bf4-4c5a-b127-2de85421d2a2 · inbound

Reviving DSP for Advanced Theorem Proving in the Era of Reasoning Models cites this paper.

Reviving DSP for Advanced Theorem Proving in the Era of Reasoning Models DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:09.031291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:09.031291Z digest=sha256:8dc759687f08b6f66cebd69aed042d5e136606d6403afb28c518e1e62eff97bc

Observation 81ce2191-54d4-409a-8c40-53578b8eaa83 · inbound

Reward Models in Deep Reinforcement Learning: A Survey cites this paper.

Reward Models in Deep Reinforcement Learning: A Survey DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T19:38:51.075704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:38:51.075704Z digest=sha256:bc4e6ecb14787b87063a399263f774abe3dcbf5862c2f1e5da9f6c5119ec2784

Observation e25544d2-3ac1-47ec-a3b3-b0cbeb20bfbf · inbound

Reasoning in machine vision by learning fast and slow thinking cites this paper.

Reasoning in machine vision by learning fast and slow thinking DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T22:16:18.677828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:16:18.677828Z digest=sha256:71a406ffca43e87e2d51cd687e4c232e7c9b684c4bc14dddc2ed97fc2e15da2a

Observation b8f3c95e-e4b9-4d45-abd3-496250eac20c · inbound

Clarifying Before Reasoning: A Coq Prover with Structural Context cites this paper.

Clarifying Before Reasoning: A Coq Prover with Structural Context DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:27.641174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:31:27.641174Z digest=sha256:157c9752afce81b059bffe4ac28aa98b7b3401d9717c0478e03efc55a8a09f82

Observation 17991571-cd2a-444c-a667-59dc82f1846b · inbound

Bourbaki: Self-Generated and Goal-Conditioned MDPs for Theorem Proving cites this paper.

Bourbaki: Self-Generated and Goal-Conditioned MDPs for Theorem Proving DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T20:27:36.371313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:27:36.371313Z digest=sha256:2f06e8fc8514e5b842d9c16486f731c679ace75a2e58850c9ea7a568f520bcdf

Observation 11e8fc58-59b8-4021-aa85-c1d3d3da460f · inbound

Advocate for Complete Benchmarks for Formal Reasoning with Formal/Informal Statements and Formal/Informal Proofs cites this paper.

Advocate for Complete Benchmarks for Formal Reasoning with Formal/Informal Statements and Formal/Informal Proofs DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T19:45:12.454495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:45:12.454495Z digest=sha256:54365fc26f002a073314bc7e9c4d0589681cb04c5e1cc78ab0c17b9afec82aa7

Observation ae4628e3-3ea9-4fca-92cf-4ce7715b3919 · inbound

CriticLean: Critic-Guided Reinforcement Learning for Mathematical Formalization cites this paper.

CriticLean: Critic-Guided Reinforcement Learning for Mathematical Formalization DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T19:14:20.435556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:14:20.435556Z digest=sha256:80301067f9a05a039162fbcf54e2ce3a90525a2a4aa35628c8dc5f1ccf59833a

Observation a890d5cb-06cf-41ab-b7c4-9cf0e25be14e · inbound

Towards Solving More Challenging IMO Problems via Decoupled Reasoning and Proving cites this paper.

Towards Solving More Challenging IMO Problems via Decoupled Reasoning and Proving DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T19:30:50.388008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:30:50.388008Z digest=sha256:e3fbc19558f88957fc6171b7b9e3834b62723e7310b15c38da6e9222726718c1

Observation b592f4b3-63e6-4fdf-8045-4b2ed7246c7a · inbound

Leanabell-Prover-V2: Verifier-integrated Reasoning for Formal Theorem Proving via Reinforcement Learning cites this paper.

Leanabell-Prover-V2: Verifier-integrated Reasoning for Formal Theorem Proving via Reinforcement Learning DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T18:18:46.956668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:18:46.956668Z digest=sha256:e5cdd5a170e40a50dc779e9e118ac6181e290fd7781787f659a06858062ec996

Observation 9c665c9f-b7e4-4dc8-8d47-e0efc2e95bb7 · inbound

It's Not That Simple. An Analysis of Simple Test-Time Scaling cites this paper.

It's Not That Simple. An Analysis of Simple Test-Time Scaling DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:06.494941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:06.494941Z digest=sha256:29cc70deb38ff145507026e19f0bd6106c77d5007b91ba45fb12dbb7463655a6

Observation d0835ab6-7595-4193-8269-c7e79b982a38 · inbound

LeanTree: Accelerating White-Box Proof Search with Factorized States in Lean 4 cites this paper.

LeanTree: Accelerating White-Box Proof Search with Factorized States in Lean 4 DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T15:54:12.937920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:54:12.937920Z digest=sha256:ba79540329b33cd2b358c31cfc678c2c894120f3cc31ac5b4fe1d7e19916ad32

Observation 08aa81db-f101-429f-9659-fd3c8cfb03d9 · inbound

StepFun-Prover Preview: Let's Think and Verify Step by Step cites this paper.

StepFun-Prover Preview: Let's Think and Verify Step by Step DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:37.354517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:37.354517Z digest=sha256:461ec38ba6ae89cb61c27af460fe0d79286985d7d593a41cf9fcf71115aedb76

Observation e2641853-d88e-4123-92cc-f015eff25bff · inbound

FormaRL: Enhancing Autoformalization with no Labeled Data cites this paper.

FormaRL: Enhancing Autoformalization with no Labeled Data DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T16:11:23.256524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:11:23.256524Z digest=sha256:7d710e3a5023a2985fa6dc73c054624f7ba82aedfe355de7dc8b77c3a9b823af

Observation 038c7cfd-b0ef-4822-afaf-c0087f287699 · inbound

Ax-Prover: A Deep Reasoning Agentic Framework for Theorem Proving in Mathematics and Quantum Physics cites this paper.

Ax-Prover: A Deep Reasoning Agentic Framework for Theorem Proving in Mathematics and Quantum Physics DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-25T07:46:42.257511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-25T07:46:30.936002Z digest=sha256:5e36bb089eac6819745e18cd3e5bee423b46d43e73cdf1f1f69b2aa00c6456a5

Observation 58982c8d-fb8f-4747-8deb-1779904b6727 · inbound

Attention Illuminates LLM Reasoning: The Preplan-and-Anchor Rhythm Enables Fine-Grained Policy Optimization cites this paper.

Attention Illuminates LLM Reasoning: The Preplan-and-Anchor Rhythm Enables Fine-Grained Policy Optimization DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-04T09:48:10.345855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T09:48:10.345855Z digest=sha256:863717b386225fb0b1cf0ae8616ea1f4872af3fff0ccc094019cdd90bde28d0a

Observation 67f4891e-ace1-4cfe-a0e9-efa6fee65cb3 · inbound

The Search for Constrained Random Generators cites this paper.

The Search for Constrained Random Generators DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:15:21.810561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-17T22:14:38.898617Z digest=sha256:890d55e106a3bbf1dc2862a9946b9674fc1ba775a1aeb983da9c3e339302d8d9

Observation 383b457d-6862-40b6-be22-445f1b488469 · inbound

Geometry of Reason: Spectral Signatures of Valid Mathematical Reasoning cites this paper.

Geometry of Reason: Spectral Signatures of Valid Mathematical Reasoning DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T13:03:23.184840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:03:23.184840Z digest=sha256:f340635681f9424f6611ee863e36830840c14dfe088ec2e92d71bf529033ec5d

Observation f95f7968-3759-442c-9096-7548cdb8b70c · inbound

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula cites this paper.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:37.574813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:37.574813Z digest=sha256:a5d5f893f7cc912501974ce51a4fddc4a0a3bc21fd9fff4bf2ae2ccb22eb3e7a

Observation 5f181f44-d589-4c21-a3e6-33112643c945 · inbound

A Minimal Agent for Automated Theorem Proving cites this paper.

A Minimal Agent for Automated Theorem Proving DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-15T18:46:29.322796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T18:44:35.600033Z digest=sha256:72aadd33dbbf2e6d646419368b80a5b858bea86427f3e54eddb6d4ed5502ee67

Observation 5a286a8d-a65d-482a-8764-e050a284fc7b · inbound

How Your Credentials Are Leaked by LLM Agent Skills: An Empirical Study cites this paper.

How Your Credentials Are Leaked by LLM Agent Skills: An Empirical Study DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-13T13:34:41.548490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:34:41.548490Z digest=sha256:d57115707c621724341b1fcfe971160960c4d7b6695d7c5f7710c8a606cc459c

Observation eda0421a-7286-46b2-95a5-efc298ac2970 · inbound

Automatic Textbook Formalization cites this paper.

Automatic Textbook Formalization DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:08:12.507942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T20:08:10.087342Z digest=sha256:be04d303b6b3863e2f99c98d8231941d527be5e0004cbfd1c5ef720f9244fb46

Observation 194727a8-828a-4c7a-bd68-1ff9457a8f08 · inbound

Understanding Performance Gap Between Parallel and Sequential Sampling in Large Reasoning Models cites this paper.

Understanding Performance Gap Between Parallel and Sequential Sampling in Large Reasoning Models DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:20:52.408031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T18:36:01.200412Z digest=sha256:e3b1af336393781a3782bbffacb939a2cefeeba354c172e47f482732862db12c

Observation c4b7b6ed-a3e2-444e-b601-eb544e0328b1 · inbound

Re$^2$Math: Benchmarking Theorem Retrieval in Research-Level Mathematics cites this paper.

Re$^2$Math: Benchmarking Theorem Retrieval in Research-Level Mathematics DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:11:16.043910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T02:07:39.648269Z digest=sha256:8f159612f853137e0bfac6328c51b4eea9a896abc49869fbea187bcc01fcbbc3

Observation 8ce3d95b-1300-4c04-aa35-7a1f6e34763f · inbound

CauSim: Scaling Causal Reasoning with Increasingly Complex Causal Simulators cites this paper.

CauSim: Scaling Causal Reasoning with Increasingly Complex Causal Simulators DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:31:29.658752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T02:33:04.549851Z digest=sha256:3179d81e0375365844e7af9bc3efea9340ae86c742352dbc7fe865acafc43550

Observation a93e1740-4a2e-4408-ac60-f1675d37b284 · inbound

Rethinking Supervision Granularity: Segment-Level Learning for LLM-Based Theorem Proving cites this paper.

Rethinking Supervision Granularity: Segment-Level Learning for LLM-Based Theorem Proving DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:12:22.950755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T06:07:29.492413Z digest=sha256:ea2415eff39c71a5f65f8d01a1a7eb38389adc58705877af086cb2bfe5361e2a

Observation 8bb3386d-ace6-4532-9af3-c49f0536c33d · inbound

CAM-Bench: A Benchmark for Computational and Applied Mathematics in Lean cites this paper.

CAM-Bench: A Benchmark for Computational and Applied Mathematics in Lean DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-20T13:38:19.322697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-20T13:35:04.729506Z digest=sha256:20f72638c4b9208aa9cb7ab1ae5977681a3ceead788c335000ee73d2d0fb365d

Observation a957e35b-fb21-421a-9241-e23bd4181a01 · inbound

OProver: A Unified Framework for Agentic Formal Theorem Proving cites this paper.

OProver: A Unified Framework for Agentic Formal Theorem Proving DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 151

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T14:48:23.540141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-20T14:43:46.517807Z digest=sha256:ebe307a74f1444edb5e859628d920c75d0afe92911b1c918575bde77a5eb4a92

Observation 9e820d5b-0f1f-4a28-8f02-a2405fbbc2ea · inbound

Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search cites this paper.

Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T08:54:05.907784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-21T08:51:59.101930Z digest=sha256:1635263ec7ccd0f210ab5f9f08a9e8023327bae7913ebd5d51fc6efff21c7007

Observation fd45d7f6-e905-4bf8-8478-ce9b3d49c872 · inbound

Less Effort, Shorter Proofs: Reinforcement Learning for Security Protocol Analysis in Tamarin cites this paper.

Less Effort, Shorter Proofs: Reinforcement Learning for Security Protocol Analysis in Tamarin DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:10:19.043597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-25T04:09:10.588205Z digest=sha256:52d69bd1130c615444be9aba8c82280a991065ce6c1fa79c74c4c1c556b661e7

Observation 2dce0e3e-20cf-4784-8b4a-de35f7be0386 · inbound

Automating Formal Verification with Reinforcement Learning and Recursive Inference cites this paper.

Automating Formal Verification with Reinforcement Learning and Recursive Inference DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T23:52:48.466037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T23:52:36.891080Z digest=sha256:3804cab43ac45d0eefaa2960325ba49d703c88bf9fefceb7d9366ee58985c885

Observation fbfd1f37-2dde-4a79-b344-867f49a65649 · inbound

LeanMarathon: Toward Reliable AI Co-Mathematicians through Long-Horizon Lean Autoformalization cites this paper.

LeanMarathon: Toward Reliable AI Co-Mathematicians through Long-Horizon Lean Autoformalization DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-02T08:36:48.809434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T05:48:56.691155Z digest=sha256:910f27d9d37bb6dfd40b28aca08a0e89e3bd7086f11629dc3b8a2e736a0f2957

Observation 0bf5baf4-a935-494f-a80f-9041b3cbc47f · inbound

TheoremBench: Evaluating LLMs on Theorem Proving in Formal Mathematics cites this paper.

TheoremBench: Evaluating LLMs on Theorem Proving in Formal Mathematics DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:47:31.552153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T16:19:11.123994Z digest=sha256:3d37c24d8b741d5c83a3e3d8e0356b067afc0f5054a2ad9fd664b458179e3a03

Observation 67954b27-223b-453a-8fee-c3c728491771 · inbound

LAMP: Lean-based Agentic framework with MCP and Proof Repair cites this paper.

LAMP: Lean-based Agentic framework with MCP and Proof Repair DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-06-30T08:44:27.799923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T08:35:40.617232Z digest=sha256:0344bbf9c9b9237af0338596b3afba6f248c7b3847ea1c97e81f893f97e0aabd

Observation 2fe0c78d-98bf-479f-abba-c9b3a7a6ae9e · inbound

Beyond the Library: An Agentic Framework for Autoformalizing Research Mathematics cites this paper.

Beyond the Library: An Agentic Framework for Autoformalizing Research Mathematics DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:05:40.783835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-01T05:54:51.200436Z digest=sha256:c1035178cbd46e21b84c3b7f1a581b407025cac98f37c03070e220fde50f5c4f

Observation 59748c15-fb55-439b-ad3c-3c571e05cdc7 · inbound

Beyond the Library: An Agentic Framework for Autoformalizing Research Mathematics cites this paper.

Beyond the Library: An Agentic Framework for Autoformalizing Research Mathematics DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-03T22:39:01.224077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-03T22:34:08.241014Z digest=sha256:7e87caa39145f2aa58a49953506e0600bf85078a3e54c03b85a7798a82d4b42d

Observation 5ca7d177-3904-4ede-adaa-05bf3a79e060 · inbound

ShannonProver: Towards Automating Formal Cryptographic Proofs cites this paper.

ShannonProver: Towards Automating Formal Cryptographic Proofs DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T06:39:03.623110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:39:03.623110Z digest=sha256:b72f540bae49695f02ba9978c398f54ed9a877f4c4c931cc6203818e940e7276

Observation b64c7d61-f431-4082-ae65-36ca537c6fe7 · inbound

AutoCedar: An Agentic Framework for Verifier-Guided Access Control Policy Synthesis cites this paper.

AutoCedar: An Agentic Framework for Verifier-Guided Access Control Policy Synthesis DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:42.929719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:42.929719Z digest=sha256:1760294e8055251cac7588e4498c50a25facddc6fc1ac66260a9252b35cbbeb4

Observation 30e699b7-0798-4495-a085-c045ca5cf4ee · inbound

From Solvers to Research: Large Language Model-Driven Formal Mathematics at the Research Frontier cites this paper.

From Solvers to Research: Large Language Model-Driven Formal Mathematics at the Research Frontier DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 268

Resolution
verified exact
local_arxiv, observed 2026-07-10T18:17:33.834916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-10T18:16:31.176239Z digest=sha256:df70790d406d0bad90d5f5fdd6c915982ff79c3aa2ad84d8f8c9fb8929066135

Observation 94f34faf-6829-4be2-a21b-42e4cb001e21 · inbound

TreeThink: A Modular Tree Search Library for Mathematical Reasoning with LLMs cites this paper.

TreeThink: A Modular Tree Search Library for Mathematical Reasoning with LLMs DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 79

Resolution
unresolved
no resolver link, observed 2026-07-14T05:57:23.399019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T05:57:23.399019Z digest=sha256:6a6ac4007663ca3db7aaf032a67c105f201976b12d8bad2433c0ea599e6f49ad

Observation fa0d14ba-0200-452d-b3c5-d89428de5e54 · inbound

Albilich: Steerable Proof-State Orchestration for LLM-Based Mathematical Research with CAS Integration cites this paper.

Albilich: Steerable Proof-State Orchestration for LLM-Based Mathematical Research with CAS Integration DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T03:02:28.586204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:02:28.586204Z digest=sha256:9e95c9607bbd919123ad8fbd409e41bb887545118a0d3082310f0113d8c91836