Pith. sign in

Paper Citation Record · LEDGER

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback

As of 7 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 1 inbound Pith citation observation for arXiv:2507.15024.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.15024 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:47:33.769274Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T13:20:32.432002Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:17:40.139105Z

Reference resolution

38 of 38 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved38
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4f90c26e-6e16-4087-8845-8faec0fc3134 · outbound

This paper cites online" 'onlinestring :=.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.497379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.497379Z digest=sha256:f934f1aa48cb084b56c3659d357918acb073c6852b6264f3374bf6c6db05e556

Observation 3591f7da-6e0f-40ca-b8ec-12e89ec7c928 · outbound

This paper cites write newline.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.538772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.538772Z digest=sha256:3a7ea5a2f57d70ac650ecd8234b1247da77f30b9825b545171650163c308c297

Observation 6f7414c9-713f-489a-a2f9-ed4b1c21aeac · outbound

This paper cites Critique-out-Loud Reward Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Critique-out-Loud Reward Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.588128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.588128Z digest=sha256:9c17a33445c4c61825c9e713fbb4132c28de5169bc58865ddb3a030cc3cf4359

Observation aa17fe7c-c25c-4f5f-917d-bd674ebb3b9c · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:47:35.491866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:47:30.626895Z digest=sha256:f59514c5496af515a8e5753718292b3a452c97a35517de9d08ccec7f381be9b3

Observation 449b5695-ac43-4588-9625-8b28fde75fd8 · outbound

This paper cites SPC: Evolving Self-Play Critic via Adversarial Games for LLM Reasoning.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback SPC: Evolving Self-Play Critic via Adversarial Games for LLM Reasoning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.674164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.674164Z digest=sha256:a5531cae06db9430a26b4dcfcd0d27b1300e78b4f4f8c5ff94c2aa827cac4890

Observation 4dc8b538-237c-412c-9832-37712ac325a3 · outbound

This paper cites Boosting of Thoughts: Trial-and-Error Problem Solving with Large Language Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Boosting of Thoughts: Trial-and-Error Problem Solving with Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.790393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.790393Z digest=sha256:e0bef0775559af10ac9b7025945dcc4df382c1066438149746d07aa93f9811e0

Observation 5fb05160-88d7-460a-91d7-1820eabc0d91 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.853821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.853821Z digest=sha256:e8ac2d99168ff4ab6c36dcfb041c34c8fd5dd4b08063cd3f44d2b078a5e5e44a

Observation db27df59-873d-4426-982d-71b0e9e39cd3 · outbound

This paper cites OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.972701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.972701Z digest=sha256:ec10100e3e954c4625b22b439fd2084229ee47a99b5c47c89d5d77dc51398a6c

Observation e33a6f0a-07a2-4590-8043-8b9129eee944 · outbound

This paper cites Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.064312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.064312Z digest=sha256:b68b97d856bde345a741938ed110813529fddc25cc39c0eee7b1d9cc7682770d

Observation 99210596-09d1-48dd-8795-647088ace763 · outbound

This paper cites Qwen2.5-Coder Technical Report.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Qwen2.5-Coder Technical Report

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.143756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.143756Z digest=sha256:a30e63751184db50c87ae67aaa77b00945c94f68890ca43d1af86fa17069c511

Observation 3a3d8ee3-725a-4d21-be1e-510383e8109a · outbound

This paper cites OpenAI o1 System Card.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback OpenAI o1 System Card

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.229548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.229548Z digest=sha256:dc1ae1cf9731fc3ddade4d165a6c7676dda6ab72b43748b1efe7f93d1cdaa6bc

Observation c84370b2-5226-4678-baf5-c5b3e66f8c6c · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.341295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.341295Z digest=sha256:418ef324bb58daa4bb0e1e9b2003bc04f447c78afc4ebe78d9d6c6f7549765a8

Observation 69fa5b1c-27e6-489f-a8fe-bd2913737502 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.420751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.420751Z digest=sha256:5d6f78a10b5bcfed8fa7bbc9035525887b903935861415f91db728a46987ab0a

Observation f3b07d0f-9fef-4dcc-bd58-03d9d0aded52 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:47:35.271981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:47:31.510699Z digest=sha256:b16ba7aea04a0be5ffd2cfc95bed8351dad65953b4cd76289c0568b04c9bd88b

Observation 7ecf92e4-9072-4f79-b15c-b1096c2eaa18 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:47:34.996452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:47:31.620920Z digest=sha256:03b34da985df91479b0f5377bcc68018ddc462de02ba9cfe390043382975248b

Observation 1ddfd3e9-9406-46be-83f0-477ab0a1bfc5 · outbound

This paper cites Let's Verify Step by Step.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Let's Verify Step by Step

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.690472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.690472Z digest=sha256:5cbbd0259aa15378aec8a55917efa48e6e7b960bc70a53f9b3049b178c68a18d

Observation 6d143611-5771-45c2-b162-1a9dcf1eea7c · outbound

This paper cites CriticBench: Benchmarking LLMs for Critique-Correct Reasoning.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback CriticBench: Benchmarking LLMs for Critique-Correct Reasoning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.805450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.805450Z digest=sha256:04eab4c13d4f6283e365571321db55271c2027bffb064592fc0f8cac438fe12d

Observation aed862b1-c602-44a4-82f4-8ebca8615d41 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.927640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.927640Z digest=sha256:58a5ab0067e75222ef643133fcd28b159cb14ffdddb0d7052f96415c15ae95bf

Observation 6ce12686-ffac-4a00-95d6-eb6f36f0b990 · outbound

This paper cites Generative Reward Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Generative Reward Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.042776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.042776Z digest=sha256:df26cdae48cd717d36e171b48e682cbb96f8902c21cd713612cf2a1682139bf8

Observation 395ff2a7-3a94-42ad-afc1-0e52bc8d9e71 · outbound

This paper cites LLM Critics Help Catch LLM Bugs.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback LLM Critics Help Catch LLM Bugs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.141323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.141323Z digest=sha256:fde1b34fc667a694844094f03ab029c5d9f92491a940ff6b2344af0172a29872

Observation fe3c777f-a9a9-482b-8ac1-7bc38453e2d4 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.251144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.251144Z digest=sha256:f9ee3ba1610b2460329693471b3ce2f2b4df24d2d08874b99b2ed33dec9b2578

Observation e32f7fd4-7b89-49a2-89d4-06f1fe121628 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.354499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.354499Z digest=sha256:70b531db818f38ce674b3abdb498ddcb77edad9d3e960b111d6235347e757d96

Observation 512245a2-62d9-4b47-845f-d8d2968ee0fb · outbound

This paper cites Heimdall: test-time scaling on the generative verification.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Heimdall: test-time scaling on the generative verification

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.464146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.464146Z digest=sha256:27244bd4db5e05582bab35e21969a26a7837405f128d6eec736592dc8e29ac13

Observation cfd9916b-b335-47b2-a4d8-ec9c1f4fb40d · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.548031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.548031Z digest=sha256:17826f037fed1d366c00057bbe19e9477e3897bd5e2b18a313a3dec0dbdd1230

Observation 1e1cb7b9-361f-423f-b192-1d6d3dcc962d · outbound

This paper cites Self-Evolving Critique Abilities in Large Language Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Self-Evolving Critique Abilities in Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.650831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.650831Z digest=sha256:8e2543d254514f059e77fe3b087ee11af5af651a3a55b0cfdd03e1d15d44ea64

Observation d81f6338-9e43-4d73-bece-a415c274ad59 · outbound

This paper cites RealCritic: Towards Effectiveness-Driven Evaluation of Language Model Critiques.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback RealCritic: Towards Effectiveness-Driven Evaluation of Language Model Critiques

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.729477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.729477Z digest=sha256:583843d59bc7d73d374f4ab0ebefe48cdad7056f43f7f4e7ab2c8eb6295ebf55

Observation 899a2785-16bc-4678-b659-9736ec3457ce · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.817233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.817233Z digest=sha256:378696d8a3c2f1decf7170d6e0962af01de158a1f1b759c1ca1475e50693dbe2

Observation cbfad0d3-5e83-4ea7-a463-291d5f253719 · outbound

This paper cites Solving math word problems with process- and outcome-based feedback.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Solving math word problems with process- and outcome-based feedback

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.904058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.904058Z digest=sha256:48d9cc87de303dc7fd7f207388d8b6072fdde79469594b31eba3f7aa6eb9b7de

Observation 1f082e5a-6744-4873-b8e5-0330a6545774 · outbound

This paper cites Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.007036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.007036Z digest=sha256:d21b935ed4810a25fb880db86383e431f2bb69fd75b70c856e97ffbb58da42e4

Observation 81de2408-7c57-4eb4-a988-12b937d18fe3 · outbound

This paper cites Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.068552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.068552Z digest=sha256:db374acce2e71f4ace7ef3d1ca57368e2e8bd94f25c685d87a64dfccf9369743

Observation 91982f38-b9ba-4dad-9675-4ff95486070d · outbound

This paper cites Qwen3 Technical Report.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Qwen3 Technical Report

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.161410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.161410Z digest=sha256:2e530501b3565ba1b76620f2cb053b98c994aa86ba9196fbb7687f5f852af1e8

Observation 294f5c9b-2f2c-4f5c-84ac-77f27ec6074b · outbound

This paper cites DeepCritic: Deliberate Critique with Large Language Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback DeepCritic: Deliberate Critique with Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.218455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.218455Z digest=sha256:bf3ec885a1fa275703e1526f614e483050abeb2f30986d22b6961e9dadea3a1e

Observation b7c1ccc2-b88c-4715-a1ec-9f04bc5b4cb7 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.334831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.334831Z digest=sha256:1f74f8156a5355f947931a330f80caff1e2704f68e94b56f59f2dcb79b7903ba

Observation 8d7389d0-9cca-4d9a-929a-fe97407a60de · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:47:34.703734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:47:33.411742Z digest=sha256:d9616fae3c82487a82d0cc5fad822d0293624ca5d869fa5602d7428f3b54f807

Observation 21395f0b-6e69-4a79-8e1f-e50c6f89ba4c · outbound

This paper cites The Lessons of Developing Process Reward Models in Mathematical Reasoning.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback The Lessons of Developing Process Reward Models in Mathematical Reasoning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.515460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.515460Z digest=sha256:858521f3a7ce1db6bd83792fdb0d12183c2ef31e858679abf602cf2b49f90177

Observation 0d914c5f-701e-485f-a7f2-c7e916999e3c · outbound

This paper cites ProcessBench: Identifying Process Errors in Mathematical Reasoning.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback ProcessBench: Identifying Process Errors in Mathematical Reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.591244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.591244Z digest=sha256:f3236b0e95157ca1fe386ed56bce16bed582baff7b68efa2944cb90adedd0199

Observation 77882341-f8f7-4bfe-adf2-7c2cf0f6da3b · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.690552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.690552Z digest=sha256:aeb736ca834788ccaee74d5f37ac9ca7e8d5b266c78045fe478565f0d9a9783e

Observation 60bd0749-61dd-470d-80bf-726a22ebc83b · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:47:34.475927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:47:33.769274Z digest=sha256:09caa1b911e1971f96b54e59bd89a0b35d988b8bea71a98708fb79780661e93d

Pith citing papers

Observation e14f3855-713e-4386-ae59-1a6f75f5b28f · inbound

A History-Aware Visually Grounded Critic for Computer Use Agents cites this paper.

A History-Aware Visually Grounded Critic for Computer Use Agents RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:17:40.140375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T13:20:32.432002Z digest=sha256:f2af0cf95cc4c4bc6c0f29c6671f38000bd1d7c3f844784c378b51050db73389