Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T07:56:48.047690Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 0 inbound Pith citation observations for arXiv:2510.24803.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T07:56:48.047690Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
27 of 27 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b7692c4d-85f9-4910-a5a7-d62e1a7e7240 · outbound
MASPRM: Multi-Agent System Process Reward Model online" 'onlinestring :=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 684fef4c-a351-48f4-868c-d894a4bc037f · outbound
MASPRM: Multi-Agent System Process Reward Model write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3f51e71-208f-4874-9e3b-34949f3ee422 · outbound
MASPRM: Multi-Agent System Process Reward Model Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb73fb84-8b1e-4f77-bbaf-95cbc41138c7 · outbound
MASPRM: Multi-Agent System Process Reward Model Scaling Autonomous Agents via Automatic Reward Modeling And Planning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69ce9a63-a132-439c-b4ef-6b5382d7eb0c · outbound
MASPRM: Multi-Agent System Process Reward Model Process Reward Models for LLM Agents: Practical Framework and Directions
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53283494-a6e1-490e-9314-547247250b80 · outbound
MASPRM: Multi-Agent System Process Reward Model Training Verifiers to Solve Math Word Problems
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 881e3b73-1110-4772-9f13-8601f252e676 · outbound
MASPRM: Multi-Agent System Process Reward Model Process Reinforcement through Implicit Rewards
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43d7f39d-9f5c-48a5-833c-98136f934a23 · outbound
MASPRM: Multi-Agent System Process Reward Model Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c1c197b-2910-421a-b4f6-af84b593fac8 · outbound
MASPRM: Multi-Agent System Process Reward Model rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b70fef98-0b36-4922-815d-b009d1471623 · outbound
MASPRM: Multi-Agent System Process Reward Model Measuring Mathematical Problem Solving With the MATH Dataset
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 021f8d04-102a-46b9-8d57-c22f8d372532 · outbound
MASPRM: Multi-Agent System Process Reward Model Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a8fa645-9e8a-43c9-b9ee-0c03bf070808 · outbound
MASPRM: Multi-Agent System Process Reward Model Six Challenges for Neural Machine Translation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4c744a8-7fc4-4fc7-a50f-d1ac819ee512 · outbound
MASPRM: Multi-Agent System Process Reward Model MARFT: Multi-Agent Reinforcement Fine-Tuning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccd2f73c-dac3-476f-90d9-542b1c7f04b8 · outbound
MASPRM: Multi-Agent System Process Reward Model Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61540494-045a-46b4-b6e9-baa27c71ad69 · outbound
MASPRM: Multi-Agent System Process Reward Model Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48877fb9-e9cd-4017-825f-afe11f9ab73f · outbound
MASPRM: Multi-Agent System Process Reward Model Improve Mathematical Reasoning in Language Models by Automated Process Supervision
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ef2ecb1-3502-4295-9f9a-d9a99d76b874 · outbound
MASPRM: Multi-Agent System Process Reward Model Leveraging Large Language Models for Effective and Explainable Multi-Agent Credit Assignment
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa185488-4eb0-48e8-8ae1-8d51f677c548 · outbound
MASPRM: Multi-Agent System Process Reward Model Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 479e194e-4b08-45f6-b0b4-2035b59a8ee8 · outbound
MASPRM: Multi-Agent System Process Reward Model Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e81255b8-ca8b-4fa4-a25f-17e737f2a235 · outbound
MASPRM: Multi-Agent System Process Reward Model SPA-RL: Reinforcing LLM Agents via Stepwise Progress Attribution
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf0f3b42-89d9-4fe8-b040-d0d9f591fd22 · outbound
MASPRM: Multi-Agent System Process Reward Model Self-Consistency Improves Chain of Thought Reasoning in Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1525c82c-fdbb-405e-a361-e760530203a1 · outbound
MASPRM: Multi-Agent System Process Reward Model Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f56fec8-7d61-4596-9512-7fee015ef8da · outbound
MASPRM: Multi-Agent System Process Reward Model Breaking the Beam Search Curse: A Study of (Re-)Scoring Methods and Stopping Criteria for Neural Machine Translation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4314658c-e392-409d-8aa4-778ee98c19c7 · outbound
MASPRM: Multi-Agent System Process Reward Model Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4f44dfe-ee49-448d-9142-681886b99854 · outbound
MASPRM: Multi-Agent System Process Reward Model VersaPRM: Multi-Domain Process Reward Model via Synthetic Reasoning Data
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 525660d8-24af-4269-b8ea-866c6bd7e667 · outbound
MASPRM: Multi-Agent System Process Reward Model G-Designer: Architecting Multi-agent Communication Topologies via Graph Neural Networks
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efbc73d8-16ad-4a69-b59b-a06923dfcd0d · outbound
MASPRM: Multi-Agent System Process Reward Model Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.