Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:59:15.598200Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 2 inbound Pith citation observations for arXiv:2506.12307.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:59:15.598200Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-10T18:50:22.827472Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T18:57:31.625203Z
34 of 34 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8401f3de-baca-447b-856d-a4ace2fbfc4b · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 056a1ff2-d7cc-4709-ac9a-4ec2702727da · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04de581b-a237-43cd-9123-0c3a50d17f75 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 31c6dc5b-2677-45d6-9443-ebba305728e6 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99af1e2b-ddad-407c-9840-1d3933616316 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1e43fd60-59c2-4faf-b2ea-d4f8eb9dcba4 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb78a099-e268-46bb-9a19-9a1792162fef · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46322889-addc-4d97-9840-264a60c2d1f8 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84c77575-cc46-4928-8c8f-384b66d037ee · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning O1 Replication Journey -- Part 3: Inference-time Scaling for Medical Reasoning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f783802-b219-495c-8450-3daf0ed02b0e · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning OpenAI o1 System Card
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e8a4fb8-8f3b-430a-9ee1-a147d52dc6c2 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Med-MoE: Mixture of Domain-Specific Experts for Lightweight Medical Vision-Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7a337e5-a406-4d6d-9a74-1117d2153a68 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f2db5fb0-5f05-4ff8-89c2-03d53c708fd5 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f49c605-4ba8-4fb4-a4b2-360e9411bd1d · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fc937c5d-b53c-43e1-a2de-5cce7ee5bb53 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb20ba7a-af7f-4a4d-a0b5-5d4b83a6c639 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac93fe8b-42b7-47d1-9945-99ab4958ccc8 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning MedCoT: Medical Chain of Thought via Hierarchical Expert
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe447467-8de9-422f-aeff-768cf53608a1 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning s1: Simple test-time scaling
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3687db58-cecc-43ed-9611-cae57eb00cdf · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51bedf58-c688-4b4a-bf4c-49668a85f01e · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5688dbab-e7b8-4dbe-9205-a259ecfb5437 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b442089-496d-4b5a-98ed-c8938c82a44c · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning A Long Way to Go: Investigating Length Correlations in RLHF
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bd87863-1112-4f2b-bd04-c846b5b0315b · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d0c47de-a844-4976-bfd7-516e04eb9b0e · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7cdb5a5a-7032-4b64-a3a9-134e3745bd11 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3313ee2-93db-4c35-9ba7-ea55de94f989 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Qwen2.5 Technical Report
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41080bab-eecf-4ffd-a170-102d4a00c6d1 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning FineMedLM-o1: Enhancing Medical Knowledge Reasoning Ability of LLM from Supervised Fine-Tuning to Test-Time Training
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae500814-b1e4-4faa-bb52-f6f79ade8617 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Following Length Constraints in Instructions
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f95b5e92-aa72-4268-aa69-ba3242d98517 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6e31e4cb-653f-425f-867c-c7a1b88411bb · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning Med-RLVR: Emerging Medical Reasoning from a 3B base model via reinforcement Learning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85f4c374-610f-4f6c-b165-40700cb6eb05 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b21cae6e-9d43-4d4a-81f2-714733e39145 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ce861ae-a985-4971-9f30-cdaa5fb58be6 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning online" 'onlinestring :=
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 092f3c12-39f9-4ef5-9958-81a71dc70da8 · outbound
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning write newline
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d7bdbbb-83c5-4a6f-9d9d-56823e545fba · inbound
Scalable Stewardship of an LLM-Assisted Clinical Benchmark with Physician Oversight Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cfa7d1cf-2e59-4fb1-8a59-84a31d1fb65d · inbound
Aligning Clinical Needs and AI Capabilities: A Survey on LLMs for Medical Reasoning Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning
Reference 104
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.