Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T11:08:09.510213Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 5 of 5 outbound references and 6 inbound Pith citation observations for arXiv:2601.07376.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T11:08:09.510213Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T12:27:08.805309Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T07:56:47.922960Z
5 of 5 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 70d7ed3e-a225-46ab-9638-255e8247c743 · outbound
OpenTinker: Separating Concerns in Agentic Reinforcement Learning AReaL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 842e9f74-e1b6-4333-ae05-237dacb7c4f8 · outbound
OpenTinker: Separating Concerns in Agentic Reinforcement Learning OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 163af5e1-8833-4886-9f02-b695ce74d492 · outbound
OpenTinker: Separating Concerns in Agentic Reinforcement Learning Tinker, 2025
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa5821be-347f-433a-8119-8b008824ceb0 · outbound
OpenTinker: Separating Concerns in Agentic Reinforcement Learning Agent Lightning: Train ANY AI Agents with Reinforcement Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec43eaac-95a9-4c57-aeb9-07ecefb38d4d · outbound
OpenTinker: Separating Concerns in Agentic Reinforcement Learning Hybridflow: A flexible and efficient rlhf framework
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 952ab6ce-a753-4394-acb0-ca502f3766bc · inbound
Agentic AI Systems Should Be Designed as Marginal Token Allocators OpenTinker: Separating Concerns in Agentic Reinforcement Learning
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a5b811a2-f57d-41b7-8258-3bd9938189e2 · inbound
MinT: Managed Infrastructure for Training and Serving Millions of LLMs OpenTinker: Separating Concerns in Agentic Reinforcement Learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 899154d4-fdd6-40d0-890c-1af55d0dc941 · inbound
MinT: Managed Infrastructure for Training and Serving Millions of LLMs OpenTinker: Separating Concerns in Agentic Reinforcement Learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 34186f12-8aa4-4978-9cdd-21d18e70a184 · inbound
AgentJet: A Distributed Swarm Training Framework for Agentic Reinforcement Learning OpenTinker: Separating Concerns in Agentic Reinforcement Learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f1932882-5d1f-4461-8d10-0ce09ed13d4d · inbound
AgentJet: A Distributed Swarm Training Framework for Agentic Reinforcement Learning OpenTinker: Separating Concerns in Agentic Reinforcement Learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91d2c092-7bfc-4e20-a5a3-0b80ad0888f9 · inbound
JoyNexus: Service-Oriented Multi-Tenant Post-Training for VLA Models OpenTinker: Separating Concerns in Agentic Reinforcement Learning
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.