Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:27:52.808169Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2505.24500.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:27:52.808169Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 19b10a2d-5766-43e4-baf5-9eeb9d2a81c2 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 308dc582-050e-4727-b741-33d9c09ef9e0 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39d74bd7-4134-4130-aef8-714800e24531 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fb164ae-8e0c-4378-893d-f0e5e985c877 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1093a78b-dfa2-4b22-995a-1f9a9b8871c8 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence QueryAgent: A Reliable and Efficient Reasoning Framework with Environmental Feedback-based Self-Correction
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f3181d8-1738-40d3-9e36-d59e4c93e32f · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence OpenAI o1 System Card
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7a8e223-9433-4d61-ad96-5bda5d2fa32a · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d54a5a6-4a1b-4af1-a6cd-1ce9beb6e125 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Perceptions to beliefs: Exploring precursory inferences for theory of mind in large language models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 86bff7f8-0ddb-470d-8ee2-25c119f40e4e · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence s1: Simple test-time scaling
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c307fdb2-94fb-47f7-965b-c555ddcb52af · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Explore Theory of Mind: Program-guided adversarial data generation for theory of mind reasoning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3fdab7c-ad12-4103-b3fe-b67b516ae7e2 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7eacdde7-166a-4f7c-9059-def5e592d386 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence HybridFlow: A Flexible and Efficient RLHF Framework
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e809d746-7890-44c6-bfc8-5b44b0404e16 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence ToMATO: Verbalizing the Mental States of Role-Playing LLMs for Benchmarking Theory of Mind
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5df7549-894f-46fc-8577-1c81973ef424 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9edca75c-ed6c-4a03-ac76-063551775524 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Climbing the ladder of reasoning: What llms can-and still can’t-solve after sft? arXiv preprint arXiv:2504.11741,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e57198dc-bc3b-4e25-9d4e-e0450c9c8766 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence HelpSteer2: Open-source dataset for training top-performing reward models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b5f4533-fe3b-4e24-8961-ccccea2f8de8 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7474caf0-d8fe-4815-8c1d-71db8aaaefec · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Hi-tom: A benchmark for evaluating higher-order theory of mind reasoning in large language models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2c95a59f-4855-449a-a573-10761728cc9d · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1634f636-6489-4528-98a5-5373a857391d · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28372f1e-66ef-40a4-a453-95af9b8631b5 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Qwen2.5 Technical Report
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4451955-6d9b-420b-85b6-d8f8577955af · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence LIMO: Less is More for Reasoning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f316624-31ca-4996-a14e-5958a5b4f105 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Z1: Efficient Test-time Scaling with Code
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 585c6ccc-11bf-434b-b600-bd8655e19a10 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f36ab083-ece6-43d4-b9c9-3f3f29937d35 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Autotom: Automated bayesian inverse planning and model discovery for open-ended theory of mind
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd96bd9b-3e57-4aa4-80ac-fe5d99d7b201 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f807b02-017d-4bed-a1be-c32a4eab9262 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence SWEET-RL: Training Multi-Turn LLM Agents on Collaborative Reasoning Tasks
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb1d23b2-1d90-41b0-b60f-9e45ef984190 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence The comparison results are shown in Table
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 98b817c3-08ca-43c5-9bba-9f4ad29f3d62 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence What LLMs Can—and Still Can’t—Solve after SFT?
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dd47479f-7f0e-451e-a765-43ba4198eef4 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence budget forcing
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 43b977b2-7bc4-4a82-8d8b-6f534d1c1dde · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 1996
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8dc670e-9460-401a-9bd4-a8a430bd1ccf · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Revisiting the evaluation of theory of mind through question answering
Reference 2011
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c181eec8-6d4f-42e4-a4f9-56e8da5edba2 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Video-R1: Reinforcing Video Reasoning in MLLMs
Reference 2013
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 257d3a63-41ab-44bc-afab-438bd05e1d94 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Relation-r1: Cognitive chain-of-thought guided reinforcement learning for unified relational comprehension
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f1ef408-6abe-4356-bada-9bfd8f90590e · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Social iqa: Common- sense reasoning about social interactions
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ae8e3374-ad35-4ae9-bc01-3de7c29f3cb3 · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f67b7fdd-adbd-4450-a8c8-0c4c883b535a · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Large Language Monkeys: Scaling Inference Compute with Repeated Sampling
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e6e6052-e199-4952-ad4f-924d4dfd415b · outbound
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence GenCLS++: Pushing the Boundaries of Generative Classification in LLMs Through Comprehensive SFT and RL Studies Across Diverse Datasets
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.