Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:17:44.447421Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 3 inbound Pith citation observations for arXiv:2505.19475.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:17:44.447421Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-21T11:21:30.867480Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-21T11:24:08.689134Z
18 of 18 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 807d5c9c-325e-4275-b43c-c5fb24b057f5 · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection Measuring Mathematical Problem Solving With the MATH Dataset
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4a35f45-02bc-4213-9e61-4aff232acf98 · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection SuperHF: Supervised Iterative Learning from Human Feedback
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0b18e1d-6719-47ea-b3fb-c519270c0560 · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection The Entropy Enigma: Success and Failure of Entropy Minimization
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f1808f0-f761-4bde-b22c-8a0385d9eb8c · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection The effect of sampling temperature on problem solving in large language models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 693ddb9e-c624-41dc-aaee-bdf452aec13a · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection Proximal Policy Optimization Algorithms
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93af7555-85bd-4448-88db-01c371e75911 · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2549d4f-86c9-4f5e-bba8-49617eedade3 · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection Continual Learning for Large Language Models: A Survey
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69a951f2-45b4-4564-9711-2c0316699745 · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection Beyond Model Adaptation at Test Time: A Survey
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06c5c7e8-624a-4d61-877a-f1af05548987 · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection TTRL: Test-Time Reinforcement Learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f36fc033-e6de-4467-b241-d481a7ef774f · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection Training Verifiers to Solve Math Word Problems
Reference 1988
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 159f8971-b300-4993-9824-c833f327f7fd · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak Supervision
Reference 1992
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38e18a3b-642c-459f-8caa-ccc9876c8e26 · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08e01256-2de9-41f1-9149-40516eca05f1 · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection Tent: Fully Test-time Adaptation by Entropy Minimization
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b85c2f8-e47f-4d9d-bd4a-e8f099e93c21 · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection Reinforced Self-Training (ReST) for Language Modeling
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a4e6df7-7e9a-421e-964f-fbba1689d469 · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d435c54c-67a7-453a-8e56-f683ca89d209 · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection Test-Time Training on Nearest Neighbors for Large Language Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58da377b-4834-4a5c-bc68-60066dce9c07 · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection A Survey of Test-Time Compute: From Intuitive Inference to Deliberate Reasoning
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1ea5909-4513-4eaf-87c1-a98325cd8196 · outbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection Efficiently Learning at Test-Time: Active Fine-Tuning of LLMs
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c8461f9-393c-4608-9713-a5662f24200b · inbound
Training LLM Agents for Spontaneous, Reward-Free Self-Evolution via World Knowledge Exploration Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9490c672-7ce5-4cd1-a42a-9df04c413960 · inbound
Epistemic Uncertainty for Test-Time Discovery Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1a97d93e-2d8f-4ce8-9198-cdd718d374a2 · inbound
SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.