Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 26 inbound Pith citation observations for arXiv:2502.12118.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:17:44.275296Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T20:38:56.081466Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 8bbf0e62-8788-46db-b714-22f2cd90189d · inbound
OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 58108822-4319-4b0e-9fc6-2ea880285d84 · inbound
Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 211
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d61edaaa-bdcf-42ae-ab77-16a2fcdc166a · inbound
The Unreasonable Effectiveness of Entropy Minimization in LLM Reasoning Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 38e18a3b-642c-459f-8caa-ccc9876c8e26 · inbound
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00262278-648d-401e-83b7-43a01aac0f27 · inbound
Faster and Better LLMs via Latency-Aware Test-Time Scaling Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ca12fc2-9a90-4756-9d71-03f3a2ca862b · inbound
Scaling over Scaling: Exploring Test-Time Scaling Plateau in Large Reasoning Models Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fd36060-3c10-44c7-b923-e4aacd158829 · inbound
LoVeC: Reinforcement Learning for Better Verbalized Confidence in Long-Form Generations Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 55ac6afc-efc0-4230-bf21-ea854c17561c · inbound
Generalizable LLM Learning of Graph Synthetic Data with Post-training Alignment Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5cb6b72-cbe8-4b35-8f4d-b19e3acc3ea8 · inbound
Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0e4fd3c-792e-4adf-918e-215295894e38 · inbound
Sample Complexity and Representation Ability of Test-time Scaling Paradigms Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f20f4c7f-9d7b-48aa-9f8d-bad1bab99af3 · inbound
Thinking vs. Doing: Agents that Reason by Scaling Test-Time Interaction Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d155d0d9-b405-46c5-9231-8f57d771e0eb · inbound
e3: Learning to Explore Enables Extrapolation of Test-Time Compute for LLMs Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3185151-7874-4f06-bdd5-076ba5a0224f · inbound
Risk-Guided Diffusion: Toward Deploying Robot Foundation Models in Space, Where Failure Is Not An Option Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db2f22e1-3716-4ea5-b409-8dafbf119d55 · inbound
OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd3a6297-62fc-40ba-8849-979f98675793 · inbound
Learn from What We HAVE: History-Aware VErifier that Reasons about Past Interactions Online Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3121a39-00b2-46a5-a1ab-654f921f914f · inbound
Why Does Reasoning Length Converge? Unveiling the Underfitting-Overfitting Trade-off in Chain-of-Thought Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a0040a0-e9bb-426a-964a-23a9867f6fc7 · inbound
RaC: Robot Learning for Long-Horizon Tasks by Scaling Recovery and Correction Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cd3c06e-92ef-46b1-b172-6c3976ea4406 · inbound
Asking LLMs to Verify First is Almost Free Lunch Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93315f33-d916-44be-ae74-e752e30de6b1 · inbound
What Does Flow Matching Bring To TD Learning? Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7eaa64f9-baa4-4f34-8c4d-6dd271c40469 · inbound
Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation deee3ea0-b457-41b0-a510-eddef6dc0fa7 · inbound
CAPS: Cascaded Adaptive Pairwise Selection for Efficient Parallel Reasoning Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0c548352-46e2-484a-bb5d-e726d873864a · inbound
A Predictive Law for On-Policy Self-Distillation From World Feedback Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 959bee1d-0987-4be4-adfc-c3212bae2d98 · inbound
On the Generalization Gap in Self-Evolving Language Model Reasoning Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 49abea7f-0bf0-4190-a9ec-c0beb200826f · inbound
From Reasoning Traces to Reusable Modules: Understanding Compositional Generalization in Language Model Reasoning Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ec8e1e04-02ca-48cf-8566-aa37402a0fb1 · inbound
Test-Time Scaling for Small VLMs on Multilingual Visual MCQ Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86aa1bbb-a3b7-404f-b30e-bee62b08caac · inbound
Oracle Gap and Signal Fidelity: A Fixed-Pool Diagnostic for Test-Time Collaboration Scaling Test-Time Compute Without Verification or RL is Suboptimal
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.