Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:08:12.800621Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 1 inbound Pith citation observation for arXiv:2508.02186.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:08:12.800621Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-28T17:17:44.826139Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-28T17:22:25.112597Z
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 521ba75a-1329-4cd3-8f8d-ac40f3e50fdb · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training OpenAI o1 System Card
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40eaebf6-edc1-49c4-91ca-0bf36c606893 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1eaec9f-8804-40ec-999b-ae1ddea56737 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a327b30d-25c0-410c-ac2a-6739cbc9a188 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2230aad-aa95-4a74-98f9-df858b7527ff · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Efficient Reasoning Through Suppression of Self-Affirmation Reflections in Large Reasoning Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fab89506-7469-4de6-ab70-cddda240ce39 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training On reasoning strength planning in large reasoning models,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5ed0d81d-c867-47f4-adfa-3c9d21dd653a · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8858b52a-37b7-43a8-8586-e097cd95e7ea · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 485334e8-46cf-4e65-9298-00594d41d28e · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training O1-Pruner: Length-Harmonizing Fine-Tuning for O1-Like Reasoning Pruning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d05e6839-a82d-400a-8666-1d909cd41412 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Training language models to reason effi- ciently,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9c2925d9-1ba6-46d9-beec-995c06c1c312 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Chain-of-thought prompting elicits reasoning in large language models,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fb219ad-9146-41ce-8b08-f42b54d63bdd · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Reasoning with language model prompting: A survey,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 88ebd071-3832-41d0-b71f-3cda71f6fc44 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Dast: Difficulty-adaptive slow-thinking for large reasoning models,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ded8afa0-bb66-4199-8e42-bfb85be62ee5 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Think When You Need: Self-Adaptive Chain-of-Thought Learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7eca2c49-f8f5-4cb1-8350-007051668f10 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Token-budget- aware llm reasoning,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation cd454ec8-313d-4a27-aefd-16d204ee7e9e · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Can language models learn to skip steps?
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f6c42539-1989-485b-9672-74ad5a52148b · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Cot-valve: Length- compressible chain-of-thought tuning,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation dc59d6e1-9e2a-43e6-a6ce-631fc02fa798 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Data-efficient rein- forcement learning for complex nonlinear systems,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c2c3d6d9-8ffb-4831-a490-47ee413b30a5 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90cb4fe0-b8d8-4a98-adc4-35de0af72969 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training ThinkPrune: Pruning Long Chain-of-Thought of LLMs via Reinforcement Learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4466a45-860c-4843-ac05-09dab89649c4 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Optimizing Length Compression in Large Reasoning Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6915e61-61a0-4f32-a12b-f985863a2d14 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Learn to reason efficiently with adaptive length-based reward shaping,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9801a469-2c31-4c57-9b23-11cecbb294cf · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training A survey on reinforcement learning for recommender systems,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation adcddd72-8f26-4a80-82e7-625eca13fe55 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training GPQA: A Graduate-Level Google-Proof Q&A Benchmark
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a138f5a-c56c-491b-9adc-c90a4f345df9 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Livecodebench: Holistic and contamination free evaluation of large language models for code,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3b018bbb-1ba5-45df-a54a-81943378ccf4 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Hybridflow: A flexible and efficient rlhf framework,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d80eb73d-4078-4f53-a4f7-a30aba153225 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31478fc7-1122-402e-a4c7-8698150747b5 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Deepscaler: Surpassing o1-preview with a 1.5 b model by scaling rl,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 63786440-f1d6-4606-b1ff-0d1a3c1875ed · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Training Verifiers to Solve Math Word Problems
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e2aa0cd-210c-4cb7-a7ed-892c32d24fc0 · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Measuring Mathematical Problem Solving With the MATH Dataset
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0524808-5450-43a3-a900-f615b561182c · outbound
Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training Mind the gap: Bridging thought leap for improved chain-of-thought tuning,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 351497bb-0519-4a31-bf97-51810c72946f · inbound
CEAR: Certified Ensemble Adversarial Robustness in DNNs Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.