Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:46:19.068265Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 6 inbound Pith citation observations for arXiv:2505.17697.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:46:19.068265Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T16:46:01.206824Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
58 of 58 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 36d688b8-fbc4-4866-b1d6-b43e0f961024 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 701cc41c-f125-4704-a7bc-2d5fd15515eb · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Qwq: Reflect deeply on the boundaries of the unknown
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d9a3552a-6555-442a-8abc-c1ed9e71a78c · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models OpenAI o1 System Card
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9f00f70-0be5-4c96-a0a4-5e869a232ef2 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf57bf20-1cae-4bb9-bec5-79632e29f28a · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Chain-of-thought prompting elicits reasoning in large language models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e35523a1-55af-462e-9760-58367039a88f · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models RedStar: Does Scaling Long-CoT Data Unlock Better Slow-Reasoning Systems?
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d59b155b-ab1b-4502-8b9e-7c452f59feca · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Demystifying Long Chain-of-Thought Reasoning in LLMs
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d19e77e3-b165-4f9c-935d-a732ad3f57ad · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Distilling System 2 into System 1
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6778c7a6-ea67-44ee-9d54-63587336a030 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models RLHF Deciphered: A Critical Analysis of Reinforcement Learning from Human Feedback for LLMs
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2caefea0-b55d-452a-8080-eaa84af2ac0d · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models RRHF: Rank Responses to Align Language Models with Human Feedback without tears
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e98105a-f30d-421a-848d-ea70f25e3e97 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Language Models are Multilingual Chain-of-Thought Reasoners
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8f2f4c1-817e-4334-be8b-6206c7404489 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Activation addition: Steering language models without optimization
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2781fba5-15c9-4c84-a301-30e30491663f · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Transformer Feed-Forward Layers Build Predictions by Promoting Concepts in the Vocabulary Space
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76685b9a-5938-436a-b38b-df1d8f68c71b · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Measuring mathematical problem solving with the math dataset
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 238c0110-8952-468a-81f8-8b1d5f942611 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Qwen2.5: A party of foundation models, September 2024
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e5e0c71-5c6d-4352-acfc-10689d82e566 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Overtrained Language Models Are Harder to Fine-Tune
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45c2478c-69d3-497f-aa5a-a2f9123d7010 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Gpqa: A graduate-level google-proof q&a benchmark
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd34cbf1-6d0f-44d7-ada5-35e0c1095c0b · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ab3989a-e047-4c12-957f-fa7cb8dbd755 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models LIMO: Less is More for Reasoning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d7efaae-f87d-4c6f-a382-ff80d6cabbca · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Evaluating Large Language Models Trained on Code
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f288f5b7-ea06-464d-9600-ba29aff74b18 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Efficient memory management for large language model serving with pagedattention
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab6d0c12-ebd9-439e-8b5a-2901f19f96c6 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models GPT-4 Technical Report
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18be4a2f-44f8-4893-a669-79215528f10b · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models The claude 3 model family: Opus, sonnet, haiku
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d16fd9b0-42e2-4641-8ce9-02e041e43145 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74687a73-7780-467c-b8bb-ba7342c4fe3d · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Qwen3 Technical Report
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a3f8a75-d6d7-495f-847e-ebb409b358d8 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models The claude 3 model family: Opus, sonnet, haiku
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a52540cb-3136-40f1-9737-d157728bb96e · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Gemini 2.5: Our most intelligent ai model
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d58b01ac-6ffd-40bc-99e7-ddeb6c51d6a1 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Training language models to follow instructions with human feedback.Advances in neural information processing systems, 35:27730–27744, 2022
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 882703b5-905e-4b4c-930c-1981a69dabf2 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Gonzalez, Ion Stoica, and Eric P
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a8d7e7c-9318-410a-aa65-476247317808 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Tree of thoughts: Deliberate problem solving with large language models.Ad- vances in neural information processing systems, 36:11809–11822, 2023
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ec49cccf-7992-40b4-a64d-5ecc9d347567 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Solving Quantitative Reasoning Problems with Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5d9e47d-c398-45cc-954c-412e79e193a6 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Llemma: An Open Language Model For Mathematics
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbc6d12e-8114-4a66-9b0d-888f4d731bb3 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a95c3f4c-3b47-43b8-8326-63362f7f26a0 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Singhal, Shekoofeh Azizi, Tao Tu, Said Mahdavi, Jason Wei, Hyung Won Chung, Nathan Scales, Ajay Kumar Tanwani, Heather J
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 91e36603-5648-411d-9b06-5a3191e0e7d6 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Galactica: A Large Language Model for Science
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24585f50-901c-4667-a70e-ff14201be228 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models PAL: Program-aided Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e4a432b-3084-432d-86e8-53c2b39fbcdc · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Contrastive decoding: Open-ended text generation as optimization
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e3072380-ce34-40b6-b81c-d04493c30428 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Reasoning with Language Model is Planning with World Model
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8694828c-2bb2-4f8b-acde-9d3ae93b6fb4 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a604b63f-412b-4c28-8665-61ba43c4067f · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Self-Consistency Improves Chain of Thought Reasoning in Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df669bc7-cd76-47eb-a157-6b28c25709b5 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Improve Mathematical Reasoning in Language Models by Automated Process Supervision
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2aaea01-99ee-49a8-a32a-c7a4b00cad30 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models An Investigation of Neuron Activation as a Unified Lens to Explain Chain-of-Thought Eliciting Arithmetic Reasoning of LLMs
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a9080bb-fe0a-4a60-b29a-2bdaf9a9141c · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Unlocking General Long Chain-of-Thought Reasoning Capabilities of Large Language Models via Representation Engineering
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 280d9557-3591-4f8e-bb70-b23d8bb717bd · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Adaptive group policy optimization: Towards stable training and token-efficient reasoning.arXiv preprint arXiv:2503.15952, 2025
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7818414-6f38-458f-88c9-6ff2bb6c12b0 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Hybrid Group Relative Policy Optimization: A Multi-Sample Approach to Enhancing Policy Optimization
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73c61492-c207-493b-b743-8d840017c1c7 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56129202-155c-4b44-b60c-5f814e81e76f · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models OpenCodeReasoning: Advancing Data Distillation for Competitive Coding
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1e2b873-a893-4c88-b3fd-dbeb39faed09 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08a11087-c695-4a85-9fc5-b371f3689e75 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Generative Verifiers: Reward Modeling as Next-Token Prediction
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de830069-0ffe-4f12-9223-cbcb7e7b65ad · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Code to Think, Think to Code: A Survey on Code-Enhanced Reasoning and Reasoning-Driven Code Intelligence in LLMs
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 687e9e94-984b-4701-8346-dee63e3cf51e · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Virgo: A Preliminary Exploration on Reproducing o1-like MLLM
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b8f1d57-e4d7-438e-b76e-6815bbb8b4cb · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Towards Best Practices of Activation Patching in Language Models: Metrics and Methods
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc743eb7-19c7-462e-849b-906f7c8757b2 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Transformer Feed-Forward Layers Are Key-Value Memories
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd8655f5-1f3f-47ba-b36d-2ec22f393a16 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Knowledge Neurons in Pretrained Transformers
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65ded547-b1db-4eed-9e2a-beba40408413 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Discovering Latent Knowledge in Language Models Without Supervision
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97eb0a9c-ef87-4b91-82f6-c5293fd136d3 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Qwen2.5-VL Technical Report
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6f0ae5f-8eb5-4980-9248-8ede6b81bc37 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Osworld: Benchmarking multimodal agents for open-ended tasks in real computer environments, 2024
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcc5a936-bb88-4679-b1aa-b062bec97917 · outbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Mamba: Linear-Time Sequence Modeling with Selective State Spaces
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35d16542-ddcf-4da5-ab10-74fa647dbe5e · inbound
Logit Arithmetic Elicits Long Reasoning Capabilities Without Training Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c83d2c6-46e2-487f-8a52-dcfd307ef04f · inbound
How Do Answer Tokens Read Reasoning Traces? Self-Reading Patterns in Thinking LLMs for Quantitative Reasoning Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cc8530d0-91bb-4a13-94f0-af3e59b18185 · inbound
Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 34b3be45-69f3-4bf4-85d5-579ae8baf3e0 · inbound
Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cd84f755-6010-41a1-bfe5-ef48de2732e7 · inbound
The Tell-Tale Norm: $\ell_2$ Magnitude as a Signal for Reasoning Dynamics in Large Language Models Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation daf9a472-aa0c-4a06-87e8-08ac9422204f · inbound
From Reasoning Traces to Reusable Modules: Understanding Compositional Generalization in Language Model Reasoning Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.