Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:05:19.897437Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 0 inbound Pith citation observations for arXiv:2505.17153.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:05:19.897437Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
50 of 50 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation edcb3ecf-1ee1-4680-b786-1a5fbc0e4f94 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20594c8e-cf54-4b7b-ab41-0c971e689852 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8411c99-df34-483a-89a1-76374c5d0616 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2451d7b6-0fb5-48dc-b271-06680e0039ec · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Transformer Feed-Forward Layers Build Predictions by Promoting Concepts in the Vocabulary Space
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fee41cc-b67c-4220-a614-8d07141047c6 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Transformer Feed-Forward Layers Are Key-Value Memories
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3da8ab7f-1a4b-4b33-9fc6-607cd275d8a9 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN The Llama 3 Herd of Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27bac971-3325-41f3-a048-39aeb616320d · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Parameter-Efficient Fine-Tuning for Large Models: A Comprehensive Survey
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 091fe990-b33c-4595-91bd-3a51ae940bd7 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b369802f-7514-4d8c-bc95-d06f81b00da0 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN SparseAdapter: An Easy Approach for Improving the Parameter-Efficiency of Adapters
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44a2a934-5672-4d7b-8b2e-0756933028df · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Measuring Mathematical Problem Solving With the MATH Dataset
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66c33497-9e18-4736-96d2-6e9e11923946 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Parameter-efficient transfer learning for nlp
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c09dca73-b644-4614-87f3-03714e94a006 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b618ca45-392e-4a07-9b72-412fa770fcb1 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN OpenAI o1 System Card
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4b4d9fd-3564-4e50-8f30-2b84d5db40ff · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN The Power of Scale for Parameter-Efficient Prompt Tuning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34f5e305-0b0e-4771-9a15-bea1284fcc89 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68eeb38e-a341-4639-bb0b-0b8ce748747e · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Numinamath
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46ac53ae-08cc-4531-81a5-3f28b362e45f · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Prefix-Tuning: Optimizing Continuous Prompts for Generation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cd70891-dba0-49ca-8880-a9e4d17bf1ab · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Alpacaeval: An automatic evaluator of instruction-following models, 2023
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff4bc00a-3230-4281-ab2e-6656d2c1f2bb · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Forgetting Transformer: Softmax Attention with a Forget Gate
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 230253f7-1030-4595-8aa5-ef821a7fa717 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN In-context Vectors: Making In Context Learning More Effective and Controllable Through Latent Space Steering
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad3441ee-d468-4673-b568-576ea0b8e971 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Parameter-Efficient Orthogonal Finetuning via Butterfly Factorization
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 733fb8d7-ec21-433b-b1a7-dec171a681e6 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Aligning Large Language Models with Human Preferences through Representation Engineering
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c720b9a-7f7e-4f4a-8683-72bde472d96a · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Deconstructing Long Chain-of-Thought: A Structured Reasoning Optimization Framework for Long CoT Distillation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb867424-b9c0-40a4-91f1-230646a0278e · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN RWKV: Reinventing RNNs for the Transformer Era
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 125abb68-a687-4ad8-acb7-69af4274630c · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Fast Transformer Decoding: One Write-Head is All You Need
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba14f151-1384-4bf2-8424-f341d9f5056b · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN GLU Variants Improve Transformer
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f37b9ff8-a8e2-4dad-b2d9-4f044e4e8287 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN LLM Braces: Straightening Out LLM Predictions with Relevant Sub-Updates
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 91217912-0b99-41be-a8ce-fef01a32ef9d · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Extracting Latent Steering Vectors from Pretrained Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 503f1712-bfa0-4fae-a5b4-ac58094778da · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Augmenting Self-attention with Persistent Memory
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3b9e273-3a2e-4d4a-9d54-87aee7d81162 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN End-to-end memory networks.Advances in neural information processing systems , 28, 2015
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9d3f7704-a80b-4415-81ab-03d1b3823450 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Unlocking General Long Chain-of-Thought Reasoning Capabilities of Large Language Models via Representation Engineering
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 320dd58c-efa1-4828-bc2d-5b2f5235d66b · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Open Thoughts
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e31af5f-ba6e-45b9-98da-6b8ec303106e · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Attention Is All You Need
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3e006d8-133f-4a32-8f27-dc6c2ae53af5 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN AdaMix: Mixture-of-Adaptations for Parameter-efficient Model Tuning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0017b253-c6cf-4eef-91a7-27d7029881db · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Latent Space Chain-of-Embedding Enables Output-free LLM Self-Evaluation
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dadb16df-3acd-428a-ae37-a59c1aa4f0ad · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Wong, Zhuosheng Zhang, and Rui Wang
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation be08691d-74b4-4e02-8710-feccd04bb026 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a891a60-d82e-4554-820c-e3eec9d5f64f · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Light-R1: Curriculum SFT, DPO and RL for Long COT from Scratch and Beyond
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2ac8587-3ae2-4ad3-811c-031ed8cceb80 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Advancing Parameter Efficiency in Fine-tuning via Representation Editing
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 344a521c-fd11-4e3d-beb1-43cd82ff1f3f · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Manning, and Christopher Potts
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4abe153e-e8b6-4594-87b5-07c32f2a92a7 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN KV Shifting Attention Enhances Language Modeling
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68cc455c-744e-4eb2-b860-f4fe8f6ba44a · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN A Survey on Knowledge Distillation of Large Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50c6d867-dd34-4e29-b1ff-b321fb8aa9d0 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN ReFT: Representation Finetuning for Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8271824a-779a-48a6-9695-a63abe4ecf7c · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN LIMO: Less is More for Reasoning
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 037e2569-19b9-4245-8b92-657cd2b491a9 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Root mean square layer normalization
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7a9876b-4c3f-4e46-a80a-426c9a5170fa · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Qwen2.5 Technical Report
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dea9049-7e97-4eee-b831-8dcb2bf744cd · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f93021cc-42b0-443d-91f0-902852f32316 · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN AdaLoRA: Adaptive Budget Allocation for Parameter-Efficient Fine-Tuning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24cc940b-bbd8-451f-a97a-c0a53707b18a · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN LoRA: Low-Rank Adaptation of Large Language Models
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26176ccd-d43a-4b24-b6a3-2ac556b85e9f · outbound
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Embedding Trajectory for Out-of-Distribution Detection in Mathematical Reasoning
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
No inbound Pith citation observations are available.