Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:24:00.072659Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2507.02851.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:24:00.072659Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
21 of 21 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d1ab79bf-bfa9-4f86-8244-f623cc80202b · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs Training Verifiers to Solve Math Word Problems
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bbb6bad-8780-49ae-8cff-76d35cbd5203 · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10d88f10-faae-461e-92be-b5b33ae3038e · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs Measuring Mathematical Problem Solving With the MATH Dataset
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eeac5b9b-b6d6-4df1-8850-e86650132768 · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs Lost in the Middle: How Language Models Use Long Contexts
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25bb37db-d3ba-485c-b145-102e33531c4c · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs Distributed Mixture-of-Agents for Edge Inference with Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13df14a6-0ba3-4bfd-9742-b2bafe2e196f · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs s1: Simple test-time scaling
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e128e2a-9d15-41db-8ed4-447e90630a8e · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54dd50c3-df24-42f5-b8f8-ec4744002cf9 · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs Not All Thoughts are Generated Equal: Efficient LLM Reasoning via Multi-Turn Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f21235f-aaf2-4b8b-93d7-aa0cc0404c4c · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs Iteration of Thought: Leveraging Inner Dialogue for Autonomous Large Language Model Reasoning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5624c9c5-cef8-4112-969c-fec3d4d8411b · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4acbca54-faff-4b52-9264-555fd99b6fe8 · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec8f81c2-c71e-48f4-8be7-a5f134b74053 · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs Gemini: A Family of Highly Capable Multimodal Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c107b09-0a26-4ad1-b454-e6d62fdca359 · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs Think Twice: Enhancing LLM Reasoning by Scaling Multi-round Test-time Thinking
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5b6da0e-5c72-4e63-9fd5-00f508c0bc21 · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs Mixture-of-Agents Enhances Large Language Model Capabilities
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e115bb4-420d-4c02-b88a-a10199752927 · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs Inftythink: Breaking the length limits of long-context reasoning in large language models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19b1f7ff-1647-42a6-8aef-00e65e33485f · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs Infinite Retrieval: Attention Enhanced LLMs in Long-Context Processing
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d697b007-2cf2-4794-be4f-2283440f5ce2 · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1673fa13-dfda-421b-8440-b4887d3289ba · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs OpenAI o1 System Card
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd220da0-e9ad-43e9-b6fd-e1996fb0e4de · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs Synthetic Data Generation & Multi-Step RL for Reasoning & Tool Use
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f67615ef-5366-4947-a05b-f5a7327e10e0 · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs SMoA: Improving Multi-agent Large Language Models with Sparse Mixture-of-Agents
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 127a6973-4793-4e5f-bb46-1190510a15f9 · outbound
MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs Evaluating Large Language Models Trained on Code
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.