Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-17T03:40:25.706499Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 48 inbound Pith citation observations for arXiv:2504.21318.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-17T03:40:25.706499Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T00:48:28.952743Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T12:15:01.137692Z
64 of 64 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 884eb5c4-2c6f-4fcb-8561-c79b42985f6b · outbound
Phi-4-reasoning Technical Report Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 70f4f4a0-2245-4405-8cbe-5b2ae0de7fa4 · outbound
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0c522f1b-f36d-4fed-8c6f-7a155b72c24a · outbound
Phi-4-reasoning Technical Report KITAB: evaluating llms on constraint satisfaction for information retrieval
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 405b4788-a173-499f-9cfe-f84552267446 · outbound
Phi-4-reasoning Technical Report Aime 83-24
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b3c289d0-e7da-4350-b714-d80962887c12 · outbound
Phi-4-reasoning Technical Report Aime 2025
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1902c78c-a273-469b-a505-0c7b317d526d · outbound
Phi-4-reasoning Technical Report Concrete Problems in AI Safety
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f20ae402-9939-4ef4-8cab-0cfeeb0f4147 · outbound
Phi-4-reasoning Technical Report Claude 3.7 sonnet.https://www.anthropic.com/news/claude-3-7-sonnet
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9194b23c-8816-44fa-a0e8-38533ac977b9 · outbound
Phi-4-reasoning Technical Report Chain-of-Thought Reasoning In The Wild Is Not Always Faithful
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 14ac81e6-d67c-4614-990e-1fb3a65762d9 · outbound
Phi-4-reasoning Technical Report Eureka: Evaluating and Understanding Large Foundation Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 59c63ade-c704-4bfb-a7bb-4bcd342082f9 · outbound
Phi-4-reasoning Technical Report Inference-Time Scaling for Complex Tasks: Where We Stand and What Lies Ahead
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5e85f7d8-de6a-40db-82a5-23f435a81586 · outbound
Phi-4-reasoning Technical Report Matharena: Evaluating llms on uncontaminated math competitions, February 2025
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cbe872af-1240-4f64-9c6e-3f24462e89c1 · outbound
Phi-4-reasoning Technical Report Designing disaggregated evaluations of ai systems: Choices, considera- tions, and tradeoffs
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4b9ba938-2cae-4a2f-92a9-1091500c1dd9 · outbound
Phi-4-reasoning Technical Report Benchagents: Automated benchmark creation with agent interaction
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 975589be-8544-4620-88b4-f511fbc5888c · outbound
Phi-4-reasoning Technical Report Extending Context Window of Large Language Models via Positional Interpolation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation aa6524b6-d3ce-4db7-8f1a-ed4512309cd7 · outbound
Phi-4-reasoning Technical Report Reinforcement learning for reasoning in small llms: What works and what doesn’t
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 26d8073d-d36d-46a8-a2b1-0b2749b72f0c · outbound
Phi-4-reasoning Technical Report Reinforcement learning for reasoning in small llms: What works and what doesn’t
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a0d419ea-9ae3-4750-a7f9-0201382ee9fd · outbound
Phi-4-reasoning Technical Report Omni-math: A universal olympiad level mathematic benchmark for large language models.ICLR
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b9aef3d6-f435-4f7a-9fc4-0d4910ffe383 · outbound
Phi-4-reasoning Technical Report Scaling laws for reward model overoptimization
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 708c5db8-b9a9-4c31-8b02-4d88848d1fe1 · outbound
Phi-4-reasoning Technical Report Gemini flash thinking
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 14c552b0-c3f1-4f16-aa6b-e44d33d2395b · outbound
Phi-4-reasoning Technical Report rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d2038791-e7e6-475e-aa8e-2a79b644e4f3 · outbound
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation caaa20f3-36fb-4418-b8d3-d12c69e68d78 · outbound
Phi-4-reasoning Technical Report DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e1d81e83-35bb-4a4b-a9df-4b5d3e4ed44b · outbound
Phi-4-reasoning Technical Report Computers and intractability: a guide to the theory of np-completeness (michael r
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4af62f25-e51d-447f-b38d-9b1c5f2f522a · outbound
Phi-4-reasoning Technical Report ToxiGen: A large- scale machine-generated dataset for adversarial and implicit hate speech detection
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0c7d2f1c-30bd-4bd3-aa73-70f7fecdc5c2 · outbound
Phi-4-reasoning Technical Report Measuring Massive Multitask Language Understanding
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b3cdae9f-de2f-470d-a3ca-7c7a8245630f · outbound
Phi-4-reasoning Technical Report A sober look at progress in language model reasoning: Pitfalls and paths to repro- ducibility.arXiv preprint arXiv:2504.07086
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 97fc3336-691c-4ecb-9bfe-058c13bf4aeb · outbound
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 879aa7e9-f46d-4dd6-891b-6f41031ca8e6 · outbound
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 77fb31cf-ce3f-4c8f-a2c5-06955ff09624 · outbound
Phi-4-reasoning Technical Report Phi-2: The surprising power of small language models.Microsoft Research Blog
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 23a33d8e-2625-4bff-9419-3d0866e4a02e · outbound
Phi-4-reasoning Technical Report SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 650992ac-952a-4c66-bf10-937a05982ffe · outbound
Phi-4-reasoning Technical Report Same task, more tokens: the impact of input length on the reasoning performance of large language models
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1ae656e7-6a29-4e3b-9f02-e37a610bfacb · outbound
Phi-4-reasoning Technical Report Exaone deep: Reasoning enhanced language models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 06acca0a-4a6b-4ef4-adf2-8e2a0bccc721 · outbound
Phi-4-reasoning Technical Report Functional Interpolation for Relative Positions Improves Long Context Transformers
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 09202a5b-c0cc-48c2-9a0a-a4cd1f7c70c5 · outbound
Phi-4-reasoning Technical Report From Crowdsourced Data to High-Quality Benchmarks: Arena-Hard and BenchBuilder Pipeline
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 22d88ba6-9fb2-4916-9253-2b9225e41f06 · outbound
Phi-4-reasoning Technical Report Limr: Less is more for rl scaling
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 067e08a1-e275-43f4-b043-771bcfa4cf1f · outbound
Phi-4-reasoning Technical Report Is your code generated by chatgpt really correct? rigorous evaluation of large language models for code generation
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation caa1ec0a-9bce-413f-a700-0b243edd05ba · outbound
Phi-4-reasoning Technical Report Tang, Manan Roongta, Colin Cai, Jeffrey Luo, Li Erran Li, Raluca Ada Popa, and Ion Stoica
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8b4f0a72-814c-4c09-bcf3-d5361d44a8f9 · outbound
Phi-4-reasoning Technical Report A framework for automated measurement of responsible ai harms in generative ai applications
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 05cb8000-2abd-43a7-a6f0-9ed37764ac97 · outbound
Phi-4-reasoning Technical Report A Framework for Automated Measurement of Responsible AI Harms in Generative AI Applications
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bcab39d6-4e6a-41b1-830e-d5d23cf06dd6 · outbound
Phi-4-reasoning Technical Report Orca 2: Teaching Small Language Models How to Reason
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 48bba3f0-f6db-4db6-a44c-0abcb5d01cab · outbound
Phi-4-reasoning Technical Report AgentInstruct: Toward Generative Teaching with Agentic Flows
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 62b830a3-b5e2-40bc-92a7-693df01b60cb · outbound
Phi-4-reasoning Technical Report Unearthing skill-level insights for understanding trade-offs of foundation models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c3ce1455-2c71-4e32-9e01-7a65a732e463 · outbound
Phi-4-reasoning Technical Report Orca: Progressive Learning from Complex Explanation Traces of GPT-4
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b9e2937a-24b1-4aaf-870d-206decf3c76a · outbound
Phi-4-reasoning Technical Report Towards accountable ai: Hybrid human-machine analyses for character- izing system failure
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3bb89c3a-e54e-48bf-ac37-2afb70abcbc3 · outbound
Phi-4-reasoning Technical Report Openai o3-mini system card
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation eda276f5-855e-4d1e-abfc-7d7e3999e968 · outbound
Phi-4-reasoning Technical Report Computational complexity
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2935f4e6-a434-4fec-a593-134d636562c4 · outbound
Phi-4-reasoning Technical Report Overreliance on ai literature review.Microsoft Research, 339:340
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2ef8e4b6-d6f6-4472-a70a-31803eb8d840 · outbound
Phi-4-reasoning Technical Report Proof or Bluff? Evaluating LLMs on 2025 USA Math Olympiad
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c6d2135e-2de7-4c80-98ca-708e03964762 · outbound
Phi-4-reasoning Technical Report Gpqa: A graduate-level google-proof q&a benchmark
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 07f4864d-994d-4a8c-bde5-534c44c89efe · outbound
Phi-4-reasoning Technical Report DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dae5687a-7da1-497f-b6c0-301c4720db1e · outbound
Phi-4-reasoning Technical Report HybridFlow: A Flexible and Efficient RLHF Framework
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d765b8ae-949d-4062-a282-f98a624ff0c2 · outbound
Phi-4-reasoning Technical Report Language Models are Multilingual Chain-of-Thought Reasoners
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 702cef1a-b9e6-492f-8b63-0a6cea967aed · outbound
Phi-4-reasoning Technical Report RoFormer: Enhanced Transformer with Rotary Position Embedding
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7d51f35e-89ac-4b36-ade0-9ed494a3c128 · outbound
Phi-4-reasoning Technical Report Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation af2cc9d9-95e0-46d5-98f6-78193ca15954 · outbound
Phi-4-reasoning Technical Report Open Thoughts
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2b4ff737-13ae-4ac6-a000-624913b1a9c2 · outbound
Phi-4-reasoning Technical Report Qwq-32b: Embracing the power of reinforcement learning, March 2025
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 41c0b3e2-e67e-4c59-800d-c2526b0c2d1d · outbound
Phi-4-reasoning Technical Report Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ed7a1c04-96b3-4eb0-a788-2bce9ef46613 · outbound
Phi-4-reasoning Technical Report Is a picture worth a thousand words? delving into spatial reasoning for vision language models.Advances in Neural Information Processing Systems, 37:75392–75421
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 21a94ca9-336c-4565-b90e-f089ea37fe87 · outbound
Phi-4-reasoning Technical Report MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8307ea11-0e5e-435d-a947-f20e74869e07 · outbound
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9c08c4c8-1eff-47cd-8605-016f003dce04 · outbound
Phi-4-reasoning Technical Report On the Emergence of Thinking in LLMs I: Searching for the Right Intuition
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e28cb773-4c87-4c67-a3e1-afcfa5363cf0 · outbound
Phi-4-reasoning Technical Report Demystifying Long Chain-of-Thought Reasoning in LLMs
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e6fd1974-f1f3-4512-9ee8-bb2b60627811 · outbound
Phi-4-reasoning Technical Report Dapo: An open-source llm reinforcement learning system at scale
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 06a25481-bea5-48ac-a61f-9144bad29121 · outbound
Phi-4-reasoning Technical Report Instruction-Following Evaluation for Large Language Models
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation df745f2f-0866-4f33-b7f1-cbe7fbe5e9fa · inbound
Reinforcement Learning from Human Feedback Phi-4-reasoning Technical Report
Reference 170
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3f78d8d7-b2a3-4632-9083-f15a89340d50 · inbound
MathArena: Evaluating LLMs on Uncontaminated Math Competitions Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8a857573-5dcc-4e4a-8842-ec8372391689 · inbound
Do LLMs Overthink Basic Math Reasoning? Benchmarking the Accuracy-Efficiency Tradeoff in Language Models Phi-4-reasoning Technical Report
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ca875656-cd14-4e80-81b1-8a0755f4523c · inbound
ReasonCache: Accelerating Large Reasoning Model Serving through KV Cache Sharing Phi-4-reasoning Technical Report
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a4710964-eb33-481d-ab96-754f1a9cb7f5 · inbound
A Survey of Reinforcement Learning for Large Reasoning Models Phi-4-reasoning Technical Report
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4c1e6eb3-6977-47c4-9c41-5057cd7f5a15 · inbound
Visual Reasoning Agent: Robust Vision Systems in Remote Sensing via Inference-Time Scaling Phi-4-reasoning Technical Report
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8a051f55-631c-4fe2-ba5a-ca5b49fa7303 · inbound
Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation Phi-4-reasoning Technical Report
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5d21b7cb-1dac-4395-ba9a-d8826d8c2876 · inbound
Who Endorsed It? Measuring Authority Bias Across Expertise Levels in Language Models Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43239800-29bb-4948-ab42-335f3500ecd2 · inbound
Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Alignment Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 99465e2b-03fe-4e2a-8ef5-ac45a23b13bb · inbound
Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Alignment Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62bd258a-3b18-4a7d-9287-0db06f442bc8 · inbound
The Geometric Reasoner: Manifold-Informed Latent Foresight Search for Long-Context Reasoning Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1e350365-2b6d-4b50-8e89-17be021f38e5 · inbound
Do Not Waste Your Rollouts: Recycling Search Experience for Efficient Test-Time Scaling Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation aad44ffe-f341-491e-aca1-c56bf64094cf · inbound
Flexible Entropy Control in RLVR with a Gradient-Preserving Perspective Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 313f32af-5eab-45e7-b000-c5137af0bc82 · inbound
Learning to Evict from Key-Value Cache Phi-4-reasoning Technical Report
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba6c202a-3788-4db8-9695-0408dfbc8dca · inbound
Author-in-the-Loop Response Generation and Evaluation: Integrating Author Expertise and Intent in Responses to Peer Review Phi-4-reasoning Technical Report
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a55d5afc-cbfb-46a5-be08-5fe0e1c730f4 · inbound
Ranking Reasoning LLMs under Test-Time Scaling Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c2f5d83b-9b0d-4ab0-a196-c9f7d925ff05 · inbound
Contrastive Reasoning Alignment: Reinforcement Learning from Hidden Representations Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 17b61778-41a9-4fb5-bb1b-cfbda779c868 · inbound
CoME-VL: Scaling Complementary Multi-Encoder Vision-Language Learning Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7c3d0251-d398-4903-ad22-b42f4609d85d · inbound
Unified Deployment-Aware Evaluation of Open Reasoning Language Models Phi-4-reasoning Technical Report
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6c6f73d2-f14d-4bed-b812-4956cf574142 · inbound
Unified Deployment-Aware Evaluation of Open Reasoning Language Models Phi-4-reasoning Technical Report
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 31225b90-44ea-47b4-bbf4-f7091b96764d · inbound
ZeroCoder: Can LLMs Improve Code Generation Without Ground-Truth Supervision? Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 373706ed-23ef-412c-a036-4250ad0f5a2b · inbound
SeLaR: Selective Latent Reasoning in Large Language Models Phi-4-reasoning Technical Report
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b5991c9b-da4e-4930-94de-065ab2571ab5 · inbound
When AI Models Become Dependencies: Studying the Evolution of Pre-Trained Model Reuse in Downstream Software Systems Phi-4-reasoning Technical Report
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9b53877c-5c07-4ecd-95b9-a1dc1add3399 · inbound
VLM Judges Can Rank but Cannot Score: Task-Dependent Uncertainty in Multimodal Evaluation Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 53a9fc33-c0a5-4561-aaa9-d6953f4530d4 · inbound
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2e4d2d91-0987-4f84-b766-cc5cca0c8ffb · inbound
Beyond Benchmarks: MathArena as an Evaluation Platform for Mathematics with LLMs Phi-4-reasoning Technical Report
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ef6863f1-78cd-43e5-ade2-5797199b0003 · inbound
Beyond Benchmarks: MathArena as an Evaluation Platform for Mathematics with LLMs Phi-4-reasoning Technical Report
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9c0f06a0-e1fe-49c9-9cc7-5f7779f47031 · inbound
Distilling Long-CoT Reasoning through Collaborative Step-wise Multi-Teacher Decoding Phi-4-reasoning Technical Report
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 79a9a8d6-b714-494b-bb18-01b84c5dda6c · inbound
Chain-of-Thought Reasoning Enhances In-Context Learning for LLM-Based Mobile Traffic Prediction Phi-4-reasoning Technical Report
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6acf9e8c-0c6a-46a8-87b0-0b2e76b51ec4 · inbound
TwiSTAR:Think Fast, Think Slow, Then Act,Generative Recommendation with Adaptive Reasoning Phi-4-reasoning Technical Report
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b38ec3d3-e27d-4ceb-9e5d-0d87f900f57f · inbound
Artificial Intolerance: Stigmatizing Language in Clinical Documentation Skews Large Language Model Decision-Making Phi-4-reasoning Technical Report
Reference 107
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 26af29a6-5d40-4cc0-8069-809003dc0b7c · inbound
TRACE: Trajectory Correction from Cross-layer Evidence for Hallucination Reduction Phi-4-reasoning Technical Report
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1f7a6f04-65d1-4fd2-b627-b0c4b1db699a · inbound
Diagnosing Multi-step Reasoning Failures in Black-box LLMs via Stepwise Confidence Attribution Phi-4-reasoning Technical Report
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fd29940f-c883-473c-98cf-fbdb501a3432 · inbound
CopT: Contrastive On-Policy Thinking with Continuous Spaces for General and Agentic Reasoning Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9f0cc0ec-4644-405c-ae32-c1e8124165df · inbound
OPPO: Bayesian Value Recursion for Token-Level Credit Assignment in LLM Reasoning Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3c9db0b1-dd96-4e22-bbf4-4e04e3685f9f · inbound
OPPO: Bayesian Value Recursion for Token-Level Credit Assignment in LLM Reasoning Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7dbd29ac-5ccb-466c-8c22-9e5bdf90b7a5 · inbound
Trust Region On-Policy Distillation Phi-4-reasoning Technical Report
Reference 235
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0dab8cde-a0a2-4e65-ba22-f98d44567788 · inbound
Rethinking Molecular Text Representations for LLMs: An Empirical Study Phi-4-reasoning Technical Report
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 731654d1-56e4-40de-9631-570d2e7fbbf6 · inbound
When to Think Deeply: Inhibitory Deliberation for LLM Reasoning Phi-4-reasoning Technical Report
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation aa0a2c45-257b-469e-9752-70b6c9fd9a24 · inbound
RealMath-Eval: Why SOTA Judges Struggle with Real Human Reasoning Phi-4-reasoning Technical Report
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fc87d099-da85-4bec-8d63-b1f8c266294b · inbound
Attention Amnesia in Hybrid LLMs: When CoT Fine-Tuning Breaks Long-Range Recall, and How to Fix It Phi-4-reasoning Technical Report
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7d1d8372-8a79-4d8e-9e06-6e1433bf4fa3 · inbound
Every Act Has Its Price: Compressed Moral Composition in Frontier LLMs Phi-4-reasoning Technical Report
Reference 109
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 23584aa2-013b-4936-9d59-1e74d3f2c407 · inbound
From Trainee to Trainer: LLM-Designed Training Environment for RL with Multi-Agent Reasoning Phi-4-reasoning Technical Report
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ab31bb19-93ae-4bcc-a5bc-98c156ae0781 · inbound
ThinkProbe: Beyond Accuracy -- Structural Profiling of Open-Ended LLM Reasoning Traces via Non-Generative Thought Graphs Phi-4-reasoning Technical Report
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 71834db7-0c7d-4d5c-b9db-c7e3ed0bd1ab · inbound
Benchmarking Large Language Models on Floating-Point Error Classification Phi-4-reasoning Technical Report
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a40a5c98-377b-44d0-ad5a-56ecf542cd22 · inbound
On the Systematic Challenges of Culturally Loaded Machine Translation: Dream of the Red Chamber as the Cultural Lens Phi-4-reasoning Technical Report
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a1b21e5-e5f4-4521-aeb6-f6de3eb6b26f · inbound
SmartGen: Seamless Disaggregated LLM Inference with Selective KV Cache Transfer Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35da46c7-9b11-403c-b1cb-b2cdd374ecdf · inbound
Learning to Coordinate Symbolic Tools: LLM Agents for Verified Sum-of-Squares Certificates Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.