Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:20:45.911627Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2507.03340.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:20:45.911627Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
13 of 13 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 388c28c0-eb93-44e9-99a8-8cdb77e13fa6 · outbound
Degrees of Freedom for Linear Attention: Distilling Softmax Attention with Optimal Feature Efficiency An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f861ce85-a9de-4744-a2b0-b7debd850260 · outbound
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 78651683-8be8-495c-9bc2-70f2f9278298 · outbound
Degrees of Freedom for Linear Attention: Distilling Softmax Attention with Optimal Feature Efficiency direct” loss. This is natural because the cross entropy loss for next-token prediction is used in “direct
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9bcdbe42-d596-49eb-b26f-969247eb8437 · outbound
Degrees of Freedom for Linear Attention: Distilling Softmax Attention with Optimal Feature Efficiency Linformer: Self-Attention with Linear Complexity
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 643d0de8-beb3-415d-bac1-0a91e45281a8 · outbound
Degrees of Freedom for Linear Attention: Distilling Softmax Attention with Optimal Feature Efficiency • K : Rd × Rd → R is the positive definite kernel given byK(x, y) = Ez∼τ [ϕ(x; z)ϕ(y; z)], where τ is a probability measure on a measurable setZ, and ϕ : Rd × Z →R is a feature map
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 989f48a6-8f41-469b-9540-2095b8d385eb · outbound
Degrees of Freedom for Linear Attention: Distilling Softmax Attention with Optimal Feature Efficiency Scavenging Hyena: Distilling Transformers into Long Convolution Models
Reference 2007
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96359c09-3a7e-4c60-820f-940e4e64484c · outbound
Degrees of Freedom for Linear Attention: Distilling Softmax Attention with Optimal Feature Efficiency LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning
Reference 2011
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9165664-45c3-4da1-9cbd-34d955ad0e27 · outbound
Degrees of Freedom for Linear Attention: Distilling Softmax Attention with Optimal Feature Efficiency The Mamba in the Llama: Distilling and Accelerating Hybrid Models
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cd2de6b-795d-4078-a062-e6e4f1d3cbac · outbound
Degrees of Freedom for Linear Attention: Distilling Softmax Attention with Optimal Feature Efficiency Unresolved cited work
Reference 2018
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fba1e96d-fd8e-4112-97a3-ba5626d73da2 · outbound
Degrees of Freedom for Linear Attention: Distilling Softmax Attention with Optimal Feature Efficiency Sakamoto and K
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4920c157-511e-4a80-ad76-b027ccf4f504 · outbound
Degrees of Freedom for Linear Attention: Distilling Softmax Attention with Optimal Feature Efficiency DiJiang: Efficient Large Language Models through Compact Kernelization
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78264015-2891-49f7-8732-d3f03fdaa226 · outbound
Degrees of Freedom for Linear Attention: Distilling Softmax Attention with Optimal Feature Efficiency Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70cf876f-4c6d-4a79-a797-30c1114bd463 · outbound
Degrees of Freedom for Linear Attention: Distilling Softmax Attention with Optimal Feature Efficiency Ravichandran, A
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
No inbound Pith citation observations are available.