Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T22:32:56.444150Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 1 inbound Pith citation observation for arXiv:2606.07713.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T22:32:56.444150Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T13:08:09.010538Z
A source-named dated measurement, never combined with another source.
Source: cited_works
26 of 26 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3cce840d-df48-4bb0-a082-325f5c2d61b8 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels FATHOM: Fast attention through optimizing memory
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a35b2311-1aa3-49ac-a113-eb678af4a77a · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels MOSA: Matrix optimized self-attention hardware accelerator for mobile device
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ff700ae-f76d-4b9c-b841-57e88d6cbaff · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9e6bbb37-4863-4ed9-a3e4-4f72e56f162b · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6605622a-6c48-4fd0-b93a-e5e1ab789318 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels Hardware considerations for tensor implementation and analysis using the field programmable gate array
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bc413ab4-7cbb-4e9f-86cf-7049052967c2 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels Realizing mathematics of arrays operations as custom architecture hardware-software co-design solutions
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 16216584-1b2d-4b85-8a07-c0fa99c83335 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels Processing in memory for mathematics of arrays operations
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 892184b8-f20f-4656-958a-3a9e40760422 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels A Fast Optimization View: Reformulating Single Layer Attention in LLM Based on Tensor and SVM Trick, and Solving It in Matrix Multiplication Time
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1a516b0c-d999-4a5d-a791-1a5bb6b91bc9 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1055bb5-dbff-4801-914b-ad13cf3a9eb9 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels A ^3 : Accelerating attention mechanisms in neural networks with approximation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 318cae26-df65-49f6-b0e8-cce56e9e9e3a · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels Acceleration of fully connected layers on FPGA using the Strassen matrix multiplication
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b2f47cc-08a3-4c42-acf9-80b3ac870984 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels Design and Implementation of an FPGA-Based Hardware Accelerator for Transformer
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f7372ef7-b175-4c53-93e5-2bf4e08abd07 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels Research on matrix multiplication optimization and deployment method for heterogeneous platforms
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a5736dd-15d8-49d6-86c5-9d5b4fb35d16 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels Array access and performance regarding numerical algorithms
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e41d9201-4449-4f98-9682-a16cd1455ac3 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels A Mathematics of Arrays
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de69877b-f1c1-409c-9965-6f0457587a40 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels From array algebra to energy efficiency on GPUs: Data and hardware shapes with dimension-lifting to optimize memory-processor layouts
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dc789cdc-e395-4774-b16e-ecbd46091da6 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels Towards automatic, predictable and high-performance parallel code generation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82f8b2be-9d72-4e49-927f-43f4e20e1efe · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels New mathematics for computer performance: array algebra and cost functions
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 288b6404-5124-4727-8712-cc421e9ab0a0 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels OPTIMUS: Optimized matrix multiplication structure for transformer neural network accelerator
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 411e3766-809e-4e95-b3ab-8e7a5647c639 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels FACT: FFN-attention co-optimized transformer architecture with eager correlation prediction
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebbeb976-a8bf-4c79-a53a-404cc2581dbf · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels High-performance Gemmini-based matrix multiplication accelerator for deep learning workloads
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b09d5781-1cb5-4a9b-9b76-5dc65048c819 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels Design and implementation of a BRAM-banked double-buffered matrix multiplication accelerator for transformer models on edge FPGAs
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b706a879-9be6-464d-a411-9e548f1f4de6 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels Improving the performance of DGEMM with MoA and cache-blocking
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cedba33-dad1-4d9a-9f92-2e71ebd5d115 · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels Threaded multicore GEMM with MoA and cache-blocking
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01cf1fa2-2842-4610-afb9-c30081b9588a · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels Attention is all you need
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 357f9832-6c53-4e8a-ad71-fc104d91e53f · outbound
Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels Hardware friendly transformer optimization with dynamic attention matrix fusion
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 487c3671-2490-403b-b5b6-ecbcad1095be · inbound
MoA-Structured Decode Attention DNF Derivation, KV-Cache Accumulation, GQA/MQA, and OpenACC Kernel Attention at the Theoretical Minimum: A Mathematics of Arrays Framework for Memory-Optimal Transformer Kernels
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.