Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T06:03:19.387482Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 1 inbound Pith citation observation for arXiv:2506.11104.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T06:03:19.387482Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:13:27.644720Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T05:13:28.150379Z
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 39b75541-84d0-436c-a687-1d832d3d8a51 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fd441fa5-0ea3-4132-a475-96c32bb8aacd · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3260fa7d-8bec-493d-a95d-2eb25927ded6 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0aa85b62-32dd-445b-8034-ac4fda8bcde0 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Longformer: The Long-Document Transformer
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a75ea7a-6630-4698-aa96-2d9059574388 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 109f8c1b-18d4-40e9-9e3d-309081fc6931 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration NACL: A General and Effective KV Cache Eviction Framework for LLMs at Inference Time
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5e93402-0bcf-4b35-aa7d-d1b0f66e2241 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Generating Long Sequences with Sparse Transformers
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ace05112-41fc-445a-8fff-53ea8b05380c · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Masked Language Modeling for Proteins via Linearly Scalable Long-Context Transformers
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ed08cd9-9e48-4ab4-a16c-a33e2b3f3383 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Adaptively Sparse Transformers
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72b51fed-4cae-4cf4-a006-764cc2c0b555 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f462fe05-3d87-4621-b1b1-f213bc3d6e00 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Multi-News: a Large-Scale Multi-Document Summarization Dataset and Abstractive Hierarchical Model
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a38fa083-ed9b-4ed8-9083-7a3cbef28524 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b593f040-8eea-4bca-9588-07bbd603a0e1 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Semsa: Semantic sparse attention is hidden in large language models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 38126815-345c-449c-97d8-caf3610bff33 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 11123903-98f1-4399-8696-71a83d51e448 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dfb9399-0a44-4a50-9657-3da18bb4e273 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration A Dynamic Head Importance Computation Mechanism for Neural Machine Translation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a9404664-de59-48ed-99fe-06cc174f009e · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Axial Attention in Multidimensional Transformers
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 964a02ee-e3b0-467c-b77d-b3cd16ae2297 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60a43518-9018-4fbb-a493-0f4020fed5a6 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Reformer: The Efficient Transformer
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77f0c218-0e94-4bb1-80cc-3958813350b3 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration LongEval: Guidelines for Human Evaluation of Faithfulness in Long-form Summarization
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dff2d4cb-47cf-4d2b-82dc-4e4ec51871a2 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration SnapKV: LLM Knows What You are Looking for Before Generation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5eacb4a-1b9d-4430-91be-00d18496f4a5 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Global Attention Mechanism: Retain Information to Enhance Channel-Spatial Interactions
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8386c03-f072-4bf1-82c1-b59efe1bbb23 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 47463b5e-831e-45b2-a3ba-4524dcb270d4 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b849773a-ef26-49c0-b491-92f618feca64 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 723e7c91-852a-4b51-baeb-20f365ac3591 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Lightweight and Efficient Neural Natural Language Processing with Quaternion Networks
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aee541e8-2104-418f-88f2-f80226d18deb · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Unresolved cited work
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53206f46-2ae0-49ae-9df8-68285976e6a3 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Multi-Head Self-Attention with Role-Guided Masks
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 78a2e262-6928-4375-ae3e-d37f98c80da9 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Improving Transformers with Dynamically Composable Multi-Head Attention
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ad6a6f0-2888-4cc0-bc53-f5b9ee758c26 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Efficient Streaming Language Models with Attention Sinks
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f770c5da-57e3-4456-a076-a60800a339bf · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration LayerKV: Optimizing Large Language Model Serving with Layer-wise KV Cache Management
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d96aaaba-19ec-46e6-b609-794c52746810 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration ChunkAttention: Efficient Self-Attention with Prefix-Aware KV Cache and Two-Phase Partition
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 582542cf-e636-4fcf-a262-cdc2a48eb119 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Unresolved cited work
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92cf65c7-5490-460f-8c4f-d8e12ac71434 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6af78f03-c580-41bb-a8e2-244859840541 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5814662f-8955-4c90-9230-7afdb02297a1 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration DiffKV: Differentiated Memory Management for Large Language Models with Parallel KV Compaction
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e15d8fcf-d490-405b-bfa2-f6f4f5218f1c · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac3977f9-50d3-4bac-a657-5d6f08e176a4 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0fa2375a-4a90-4325-82ad-523cb12c7c76 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration BUZZ: Beehive-structured Sparse KV Cache with Segmented Heavy Hitters for Efficient LLM Inference
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ad58d27-eea1-4510-8dc9-59d415f56a26 · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration SGLang: Efficient Execution of Structured Language Model Programs
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6411a1e-082c-4e41-a447-5238a7d44b9c · outbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration BatchLLM: Optimizing Large Batched LLM Inference with Global Prefix Sharing and Throughput-oriented Token Batching
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d0b4a51-b3ee-462f-9e79-f2da65fab369 · inbound
An Overview of Algorithms for Contactless Cardiac Feature Extraction from Radar Signals: Advances and Challenges DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.