Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T14:11:39.748585Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 0 inbound Pith citation observations for arXiv:2412.12465.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T14:11:39.748585Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
27 of 27 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f7a503cc-f96d-407b-9d26-69efd1364aca · outbound
Core Context Aware Transformers for Long Context Language Modeling Same as demonstrated in existing methods (Beltagy et al., 2020; Xiao et al., 2024b)
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b396aef2-ebad-4247-940c-1fdd7b936e8a · outbound
Core Context Aware Transformers for Long Context Language Modeling Extending Context Window of Large Language Models via Positional Interpolation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6846c3a-590f-4ab2-be44-6dd2f6cead4c · outbound
Core Context Aware Transformers for Long Context Language Modeling Masked Language Modeling for Proteins via Linearly Scalable Long-Context Transformers
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d528daf-9b2b-4e23-8339-e90b4599b224 · outbound
Core Context Aware Transformers for Long Context Language Modeling LongNet: Scaling Transformers to 1,000,000,000 Tokens
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f06184fe-5287-4e72-8eca-abde9a2d60e2 · outbound
Core Context Aware Transformers for Long Context Language Modeling Data Engineering for Scaling Language Models to 128K Context
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5aa06aaf-5711-4cca-85b1-fdfbe526fe3f · outbound
Core Context Aware Transformers for Long Context Language Modeling DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64ff5deb-44bb-45a3-beb8-45b7e5b24b31 · outbound
Core Context Aware Transformers for Long Context Language Modeling s 256 512 1024 2048 4096 PPL ↓ 2.98 2.92 2.86 2.79 2.73 Latency ↓ (ms) 457.4 460.1 461.4 462.8 473.1 Effect of Different Updating Strategies
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c2fed794-ad54-46dc-85a7-6ccf3308cfd2 · outbound
Core Context Aware Transformers for Long Context Language Modeling Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1601e86a-3514-4c59-8df0-9c6a8056f98f · outbound
Core Context Aware Transformers for Long Context Language Modeling GPT-4 Technical Report
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ff27d7e-ce5e-4b83-9e6b-e3a11d3f3680 · outbound
Core Context Aware Transformers for Long Context Language Modeling LLaMA: Open and Efficient Foundation Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3935e5ac-b689-4170-93a0-c63a56593019 · outbound
Core Context Aware Transformers for Long Context Language Modeling Qwen2.5 Technical Report
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a55eb01a-948a-456f-8c96-6ffa2b6848ba · outbound
Core Context Aware Transformers for Long Context Language Modeling The MMLU benchmark spans 57 diverse subjects, ranging from elementary mathematics to professional law
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d37893a3-37fc-46b9-acb0-76cc17b11a98 · outbound
Core Context Aware Transformers for Long Context Language Modeling base frequency
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 064f819b-70ea-4a43-9ac2-460d48dc0ec2 · outbound
Core Context Aware Transformers for Long Context Language Modeling This enables us to integrate our CCA-Attention as a standalone, cache-friendly operator, effectively eliminating redundant computations
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3518ab91-f9e6-4a8b-9662-049c403156cb · outbound
Core Context Aware Transformers for Long Context Language Modeling Our experiments are based on the LLaMA-2 7B model fine-tuned on sequences of length 32K and 80K (Fu et al., 2024)
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation da97c928-c70e-4f82-b2f3-acba2bf192c7 · outbound
Core Context Aware Transformers for Long Context Language Modeling 16 Core Context Aware Transformers for Long Context Language Modeling C
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f413f8ae-88d1-4e42-88b5-ddf611e7213e · outbound
Core Context Aware Transformers for Long Context Language Modeling Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f7b4178f-21bc-4802-9862-1a136bc3393e · outbound
Core Context Aware Transformers for Long Context Language Modeling Strategy Mean Pooling Max Pooling CCA-Attention (Ours) PPL ↓ 2.99 2.99 2.85 Effect of Group Size g
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ed662da6-82bc-4bd2-8177-f90118a375c6 · outbound
Core Context Aware Transformers for Long Context Language Modeling The perplexity rapidly converges within approximately the first 100 iterations and remains stable over 1,000 iterations
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0e2e3f84-3068-4f03-9a30-9a7649d6668a · outbound
Core Context Aware Transformers for Long Context Language Modeling Unresolved cited work
Reference 2000
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 23c79c42-8f28-4294-b24e-edb9a870f05a · outbound
Core Context Aware Transformers for Long Context Language Modeling DeepSeek-V3 Technical Report
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c43a855a-abd8-40e3-a063-3ab388aad145 · outbound
Core Context Aware Transformers for Long Context Language Modeling Retentive Network: A Successor to Transformer for Large Language Models
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 970488b0-c9e3-4845-b982-5811d154f108 · outbound
Core Context Aware Transformers for Long Context Language Modeling RULER: What's the Real Context Size of Your Long-Context Language Models?
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c96ee85-bc28-43e9-b57d-72b3dbe75c84 · outbound
Core Context Aware Transformers for Long Context Language Modeling LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e5a5456-13cc-4b3b-a090-19c543b30e14 · outbound
Core Context Aware Transformers for Long Context Language Modeling Longformer: The Long-Document Transformer
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 780698b2-f707-4bcb-a74d-c9aeed4745fe · outbound
Core Context Aware Transformers for Long Context Language Modeling Chang, Y ., Wang, X., Wang, J., Wu, Y ., Yang, L., Zhu, K., Chen, H., Yi, X., Wang, C., Wang, Y ., Ye, W., Zhang, Y ., Chang, Y ., Yu, P
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 81609661-bbeb-4667-b9fa-27153bd0abb3 · outbound
Core Context Aware Transformers for Long Context Language Modeling LM-Infinite: Zero-Shot Extreme Length Generalization for Large Language Models
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.