Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:37:13.242455Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2608.12419.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:37:13.242455Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 10ecf9ea-2e27-4767-99a6-8c3b2fde81fe · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining ChineseWebText: Large-scale High-quality Chinese Web Text Extracted with Effective Evaluation Model
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ffdd637-eb79-4e81-9bfe-41a66ce05bf8 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining Generating Long Sequences with Sparse Transformers
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8e9b861-ffff-48dd-aa7c-c523e445bb0d · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining WanJuan: A Comprehensive Multimodal Dataset for Advancing English and Chinese Large Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7c9018b-0724-4c60-876b-770efc6c2e16 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining Mixtral of Experts
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1277a8b-720f-455f-93c4-adf05d28d50b · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining LM2: Large Memory Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a88d323c-1b27-44f5-9edc-13d59e00b0c0 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining Cmmlu: Measuring massive multitask language understanding in chinese
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d4f6923e-5c28-47b0-88b4-46c00825abb2 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d674f108-fe07-4546-b148-6356289e6d4f · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining YAYI 2: Multilingual Open-Source Large Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bbb0316-3e14-4ae9-b64f-1085d797167f · outbound
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbefe8db-0d73-4a7d-bc6c-a1a9659ed2ef · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining D., Man, H., Ngo, N
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3972f525-b459-4225-aedb-07acd82cc0a9 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining GPT-4 Technical Report
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 083e52af-51bf-4f5d-9672-fc0c05802e3e · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining MemGPT: Towards LLMs as Operating Systems
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5c7008d-2b31-435d-bf08-e1f8ea2e907f · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining Dense Training, Sparse Inference: Rethinking Training of Mixture-of-Experts Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e57663f-f5c3-4323-8c67-1a5822d3e045 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining Hunyuan-Large: An Open-Source MoE Model with 52 Billion Activated Parameters by Tencent
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1af0f49-2fba-435d-94f6-2d4a7a6c8f49 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc1feda2-20a5-42ac-9a25-f641f915325e · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining Skywork: A More Open Bilingual Foundation Model
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e3d6708-b9fc-4208-b9f9-f5e2fda05e47 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining Skywork-MoE: A Deep Dive into Training Techniques for Mixture-of-Experts Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c22f5d8-e225-45e1-ba7e-79e2b7b284ce · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining Qwen2.5 Technical Report
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 825b28e8-4f0f-47bc-b8b7-d613da9a8ad1 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining MAP-Neo: Highly Capable and Transparent Bilingual Large Language Model Series
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d9df4fd-26f9-4f69-aff4-d974594451fb · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a366ad2c-0efd-4a46-9fbb-8d00dbd114e0 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining It contains 7,473 training and 1,319 hand-written test questions, each requiring two to eight sequential reasoning steps
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation caef4196-23c6-4ca0-a02d-71256b7aca81 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining Newton,” “Calculus,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 6953884a-d2cc-4c52-8824-39dea6a68f0a · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining Training Verifiers to Solve Math Word Problems
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a5b64e7-082e-4341-af5e-da75f2a443b5 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 220ca2b1-b40f-44bf-a4fb-4d2a54465cd1 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining InternLM2 Technical Report
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94cda244-ea9a-499e-a5a3-27ddf39a7204 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining ChatDB: Augmenting LLMs with Databases as Their Symbolic Memory
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af8ef260-a15d-4e14-91de-e2b3839592e9 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining We organize our supplementary as follows
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 84fed58e-b464-4c71-8c58-4dac96563597 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining Evaluating Large Language Models Trained on Code
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a340e0f-f967-49f9-b52a-5820bd73ee6e · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining Longformer: The Long-Document Transformer
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40db4f9e-3cbf-42d5-a60e-9c119a3c28d1 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ff30059-21b8-4350-8253-f25a27c34997 · outbound
LoKiFormer: Locality-aware Attention with Decoupled Knowledge Memory for Efficient Large Language Model Pretraining Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.