Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:30:57.837571Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 1 inbound Pith citation observation for arXiv:2506.13996.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:30:57.837571Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T21:11:48.095184Z
A source-named dated measurement, never combined with another source.
Source: cited_works
17 of 17 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 40367eda-1cf8-4460-8d22-17d7ec9af6c2 · outbound
Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences Ring Attention with Blockwise Transformers for Near-Infinite Context
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97957d92-9f4b-4298-bd7f-1b37b5bb3d3a · outbound
Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05cf9cdb-0e18-47b3-8b67-145d5b4307f4 · outbound
Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences Reducing Activation Recomputation in Large Transformer Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9786b61f-a944-441d-8011-a71f08b31f5f · outbound
Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences Sequence Parallelism: Long Sequence Training from System Perspective
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4684dc0-5dd9-423f-b97d-1dde205943c9 · outbound
Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences DISTFLASHATTN: Distributed Memory-efficient Attention for Long-context LLMs Training
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cf2c8ee-c34e-4f71-a054-6050612b2b61 · outbound
Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences Striped Attention: Faster Ring Attention for Causal Transformers
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d9e0298-ab80-4bb6-8aba-7eadca48f60d · outbound
Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d33a8a5-0944-4486-9b2e-f980395fd95f · outbound
Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences USP: A Unified Sequence Parallelism Approach for Long Context Generative AI
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56cbb9a8-b30c-4123-8fb3-a696124d50ce · outbound
Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences LoongTrain: Efficient Training of Long-Sequence LLMs with Head-Context Parallelism
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ba52778-8e90-491f-a35d-e8bedcd57bde · outbound
Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences Liger Kernel: Efficient Triton Kernels for LLM Training
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd1ff85e-5006-4493-ac35-44d4b275336f · outbound
Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences Datasets: A community library for natural language processing,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fd638e83-7865-4f24-b9b6-1e00006d7cc6 · outbound
Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences ArcticTraining: Simplifying and accelerating post-training for large language models,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 622566c1-dd5c-448b-9a97-c4efac6fe13b · outbound
Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences Deepspeed: System optimizations enable training deep learning models with over 100 billion parameters,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca5d832f-b92b-41ae-a63e-3db5732e2929 · outbound
Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences Transformers: State-of-the-art natural language processing,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 26203125-424d-4dc8-ac48-05c8314a0a3d · outbound
Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences ZeRO: Memory Optimizations Toward Training Trillion Parameter Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19ccebd1-f9c0-45b0-91fe-236b9f080d5e · outbound
Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences BLOOM: A 176B-Parameter Open-Access Multilingual Language Model
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8d73282-b716-4528-85b6-a516858600f3 · outbound
Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences The Case for Co-Designing Model Architectures with Hardware
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f68b75ca-1d76-4926-b339-af583cf09ac8 · inbound
Untied Ulysses: Memory-Efficient Context Parallelism via Headwise Chunking Arctic Long Sequence Training: Scalable And Efficient Training For Multi-Million Token Sequences
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.