Pith. sign in

Paper Citation Record · LEDGER

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization

As of 12 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2608.09160.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.09160 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T22:26:47.452553Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact1
  • verified fuzzy8
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 34d772e3-0810-4038-bac4-75a80fa86989 · outbound

This paper cites Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T22:26:47.350167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:26:47.350167Z digest=sha256:28aea91b4f355987b9f9246fa8d44b47418e8218f1774ed75c803fcc143b3799

Observation d0006a06-aef3-4d8a-8c54-17bf8aa7f4e7 · outbound

This paper cites Efficient memory management for large language model serving with pagedattention,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Efficient memory management for large language model serving with pagedattention,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.918531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.362615Z digest=sha256:a4fafbb9f951c82905e939f819baa9cdbf59e40935a8c1902a58f44cc8acc066

Observation f74c1e6c-f28a-4de2-8e5c-1af97a64e03a · outbound

This paper cites FLUX: Fast Software-based Communication Overlap On GPUs Through Kernel Fusion.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization FLUX: Fast Software-based Communication Overlap On GPUs Through Kernel Fusion

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T22:26:47.368308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:26:47.368308Z digest=sha256:dbccda91eefad78312744b7d18169a35640e12e526924f4bcb74067508cf9d0d

Observation 2cf554a0-4d14-45a4-926d-5e995d619ba1 · outbound

This paper cites Flashoverlap: A lightweight design for efficiently overlapping communication and computation,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Flashoverlap: A lightweight design for efficiently overlapping communication and computation,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.894254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.374690Z digest=sha256:a02defa8a287a042e418586fa052d9024192401587ce15a09d5a3f88748327fb

Observation 9c524e33-19a0-4916-b461-40bddfded5b8 · outbound

This paper cites Scaling vision transformers to 22 billion parame- ters,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Scaling vision transformers to 22 billion parame- ters,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.874911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.381371Z digest=sha256:ef716c764d9e6c71b3b7e1d7918800287b01b8cd30fb3a27873266df31950ddd

Observation c08adcf2-244e-465f-8c08-cbcd21b1c191 · outbound

This paper cites 2 OLMo 2 Furious.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization 2 OLMo 2 Furious

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T22:26:47.388569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:26:47.388569Z digest=sha256:6f89008a9e038b88c622be55c07b43fdc62ae9be9cd06e267fafb6b071eba283

Observation be5b1ea6-bfe2-453f-8ffa-8188c1d28afb · outbound

This paper cites Olmoe: Open mixture-of-experts language models,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Olmoe: Open mixture-of-experts language models,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.856811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.397767Z digest=sha256:f7c9d3af6f78d9749a21037a0a669f6f7f08927782db6aa78f8cf86c6f751bd9

Observation 59df8745-5479-4aae-af06-4947d9697124 · outbound

This paper cites Olmo 3.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Olmo 3

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T22:26:47.409165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:26:47.409165Z digest=sha256:568b6deb87989adbd8046d571782333607ad65f7bd1609e8b092c239e11d96b4

Observation 57198c9f-e150-493f-a4e0-577bf2bb7b1f · outbound

This paper cites Root mean square layer normalization,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Root mean square layer normalization,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T22:26:47.417005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:26:47.417005Z digest=sha256:2a7af0610ad8ffb93572d0a5a3a6a6a838070840155433c42da1d32e5c32f178

Observation ba6b7a54-91b3-4550-b68c-25f2b8952c90 · outbound

This paper cites Gpipe: Efficient training of giant neural networks using pipeline parallelism,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Gpipe: Efficient training of giant neural networks using pipeline parallelism,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.817503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.421704Z digest=sha256:c220453da2d4b93f15358d878a58590b3db4a5ab518d124fe1040d11095e8eea

Observation 42d54c11-bb7e-489f-a017-121d854e70e8 · outbound

This paper cites ShareGPT Dataset,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization ShareGPT Dataset,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.794484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.432274Z digest=sha256:b59aebda8ec7c6c0b6104e9b276ba1f2e8bf90cb2556c8c92eecfa4c800588a1

Observation 8042f586-10b6-47cd-ac78-0c76ead1036e · outbound

This paper cites Olmo: Accelerating the science of language models,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Olmo: Accelerating the science of language models,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.764870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.440748Z digest=sha256:894f582d94faaf818cb2d3162b5c4f3e1e2cec3644aa54efebe87df0bcadb90c

Observation a020fbaf-f1b3-4106-88fe-674c6b455372 · outbound

This paper cites MiniMax TP RMSNorm,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization MiniMax TP RMSNorm,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.745497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.446922Z digest=sha256:aa2d775279b6f1e9f476a7f11bfaa392d17dcdda576dd1966d1c69ed97cb8ad1

Observation 394b9cdf-c0c2-4376-8592-aaf7eb64f993 · outbound

This paper cites Simplegpt: Improving gpt via a simple normalization strategy,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Simplegpt: Improving gpt via a simple normalization strategy,

Reference 14

Resolution
verified exact
raw_fallback, observed 2026-08-11T22:26:47.630204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.452553Z digest=sha256:5d652bcadac3975488941b625b1c9ca75700210adbe7a4179932051152e5f108

Pith citing papers

No inbound Pith citation observations are available.