Pith. sign in

Paper Citation Record · LEDGER

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization

As of 20 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2608.09160.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.09160 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T22:26:47.452553Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact1
  • verified fuzzy8
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 34d772e3-0810-4038-bac4-75a80fa86989 · outbound

This paper cites Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T22:26:47.350167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:26:47.350167Z digest=sha256:a8aacae9d21aafe98013827b637bba6f729ddcf480e1bc6d49321604dd261200

Observation d0006a06-aef3-4d8a-8c54-17bf8aa7f4e7 · outbound

This paper cites Efficient memory management for large language model serving with pagedattention,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Efficient memory management for large language model serving with pagedattention,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.918531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T22:26:47.362615Z digest=sha256:6e92407681f79ba67d47f890af14eaafa9e324b6e59c179eb589dec2a601c0a3

Observation f74c1e6c-f28a-4de2-8e5c-1af97a64e03a · outbound

This paper cites FLUX: Fast Software-based Communication Overlap On GPUs Through Kernel Fusion.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization FLUX: Fast Software-based Communication Overlap On GPUs Through Kernel Fusion

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T22:26:47.368308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:26:47.368308Z digest=sha256:87be6d7eec1ebaa8570106d03fd62183ac64170cc8868c179aa2aba0441804b3

Observation 2cf554a0-4d14-45a4-926d-5e995d619ba1 · outbound

This paper cites Flashoverlap: A lightweight design for efficiently overlapping communication and computation,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Flashoverlap: A lightweight design for efficiently overlapping communication and computation,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.894254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T22:26:47.374690Z digest=sha256:77563028b440f7eef70abd0a18ff4e286de0cd7749df871705b9374804c5fb1a

Observation 9c524e33-19a0-4916-b461-40bddfded5b8 · outbound

This paper cites Scaling vision transformers to 22 billion parame- ters,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Scaling vision transformers to 22 billion parame- ters,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.874911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T22:26:47.381371Z digest=sha256:6fef12155a73575ee4c788480b22737616d738b9ab7ebfa028369e49d8de5af0

Observation c08adcf2-244e-465f-8c08-cbcd21b1c191 · outbound

This paper cites 2 OLMo 2 Furious.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization 2 OLMo 2 Furious

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T22:26:47.388569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:26:47.388569Z digest=sha256:007c32a36d52cd73edb55e780c23643602b34816f812222fc6981eafccefc48e

Observation be5b1ea6-bfe2-453f-8ffa-8188c1d28afb · outbound

This paper cites Olmoe: Open mixture-of-experts language models,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Olmoe: Open mixture-of-experts language models,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.856811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T22:26:47.397767Z digest=sha256:71acb97f7e555642e7373042c8e2ceac25ebba693cd91166205e509f3be97c0b

Observation 59df8745-5479-4aae-af06-4947d9697124 · outbound

This paper cites Olmo 3.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Olmo 3

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T22:26:47.409165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:26:47.409165Z digest=sha256:f0dced0c32459c9332cfd1a002fab85ac39c02cecbb5c3b7bb2618c7aba46ebd

Observation 57198c9f-e150-493f-a4e0-577bf2bb7b1f · outbound

This paper cites Root mean square layer normalization,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Root mean square layer normalization,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T22:26:47.417005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:26:47.417005Z digest=sha256:ef1d4e26350aacd5a62649bbcde7af9eb6cb9fb977402c1f30fadddeb0ef8440

Observation ba6b7a54-91b3-4550-b68c-25f2b8952c90 · outbound

This paper cites Gpipe: Efficient training of giant neural networks using pipeline parallelism,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Gpipe: Efficient training of giant neural networks using pipeline parallelism,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.817503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T22:26:47.421704Z digest=sha256:64c1149a74cf6b80dceaaecd4337f92be21a871ed1623fcae82dfce34e70f57a

Observation 42d54c11-bb7e-489f-a017-121d854e70e8 · outbound

This paper cites ShareGPT Dataset,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization ShareGPT Dataset,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.794484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T22:26:47.432274Z digest=sha256:019fd112412d0cf45520e173b35d27bd04298798733f871546153103a1a52fdc

Observation 8042f586-10b6-47cd-ac78-0c76ead1036e · outbound

This paper cites Olmo: Accelerating the science of language models,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Olmo: Accelerating the science of language models,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.764870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T22:26:47.440748Z digest=sha256:8a9cbe5a25ca568b51c2b41c40a0330e9b896a2d3576ae711513465dad7aa7ac

Observation a020fbaf-f1b3-4106-88fe-674c6b455372 · outbound

This paper cites MiniMax TP RMSNorm,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization MiniMax TP RMSNorm,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.745497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T22:26:47.446922Z digest=sha256:5fb379c973723c448ee232f023b7a1db49322c4c83abfa76551a1a6385de8207

Observation 394b9cdf-c0c2-4376-8592-aaf7eb64f993 · outbound

This paper cites Simplegpt: Improving gpt via a simple normalization strategy,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Simplegpt: Improving gpt via a simple normalization strategy,

Reference 14

Resolution
verified exact
raw_fallback, observed 2026-08-11T22:26:47.630204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T22:26:47.452553Z digest=sha256:56a8c779987670b8ab6930dd3ed9bc63d6a45a92db578b260d92a5f2cf2a8e57

Pith citing papers

No inbound Pith citation observations are available.