Pith. sign in

Paper Citation Record · LEDGER

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization

As of 12 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2608.09160.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.09160 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T22:26:47.452553Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact1
  • verified fuzzy8
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 34d772e3-0810-4038-bac4-75a80fa86989 · outbound

This paper cites Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T22:26:47.350167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:26:47.350167Z digest=sha256:18a2214ce2c69598009f6e5f680a13f54acf644fd09baabdc9f2c88b0e064c6c

Observation d0006a06-aef3-4d8a-8c54-17bf8aa7f4e7 · outbound

This paper cites Efficient memory management for large language model serving with pagedattention,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Efficient memory management for large language model serving with pagedattention,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.918531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.362615Z digest=sha256:088b408828ff79f7009e83dba919968ba3f5d9e2a885b6c5f3b6ffc575c9813a

Observation f74c1e6c-f28a-4de2-8e5c-1af97a64e03a · outbound

This paper cites FLUX: Fast Software-based Communication Overlap On GPUs Through Kernel Fusion.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization FLUX: Fast Software-based Communication Overlap On GPUs Through Kernel Fusion

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T22:26:47.368308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:26:47.368308Z digest=sha256:9c9ad68e89f3fa5db56b64e0c834c4d2cd0f1442bc41de7ee80aa0ccb7b1c8fc

Observation 2cf554a0-4d14-45a4-926d-5e995d619ba1 · outbound

This paper cites Flashoverlap: A lightweight design for efficiently overlapping communication and computation,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Flashoverlap: A lightweight design for efficiently overlapping communication and computation,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.894254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.374690Z digest=sha256:9e3ab9cf81515b6476c7f5eb0f4a7d8727e81b0782de3522ab998aa3d45db82d

Observation 9c524e33-19a0-4916-b461-40bddfded5b8 · outbound

This paper cites Scaling vision transformers to 22 billion parame- ters,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Scaling vision transformers to 22 billion parame- ters,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.874911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.381371Z digest=sha256:f72c56f140c68d04e56eee45f934297b67ad9da2a7522466db8aa9fa7751427b

Observation c08adcf2-244e-465f-8c08-cbcd21b1c191 · outbound

This paper cites 2 OLMo 2 Furious.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization 2 OLMo 2 Furious

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T22:26:47.388569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:26:47.388569Z digest=sha256:027e1c4ec87470c6fe35aa8089e6b5fe252139fb99721c40e9295e7d311939ff

Observation be5b1ea6-bfe2-453f-8ffa-8188c1d28afb · outbound

This paper cites Olmoe: Open mixture-of-experts language models,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Olmoe: Open mixture-of-experts language models,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.856811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.397767Z digest=sha256:204e3a4f43f271edb05c21a326db1c1c2f94d7c8211863e8509da4da5052406b

Observation 59df8745-5479-4aae-af06-4947d9697124 · outbound

This paper cites Olmo 3.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Olmo 3

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T22:26:47.409165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:26:47.409165Z digest=sha256:62a34012b2fd4e9ef01a1b48ee37192ca6e659522de335d64025dc4f695732c5

Observation 57198c9f-e150-493f-a4e0-577bf2bb7b1f · outbound

This paper cites Root mean square layer normalization,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Root mean square layer normalization,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T22:26:47.417005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:26:47.417005Z digest=sha256:3f394272c521ebad16ce2dc120f3d4e22cc0cda749e2d7ff8a8661b790c473bf

Observation ba6b7a54-91b3-4550-b68c-25f2b8952c90 · outbound

This paper cites Gpipe: Efficient training of giant neural networks using pipeline parallelism,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Gpipe: Efficient training of giant neural networks using pipeline parallelism,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.817503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.421704Z digest=sha256:68bb31e79661cda95dddfa440eafc273ed1fa7c1d0c72499e71f4200865f18df

Observation 42d54c11-bb7e-489f-a017-121d854e70e8 · outbound

This paper cites ShareGPT Dataset,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization ShareGPT Dataset,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.794484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.432274Z digest=sha256:3ae6aa59304219734a10caeba1924282153f71cde2462b09db3bed54b839450a

Observation 8042f586-10b6-47cd-ac78-0c76ead1036e · outbound

This paper cites Olmo: Accelerating the science of language models,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Olmo: Accelerating the science of language models,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.764870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.440748Z digest=sha256:870dc7183b1effb4e1b5404e9c8642f4712a2a8f87e507454cc8f6e250060246

Observation a020fbaf-f1b3-4106-88fe-674c6b455372 · outbound

This paper cites MiniMax TP RMSNorm,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization MiniMax TP RMSNorm,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:26:47.745497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.446922Z digest=sha256:6bc203310f35b2e5018d6a5e2a33484e24db6ba5d19f9b14231e3269a69bc5e4

Observation 394b9cdf-c0c2-4376-8592-aaf7eb64f993 · outbound

This paper cites Simplegpt: Improving gpt via a simple normalization strategy,.

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization Simplegpt: Improving gpt via a simple normalization strategy,

Reference 14

Resolution
verified exact
raw_fallback, observed 2026-08-11T22:26:47.630204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T22:26:47.452553Z digest=sha256:6789741df8d5abc4d699c6f8457821e5eb1a2a476b68475c4a96db1c057b0cea

Pith citing papers

No inbound Pith citation observations are available.