Pith. sign in

Paper Citation Record · LEDGER

GQKVA: Efficient Pre-training of Transformers by Grouping Queries, Keys, and Values

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 3 inbound Pith citation observations for arXiv:2311.03426.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.03426 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T00:38:48.908718Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1375e9ee-ce70-49f8-a0eb-0e292b1ed044 · inbound

A Survey on Large Language Model Acceleration based on KV Cache Management cites this paper.

A Survey on Large Language Model Acceleration based on KV Cache Management GQKVA: Efficient Pre-training of Transformers by Grouping Queries, Keys, and Values

Reference 222

Resolution
unresolved
no resolver link, observed 2026-08-11T00:38:48.908718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:38:48.908718Z digest=sha256:97109bae8171950a9ee0abef07160f922e2ee5fdf83cce710fdd23e2933500fb

Observation b4a92ba7-73ed-4f55-bb57-cdf509c897df · inbound

ECHO-LLaMA: Efficient Caching for High-Performance LLaMA Training cites this paper.

ECHO-LLaMA: Efficient Caching for High-Performance LLaMA Training GQKVA: Efficient Pre-training of Transformers by Grouping Queries, Keys, and Values

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:21.260354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:21.260354Z digest=sha256:ec108c33f624ce4ad4ca3fcfb9778435ce588d6fe6d971e362867c7309e4256f

Observation 4a27a18d-f517-4a68-aa6d-d339d2d0adf5 · inbound

ConSA: Controllable Sparsity in Hybrid Attention via Learnable Allocation cites this paper.

ConSA: Controllable Sparsity in Hybrid Attention via Learnable Allocation GQKVA: Efficient Pre-training of Transformers by Grouping Queries, Keys, and Values

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-27T00:40:18.528573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-27T00:33:18.513616Z digest=sha256:2976bdcad97704c0bde6b3a7d8105919c95bf1a5a046dd2c6aaadd375ab2f411