Pith. sign in

Paper Citation Record · LEDGER

QCQA: Quality and Capacity-aware grouped Query Attention

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2406.10247.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.10247 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:51:04.694089Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T17:47:40.626505Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a30ca115-0565-4b76-a3cb-77e97a652df7 · inbound

A Survey on Large Language Model Acceleration based on KV Cache Management cites this paper.

A Survey on Large Language Model Acceleration based on KV Cache Management QCQA: Quality and Capacity-aware grouped Query Attention

Reference 220

Resolution
unresolved
no resolver link, observed 2026-08-11T00:38:48.868762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:38:48.868762Z digest=sha256:d35440cf0f920fc2e82eda31cbb5654d146c8c9050e16315452dae1ea2cf5143

Observation f73038ca-a15f-41d5-a349-7489303101ba · inbound

The Rise of Small Language Models in Healthcare: A Comprehensive Survey cites this paper.

The Rise of Small Language Models in Healthcare: A Comprehensive Survey QCQA: Quality and Capacity-aware grouped Query Attention

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-16T10:51:04.694089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:51:04.694089Z digest=sha256:58c8ac33c2cfcf5fcce0d63acab765d9af51de7a12e657686bf171d6c3f26811

Observation bce860d6-de4d-486c-8b27-6cfdfd0c77de · inbound

Opt-GPTQ: An Optimized GPTQ Combining Sparse Attention and Quantization Techniques cites this paper.

Opt-GPTQ: An Optimized GPTQ Combining Sparse Attention and Quantization Techniques QCQA: Quality and Capacity-aware grouped Query Attention

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T00:59:42.112115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:59:42.112115Z digest=sha256:669b4a72f706141336d0dbaf9d039917ba7be0dc3af64a555f72b1cfbc6a0171

Observation dac99d7f-cb4a-4bca-8a6d-92b8b0ae33c7 · inbound

TaDA: Training-free recipe for Decoding with Adaptive KV Cache Compression and Mean-centering cites this paper.

TaDA: Training-free recipe for Decoding with Adaptive KV Cache Compression and Mean-centering QCQA: Quality and Capacity-aware grouped Query Attention

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T10:45:09.873091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:45:09.873091Z digest=sha256:fb640ae1c658fcaa1af4313e5e00b7f2548d039eabd94b416cb8ec9e107add2a

Observation 7f542547-6f93-451f-ad3d-e986b0c9f70f · inbound

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models cites this paper.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models QCQA: Quality and Capacity-aware grouped Query Attention

Reference 103

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:47:40.629757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:47:40.266730Z digest=sha256:6355290738a812806ab2ba3cddd1dbc7c74118067b71bbfa2b108913e65a51dd