Pith. sign in

Paper Citation Record · LEDGER

TransNormerLLM: A Faster and Better Large Language Model with Improved TransNormer

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2307.14995.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.14995 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T12:25:42.707448Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

6
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4bcbbabb-b649-4df3-881c-07e853b6440b · inbound

Gated Linear Attention Transformers with Hardware-Efficient Training cites this paper.

Gated Linear Attention Transformers with Hardware-Efficient Training TransNormerLLM: A Faster and Better Large Language Model with Improved TransNormer

Reference 72

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T01:15:14.175776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-15T01:15:13.991219Z digest=sha256:149c583304d61bc4fdf7d4948867829334bbf972a0c811c7b1439f6b282c49dc

Observation eba4f707-d94c-48f4-a193-c949a550bee0 · inbound

Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality cites this paper.

Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality TransNormerLLM: A Faster and Better Large Language Model with Improved TransNormer

Reference 82

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:16:25.915160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-11T12:16:25.390683Z digest=sha256:7fe05b9e973a4f96fccd80e0f3ca1e81862ee784ba63c7aa64f498a112f1a1d9

Observation 7b8e93ba-727c-40d1-abd2-01c8cd362928 · inbound

LASP-2: Rethinking Sequence Parallelism for Linear Attention and Its Hybrid cites this paper.

LASP-2: Rethinking Sequence Parallelism for Linear Attention and Its Hybrid TransNormerLLM: A Faster and Better Large Language Model with Improved TransNormer

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T12:25:42.707448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:25:42.707448Z digest=sha256:125dc17d73a04f404a666ef2ad984450af1d06f314a82e7411506a569a5810ac

Observation b1708fc1-335d-4b45-969c-66472c17ffc4 · inbound

MesaNet: Sequence Modeling by Locally Optimal Test-Time Training cites this paper.

MesaNet: Sequence Modeling by Locally Optimal Test-Time Training TransNormerLLM: A Faster and Better Large Language Model with Improved TransNormer

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T10:30:13.833907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:30:13.833907Z digest=sha256:d70d34f7c2d0fb52d45645143bdb96204c45a493f574f213f148973aeed740e8

Observation de10f394-71f8-4b84-abee-5b99a3ced245 · inbound

Autoregressive Image Generation with Linear Complexity: A Spatial-Aware Decay Perspective cites this paper.

Autoregressive Image Generation with Linear Complexity: A Spatial-Aware Decay Perspective TransNormerLLM: A Faster and Better Large Language Model with Improved TransNormer

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T20:53:51.424146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:53:51.424146Z digest=sha256:105352e703446e5bc6338f8134dd8a0ed5646bb4ec0d90e214a251d84728f053

Observation bc6c667a-1a6c-476c-84e7-a206061178b5 · inbound

Lizard: An Efficient Linearization Framework for Large Language Models cites this paper.

Lizard: An Efficient Linearization Framework for Large Language Models TransNormerLLM: A Faster and Better Large Language Model with Improved TransNormer

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-19T04:42:04.610884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-19T04:37:55.034479Z digest=sha256:1bd8679fd5a0f8f11a70b33d11bfd613b7920899b45e4c174317fa3ed53e3379

Observation 28c342ae-0647-47e7-a519-eac408e4d8e7 · inbound

DistrAttention: An Efficient and Flexible Self-Attention Mechanism on Modern GPUs cites this paper.

DistrAttention: An Efficient and Flexible Self-Attention Mechanism on Modern GPUs TransNormerLLM: A Faster and Better Large Language Model with Improved TransNormer

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T14:59:20.061853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:59:20.061853Z digest=sha256:2ce63f99a2c82a3beeb79e6ad52b318d64a396e34c0b9f6b987d713a2f6ad2d8

Observation 219a30f7-8592-4f92-b4d1-3df23f1fa5c5 · inbound

Elucidating the Design Space of Decay in Linear Attention cites this paper.

Elucidating the Design Space of Decay in Linear Attention TransNormerLLM: A Faster and Better Large Language Model with Improved TransNormer

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T05:29:21.559505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T05:29:21.559505Z digest=sha256:4636e7dd3c0e9534b999edc3fe93a465dc9532a50bbf31eabe38594815b1f720

Observation b16d595b-1101-4eb5-a673-a1bd80c3aa41 · inbound

Kimi Linear: An Expressive, Efficient Attention Architecture cites this paper.

Kimi Linear: An Expressive, Efficient Attention Architecture TransNormerLLM: A Faster and Better Large Language Model with Improved TransNormer

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.761500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:61c004cb1064da2e7a55cb5a23ff050e09aab0e25ecea9483237dd8ae2aa48d8

Observation 1d072650-1e46-48ef-99d8-288ad88969d2 · inbound

Higher-order Linear Attention cites this paper.

Higher-order Linear Attention TransNormerLLM: A Faster and Better Large Language Model with Improved TransNormer

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T03:05:47.397652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-18T03:05:35.823369Z digest=sha256:abe15085c832d551471821c30860ba8ae56dafbf12bbcf9bedf80a9d76b08dc0

Observation 3a96ce63-40a6-4fcc-9d35-d07df23797f1 · inbound

Ling and Ring 2.6 Technical Report: Efficient and Instant Agentic Intelligence at Trillion-Parameter Scale cites this paper.

Ling and Ring 2.6 Technical Report: Efficient and Instant Agentic Intelligence at Trillion-Parameter Scale TransNormerLLM: A Faster and Better Large Language Model with Improved TransNormer

Reference 229

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:17:24.859282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-07-02T22:10:59.568675Z digest=sha256:fca464c0fe1881adf1db5a66a52e12a1ee6fc517319d547c3b93045c13f7fa2a