Pith. sign in

Paper Citation Record · LEDGER

PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2104.12369.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2104.12369 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:03:43.888122Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T13:39:31.694945Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3948bc8f-8d41-419a-a22c-3a773bde82fe · inbound

Deduplicating Training Data Makes Language Models Better cites this paper.

Deduplicating Training Data Makes Language Models Better PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-24T13:39:31.698342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-24T13:36:55.210708Z digest=sha256:b54c0fcb9e7a8a0d355c7a8c524fb64e94cc9f2cd58b9dd6485166cf35b8965a

Observation a29f3d72-3466-4234-b011-c3164ddc53b9 · inbound

Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model cites this paper.

Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-24T12:14:26.701333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-24T12:10:49.690618Z digest=sha256:cab8b5e11ab7b1f729b67f4e21b30de59639c33cd58c0b73c01922c0b76d5f77

Observation c79f795c-c85d-4ac1-8e6d-def2b55d9a2a · inbound

PaLM: Scaling Language Modeling with Pathways cites this paper.

PaLM: Scaling Language Modeling with Pathways PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 174

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:45:07.480568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-10T23:45:06.755839Z digest=sha256:4379bc8632965b25cac4fdbf33a67afa9dc244a8db0f570a5504307b356d4d6c

Observation 3c96979c-8653-459f-b5a4-4d261968a1a1 · inbound

GPT-NeoX-20B: An Open-Source Autoregressive Language Model cites this paper.

GPT-NeoX-20B: An Open-Source Autoregressive Language Model PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 107

Resolution
verified exact
arxiv_id, observed 2026-05-24T12:34:28.278379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-24T12:33:37.701655Z digest=sha256:15bd06324db668a4f244062f5d0e3cafa85e9bdaaad110e79882dda44bb19a86

Observation 23f2383e-42b8-4b43-893d-ee953bc9b09b · inbound

OPT: Open Pre-trained Transformer Language Models cites this paper.

OPT: Open Pre-trained Transformer Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 212

Resolution
verified exact
arxiv_id, observed 2026-05-10T20:53:17.723733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-10T20:53:16.720145Z digest=sha256:dffabcba7ae96bb37f64447d13ea6913206ce41e7450fcd0f1abfdd0babd02ab

Observation 017d2d79-bbb9-4e47-84a0-761718f17aab · inbound

A Survey of Large Language Models cites this paper.

A Survey of Large Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 86

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:46:39.656602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T22:46:39.268353Z digest=sha256:b62086758447e3aa9c79edddef568c8dbaffc7ff0241998dc81c1a5fb32241ea

Observation f2209cc2-3b08-4f52-8ac9-4e67ec784709 · inbound

Scaling Data-Constrained Language Models cites this paper.

Scaling Data-Constrained Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 139

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:35:21.518252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T01:35:21.150772Z digest=sha256:1257fa50e10e0fb4834a4544e64740dcebb4981de95e4418e292d872ca43bfaf

Observation e1e2dd58-f14e-420c-9abd-2db09e56e428 · inbound

The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only cites this paper.

The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:43:45.890773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T20:43:45.770157Z digest=sha256:d8ec2b04692327bc848f1a1b92b7090ae0d61f933f33fb89691daa3cbe8a8606

Observation d4940b65-64ec-42ba-ba6c-8de817f6569e · inbound

A Comprehensive Overview of Large Language Models cites this paper.

A Comprehensive Overview of Large Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 108

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:28:39.371703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T20:28:38.900026Z digest=sha256:0ddf99c4b47e0dad7920ae23522696b2a76be4e7e12a759e2a0302cdba0d3803

Observation 42886ca4-78e6-47dc-9cee-691e7c3a9df0 · inbound

The Falcon Series of Open Language Models cites this paper.

The Falcon Series of Open Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 82

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T09:46:10.086221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-16T09:46:09.701440Z digest=sha256:973c5aa94b0fe29970f8bb0be13007294694bc6929dc63a5c0daa66f9227edb0

Observation ef2c9a37-9d2a-408b-9eab-dfcbe8c31983 · inbound

LIBER: Lifelong User Behavior Modeling Based on Large Language Models cites this paper.

LIBER: Lifelong User Behavior Modeling Based on Large Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T15:03:43.888122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:03:43.888122Z digest=sha256:6fd6355c8b1d9e61ab26ca1c985d90c61290e6da7a45c0ca8aaf6217e2162e74

Observation ea10a133-79cb-49cf-87d7-a5a2c1a4f68a · inbound

Survey of different Large Language Model Architectures: Trends, Benchmarks, and Challenges cites this paper.

Survey of different Large Language Model Architectures: Trends, Benchmarks, and Challenges PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T22:41:12.804749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:41:12.804749Z digest=sha256:5d75e43f5e823c7bdaf50c3ea4406fda69c21458900524a08fbeb102ccd6118f

Observation 64ed0ea6-16c9-46a1-a203-05816e930755 · inbound

An Automatic Graph Construction Framework based on Large Language Models for Recommendation cites this paper.

An Automatic Graph Construction Framework based on Large Language Models for Recommendation PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-11T04:58:31.887938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:58:31.887938Z digest=sha256:296d60156fc120249541898845056dc8767c592d7f1bc92fb3cd2c74ca5768a1

Observation 02f0e9ed-2255-4b81-8a61-820dba61854f · inbound

A Survey on Large Language Models with some Insights on their Capabilities and Limitations cites this paper.

A Survey on Large Language Models with some Insights on their Capabilities and Limitations PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 139

Resolution
unresolved
no resolver link, observed 2026-08-10T22:17:55.915702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:17:55.915702Z digest=sha256:afea364c49e726c2ba93a43f57cc0622d919e0e4f1e9e9ce13effaf1c2465d97

Observation eb90aaa4-10ba-4ef9-ae03-b7e001ece8c2 · inbound

Occ-LLM: Enhancing Autonomous Driving with Occupancy-Based Large Language Models cites this paper.

Occ-LLM: Enhancing Autonomous Driving with Occupancy-Based Large Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-08T15:33:48.493305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:33:48.493305Z digest=sha256:a76adaad1aff9419c02891c9e78174314b6a792999549f27946af02670dadd86

Observation 67e8d5f3-3fde-4085-ada6-97234ce1031f · inbound

EvdCLIP: Improving Vision-Language Retrieval with Entity Visual Descriptions from Large Language Models cites this paper.

EvdCLIP: Improving Vision-Language Retrieval with Entity Visual Descriptions from Large Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:32:41.114814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:32:41.114814Z digest=sha256:f4f2ef865171bffa13c57ad9557c8e90aebdd0a408877cb946c11418a32cb70e

Observation 76bc0cd3-7146-40ba-8227-3fc877d7551a · inbound

Scalable Complexity Control Facilitates Reasoning Ability of LLMs cites this paper.

Scalable Complexity Control Facilitates Reasoning Ability of LLMs PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:15.459004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:01:15.459004Z digest=sha256:6460a36905b3500102f754a5738403f4119ffac56e30f2d49d74feeaa6cc2227

Observation 29c79227-7137-4cda-a89f-732ea61f9406 · inbound

Chengyu-Bench: Benchmarking Large Language Models for Chinese Idiom Understanding and Use cites this paper.

Chengyu-Bench: Benchmarking Large Language Models for Chinese Idiom Understanding and Use PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:19.553332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:27:19.553332Z digest=sha256:9cbd880a2e403d24e64cee4adb06e372196a4848af525cc453966a9189dff80f