Pith. sign in

Paper Citation Record · LEDGER

PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2104.12369.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2104.12369 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-24T13:36:55.210708Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T13:39:31.694945Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3948bc8f-8d41-419a-a22c-3a773bde82fe · inbound

Deduplicating Training Data Makes Language Models Better cites this paper.

Deduplicating Training Data Makes Language Models Better PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-24T13:39:31.698342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-24T13:36:55.210708Z digest=sha256:be790e5413db43dcc09731fc1c3d08c5136533bc6968704132e20790fa44fe25

Observation a29f3d72-3466-4234-b011-c3164ddc53b9 · inbound

Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model cites this paper.

Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-24T12:14:26.701333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T12:10:49.690618Z digest=sha256:218ec40235350110a12e6efb18a704d4c9235e7e35d65887e037bded0e9e65e2

Observation c79f795c-c85d-4ac1-8e6d-def2b55d9a2a · inbound

PaLM: Scaling Language Modeling with Pathways cites this paper.

PaLM: Scaling Language Modeling with Pathways PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 174

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:45:07.480568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T23:45:06.755839Z digest=sha256:cc7004d0bbc063e5de73e794efabc770b33e3b35fb8cc992487c6b5024181bb7

Observation 3c96979c-8653-459f-b5a4-4d261968a1a1 · inbound

GPT-NeoX-20B: An Open-Source Autoregressive Language Model cites this paper.

GPT-NeoX-20B: An Open-Source Autoregressive Language Model PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 107

Resolution
verified exact
arxiv_id, observed 2026-05-24T12:34:28.278379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-24T12:33:37.701655Z digest=sha256:b19908ac544d74ffc650b2c7db42964c92a610bf7b1063bd2690b265dc99a78b

Observation 23f2383e-42b8-4b43-893d-ee953bc9b09b · inbound

OPT: Open Pre-trained Transformer Language Models cites this paper.

OPT: Open Pre-trained Transformer Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 212

Resolution
verified exact
arxiv_id, observed 2026-05-10T20:53:17.723733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T20:53:16.720145Z digest=sha256:6e73fb81a0c36a9f840682822ed5869b57dd10b86daf942f2c6a095185595ed6

Observation 017d2d79-bbb9-4e47-84a0-761718f17aab · inbound

A Survey of Large Language Models cites this paper.

A Survey of Large Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 86

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:46:39.656602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T22:46:39.268353Z digest=sha256:c296160e714e4183d14e368cb0306b6ef96e0a6fee86e2c9fd9073c2b1236ce9

Observation f2209cc2-3b08-4f52-8ac9-4e67ec784709 · inbound

Scaling Data-Constrained Language Models cites this paper.

Scaling Data-Constrained Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 139

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:35:21.518252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T01:35:21.150772Z digest=sha256:9592648622951edda87dd97bd96c3b476be307e42ef39adc6ca5a64a16518901

Observation e1e2dd58-f14e-420c-9abd-2db09e56e428 · inbound

The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only cites this paper.

The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:43:45.890773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T20:43:45.770157Z digest=sha256:15e278e7c1d57bfccbd610d96c648bcedf7f30e9b5708a3988ce2b557de0181d

Observation d4940b65-64ec-42ba-ba6c-8de817f6569e · inbound

A Comprehensive Overview of Large Language Models cites this paper.

A Comprehensive Overview of Large Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 108

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:28:39.371703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T20:28:38.900026Z digest=sha256:0dfb7da50d939960ef72556e52c87393dde58d18951ec01aed07ce7c2593c1f4

Observation 42886ca4-78e6-47dc-9cee-691e7c3a9df0 · inbound

The Falcon Series of Open Language Models cites this paper.

The Falcon Series of Open Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 82

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T09:46:10.086221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-16T09:46:09.701440Z digest=sha256:c4342350ef5ac8cf4c3c1b804f5bd889c99863d8789c8d1eafc891d87dfdb82a