Pith. sign in

Paper Citation Record · LEDGER

PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2104.12369.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2104.12369 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T15:33:48.493305Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T13:39:31.694945Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3948bc8f-8d41-419a-a22c-3a773bde82fe · inbound

Deduplicating Training Data Makes Language Models Better cites this paper.

Deduplicating Training Data Makes Language Models Better PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-24T13:39:31.698342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-24T13:36:55.210708Z digest=sha256:a351ec60c4cd5fb5be8ac00183ff36ee85ec3d28214e45f325d065a281f54716

Observation a29f3d72-3466-4234-b011-c3164ddc53b9 · inbound

Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model cites this paper.

Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-24T12:14:26.701333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-24T12:10:49.690618Z digest=sha256:2346c452247b348749f154328ac170dd692cb2ac665c023bda5f90fa38a281ab

Observation c79f795c-c85d-4ac1-8e6d-def2b55d9a2a · inbound

PaLM: Scaling Language Modeling with Pathways cites this paper.

PaLM: Scaling Language Modeling with Pathways PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 174

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:45:07.480568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T23:45:06.755839Z digest=sha256:19ea8743c3cf478c1dd3222be04d8f374da1e1a1e2ba4f4c3e91cd652c50a7f9

Observation 3c96979c-8653-459f-b5a4-4d261968a1a1 · inbound

GPT-NeoX-20B: An Open-Source Autoregressive Language Model cites this paper.

GPT-NeoX-20B: An Open-Source Autoregressive Language Model PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 107

Resolution
verified exact
arxiv_id, observed 2026-05-24T12:34:28.278379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-24T12:33:37.701655Z digest=sha256:ba22e8fba89894dea7b070bac7fa216a05d53b6da1c2a62f55b5592efa4e0d1b

Observation 23f2383e-42b8-4b43-893d-ee953bc9b09b · inbound

OPT: Open Pre-trained Transformer Language Models cites this paper.

OPT: Open Pre-trained Transformer Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 212

Resolution
verified exact
arxiv_id, observed 2026-05-10T20:53:17.723733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T20:53:16.720145Z digest=sha256:e8c84b18cab8cfcf12237eb211aa7f8c73ef9e20cc31b42a88cdcf9a6250ea6c

Observation 017d2d79-bbb9-4e47-84a0-761718f17aab · inbound

A Survey of Large Language Models cites this paper.

A Survey of Large Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 86

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:46:39.656602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T22:46:39.268353Z digest=sha256:9876e3991d0ffee8ffaf9153a8246c58c7106fa480edad258a027290de04e97b

Observation f2209cc2-3b08-4f52-8ac9-4e67ec784709 · inbound

Scaling Data-Constrained Language Models cites this paper.

Scaling Data-Constrained Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 139

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:35:21.518252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T01:35:21.150772Z digest=sha256:d9bc7b2fe9b7996683c7148c0e0cd79b2b2db13aec33b7f73f480afe87a0e367

Observation e1e2dd58-f14e-420c-9abd-2db09e56e428 · inbound

The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only cites this paper.

The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:43:45.890773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T20:43:45.770157Z digest=sha256:57cfdffa6aa526d7cf9a067ab712b5b7b2441138e52a1f12ce6e24ffd2a9c19b

Observation d4940b65-64ec-42ba-ba6c-8de817f6569e · inbound

A Comprehensive Overview of Large Language Models cites this paper.

A Comprehensive Overview of Large Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 108

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:28:39.371703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T20:28:38.900026Z digest=sha256:571c57fa849211825267aa378c2747a0bae8311fbc5607af07d39e916142755e

Observation 42886ca4-78e6-47dc-9cee-691e7c3a9df0 · inbound

The Falcon Series of Open Language Models cites this paper.

The Falcon Series of Open Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 82

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T09:46:10.086221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-16T09:46:09.701440Z digest=sha256:28ab9f3504f885cb61586278b9a9fd616db270321a38dbe80c8f49697240243a

Observation eb90aaa4-10ba-4ef9-ae03-b7e001ece8c2 · inbound

Occ-LLM: Enhancing Autonomous Driving with Occupancy-Based Large Language Models cites this paper.

Occ-LLM: Enhancing Autonomous Driving with Occupancy-Based Large Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-08T15:33:48.493305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:33:48.493305Z digest=sha256:09deb520a68b7763c3af10c57598f8cbfd91c17c4cae25e1a9626ad2cc45fae6

Observation 67e8d5f3-3fde-4085-ada6-97234ce1031f · inbound

EvdCLIP: Improving Vision-Language Retrieval with Entity Visual Descriptions from Large Language Models cites this paper.

EvdCLIP: Improving Vision-Language Retrieval with Entity Visual Descriptions from Large Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:32:41.114814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:32:41.114814Z digest=sha256:1f879a496a5e958913331047858e065bf169e02a2dc02b6e6637a86f5db55342

Observation 76bc0cd3-7146-40ba-8227-3fc877d7551a · inbound

Scalable Complexity Control Facilitates Reasoning Ability of LLMs cites this paper.

Scalable Complexity Control Facilitates Reasoning Ability of LLMs PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:15.459004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:01:15.459004Z digest=sha256:e89b112ab338c42b75a33436e491f478d58eb95c1f6c9a38bff6e0bb61575d35

Observation 29c79227-7137-4cda-a89f-732ea61f9406 · inbound

Chengyu-Bench: Benchmarking Large Language Models for Chinese Idiom Understanding and Use cites this paper.

Chengyu-Bench: Benchmarking Large Language Models for Chinese Idiom Understanding and Use PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:19.553332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:27:19.553332Z digest=sha256:3925dae643609f5b82d8a96409294a0c133531560d21778f8caa662d7ee87794