Pith. sign in

Paper Citation Record · LEDGER

PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2104.12369.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2104.12369 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T15:33:48.493305Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T13:39:31.694945Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3948bc8f-8d41-419a-a22c-3a773bde82fe · inbound

Deduplicating Training Data Makes Language Models Better cites this paper.

Deduplicating Training Data Makes Language Models Better PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-24T13:39:31.698342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-24T13:36:55.210708Z digest=sha256:a153740da74315906ab218e72d10d19ee21695f1703ddf7a81b4788b3f8630fc

Observation a29f3d72-3466-4234-b011-c3164ddc53b9 · inbound

Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model cites this paper.

Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-24T12:14:26.701333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-24T12:10:49.690618Z digest=sha256:32a5e80a8d2b3600e191282011cb1dd8fe99ca2d7163aa622ff996c34f9bf68f

Observation c79f795c-c85d-4ac1-8e6d-def2b55d9a2a · inbound

PaLM: Scaling Language Modeling with Pathways cites this paper.

PaLM: Scaling Language Modeling with Pathways PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 174

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:45:07.480568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T23:45:06.755839Z digest=sha256:90071aabde0a382cec7a3e380ebd60d6151adb6f038f8328909e3d9b3cb772b2

Observation 3c96979c-8653-459f-b5a4-4d261968a1a1 · inbound

GPT-NeoX-20B: An Open-Source Autoregressive Language Model cites this paper.

GPT-NeoX-20B: An Open-Source Autoregressive Language Model PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 107

Resolution
verified exact
arxiv_id, observed 2026-05-24T12:34:28.278379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-24T12:33:37.701655Z digest=sha256:ae22824eaa6c391bcee3937e39d7a3260c1cdae18682a7229bb425c33b21078b

Observation 23f2383e-42b8-4b43-893d-ee953bc9b09b · inbound

OPT: Open Pre-trained Transformer Language Models cites this paper.

OPT: Open Pre-trained Transformer Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 212

Resolution
verified exact
arxiv_id, observed 2026-05-10T20:53:17.723733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T20:53:16.720145Z digest=sha256:ed4b0b3ea6026d8c9d01f36ef25716ba75ec331d90cc6b350a8d6c31b3e2aab1

Observation 017d2d79-bbb9-4e47-84a0-761718f17aab · inbound

A Survey of Large Language Models cites this paper.

A Survey of Large Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 86

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:46:39.656602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T22:46:39.268353Z digest=sha256:2dd62ded2fbdaac5f43f9793d68032b8329380c6653d05fd13a9f3c555ae8f51

Observation f2209cc2-3b08-4f52-8ac9-4e67ec784709 · inbound

Scaling Data-Constrained Language Models cites this paper.

Scaling Data-Constrained Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 139

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:35:21.518252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T01:35:21.150772Z digest=sha256:a7678b5478ed40a153384e603d20b4cf63a819f26065fb16fc90019c5de62745

Observation e1e2dd58-f14e-420c-9abd-2db09e56e428 · inbound

The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only cites this paper.

The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:43:45.890773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T20:43:45.770157Z digest=sha256:9860e7da9a319e724c3f2983a48b7452f097c8ac5085d46c26ac305245e6a415

Observation d4940b65-64ec-42ba-ba6c-8de817f6569e · inbound

A Comprehensive Overview of Large Language Models cites this paper.

A Comprehensive Overview of Large Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 108

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:28:39.371703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T20:28:38.900026Z digest=sha256:67eed43a01576283b643c04d1b5f005ea9b2e42598f99652844be02d8b5e5380

Observation 42886ca4-78e6-47dc-9cee-691e7c3a9df0 · inbound

The Falcon Series of Open Language Models cites this paper.

The Falcon Series of Open Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 82

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T09:46:10.086221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-16T09:46:09.701440Z digest=sha256:29fa42c371ef1b2a851ae381b30e58529789bfab76181653603a909e94ae5e3e

Observation eb90aaa4-10ba-4ef9-ae03-b7e001ece8c2 · inbound

Occ-LLM: Enhancing Autonomous Driving with Occupancy-Based Large Language Models cites this paper.

Occ-LLM: Enhancing Autonomous Driving with Occupancy-Based Large Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-08T15:33:48.493305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:33:48.493305Z digest=sha256:4f85973821709f6018c30a17bb2372ea20a3f0d30dde5c91f710179b33d4dfbf

Observation 67e8d5f3-3fde-4085-ada6-97234ce1031f · inbound

EvdCLIP: Improving Vision-Language Retrieval with Entity Visual Descriptions from Large Language Models cites this paper.

EvdCLIP: Improving Vision-Language Retrieval with Entity Visual Descriptions from Large Language Models PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:32:41.114814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:32:41.114814Z digest=sha256:1f879a496a5e958913331047858e065bf169e02a2dc02b6e6637a86f5db55342

Observation 76bc0cd3-7146-40ba-8227-3fc877d7551a · inbound

Scalable Complexity Control Facilitates Reasoning Ability of LLMs cites this paper.

Scalable Complexity Control Facilitates Reasoning Ability of LLMs PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:15.459004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:01:15.459004Z digest=sha256:e89b112ab338c42b75a33436e491f478d58eb95c1f6c9a38bff6e0bb61575d35

Observation 29c79227-7137-4cda-a89f-732ea61f9406 · inbound

Chengyu-Bench: Benchmarking Large Language Models for Chinese Idiom Understanding and Use cites this paper.

Chengyu-Bench: Benchmarking Large Language Models for Chinese Idiom Understanding and Use PanGu-$\alpha$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:19.553332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:27:19.553332Z digest=sha256:124754bfeb76b65ae6aebd1ddc41539a0cbbd1c61f8ac00b0549d74a03f67045