Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 26 inbound Pith citation observations for arXiv:2304.03208.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T13:44:00.474327Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
24
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 72709979-5cf3-409f-93c8-87bf672b1828 · inbound
The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 11ce8e20-346e-4ee5-9948-e01920c03746 · inbound
The Falcon Series of Open Language Models Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 273
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7ded08bb-9350-4ace-a73c-fa0057af11eb · inbound
MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7f2cbd0f-da97-422e-8f01-f1548ff192de · inbound
Small Language Models (SLMs) Can Still Pack a Punch: A survey (updated 2026) Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8d69a650-a8bd-4bfe-bada-f1598e01d261 · inbound
SHARP: Accelerating Language Model Inference by SHaring Adjacent layers with Recovery Parameters Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d10b42b-515c-40a5-b47a-0fafba6430f3 · inbound
Get Experience from Practice: LLM Agents with Record & Replay Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df1ccc47-c817-46e5-b64c-4ed0e368cc4e · inbound
MuLoCo: Muon is a practical inner optimizer for DiLoCo Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b19a67a-b969-4e9b-bd05-f3e374efaaa7 · inbound
Basis Transformers for Multi-Task Tabular Regression Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c52a78a-a6e0-4ad8-a59b-a3e36ad36ad0 · inbound
Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 696b93b9-96f5-49dc-af04-c4a478f741f8 · inbound
Similarity Field Theory: A Mathematical Framework for Intelligence Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f6d2d069-c6f6-451d-9f9d-23f49452f5d0 · inbound
SpaDA: A Spatial Dataflow Architecture Programming Language Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1c9812f9-a439-4926-ad6d-d58639597ea3 · inbound
Spectral Condition for $\mu$P under Width-Depth Scaling Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3030597b-90d3-49d4-b542-8bcf4d65b7e0 · inbound
Learning the Signature of Memorization in Autoregressive Language Models Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 397fe58b-91bd-4d8c-832d-adabb3ac28a1 · inbound
Don't Waste Bits! Adaptive KV-Cache Quantization for Lightweight On-Device LLMs Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a5243836-dd44-4be2-b867-29d28b90e009 · inbound
Better and Worse with Scale: How Contextual Entrainment Diverges with Model Size Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e53f190d-6c90-453b-8425-d49eb776e48f · inbound
GQA-{\mu}P: The maximal parameterization update for grouped query attention Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 33fd6283-7502-4d89-822b-da9af1e96e67 · inbound
Transformer Scalability Crisis: The First Comprehensive Empirical Analysis of Performance Walls in Modern Language Models Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e9e163fd-e946-4b77-a4fc-472196530f2d · inbound
How Much Is a Dataset Worth? Scaling Laws, the Vendi Score, and Matrix Spectral Functions Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fd334a5c-971e-46d0-917f-5e35d581e4cb · inbound
How Much Is a Dataset Worth? Scaling Laws, the Vendi Score, and Matrix Spectral Functions Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4d79cee-fbab-475c-a54c-8e6d81140666 · inbound
Validity Threats for Foundation Model Research Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 83636f29-a75c-4049-8d5e-065e442f67a7 · inbound
Predictable Scaling Laws of Optimal Hyperparameters for LLM Continued Pre-training Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 345089cb-143d-4052-8b8a-99f84689ca1e · inbound
Statistical Properties of Training & Generalization Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4b39796f-3374-4299-852e-1e98a5359572 · inbound
Statistical Properties of Training & Generalization Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation be861fad-ea9a-493a-834d-189e4735834a · inbound
SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7eccc6d1-5602-4853-ae63-1f8ab990bf7f · inbound
SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 970f52ba-dd88-4742-ad29-5b29c5482ca2 · inbound
Sentence-Level Contextual Entrainment in Large Language Models Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.