Pith. sign in

Paper Citation Record · LEDGER

Large scale distributed neural network training through online distillation

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:1804.03235.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1804.03235 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:48:56.719676Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

152
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b4290041-96f0-4152-9a82-96f990142131 · inbound

Emerging Properties in Self-Supervised Vision Transformers cites this paper.

Emerging Properties in Self-Supervised Vision Transformers Large scale distributed neural network training through online distillation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:04:51.498387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-16T14:04:51.458382Z digest=sha256:ea52c797c964790afd832b8ee5e66d7b519f5170ea2c27dc563d0e89a2b76b73

Observation b717688e-bec4-43c8-9fe5-b5d3e8375e3b · inbound

Vision Transformers Need Registers cites this paper.

Vision Transformers Need Registers Large scale distributed neural network training through online distillation

Reference 182

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T09:41:38.167514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-13T09:41:37.937046Z digest=sha256:58cdc072e14f9845a39248f235088e274ecdc2f447d5ab873c3baf93947ab800

Observation ce14800f-1557-42fb-87fc-dd9ac110c15d · inbound

Gemma 2: Improving Open Language Models at a Practical Size cites this paper.

Gemma 2: Improving Open Language Models at a Practical Size Large scale distributed neural network training through online distillation

Reference 159

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T12:11:16.544210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-10T12:11:16.326752Z digest=sha256:73980ae8ce379ee7e2853b12e28af65ed05edb4cf6f7c477867c8a843b8cad21

Observation 1b6250d1-edef-448b-a617-b70536537829 · inbound

Beyond Model Scale Limits: End-Edge-Cloud Federated Learning with Self-Rectified Knowledge Agglomeration cites this paper.

Beyond Model Scale Limits: End-Edge-Cloud Federated Learning with Self-Rectified Knowledge Agglomeration Large scale distributed neural network training through online distillation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T22:48:56.719676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:48:56.719676Z digest=sha256:387ec5e0c22d1478e7b44226cda60c63556cec79663ccefd48036635ad1dcac7

Observation 23fca9e2-d940-40cc-919f-24fd1e71bf8e · inbound

Gemma 3 Technical Report cites this paper.

Gemma 3 Technical Report Large scale distributed neural network training through online distillation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-22T22:22:12.260167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T22:18:55.976503Z digest=sha256:0801291e5ceb3d49e4fa69b54279f749dcbf339420d2f66b4b2793c24d932295

Observation d8b87f14-6a0b-4696-b8be-4cb61f993b47 · inbound

Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities cites this paper.

Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities Large scale distributed neural network training through online distillation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:52:07.665258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-19T05:48:02.828938Z digest=sha256:e04c5298ecf6b3c85071bd7e321d075494ff6b03c67d6d8dcfa107954ed82c9e

Observation c4c0afed-2076-466a-a5b2-92dd9baf6b6c · inbound

Optimal Transceiver Design in Over-the-Air Federated Distillation cites this paper.

Optimal Transceiver Design in Over-the-Air Federated Distillation Large scale distributed neural network training through online distillation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:43:46.990439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:43:46.990439Z digest=sha256:c4cb4c978ab46ea521d200ea342cd0ba7b7f03538aff99e1ebf4d1dcb03672f2

Observation 2b55ff7a-7fd7-4257-9a7e-274e60bbda3f · inbound

The Ratchet Effect in Silico: How Interaction Drives Cumulative Intelligence in Large Language Models cites this paper.

The Ratchet Effect in Silico: How Interaction Drives Cumulative Intelligence in Large Language Models Large scale distributed neural network training through online distillation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-19T02:42:00.002358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-19T02:38:27.510872Z digest=sha256:704069b916a60bd982c57faf09f16102e83875f022be0811c243e6aaf0ade1fe

Observation 95e477ed-9046-42c0-96bd-44e82e0536af · inbound

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering cites this paper.

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering Large scale distributed neural network training through online distillation

Reference 178

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T23:54:45.166952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-09T23:51:47.724033Z digest=sha256:b9b9146f0551a956e7154d14d8f0bf4a50dae1d1d4b9fd5632ac7939c5c89a4b

Observation 2c0d22fa-eedf-4fff-adbc-f7a777f31145 · inbound

LatentBurst: A Fast and Efficient Multi Frame Super-Resolution for Hexadeca-Bayer Pattern CIS images cites this paper.

LatentBurst: A Fast and Efficient Multi Frame Super-Resolution for Hexadeca-Bayer Pattern CIS images Large scale distributed neural network training through online distillation

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:31:11.968340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T08:48:09.398221Z digest=sha256:dd24d2b00f6816c4f59db556d27c216ed02918ba1804616266fc4b1c96ffffc8

Observation 59aebcfc-3205-4747-9a21-72a29377f62c · inbound

Enabling Federated Inference via Unsupervised Consensus Embedding cites this paper.

Enabling Federated Inference via Unsupervised Consensus Embedding Large scale distributed neural network training through online distillation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:36:07.875551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T15:06:27.437479Z digest=sha256:b266bba79e84f0019d53994a82ac82102ccfb7b9ae9e4bec8479d9ad037136b4

Observation 65ed5b5e-5cc9-4a64-9119-4dc667e6c3ef · inbound

Function-Space ADMM for Decentralized Federated Learning: A Control Theoretic Perspective cites this paper.

Function-Space ADMM for Decentralized Federated Learning: A Control Theoretic Perspective Large scale distributed neural network training through online distillation

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:36:26.935644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-12T04:07:54.716306Z digest=sha256:59d1176acb7673b4d4c39bc6829376fe19e2d9576cddccbeaeec87d355a276ef

Observation 7ba269ad-7474-4973-a851-add08194d938 · inbound

Optimized Federated Knowledge Distillation with Distributed Neural Architecture Search cites this paper.

Optimized Federated Knowledge Distillation with Distributed Neural Architecture Search Large scale distributed neural network training through online distillation

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:43:58.955772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-21T05:42:32.050913Z digest=sha256:ce769b8d3b9467e99a5672760eef964f4d9d817d2c63cfdb9857ea82461fe182

Observation f719f77c-941f-4a0f-8cc4-af0a02c5fb75 · inbound

Scaling Laws for Task-Specific LLM Distillation cites this paper.

Scaling Laws for Task-Specific LLM Distillation Large scale distributed neural network training through online distillation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:40:00.752562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-25T23:29:49.787477Z digest=sha256:7442e3b85e302223042604b6328c5c831e52a6f6f1ad9ccda177d3af451e0f8a

Observation 82af35fa-8f92-4ab7-93cb-f7edf6ec8898 · inbound

Bridging Compute- and Data-Optimal Pretraining cites this paper.

Bridging Compute- and Data-Optimal Pretraining Large scale distributed neural network training through online distillation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-01T03:02:01.645052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:02:01.645052Z digest=sha256:24aa9814cb98049b0b49beaaad7be46555bb6dfbb61013876d8ee0b25d9347ad

Observation 82fd8473-1d5f-43ee-9503-fddb8d10444f · inbound

Slot2Text: Object-Centric Visual Tokenization for Efficient and Spatially Traceable Surgical MLLMs cites this paper.

Slot2Text: Object-Centric Visual Tokenization for Efficient and Spatially Traceable Surgical MLLMs Large scale distributed neural network training through online distillation

Reference 116

Resolution
unresolved
no resolver link, observed 2026-08-06T00:10:37.428710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:10:37.428710Z digest=sha256:f78e72d21a6d1c569f4fef190459a7effb4fbcbcccd9c6c7bf3a606f73e597ac