Pith. sign in

Paper Citation Record · LEDGER

Scaling Laws for Native Multimodal Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2504.07951.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.07951 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:35:54.984579Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T22:18:21.000835Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3e059d4e-8c73-46df-a29c-1c8a4bb510f2 · inbound

OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation cites this paper.

OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation Scaling Laws for Native Multimodal Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:17.154880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:44:17.154880Z digest=sha256:924cafaba0de27f35d9b8179c3752f834904764393b10d07b42cba7e6d2beda4

Observation a73140c5-91fa-407b-99ab-3bd6fb1bcb23 · inbound

SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics cites this paper.

SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics Scaling Laws for Native Multimodal Models

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:22:37.522985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T21:22:36.902119Z digest=sha256:ba79e5a9b973e63d9caa4624296dc5014c2657840ce4cbfca84c3249a4804a09

Observation ab65bb51-1a9c-493c-a589-2290c2ca0f6a · inbound

"GenAI Defaults to Bias!" Gamify AI Literacy Through Reflections on Prompts cites this paper.

"GenAI Defaults to Bias!" Gamify AI Literacy Through Reflections on Prompts Scaling Laws for Native Multimodal Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-04T16:29:41.438138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:29:41.438138Z digest=sha256:853efbbd81e54831685198ac85ce37994ddf2686d6808bd11dce0b6cf79a0cc8

Observation 1aa1e027-51e4-4a93-b05e-9bd8929c139e · inbound

Lifting Unlabeled Internet-level Data for 3D Scene Understanding cites this paper.

Lifting Unlabeled Internet-level Data for 3D Scene Understanding Scaling Laws for Native Multimodal Models

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:18:21.002398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T22:16:57.890955Z digest=sha256:9f0093efb1282bb39ccb7ed80ba69ec7ae4a313b306f7fee9825c3a2bb147f43

Observation 680fb1d0-593e-4c6a-ba5f-bd38ff94ee31 · inbound

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning cites this paper.

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning Scaling Laws for Native Multimodal Models

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:11:04.482942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T16:46:36.010169Z digest=sha256:6b0564e7d2c5b18a4ca7e6c627c7e56ec8e0a5840dacf143341d12cce9037103

Observation a996b843-037d-46e4-b743-3416d5c84dff · inbound

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning cites this paper.

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning Scaling Laws for Native Multimodal Models

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:01:26.725349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T04:38:06.774877Z digest=sha256:9849174c5a7610a94dfaf920946f3eaac43f1481a576438c9021b1becae358cd

Observation fc704cb3-8539-48ec-97c2-62a6e2b2617f · inbound

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning cites this paper.

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning Scaling Laws for Native Multimodal Models

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:27:29.457255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T07:26:59.917150Z digest=sha256:b08f7c37efc5acf579fbc6e19c6fc5d39d4e258aacff139ce82d5d4f1112c004

Observation bd6170f1-0226-4a3d-bea7-347a92712a4a · inbound

OmniMouse: Scaling properties of multi-modal, multi-task Brain Models on 150B Neural Tokens cites this paper.

OmniMouse: Scaling properties of multi-modal, multi-task Brain Models on 150B Neural Tokens Scaling Laws for Native Multimodal Models

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:51:04.979363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T02:43:25.842048Z digest=sha256:bbb75c806c24bc8cb255f2c3d410679424ec463ed16ce4e689cad25f30f85e94

Observation 7b91f7f1-c2a2-4df5-8000-99c61ff51fb2 · inbound

On the Invariance and Generality of Neural Scaling Laws cites this paper.

On the Invariance and Generality of Neural Scaling Laws Scaling Laws for Native Multimodal Models

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:15:56.057740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T02:34:14.087140Z digest=sha256:d798d299c05d6e44377a3a68f80cf51596cd72f29ea0c64ca11624e6050420a7

Observation 7d89982d-90d5-472a-86e0-65c1a77ebeda · inbound

ZAYA1-VL-8B Technical Report cites this paper.

ZAYA1-VL-8B Technical Report Scaling Laws for Native Multimodal Models

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:21:23.450737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T01:15:16.607346Z digest=sha256:9ec3dce2f0baec780d70320adb8f9c0b58b7101b80b9fd40219fd7e257103dea

Observation 019656ee-8ff4-4a57-809e-ddd0ef5fdde9 · inbound

What Transfers from Text to Vision? Capability Scaling Laws and Transfer Dynamics for VLMs cites this paper.

What Transfers from Text to Vision? Capability Scaling Laws and Transfer Dynamics for VLMs Scaling Laws for Native Multimodal Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T02:46:49.703170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:46:49.703170Z digest=sha256:6794b20d21a9a86b848828d640d45492c8fdad686964fbf64b4466e4c0fd725d

Observation e0ea58d5-84a9-4a15-806a-a90366ac8381 · inbound

Routing Is Least Learnable Where It Is Most Valuable: Bounds on Representation Routing for Web Agents cites this paper.

Routing Is Least Learnable Where It Is Most Valuable: Bounds on Representation Routing for Web Agents Scaling Laws for Native Multimodal Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T13:35:54.984579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:35:54.984579Z digest=sha256:8252e84c47a55c7395e22193cf5ad01d0648ba39c186d80bbd656cf40d06ce32