Pith. sign in

Paper Citation Record · LEDGER

Scaling Laws for Native Multimodal Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2504.07951.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.07951 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:35:54.984579Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T22:18:21.000835Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3e059d4e-8c73-46df-a29c-1c8a4bb510f2 · inbound

OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation cites this paper.

OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation Scaling Laws for Native Multimodal Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:17.154880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:44:17.154880Z digest=sha256:924cafaba0de27f35d9b8179c3752f834904764393b10d07b42cba7e6d2beda4

Observation a73140c5-91fa-407b-99ab-3bd6fb1bcb23 · inbound

SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics cites this paper.

SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics Scaling Laws for Native Multimodal Models

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:22:37.522985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T21:22:36.902119Z digest=sha256:a80743749c58497a6c2405739b0e30344ecb47bf46fc160e7f7dc2f1fe07cb9e

Observation ab65bb51-1a9c-493c-a589-2290c2ca0f6a · inbound

"GenAI Defaults to Bias!" Gamify AI Literacy Through Reflections on Prompts cites this paper.

"GenAI Defaults to Bias!" Gamify AI Literacy Through Reflections on Prompts Scaling Laws for Native Multimodal Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-04T16:29:41.438138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:29:41.438138Z digest=sha256:d85f0f3e87f2b9aebd7ad5b94b4b39277314862b87f315647664b8c2153a9a31

Observation 1aa1e027-51e4-4a93-b05e-9bd8929c139e · inbound

Lifting Unlabeled Internet-level Data for 3D Scene Understanding cites this paper.

Lifting Unlabeled Internet-level Data for 3D Scene Understanding Scaling Laws for Native Multimodal Models

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:18:21.002398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T22:16:57.890955Z digest=sha256:b4215c6e24a4ea58a1b33738641277b11df86cc48363dc7a5f4fdd186030ce49

Observation 680fb1d0-593e-4c6a-ba5f-bd38ff94ee31 · inbound

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning cites this paper.

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning Scaling Laws for Native Multimodal Models

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:11:04.482942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T16:46:36.010169Z digest=sha256:87249d9ebb197229288573dffa9f44ece97673509555254a958aa7bba3af3172

Observation a996b843-037d-46e4-b743-3416d5c84dff · inbound

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning cites this paper.

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning Scaling Laws for Native Multimodal Models

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:01:26.725349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T04:38:06.774877Z digest=sha256:c3f9a0ddfb185aef6fa1b05c682f57b5509873675082ad3e234e9a06ca5f0b74

Observation fc704cb3-8539-48ec-97c2-62a6e2b2617f · inbound

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning cites this paper.

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning Scaling Laws for Native Multimodal Models

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:27:29.457255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T07:26:59.917150Z digest=sha256:34b40756e6d2def81c1fe10b86f26107b8fd6c32423d3b07eddbaa74f1ae531f

Observation bd6170f1-0226-4a3d-bea7-347a92712a4a · inbound

OmniMouse: Scaling properties of multi-modal, multi-task Brain Models on 150B Neural Tokens cites this paper.

OmniMouse: Scaling properties of multi-modal, multi-task Brain Models on 150B Neural Tokens Scaling Laws for Native Multimodal Models

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:51:04.979363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T02:43:25.842048Z digest=sha256:37f60a21da84ec70a183bcfc1ec000c39c2bb74e6b8df5e84828ca20d23767ac

Observation 7b91f7f1-c2a2-4df5-8000-99c61ff51fb2 · inbound

On the Invariance and Generality of Neural Scaling Laws cites this paper.

On the Invariance and Generality of Neural Scaling Laws Scaling Laws for Native Multimodal Models

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:15:56.057740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T02:34:14.087140Z digest=sha256:c9604efe45e7ec8d93f6aad4532b7c21551c081ee463fb38069286b8d7e65696

Observation 7d89982d-90d5-472a-86e0-65c1a77ebeda · inbound

ZAYA1-VL-8B Technical Report cites this paper.

ZAYA1-VL-8B Technical Report Scaling Laws for Native Multimodal Models

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:21:23.450737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T01:15:16.607346Z digest=sha256:136f6de713f07dbbedbf3baabb96de96164a26b664dd5502709310a23bab9558

Observation 019656ee-8ff4-4a57-809e-ddd0ef5fdde9 · inbound

What Transfers from Text to Vision? Capability Scaling Laws and Transfer Dynamics for VLMs cites this paper.

What Transfers from Text to Vision? Capability Scaling Laws and Transfer Dynamics for VLMs Scaling Laws for Native Multimodal Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T02:46:49.703170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:46:49.703170Z digest=sha256:6794b20d21a9a86b848828d640d45492c8fdad686964fbf64b4466e4c0fd725d

Observation e0ea58d5-84a9-4a15-806a-a90366ac8381 · inbound

Routing Is Least Learnable Where It Is Most Valuable: Bounds on Representation Routing for Web Agents cites this paper.

Routing Is Least Learnable Where It Is Most Valuable: Bounds on Representation Routing for Web Agents Scaling Laws for Native Multimodal Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T13:35:54.984579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:35:54.984579Z digest=sha256:47a42e6bc1b32da9330a4a916c9db9950323eb8a3975c4e5c38caa9073c067d4