Pith. sign in

Paper Citation Record · LEDGER

On Path to Multimodal Generalist: General-Level and General-Bench

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2505.04620.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.04620 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:37:43.985549Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:06:21.338407Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2b532ea4-62d2-495c-80ce-b2bdeed8c3b1 · inbound

Circle-RoPE: Cone-like Decoupled Rotary Positional Embedding for Large Vision-Language Models cites this paper.

Circle-RoPE: Cone-like Decoupled Rotary Positional Embedding for Large Vision-Language Models On Path to Multimodal Generalist: General-Level and General-Bench

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-22T14:21:39.794474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T14:19:34.622854Z digest=sha256:d9db42f1ece1ed10a9dcbb21b1c580e2dcc0620ecc39459068debc8a55cbe74b

Observation 1b27fa32-f58d-4a23-b944-4dd3a2c638d1 · inbound

Mixed-R1: Unified Reward Perspective For Reasoning Capability in Multimodal Large Language Models cites this paper.

Mixed-R1: Unified Reward Perspective For Reasoning Capability in Multimodal Large Language Models On Path to Multimodal Generalist: General-Level and General-Bench

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:37:43.985549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:37:43.985549Z digest=sha256:b16e7e0d8201d685b891e8919103d53319d801858a93454fc56e8f8f6cc9a778

Observation 391054ec-1810-4d02-8d9e-4bf38407ecfb · inbound

Dense360: Dense Understanding from Omnidirectional Panoramas cites this paper.

Dense360: Dense Understanding from Omnidirectional Panoramas On Path to Multimodal Generalist: General-Level and General-Bench

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:23:48.104575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:23:48.104575Z digest=sha256:6b5445dca490b2f45c74348dd39888357190ea4c2bdfa39ebc3a295909fa7d76

Observation 43432d11-d636-49d1-ba2b-56cb6e0df3fb · inbound

AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs cites this paper.

AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs On Path to Multimodal Generalist: General-Level and General-Bench

Reference 63

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:06:21.340815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T14:36:53.295540Z digest=sha256:58d2792fe632874ea731f6c09f5b6e1fbb38859f1d8a89ed24fc45e1751242da