Pith. sign in

Paper Citation Record · LEDGER

MEGA-Bench: Scaling Multimodal Evaluation to over 500 Real-World Tasks

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2410.10563.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.10563 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:28:29.775775Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T15:09:55.301348Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation eb001b3c-1e32-447c-aac6-58d4f013fa70 · inbound

OmniGenBench: A Benchmark for Omnipotent Multimodal Generation across 50+ Tasks cites this paper.

OmniGenBench: A Benchmark for Omnipotent Multimodal Generation across 50+ Tasks MEGA-Bench: Scaling Multimodal Evaluation to over 500 Real-World Tasks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:28:29.775775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:28:29.775775Z digest=sha256:fceaddc68be1f6ad95d73cce35786187c3923737715dc868a730808401b7d23c

Observation adb9539c-bbe9-4d46-bc61-7e061c22ee24 · inbound

VFaith: Do Large Multimodal Models Really Reason on Seen Images Rather than Previous Memories? cites this paper.

VFaith: Do Large Multimodal Models Really Reason on Seen Images Rather than Previous Memories? MEGA-Bench: Scaling Multimodal Evaluation to over 500 Real-World Tasks

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T04:07:31.778408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:07:31.778408Z digest=sha256:ca8ceda80ceb76b5b3de5e4ccfa1cf4dcca5f590725f53bc58c6d129eb8bb8e1

Observation 43c1e1de-217a-495a-8173-e0d12cd02b9f · inbound

MARBLE: A Hard Benchmark for Multimodal Spatial Reasoning and Planning cites this paper.

MARBLE: A Hard Benchmark for Multimodal Spatial Reasoning and Planning MEGA-Bench: Scaling Multimodal Evaluation to over 500 Real-World Tasks

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T21:58:34.062989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:58:34.062989Z digest=sha256:5cbe8240fc375710418eacb15f04b655ec25f3b62f7cf1e93b81180785af165b

Observation 1c8db78e-3087-448a-b247-82dc018016e8 · inbound

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation cites this paper.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation MEGA-Bench: Scaling Multimodal Evaluation to over 500 Real-World Tasks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:38.097897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:38.097897Z digest=sha256:99fcacea684990e4b484a2f7971548fcbc9552eeb3edf1f3f6aadeefc93e5647

Observation 56ca49fb-875b-40e5-a123-921594857936 · inbound

Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation cites this paper.

Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation MEGA-Bench: Scaling Multimodal Evaluation to over 500 Real-World Tasks

Reference 191

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:18:43.747618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T04:19:26.332718Z digest=sha256:05df29c62bac88ffed7bc3dc7ebbc398ff2567c465e290a86cd97209c1c705c1

Observation b9a431f3-7681-4989-872e-0e404a3a7d47 · inbound

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models cites this paper.

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models MEGA-Bench: Scaling Multimodal Evaluation to over 500 Real-World Tasks

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:09:55.303282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T01:50:54.242508Z digest=sha256:002b7fa08670f0e8e2beb1d4304237de02653771b77245404079226a516623cd