Pith. sign in

Paper Citation Record · LEDGER

MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2402.04788.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.04788 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T04:23:01.041270Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

5
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cb499709-1d31-4e76-9d12-ef6c59a6eaeb · inbound

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods cites this paper.

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:08:36.075491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T23:08:34.312466Z digest=sha256:eb374770ba54bbeb9012f20473f1113b0909d217f9d41de638c76f204a8e5e21

Observation 85009f16-0f8c-42a4-9efe-30768a7fe44c · inbound

REALEDIT: Reddit Edits As a Large-scale Empirical Dataset for Image Transformations cites this paper.

REALEDIT: Reddit Edits As a Large-scale Empirical Dataset for Image Transformations MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T04:23:01.041270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:23:01.041270Z digest=sha256:199bf05fb4fa0c59ed298f0867c394023e8944be93e0924aae9f88c77a620b5c

Observation 482411d4-fa0d-4e8a-8461-2d9a13dd25dd · inbound

Survey on AI-Generated Media Detection: From Non-MLLM to MLLM cites this paper.

Survey on AI-Generated Media Detection: From Non-MLLM to MLLM MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 196

Resolution
unresolved
no resolver link, observed 2026-08-08T21:12:23.258498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:12:23.258498Z digest=sha256:5ed1ee050d2427fef3e5af602db233d91ff8f7cf70364b58fdfbbfa78a474079

Observation e68f6283-fb80-43fa-ac1c-763f9e4bb065 · inbound

Towards an AI co-scientist cites this paper.

Towards an AI co-scientist MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:02:44.420423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-11T13:02:43.571234Z digest=sha256:ccc2d14aa62f16109f662c069894ef39fb89cfac3d0e597ee10b7bccdc350313

Observation b34f1779-4abc-4d04-83ea-91e003946f69 · inbound

$I^2G$: Generating Instructional Illustrations via Text-Conditioned Diffusion cites this paper.

$I^2G$: Generating Instructional Illustrations via Text-Conditioned Diffusion MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:03:36.345971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:03:36.345971Z digest=sha256:24142d39d72d1c44e9e6877144bb32a69f51ef598e14c9c4581cceb786350073

Observation 983eec7b-6599-4169-9aea-8f3adc1d6000 · inbound

Chart-to-Experience: Benchmarking Multimodal LLMs for Predicting Experiential Impact of Charts cites this paper.

Chart-to-Experience: Benchmarking Multimodal LLMs for Predicting Experiential Impact of Charts MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:53:05.448409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:53:05.448409Z digest=sha256:209e93fcb7146b46f684bef78a7232d88bebc0c0d599eecdd9348db98acd05dc

Observation 9b0b113d-24da-4dca-943b-db2765bbe411 · inbound

HiGarment: Cross-modal Harmony Based Diffusion Model for Flat Sketch to Realistic Garment Image cites this paper.

HiGarment: Cross-modal Harmony Based Diffusion Model for Flat Sketch to Realistic Garment Image MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:55:50.251731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:55:50.251731Z digest=sha256:57dcb2ed2c6635350e64ac8f0bdec6ec7e284153a147c02d8bdb9942b8de9366

Observation 206b4e58-b2e4-4ecb-9bce-193702e95578 · inbound

Fast or Slow? Integrating Fast Intuition and Deliberate Thinking for Enhancing Visual Question Answering cites this paper.

Fast or Slow? Integrating Fast Intuition and Deliberate Thinking for Enhancing Visual Question Answering MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:02:10.486815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:02:10.486815Z digest=sha256:107efc44135770f931f47a3921f276a37f7546581461eb32b2c82256a0662bff

Observation fa370235-2ae2-4352-96ee-a0c211d7ee3a · inbound

IntuiTF: MLLM-Guided Transfer Function Optimization for Direct Volume Rendering cites this paper.

IntuiTF: MLLM-Guided Transfer Function Optimization for Direct Volume Rendering MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:54.855811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:54.855811Z digest=sha256:4fe2f5b81a1a9112e863cf88081cfc66b89aaa0f9d0c0fba254bf9cb2a75b9a3

Observation 1a200616-6228-4a03-a1ee-4fe5f100a1a2 · inbound

Adaptive Repetition for Mitigating Position Bias in LLM-Based Ranking cites this paper.

Adaptive Repetition for Mitigating Position Bias in LLM-Based Ranking MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T14:56:03.479513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:56:03.479513Z digest=sha256:9dfedc5664a9d255220c88b81cc604a97dccac51bb3d3eb5a0608d854b95c5f9

Observation a5e47dc2-c021-40db-aded-c3635a193cdb · inbound

Does Pass Rate Tell the Whole Story? Evaluating Design Constraint Compliance in LLM-based Issue Resolution cites this paper.

Does Pass Rate Tell the Whole Story? Evaluating Design Constraint Compliance in LLM-based Issue Resolution MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T18:45:43.535505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T18:45:06.682021Z digest=sha256:895ec2d96ba743934b128923358dfd7f5de4dbc0b69a4b945835f73a99db8963

Observation c474bfe6-1dca-4c9c-a17f-4b61f9cf7df5 · inbound

Deep Pre-Alignment for VLMs cites this paper.

Deep Pre-Alignment for VLMs MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-19T16:27:39.161323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-19T16:26:41.094936Z digest=sha256:c57462bd5fcf28a3d9a72492ddeb01412dbc857255f2ba63a0f8bca6685651b6

Observation f58c1f56-4a97-4173-a55b-6348d1112e3b · inbound

NextMotionQA: Benchmarking and Judging Human Motion Understanding with Vision-Language Models cites this paper.

NextMotionQA: Benchmarking and Judging Human Motion Understanding with Vision-Language Models MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:26:46.211206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T06:55:22.332372Z digest=sha256:67ea37733970e30bef04710b9f3b02447ccb086abf7b160742fb4e2643989a64

Observation 43aab11e-2796-4f18-aa2a-f935279c612a · inbound

Brevity is the Soul of Inference Efficiency: Inducing Concision in VLMs via Data Curation cites this paper.

Brevity is the Soul of Inference Efficiency: Inducing Concision in VLMs via Data Curation MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:40:07.108811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-25T21:05:36.836361Z digest=sha256:a2285d2a4d6ee0f35243c4484c2c85c40a69cb7b51acfa2598d2620b6e8ab681

Observation 4ae4865e-347b-4249-823b-2178a3383d0d · inbound

Brevity is the Soul of Inference Efficiency: Inducing Concision in VLMs via Data Curation cites this paper.

Brevity is the Soul of Inference Efficiency: Inducing Concision in VLMs via Data Curation MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-01T09:35:39.554250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-01T06:30:27.178950Z digest=sha256:581b9b6892828c4fdc32a97c743e6df5869e9832b81c07717375c6100b449a71

Observation 4c088528-20a4-49ec-af11-57dd9238495d · inbound

MentalThink: Shaping Thoughts in Mental SVG World cites this paper.

MentalThink: Shaping Thoughts in Mental SVG World MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 229

Resolution
unresolved
no resolver link, observed 2026-07-12T01:50:59.184754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T01:50:59.184754Z digest=sha256:e27b999fb5c647131feb7645411e00de5be95070f2f2d87342c99de656f179f8

Observation 0f92b4b2-4ed7-4543-910f-cf5899c7687a · inbound

LLM-as-a-Verifier: A General-Purpose Verification Framework cites this paper.

LLM-as-a-Verifier: A General-Purpose Verification Framework MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 70

Resolution
metadata mismatch
local_arxiv, observed 2026-07-07T12:53:50.221998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-07T12:47:29.552283Z digest=sha256:8acc770a8cae9934463f5f03502ffa0c7c5371a4a72d74bc7ca4efe221ca543c

Observation 7fe4ef8e-7ef3-4cbb-90fe-4457cebebf50 · inbound

LLM-as-a-Verifier: A General-Purpose Verification Framework cites this paper.

LLM-as-a-Verifier: A General-Purpose Verification Framework MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 70

Resolution
unresolved
no resolver link, observed 2026-07-11T07:02:51.850836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T07:02:51.850836Z digest=sha256:7a55c0ac795414dbd5fac44a5547ab28945116747b22e18b4a066adbee427f0f

Observation 50983d72-a1f0-41fc-9126-68fad673c8f0 · inbound

OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models cites this paper.

OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-31T02:18:01.617281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T02:18:01.617281Z digest=sha256:60e27d23de8e682250dd7627e81b792d1f0d9d2b4c7ed2699365d2d31eb43159

Observation 646e77bc-12dd-4ed3-8be8-7c4c6a2375ff · inbound

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges cites this paper.

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 288

Resolution
unresolved
no resolver link, observed 2026-08-03T00:55:42.558312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:55:42.558312Z digest=sha256:1ad925ba18049a4e3f916004939acabbc8156a3f88d727591327a77b3e65a226