Pith. sign in

Paper Citation Record · LEDGER

MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2402.04788.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.04788 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T04:23:01.041270Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

5
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cb499709-1d31-4e76-9d12-ef6c59a6eaeb · inbound

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods cites this paper.

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:08:36.075491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T23:08:34.312466Z digest=sha256:2653d54f70ec44d6374ac87425625b1a06b1fcf92c3d1ea867708424c3192a06

Observation 85009f16-0f8c-42a4-9efe-30768a7fe44c · inbound

REALEDIT: Reddit Edits As a Large-scale Empirical Dataset for Image Transformations cites this paper.

REALEDIT: Reddit Edits As a Large-scale Empirical Dataset for Image Transformations MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T04:23:01.041270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:23:01.041270Z digest=sha256:199bf05fb4fa0c59ed298f0867c394023e8944be93e0924aae9f88c77a620b5c

Observation 482411d4-fa0d-4e8a-8461-2d9a13dd25dd · inbound

Survey on AI-Generated Media Detection: From Non-MLLM to MLLM cites this paper.

Survey on AI-Generated Media Detection: From Non-MLLM to MLLM MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 196

Resolution
unresolved
no resolver link, observed 2026-08-08T21:12:23.258498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:12:23.258498Z digest=sha256:38026016bec75aab35fb7df273196d2ca11aeafd37e72a1086ade26a3e307dbb

Observation e68f6283-fb80-43fa-ac1c-763f9e4bb065 · inbound

Towards an AI co-scientist cites this paper.

Towards an AI co-scientist MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:02:44.420423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T13:02:43.571234Z digest=sha256:b4a6468737662aaebd15e0feeb3897d4ef6aed6ae37039daffc85b0ea985ca9c

Observation b34f1779-4abc-4d04-83ea-91e003946f69 · inbound

$I^2G$: Generating Instructional Illustrations via Text-Conditioned Diffusion cites this paper.

$I^2G$: Generating Instructional Illustrations via Text-Conditioned Diffusion MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:03:36.345971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:03:36.345971Z digest=sha256:24142d39d72d1c44e9e6877144bb32a69f51ef598e14c9c4581cceb786350073

Observation 983eec7b-6599-4169-9aea-8f3adc1d6000 · inbound

Chart-to-Experience: Benchmarking Multimodal LLMs for Predicting Experiential Impact of Charts cites this paper.

Chart-to-Experience: Benchmarking Multimodal LLMs for Predicting Experiential Impact of Charts MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:53:05.448409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:53:05.448409Z digest=sha256:209e93fcb7146b46f684bef78a7232d88bebc0c0d599eecdd9348db98acd05dc

Observation 9b0b113d-24da-4dca-943b-db2765bbe411 · inbound

HiGarment: Cross-modal Harmony Based Diffusion Model for Flat Sketch to Realistic Garment Image cites this paper.

HiGarment: Cross-modal Harmony Based Diffusion Model for Flat Sketch to Realistic Garment Image MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:55:50.251731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:55:50.251731Z digest=sha256:57dcb2ed2c6635350e64ac8f0bdec6ec7e284153a147c02d8bdb9942b8de9366

Observation 206b4e58-b2e4-4ecb-9bce-193702e95578 · inbound

Fast or Slow? Integrating Fast Intuition and Deliberate Thinking for Enhancing Visual Question Answering cites this paper.

Fast or Slow? Integrating Fast Intuition and Deliberate Thinking for Enhancing Visual Question Answering MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:02:10.486815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:02:10.486815Z digest=sha256:107efc44135770f931f47a3921f276a37f7546581461eb32b2c82256a0662bff

Observation fa370235-2ae2-4352-96ee-a0c211d7ee3a · inbound

IntuiTF: MLLM-Guided Transfer Function Optimization for Direct Volume Rendering cites this paper.

IntuiTF: MLLM-Guided Transfer Function Optimization for Direct Volume Rendering MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:54.855811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:54.855811Z digest=sha256:158334fd06cadbed0469d8e288c524c13bd467ef2b3ddd970245525980f73088

Observation 1a200616-6228-4a03-a1ee-4fe5f100a1a2 · inbound

Adaptive Repetition for Mitigating Position Bias in LLM-Based Ranking cites this paper.

Adaptive Repetition for Mitigating Position Bias in LLM-Based Ranking MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T14:56:03.479513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:56:03.479513Z digest=sha256:9dfedc5664a9d255220c88b81cc604a97dccac51bb3d3eb5a0608d854b95c5f9

Observation a5e47dc2-c021-40db-aded-c3635a193cdb · inbound

Does Pass Rate Tell the Whole Story? Evaluating Design Constraint Compliance in LLM-based Issue Resolution cites this paper.

Does Pass Rate Tell the Whole Story? Evaluating Design Constraint Compliance in LLM-based Issue Resolution MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T18:45:43.535505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:45:06.682021Z digest=sha256:486e8bf006a46208b3eeb549e7340a760cfc623185fe8929ac7f514202948d1b

Observation c474bfe6-1dca-4c9c-a17f-4b61f9cf7df5 · inbound

Deep Pre-Alignment for VLMs cites this paper.

Deep Pre-Alignment for VLMs MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-19T16:27:39.161323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-19T16:26:41.094936Z digest=sha256:5ed5ccada16d60765af93a2890a66b07f00e75af9bf12cc893a742a477219d9b

Observation f58c1f56-4a97-4173-a55b-6348d1112e3b · inbound

NextMotionQA: Benchmarking and Judging Human Motion Understanding with Vision-Language Models cites this paper.

NextMotionQA: Benchmarking and Judging Human Motion Understanding with Vision-Language Models MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:26:46.211206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T06:55:22.332372Z digest=sha256:30e2b577b5741507f7920dd296e1b7d8a02d9a6532b4a8e6362d099f3771a1b7

Observation 43aab11e-2796-4f18-aa2a-f935279c612a · inbound

Brevity is the Soul of Inference Efficiency: Inducing Concision in VLMs via Data Curation cites this paper.

Brevity is the Soul of Inference Efficiency: Inducing Concision in VLMs via Data Curation MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:40:07.108811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-25T21:05:36.836361Z digest=sha256:8562f09bfbd1e6b5866093a2c2d68759114d2493600a2d52b260277d6a33d5bd

Observation 4ae4865e-347b-4249-823b-2178a3383d0d · inbound

Brevity is the Soul of Inference Efficiency: Inducing Concision in VLMs via Data Curation cites this paper.

Brevity is the Soul of Inference Efficiency: Inducing Concision in VLMs via Data Curation MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-01T09:35:39.554250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-01T06:30:27.178950Z digest=sha256:353ec45c533457f5d45bbb6c6ea687c63bf2ca161badd5c4899c66835272b834

Observation 4c088528-20a4-49ec-af11-57dd9238495d · inbound

MentalThink: Shaping Thoughts in Mental SVG World cites this paper.

MentalThink: Shaping Thoughts in Mental SVG World MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 229

Resolution
unresolved
no resolver link, observed 2026-07-12T01:50:59.184754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T01:50:59.184754Z digest=sha256:e27b999fb5c647131feb7645411e00de5be95070f2f2d87342c99de656f179f8

Observation 0f92b4b2-4ed7-4543-910f-cf5899c7687a · inbound

LLM-as-a-Verifier: A General-Purpose Verification Framework cites this paper.

LLM-as-a-Verifier: A General-Purpose Verification Framework MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 70

Resolution
metadata mismatch
local_arxiv, observed 2026-07-07T12:53:50.221998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-07T12:47:29.552283Z digest=sha256:d68dffb2c4765a18ddba6fa3b476a54f5bca9b388bca999064fb23d6eb18fa09

Observation 7fe4ef8e-7ef3-4cbb-90fe-4457cebebf50 · inbound

LLM-as-a-Verifier: A General-Purpose Verification Framework cites this paper.

LLM-as-a-Verifier: A General-Purpose Verification Framework MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 70

Resolution
unresolved
no resolver link, observed 2026-07-11T07:02:51.850836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T07:02:51.850836Z digest=sha256:7a55c0ac795414dbd5fac44a5547ab28945116747b22e18b4a066adbee427f0f

Observation 50983d72-a1f0-41fc-9126-68fad673c8f0 · inbound

OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models cites this paper.

OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-31T02:18:01.617281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T02:18:01.617281Z digest=sha256:7a1e5672cf5bf9634dab65313bf8bf981c82043740ab7dac1265b4c902c32fed

Observation 646e77bc-12dd-4ed3-8be8-7c4c6a2375ff · inbound

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges cites this paper.

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark

Reference 288

Resolution
unresolved
no resolver link, observed 2026-08-03T00:55:42.558312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:55:42.558312Z digest=sha256:1ad925ba18049a4e3f916004939acabbc8156a3f88d727591327a77b3e65a226