Pith. sign in

Paper Citation Record · LEDGER

LLaVA-Critic: Learning to Evaluate Multimodal Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2410.02712.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.02712 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 27 of 27 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T04:46:25.443666Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-07T12:53:50.337796Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 9d0b3fe6-8060-4e87-bae6-d2bb2be78b64 · inbound

A Survey on LLM-as-a-Judge cites this paper.

A Survey on LLM-as-a-Judge LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 181

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:35:44.299511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T17:33:13.394338Z digest=sha256:7efca30ec032d8dea54279be936f686a3c0f7a4937c65a92df512fe1f36e5f70

Observation e52cd381-b29d-4417-a40f-1d6d1bf25aec · inbound

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods cites this paper.

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 262

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:08:35.755399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T23:08:34.312466Z digest=sha256:5c5f1beaf9c84c43516b757933b190497d83ab0e4fb98363202ea3b3c8e754fb

Observation 033f9bd0-0d6a-4e7c-b6ee-886421e78a90 · inbound

YINYANG-ALIGN: Benchmarking Contradictory Objectives and Proposing Multi-Objective Optimization based DPO for Text-to-Image Alignment cites this paper.

YINYANG-ALIGN: Benchmarking Contradictory Objectives and Proposing Multi-Objective Optimization based DPO for Text-to-Image Alignment LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 2001

Resolution
unresolved
no resolver link, observed 2026-08-09T04:46:25.443666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:46:25.443666Z digest=sha256:da2c2fcc09ea4340d1fd0c943ba91882e04435a202d6887dd7a16c772d8de2d5

Observation 3c26313c-026e-4330-aa1a-b358d00c869b · inbound

MM-RLHF: The Next Step Forward in Multimodal LLM Alignment cites this paper.

MM-RLHF: The Next Step Forward in Multimodal LLM Alignment LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T18:23:50.433458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T18:23:50.433458Z digest=sha256:ce069e5b89c8ead44ad588fbfb7209e2ac2b0d881116f7551d153b2631d213d6

Observation 966cb689-a729-4653-b51c-3bbaa0857b33 · inbound

Unified Reward Model for Multimodal Understanding and Generation cites this paper.

Unified Reward Model for Multimodal Understanding and Generation LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T00:44:30.605743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T00:44:30.558048Z digest=sha256:1162bdfa8f8872e7d28d3e2b05b7f24bed86471c4b068f7edeb95890f0e4f6a5

Observation 9c0703a0-d95f-4e4d-a78c-9d84baac40c0 · inbound

UniGen: Enhanced Training & Test-Time Strategies for Unified Multimodal Understanding and Generation cites this paper.

UniGen: Enhanced Training & Test-Time Strategies for Unified Multimodal Understanding and Generation LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T15:34:55.848923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:34:55.848923Z digest=sha256:5b58d27cb66fce51784475c4e9ed949813fc0e3e4f0985250ed8eb6c5bc5a9b0

Observation 4ea970ee-a6be-42d2-a670-b4131defb727 · inbound

Understanding Generative AI Capabilities in Everyday Image Editing Tasks cites this paper.

Understanding Generative AI Capabilities in Everyday Image Editing Tasks LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:24.902387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:24.902387Z digest=sha256:01d88e7d14ca228a196b304c4485832d43b0cc0e59f9b459d27411c369b38c10

Observation a30d78d2-7615-43c0-8935-7779e8671eeb · inbound

Generative RLHF-V: Learning Principles from Multi-modal Human Preference cites this paper.

Generative RLHF-V: Learning Principles from Multi-modal Human Preference LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:34:53.906318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:34:53.906318Z digest=sha256:ea9ad1ef96f48aee5f54958988061383f2f69ccc39b4ebae74358cb43e3c1c7a

Observation 50b6b785-2e1b-45ee-bdcc-8e2329ea570a · inbound

Chain-of-Zoom: Extreme Super-Resolution via Scale Autoregression and Preference Alignment cites this paper.

Chain-of-Zoom: Extreme Super-Resolution via Scale Autoregression and Preference Alignment LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-19T13:32:19.156594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T13:32:06.974109Z digest=sha256:93040412bc254d23d6ba40ef72e6e8b71f7f5d6979f9b35bc4e55bd6e086c2ea

Observation 1308ad25-3d16-4e2b-b6bf-4c6b1eff3632 · inbound

AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time cites this paper.

AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:45.496701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:45.496701Z digest=sha256:385237cfdb9edb64660f5c503bbf3a926da2980889b09d7de1154d0fa9235879

Observation 406e5081-45d4-448a-93c0-300a95aa87e7 · inbound

Evaluating MLLMs with Multimodal Multi-image Reasoning Benchmark cites this paper.

Evaluating MLLMs with Multimodal Multi-image Reasoning Benchmark LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:05:19.667340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:05:19.667340Z digest=sha256:7db180232557d7f41d262b58ba54be6e2a44448dcd2da9548a268e56f15bb013

Observation 7219ed15-c1e1-48dd-91da-592f664f44aa · inbound

LeanPO: Lean Preference Optimization for Likelihood Alignment in Video-LLMs cites this paper.

LeanPO: Lean Preference Optimization for Likelihood Alignment in Video-LLMs LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:49.885639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:49.885639Z digest=sha256:310ea8a1f1babffd989836add70ea287e212dc06e28ec6cdbf2f67a0aa136757

Observation b15cf5c1-0f16-4079-9e14-187d98653191 · inbound

ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs cites this paper.

ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T04:40:12.386906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:40:12.386906Z digest=sha256:df22abacd64d43ea5a405c7748a9b12c49f1230166b23f66bad5cf422201552e

Observation e3a2a723-284e-4cfc-bbcb-7508e1d319c5 · inbound

VL-GenRM: Enhancing Vision-Language Verification via Vision Experts and Iterative Training cites this paper.

VL-GenRM: Enhancing Vision-Language Verification via Vision Experts and Iterative Training LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T00:34:00.668832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:34:00.668832Z digest=sha256:a84dbb3058f236ead13e06dade5e59d2685aba7313914ca3501e7647bce5aee4

Observation d599e882-1779-426b-a9c6-f0a1c124310d · inbound

Dual-Stage Value-Guided Inference with Margin-Based Reward Adjustment for Fast and Faithful VLM Captioning cites this paper.

Dual-Stage Value-Guided Inference with Margin-Based Reward Adjustment for Fast and Faithful VLM Captioning LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:22.277104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:22.277104Z digest=sha256:c412e2a2a4aa80c2751e762440421e4d7649952155317cfdbf5f0a38a4e13bfb

Observation 708e58dd-3871-433a-9b07-054563273b3e · inbound

MCA-LLaVA: Manhattan Causal Attention for Reducing Hallucination in Large Vision-Language Models cites this paper.

MCA-LLaVA: Manhattan Causal Attention for Reducing Hallucination in Large Vision-Language Models LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T18:09:10.589173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:09:10.589173Z digest=sha256:01a693d21a56a8df80e5b344b58da44adb4bf7732cc177176a73f1e2c2cc1754

Observation ff5dffb5-e841-4586-a7a9-99f3cd9583dc · inbound

AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation cites this paper.

AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T16:47:58.673151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:47:58.673151Z digest=sha256:19adc8921c0ebc0cf39303edf71982f24c612c0effc789b85b25be4a0413a48c

Observation 20751638-ba9e-4f60-84aa-533399259549 · inbound

VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering cites this paper.

VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T04:49:42.003797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:49:42.003797Z digest=sha256:e455e48cf7407671133e280ab850fe2055a29bd42e1f919bb0e78ec7d4b8a4ba

Observation d55eb9da-3eb0-4479-8494-2aca89b2e0a5 · inbound

LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model cites this paper.

LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T13:24:39.903930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:24:39.903930Z digest=sha256:f147c700b86df7e1960eaa82958e58b5c8aa0f55fbe44ac872d941d1b75ff4c9

Observation 066bc149-5148-4573-8a15-86059f0bc9a7 · inbound

Improving Large Vision and Language Models by Learning from a Panel of Peers cites this paper.

Improving Large Vision and Language Models by Learning from a Panel of Peers LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-05T12:27:27.660164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:27:27.660164Z digest=sha256:2694b8e689d4a5ca3080ac140a9b01fe22531a72a803ea2dc22723a9ae7011cd

Observation f21fb5ec-3b3f-44fa-9346-49a2ae2c8446 · inbound

VCap: Hypergeometric Rewards for Weak-to-Strong Visual Captioning cites this paper.

VCap: Hypergeometric Rewards for Weak-to-Strong Visual Captioning LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:23:28.486548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T13:13:57.599970Z digest=sha256:7310553c0158a5f90607e3e676bb66bf55539020ca4e7fb200934f6d70eaa58e

Observation aed827bd-0d3e-400f-8016-c16370e66060 · inbound

PortraitGen: Exemplar-Driven GRPO with Dual-Reward Guidance for Photorealistic Portrait Generation cites this paper.

PortraitGen: Exemplar-Driven GRPO with Dual-Reward Guidance for Photorealistic Portrait Generation LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:29:50.912396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T05:14:14.053344Z digest=sha256:eaa6c723d17e155fb79f92be6bbb2f1d49f1651eb2688a2ca64811019e50f38c

Observation add352b4-c69a-46df-9dff-212897c68f86 · inbound

MentalThink: Shaping Thoughts in Mental SVG World cites this paper.

MentalThink: Shaping Thoughts in Mental SVG World LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 244

Resolution
unresolved
no resolver link, observed 2026-07-12T01:50:59.184754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T01:50:59.184754Z digest=sha256:c88ccd447d376f036a37f23750878e5a0848ddac675cc7d7cf4e7a238ee569a4

Observation d3572975-6093-40d7-8943-a47a6e29b6ce · inbound

LLM-as-a-Verifier: A General-Purpose Verification Framework cites this paper.

LLM-as-a-Verifier: A General-Purpose Verification Framework LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 72

Resolution
metadata mismatch
local_arxiv, observed 2026-07-07T12:53:50.339357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-07T12:47:29.552283Z digest=sha256:049f981831750bc74bca1c81defb9b7920f0e312aaa8004decb8906fa4a4c7e6

Observation ad819da2-2bf2-4957-82d1-61e88aafbb35 · inbound

LLM-as-a-Verifier: A General-Purpose Verification Framework cites this paper.

LLM-as-a-Verifier: A General-Purpose Verification Framework LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-07-11T07:02:51.850836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T07:02:51.850836Z digest=sha256:64e4098861bc60bde1c7303ec7bacb47463568f0247349c05150c84f49840391

Observation bb689320-823d-4dbb-b920-2895516fd83f · inbound

Gaze-to-text Generation: Beyond Categorical Decoding of Human Attention cites this paper.

Gaze-to-text Generation: Beyond Categorical Decoding of Human Attention LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 81

Resolution
unresolved
no resolver link, observed 2026-07-31T23:36:32.551429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:36:32.551429Z digest=sha256:bb31fb1bed4a9b3e1a801a3c7dce607b32db771639e4762c66067313b1d6d1ca

Observation 53b7be44-7b3f-41ff-a162-86e500666011 · inbound

Poplar: A Scalable Pipeline for Human-Centric Image Dataset Synthesis cites this paper.

Poplar: A Scalable Pipeline for Human-Centric Image Dataset Synthesis LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T01:00:07.400569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T01:00:07.400569Z digest=sha256:4b2ba1a6a7473966e65d0364d254a5545d9407999f9d7d441e030411291912a8