Pith. sign in

Paper Citation Record · LEDGER

LLaVA-Critic: Learning to Evaluate Multimodal Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 26 inbound Pith citation observations for arXiv:2410.02712.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.02712 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 26 of 26 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T18:23:50.433458Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-07T12:53:50.337796Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 9d0b3fe6-8060-4e87-bae6-d2bb2be78b64 · inbound

A Survey on LLM-as-a-Judge cites this paper.

A Survey on LLM-as-a-Judge LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 181

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:35:44.299511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T17:33:13.394338Z digest=sha256:ca0ffaad9a818fa7d34e950d10c0fa4772d36e9614d3c794c6292fbbf6071bb5

Observation e52cd381-b29d-4417-a40f-1d6d1bf25aec · inbound

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods cites this paper.

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 262

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:08:35.755399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T23:08:34.312466Z digest=sha256:20ab5e81a3ddad9bcfc1e348fe00fdfe20342a2a0f3b64d964790206e29ef226

Observation 3c26313c-026e-4330-aa1a-b358d00c869b · inbound

MM-RLHF: The Next Step Forward in Multimodal LLM Alignment cites this paper.

MM-RLHF: The Next Step Forward in Multimodal LLM Alignment LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T18:23:50.433458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T18:23:50.433458Z digest=sha256:eb1547bfec7efc2dbb40046ad835476ec2f3e1974ffeffbef9d6e7df9c853acd

Observation 966cb689-a729-4653-b51c-3bbaa0857b33 · inbound

Unified Reward Model for Multimodal Understanding and Generation cites this paper.

Unified Reward Model for Multimodal Understanding and Generation LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T00:44:30.605743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T00:44:30.558048Z digest=sha256:691275490d86313cb18f6d2fad946f6789e7bf3991e005f5538a181328b3f703

Observation 9c0703a0-d95f-4e4d-a78c-9d84baac40c0 · inbound

UniGen: Enhanced Training & Test-Time Strategies for Unified Multimodal Understanding and Generation cites this paper.

UniGen: Enhanced Training & Test-Time Strategies for Unified Multimodal Understanding and Generation LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T15:34:55.848923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:34:55.848923Z digest=sha256:5ef14c22aa88211754cafe3a4d41b3750351a1539a142b87ae60976d10e04d91

Observation 4ea970ee-a6be-42d2-a670-b4131defb727 · inbound

Understanding Generative AI Capabilities in Everyday Image Editing Tasks cites this paper.

Understanding Generative AI Capabilities in Everyday Image Editing Tasks LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:24.902387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:24.902387Z digest=sha256:6e52692b0af23b223d0816682d35d13af462b5831118babbe7ac4d3919bb6dde

Observation a30d78d2-7615-43c0-8935-7779e8671eeb · inbound

Generative RLHF-V: Learning Principles from Multi-modal Human Preference cites this paper.

Generative RLHF-V: Learning Principles from Multi-modal Human Preference LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:34:53.906318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:34:53.906318Z digest=sha256:ce7a4862d06e0c25b0b8190f3c5d612cbd7d2433afd6e78f148ff3bb9c597a33

Observation 50b6b785-2e1b-45ee-bdcc-8e2329ea570a · inbound

Chain-of-Zoom: Extreme Super-Resolution via Scale Autoregression and Preference Alignment cites this paper.

Chain-of-Zoom: Extreme Super-Resolution via Scale Autoregression and Preference Alignment LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-19T13:32:19.156594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T13:32:06.974109Z digest=sha256:fd4b140b4ab3af8eae9233123e6b310af4bfdda0043409e29c504f1b34198289

Observation 1308ad25-3d16-4e2b-b6bf-4c6b1eff3632 · inbound

AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time cites this paper.

AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:45.496701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:45.496701Z digest=sha256:8256ad09ec0136beba6dab6abf97fe109eb480c53538413eed27a6c71d936fed

Observation 406e5081-45d4-448a-93c0-300a95aa87e7 · inbound

Evaluating MLLMs with Multimodal Multi-image Reasoning Benchmark cites this paper.

Evaluating MLLMs with Multimodal Multi-image Reasoning Benchmark LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:05:19.667340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:05:19.667340Z digest=sha256:1db5dc778436e38c3f878fd01e00e94eec617e60d85aedde4ac16b9e61d6cc5f

Observation 7219ed15-c1e1-48dd-91da-592f664f44aa · inbound

LeanPO: Lean Preference Optimization for Likelihood Alignment in Video-LLMs cites this paper.

LeanPO: Lean Preference Optimization for Likelihood Alignment in Video-LLMs LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:49.885639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:49.885639Z digest=sha256:e204720cbf9728ca8cde3cf985b96777329d039846bd2099f47ad589153dbb7c

Observation b15cf5c1-0f16-4079-9e14-187d98653191 · inbound

ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs cites this paper.

ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T04:40:12.386906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:40:12.386906Z digest=sha256:74cc8310975389285ed2cdc8a7f4fb8f24999645f8fdec0828977e75310b554b

Observation e3a2a723-284e-4cfc-bbcb-7508e1d319c5 · inbound

VL-GenRM: Enhancing Vision-Language Verification via Vision Experts and Iterative Training cites this paper.

VL-GenRM: Enhancing Vision-Language Verification via Vision Experts and Iterative Training LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T00:34:00.668832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:34:00.668832Z digest=sha256:15aa19af8ccf89b20ce08f67ed83bffede3a49916e79f4d322ab1641fbcabff0

Observation d599e882-1779-426b-a9c6-f0a1c124310d · inbound

Dual-Stage Value-Guided Inference with Margin-Based Reward Adjustment for Fast and Faithful VLM Captioning cites this paper.

Dual-Stage Value-Guided Inference with Margin-Based Reward Adjustment for Fast and Faithful VLM Captioning LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:22.277104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:22.277104Z digest=sha256:46443fa0fa7f23e0122a4124d5ee929299f6b86d8f63919ee95653f87dbcd1a7

Observation 708e58dd-3871-433a-9b07-054563273b3e · inbound

MCA-LLaVA: Manhattan Causal Attention for Reducing Hallucination in Large Vision-Language Models cites this paper.

MCA-LLaVA: Manhattan Causal Attention for Reducing Hallucination in Large Vision-Language Models LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T18:09:10.589173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:09:10.589173Z digest=sha256:975775bbab79aea27653076eae5ab28931427bc63ce0d65205c53b1c90b06fbf

Observation ff5dffb5-e841-4586-a7a9-99f3cd9583dc · inbound

AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation cites this paper.

AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T16:47:58.673151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:47:58.673151Z digest=sha256:54ce0473846746afd3b049aa61356f6fe5a3acf5c2b680905efc58ca8e88a0d4

Observation 20751638-ba9e-4f60-84aa-533399259549 · inbound

VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering cites this paper.

VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T04:49:42.003797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:49:42.003797Z digest=sha256:c82a2c274fcd53dcf0d7542ee39ec00fc0df162dac181aeb72c877e61b27355a

Observation d55eb9da-3eb0-4479-8494-2aca89b2e0a5 · inbound

LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model cites this paper.

LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T13:24:39.903930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:24:39.903930Z digest=sha256:580149b21b84b6ccf9800932fb14242a0f170aca0bce9fd3b469058c870c2df1

Observation 066bc149-5148-4573-8a15-86059f0bc9a7 · inbound

Improving Large Vision and Language Models by Learning from a Panel of Peers cites this paper.

Improving Large Vision and Language Models by Learning from a Panel of Peers LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-05T12:27:27.660164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:27:27.660164Z digest=sha256:25d1d531ff3b423bfc3af979eabfb2d59518ba099fb6024bbd6c5937a32c7975

Observation f21fb5ec-3b3f-44fa-9346-49a2ae2c8446 · inbound

VCap: Hypergeometric Rewards for Weak-to-Strong Visual Captioning cites this paper.

VCap: Hypergeometric Rewards for Weak-to-Strong Visual Captioning LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:23:28.486548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T13:13:57.599970Z digest=sha256:334b0b65a51506157773b336d6d9adbc92b1d6f526b763112e5aa8e27e122ad3

Observation aed827bd-0d3e-400f-8016-c16370e66060 · inbound

PortraitGen: Exemplar-Driven GRPO with Dual-Reward Guidance for Photorealistic Portrait Generation cites this paper.

PortraitGen: Exemplar-Driven GRPO with Dual-Reward Guidance for Photorealistic Portrait Generation LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:29:50.912396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T05:14:14.053344Z digest=sha256:a808bc7499c03da3341a3a44ff034fc007202b588b387e9946ad83286c2eaaf3

Observation add352b4-c69a-46df-9dff-212897c68f86 · inbound

MentalThink: Shaping Thoughts in Mental SVG World cites this paper.

MentalThink: Shaping Thoughts in Mental SVG World LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 244

Resolution
unresolved
no resolver link, observed 2026-07-12T01:50:59.184754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T01:50:59.184754Z digest=sha256:f25ff72b03d51de6558ec8d75298015a970d193924baf37801baf75eac0fdcd4

Observation d3572975-6093-40d7-8943-a47a6e29b6ce · inbound

LLM-as-a-Verifier: A General-Purpose Verification Framework cites this paper.

LLM-as-a-Verifier: A General-Purpose Verification Framework LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 72

Resolution
metadata mismatch
local_arxiv, observed 2026-07-07T12:53:50.339357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-07T12:47:29.552283Z digest=sha256:f20a81de38ab7c728bf234c0c8c55be1fe53c7d27ee52a5faac2a11772de30cd

Observation ad819da2-2bf2-4957-82d1-61e88aafbb35 · inbound

LLM-as-a-Verifier: A General-Purpose Verification Framework cites this paper.

LLM-as-a-Verifier: A General-Purpose Verification Framework LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-07-11T07:02:51.850836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T07:02:51.850836Z digest=sha256:e01e1af7c0b7a6d7b9393a8bd3b4b3c7298b8bbc7c0499a387df416efb199115

Observation bb689320-823d-4dbb-b920-2895516fd83f · inbound

Gaze-to-text Generation: Beyond Categorical Decoding of Human Attention cites this paper.

Gaze-to-text Generation: Beyond Categorical Decoding of Human Attention LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 81

Resolution
unresolved
no resolver link, observed 2026-07-31T23:36:32.551429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:36:32.551429Z digest=sha256:60ca46b71d82d5c9fb6a2ed68dbc3fdc858268d595eafdbd16b34a1a01a111b6

Observation 53b7be44-7b3f-41ff-a162-86e500666011 · inbound

Poplar: A Scalable Pipeline for Human-Centric Image Dataset Synthesis cites this paper.

Poplar: A Scalable Pipeline for Human-Centric Image Dataset Synthesis LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T01:00:07.400569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T01:00:07.400569Z digest=sha256:46d540e87f463ac03451795dced80b2e3c0262c662227407374abe44ad2c7dd8