Pith. sign in

Paper Citation Record · LEDGER

WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2406.11069.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.11069 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T21:56:20.264700Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T22:36:16.753738Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f535af8a-a6cb-43cc-927f-a89c9d3e7798 · inbound

MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans? cites this paper.

MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans? WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:59:32.834674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T07:59:32.638758Z digest=sha256:1dbd15675374d1459ad56b951c025c9ee55d7e7f7b539f9c4c520ba78b2294e8

Observation 7624be59-6f43-4a48-b087-3d8d57ad9fc4 · inbound

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling cites this paper.

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 171

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:23:57.801367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T13:23:57.588851Z digest=sha256:e5017d66a151265057e3454fa7766cd451f7ae5b62db53b183a80f904ae6be72

Observation 9c531192-deec-4d34-a691-ed4fecd7659b · inbound

Copilot Arena: A Platform for Code LLM Evaluation in the Wild cites this paper.

Copilot Arena: A Platform for Code LLM Evaluation in the Wild WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T21:56:20.264700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:56:20.264700Z digest=sha256:87991a08cbdfb6f501a713725f4be919818535e8fa6e97cfc9fb8996b228db9d

Observation 806c8e65-cc4a-4347-88c4-7850e57ad0dd · inbound

MM-RLHF: The Next Step Forward in Multimodal LLM Alignment cites this paper.

MM-RLHF: The Next Step Forward in Multimodal LLM Alignment WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T18:23:50.181923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T18:23:50.181923Z digest=sha256:dd83d8028f8e7a21932fd16b6a57d57438de960e77278ed53467510fddbe9c43

Observation d17ec5f6-459a-455b-b4b0-3829971dd4ef · inbound

Unified Reward Model for Multimodal Understanding and Generation cites this paper.

Unified Reward Model for Multimodal Understanding and Generation WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-14T00:44:30.617150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T00:44:30.558048Z digest=sha256:5164ad73ed15f0450dad4650953cff2dbdf38fc7ce3a2610a835f925ddf9292c

Observation 2b254b4a-fbad-4a33-8c1f-dc89cd8c24f3 · inbound

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models cites this paper.

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:41:08.056502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T13:41:07.991012Z digest=sha256:91ed23a266176ad57494325343ecd7e3526e4af8d720abfddf6c0efceab0e90c

Observation 4cf8f6cb-6683-4aa0-95ff-b22f40faa089 · inbound

VL-GenRM: Enhancing Vision-Language Verification via Vision Experts and Iterative Training cites this paper.

VL-GenRM: Enhancing Vision-Language Verification via Vision Experts and Iterative Training WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T00:34:01.366014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:34:01.366014Z digest=sha256:0e5f3da6edc693d3a90473112c50dc8d1a21b6ba3cd22082c53c05a384154f3d

Observation 1b52840a-aa15-4a01-a719-f930a620946d · inbound

GenRecal: Generation after Recalibration from Large to Small Vision-Language Models cites this paper.

GenRecal: Generation after Recalibration from Large to Small Vision-Language Models WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:24.649231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:24.649231Z digest=sha256:7a2ec2fc9c38a2f17a10078bfa85f4fbbf4abb0065e03d316c1e7696beecd7f7

Observation 81477c90-a19e-45d7-9e93-13cbe146a047 · inbound

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency cites this paper.

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:58:59.198792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T11:58:58.660564Z digest=sha256:ee8999dbc1e43ec6af2666bef2897da15d35e173eb618e443c037f17bcd9b62b

Observation 13178546-7ca2-40d5-bc30-fc30b9c5f40e · inbound

Improving Large Vision and Language Models by Learning from a Panel of Peers cites this paper.

Improving Large Vision and Language Models by Learning from a Panel of Peers WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T12:27:27.424386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:27:27.424386Z digest=sha256:dae392ebe93981a363dbeb02cc78e3d056f86a8d08aacf961e99d3c988dc85f5

Observation 22678309-66f4-49ea-b1f6-3cb9ed156487 · inbound

Understanding Space Is Rocket Science -- Only Top Reasoning Models Can Solve Spatial Understanding Tasks cites this paper.

Understanding Space Is Rocket Science -- Only Top Reasoning Models Can Solve Spatial Understanding Tasks WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T11:56:10.304534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:56:10.304534Z digest=sha256:1f887209fa4d9e8fd43ee6b77726cd65b4e8cd7c2c6ba0bcd04f8ba39b8ae901

Observation f0a16358-ba02-47c4-8dd6-b1a4b9d12506 · inbound

Assessing Privacy Preservation and Utility in Online Vision-Language Models cites this paper.

Assessing Privacy Preservation and Utility in Online Vision-Language Models WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:30:50.217647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:49:04.957341Z digest=sha256:fc01df2b8f7701df48fe241da401d9e7a91dba781d415e6f32898defec4656e2

Observation f621129b-a43f-41fe-8df7-91606e63403c · inbound

When Meaning Travels: A Granular Lens on Hybrid-MoE's Role in Idiomatic Understanding for Language Models cites this paper.

When Meaning Travels: A Granular Lens on Hybrid-MoE's Role in Idiomatic Understanding for Language Models WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 71

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:36:16.756023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T15:19:26.983760Z digest=sha256:a315f40c6304c9809f3a2216e1d413b301daf4ebdb071b28f9f4ae0683845487