Pith. sign in

Paper Citation Record · LEDGER

HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2410.12381.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.12381 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T20:55:05.985112Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T11:56:55.430688Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 781a1daf-41e5-4c0d-a0d0-9b7e712c6709 · inbound

ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models cites this paper.

ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T20:55:05.985112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:55:05.985112Z digest=sha256:0e33e7948e76d6cb1302629fcfbc9e822168d1d557257303c7e9c3b18f61b489

Observation 2e2a2458-608b-4513-a8aa-b9be07b56ce9 · inbound

R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization cites this paper.

R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-16T00:19:20.545535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T00:19:20.462455Z digest=sha256:c9f8d94e4c65203b61aaaa476ff01730bd1220daed241d95a44422651c4f0ac0

Observation edb1385a-aa8f-4b6a-a67f-e6175f0dace0 · inbound

MMR-V: What's Left Unsaid? A Benchmark for Multimodal Deep Reasoning in Videos cites this paper.

MMR-V: What's Left Unsaid? A Benchmark for Multimodal Deep Reasoning in Videos HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:35.576700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:52:35.576700Z digest=sha256:a120457295c878b9e6d6dad51f4e5bed2f617a032fd9a7575ebca914adf7f22a

Observation 662d4919-af9f-4c31-a24b-d5b8e1316c15 · inbound

SlideCoder: Layout-aware RAG-enhanced Hierarchical Slide Generation from Design cites this paper.

SlideCoder: Layout-aware RAG-enhanced Hierarchical Slide Generation from Design HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:26:15.141311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:26:15.141311Z digest=sha256:c6951a7c8017175884cc42956dfd1cef52d73c19195e45d82b1da8bd28019e46

Observation a29e0267-9a48-4c26-9cce-465bded6ffc4 · inbound

Multilingual Multimodal Software Developer for Code Generation cites this paper.

Multilingual Multimodal Software Developer for Code Generation HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:58.380932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:58.380932Z digest=sha256:50b0997e4e243cfa6051ab8a1e06d44afab4578fd72fee89760932241cf79dca

Observation 6e2acc19-30d8-4d7d-a8d5-7394338b1587 · inbound

VisCodex: Unified Multimodal Code Generation via Merging Vision and Coding Models cites this paper.

VisCodex: Unified Multimodal Code Generation via Merging Vision and Coding Models HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-05T20:44:39.584064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:44:39.584064Z digest=sha256:3e56c8cf741a9ee43c28041bb7a24d1e6f164464bccb528191a4ccb63f8e090d

Observation e14d6276-bb43-4220-bc27-33d47dda05a8 · inbound

Understanding Space Is Rocket Science -- Only Top Reasoning Models Can Solve Spatial Understanding Tasks cites this paper.

Understanding Space Is Rocket Science -- Only Top Reasoning Models Can Solve Spatial Understanding Tasks HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-05T11:56:10.967394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:56:10.967394Z digest=sha256:71e7620e5bb472e25cc09350ab8012065b17b04711acca5b33ccacbef801a4cd

Observation d6088590-dd16-4f8c-ae1a-f242e0762ab0 · inbound

VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs cites this paper.

VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 116

Resolution
unresolved
no resolver link, observed 2026-08-03T20:23:09.675588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:23:09.675588Z digest=sha256:bd8dba4e7d3e7657fa82fdb2e737964f35b488026a3ab217f9be989cd9b293ab

Observation 967a24f9-866d-458e-a97c-5cb9a03ecbd1 · inbound

ChartAttack: Testing the Vulnerability of LLMs to Malicious Prompting in Chart Generation cites this paper.

ChartAttack: Testing the Vulnerability of LLMs to Malicious Prompting in Chart Generation HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T09:45:31.514666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:45:31.514666Z digest=sha256:743af333caacea5bbc56cfa59f4be94f8b0d0cf007454faa924ae0f7c72e1e4a

Observation a154fba6-4f9c-4295-a26d-e4c9b0a721cd · inbound

Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles cites this paper.

Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:51:15.396365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T07:51:13.362986Z digest=sha256:71d142815d4f7339558d3e4e48bb3e0a07a4519feaadc10b1b6f346191413109

Observation 85e648c6-9d82-47f6-90e6-c274ed0a3d76 · inbound

Imagine Before You Predict: Interleaved Latent Visual Reasoning for Video Event Prediction cites this paper.

Imagine Before You Predict: Interleaved Latent Visual Reasoning for Video Event Prediction HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T11:56:55.432310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T02:46:50.373450Z digest=sha256:da56545d681fe8c3f1e2d68fd56023aab514502a65a56a51ab2675d1c10bd931