Pith. sign in

Paper Citation Record · LEDGER

VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2504.07956.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.07956 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:12:18.435691Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T05:49:36.609405Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1a4e1cb2-2eba-4859-8076-4d0df32d53a9 · inbound

Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models cites this paper.

Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-16T05:12:18.435691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:12:18.435691Z digest=sha256:7eccad26035cfbfc438e75a46fffeeff05fa1479f44e35f99b0535c75c5f52b5

Observation 674d1b29-563d-4b34-8aa5-498a40a647db · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 132

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:20.019923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:20.019923Z digest=sha256:bd8c552147b7098c2e14cd0688c26509e8a580f1bfc6fc6cf64bd53f2a0d9f15

Observation 07105e7d-bf71-4a16-868a-f72ca768e4c7 · inbound

Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning? cites this paper.

Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning? VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:40:56.094429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-17T05:40:55.944288Z digest=sha256:db9d5d1f7421964711c68fc15225baac652e488b1d596645287991c0bf8b9c3a

Observation feb29216-6bdf-4d59-a73f-aa1be882fccc · inbound

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos cites this paper.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.390535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.390535Z digest=sha256:89e1ce63242a2f3635375f38027244a747febf408c953cba50a42f17d5698f58

Observation 4bb1c417-c44c-49eb-b3bd-7193a9df94e0 · inbound

ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts cites this paper.

ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T13:12:40.063831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:12:40.063831Z digest=sha256:1f5308cc7a47dd1d66ff53c1142b6c631569ad4f986c316fd2165aa292079354

Observation b4171d3f-c6fa-41b1-aa78-1e9f60fb1f48 · inbound

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes cites this paper.

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T19:03:06.647166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:03:06.647166Z digest=sha256:99f867951203dbbabbf31dbd94654056bb012c227c45cca441617553c0a10076

Observation 0808bec6-7856-4334-8b37-7b84434757b4 · inbound

VIDEOP2R: Video Understanding from Perception to Reasoning cites this paper.

VIDEOP2R: Video Understanding from Perception to Reasoning VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:25:22.616824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-17T22:24:41.760120Z digest=sha256:6f232f7c3dd17c62be52caa3bfbdb26a8f147c5e0045945e946921fb730e1bff

Observation 6fadac5a-4015-41a3-bd48-0e9be9b5318f · inbound

Seed1.8 Model Card: Towards Generalized Real-World Agency cites this paper.

Seed1.8 Model Card: Towards Generalized Real-World Agency VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:45:14.376813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-15T07:44:02.827006Z digest=sha256:9c33254c506b51ffebe80fcb3bff4abdb0f471d00cb8ae3deef82181e1732af4

Observation c016cd63-2897-4f16-92a7-f84038a7cae7 · inbound

Video-Oasis: Rethinking Evaluation of Video Understanding cites this paper.

Video-Oasis: Rethinking Evaluation of Video Understanding VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-13T15:38:41.390945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T15:38:41.390945Z digest=sha256:d70a03c265ad9fcc4279ffd55cd4646cd6a45ac2c425b71cb82cb53c30fe756c

Observation 78654518-e85c-4b52-81aa-8c35bb74e250 · inbound

Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding cites this paper.

Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:55:51.277971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T18:47:32.778695Z digest=sha256:7f96ca9b448c0c699aa452ba3804b0fb2b5b6508a80073c77b6e67910b35bb77

Observation f9f80566-ca13-45fd-997a-e3bacfff75b5 · inbound

EasyVideoR1: Easier RL for Video Understanding cites this paper.

EasyVideoR1: Easier RL for Video Understanding VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:47:12.639121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T07:41:27.231098Z digest=sha256:6569c8a683ba705235a98f0fae6dd9c076d5ac1047b9ef63df75177782a4528e

Observation 664a807d-d961-4046-9b91-da6887caa58a · inbound

Act2See: Emergent Active Visual Perception for Video Reasoning cites this paper.

Act2See: Emergent Active Visual Perception for Video Reasoning VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-09T05:45:22.906813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-08T19:34:53.683729Z digest=sha256:8184fff073043cb4f1b37f3077c9931bdde0c4324da07835ae668015a0b7a566

Observation cbd3bf73-c4af-40b5-9364-d0063e8024bc · inbound

VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing cites this paper.

VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:11:13.408349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-08T01:30:19.531699Z digest=sha256:fdd8309265a29e31d981f8283968d7665512f4c532883979013a98291025c647

Observation eb9960d6-39dc-45a2-9a1d-54c5718145b7 · inbound

VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing cites this paper.

VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-12T01:46:13.974098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-12T01:43:34.898639Z digest=sha256:405eeb3709b106fa2eb811fbf639065e81941ce856ebf3897845cd734a3845e4

Observation d8b2736d-495a-4c49-935c-a4edf52757d8 · inbound

VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation cites this paper.

VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:38:56.199595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-20T19:37:09.244578Z digest=sha256:c03f2af92920d645da0f0de81635623bcc95cd5ba247554357e044d96906f133

Observation 7a0ee72f-3988-4eb9-9088-b2e897cb7516 · inbound

ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning cites this paper.

ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T21:07:23.872673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-27T19:56:09.820812Z digest=sha256:557c50df4a71247d2f40398d99680a63109885a01b7aff26187fa53fec644a00

Observation 29b40ea8-f33a-4fff-82bb-8af022e8d9f0 · inbound

ELVA: Exploring Ranking-Driven Universal Multimodal Retrieval cites this paper.

ELVA: Exploring Ranking-Driven Universal Multimodal Retrieval VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T05:49:36.611612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-26T15:34:55.062016Z digest=sha256:1e5c42a4be753510c1250b3175ef8e8cb5eb388e0590ea74ac3130e37ee236b5

Observation 175b037c-ead7-402c-ad2c-8772d7a5252d · inbound

RoboGaze: Evaluating Robot World Models via Structured Vision-Language Analysis cites this paper.

RoboGaze: Evaluating Robot World Models via Structured Vision-Language Analysis VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-06-30T10:54:36.512557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-30T10:49:58.019238Z digest=sha256:bf99856866a3318189cad2cbff5f230eae0c1689653e4d84c682cfe8e5fcd5f6

Observation acc800f7-de2f-424e-aaed-24d5c04790d9 · inbound

SimpleSearch-VL: A Simple Recipe for Multimodal Agentic Deep Search cites this paper.

SimpleSearch-VL: A Simple Recipe for Multimodal Agentic Deep Search VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:55:41.066581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-07-01T06:02:48.532478Z digest=sha256:adbba4ec010500ebe057195b0448aa01451f4a0a103817b803739abd31f925ca

Observation 3b8658af-e80d-41ee-954c-72225ab03aae · inbound

Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent cites this paper.

Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T04:44:27.371880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T04:44:27.371880Z digest=sha256:955812b26eb5096c20bf467c3c56c2b5ea9b843baedba04c1fed5a2be4405039

Observation bdcbcb62-e23e-4b46-88d0-e61aca10e28e · inbound

Order Matters: LVLMs as Judges for Temporal Reasoning in Image Sequences cites this paper.

Order Matters: LVLMs as Judges for Temporal Reasoning in Image Sequences VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T14:37:30.050698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:37:30.050698Z digest=sha256:0f3d1eeb137926b6ce26d67e4f2c588fbdc6b44d5378d723ac4281883e595744