Pith. sign in

Paper Citation Record · LEDGER

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos

As of 20 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2507.16878.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.16878 v1

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:14:17.453302Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

21 of 21 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 663b8e09-9da9-42dc-a814-3449e29abc99 · outbound

This paper cites Large Language Models for Planning: A Comprehensive and Systematic Survey.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos Large Language Models for Planning: A Comprehensive and Systematic Survey

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.299355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.299355Z digest=sha256:dd7a7b88f0091838c352efb88ec0b11b4b6531e3a675271f4daa575a57d3075f

Observation c50bf9a9-667f-402c-aa4e-6cf5d2c614d2 · outbound

This paper cites Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning?.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning?

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.308897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.308897Z digest=sha256:de5cda1de8099da14d08770d8025bf8143e00e3575fa7f1a990d650cdbdb93ae

Observation 01f3cec6-94bc-4467-b9a8-00f90c6b63fb · outbound

This paper cites Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.320440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.320440Z digest=sha256:552df89887270896636b0eb4ab64ab2c313a1994e825fd1e95c940dc15547c78

Observation cf07f444-28eb-4b9c-a345-6ac8b2ac5981 · outbound

This paper cites Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.338467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.338467Z digest=sha256:7504776a63b6ec45e180c4b8494ee58fc3160cb6acd895c9e94f8f83e1865a40

Observation 423b9070-8f87-49d5-8759-bf6e0e0196cf · outbound

This paper cites FIOVA: A Multi-Annotator Benchmark for Human-Aligned Video Captioning.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos FIOVA: A Multi-Annotator Benchmark for Human-Aligned Video Captioning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.350806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.350806Z digest=sha256:35969f5b4118d984986dee16284ed3cf3e58ae09bf84c70b4cff3e9e894dae55

Observation de676c18-e6a8-4d9b-886d-ea1876e61114 · outbound

This paper cites Gemma 3 Technical Report.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos Gemma 3 Technical Report

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.358398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.358398Z digest=sha256:89638f0d7d83d9bcd82344f6eeb3a191d289f3dae433de5521b3f35feb400de7

Observation 7ba920d1-40c7-412f-9313-65550a564233 · outbound

This paper cites VerifyBench: A Systematic Benchmark for Evaluating Reasoning Verifiers Across Domains.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos VerifyBench: A Systematic Benchmark for Evaluating Reasoning Verifiers Across Domains

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.376888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.376888Z digest=sha256:92e1318b9d23a34d1bb2fc1c55d9c74c556d1d6f57da9b8b5ba89c05d7181e3f

Observation feb29216-6bdf-4d59-a73f-aa1be882fccc · outbound

This paper cites VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.390535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.390535Z digest=sha256:89e1ce63242a2f3635375f38027244a747febf408c953cba50a42f17d5698f58

Observation 5ba3d1b2-fbac-4f67-836b-d3ecceacdf07 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.396246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.396246Z digest=sha256:b555af511fe77b3296aaa6c158a42f5b1d3502dfe21b30787bd187aa604265dc

Observation 303582fa-a139-4738-9cba-d1f8a4afd374 · outbound

This paper cites LVBench: An Extreme Long Video Understanding Benchmark.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos LVBench: An Extreme Long Video Understanding Benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.409725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.409725Z digest=sha256:52d66e91ec28abff28f88cdaf7f4605cade7a3cff4193934660ecd8b0ecd395d

Observation a7023372-063e-4b50-9613-652cf79fe556 · outbound

This paper cites Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.416746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.416746Z digest=sha256:9b4168b3cb5b034ee5fc0d1799ae8cc158b876aea39877763f19512733f91d06

Observation 701d3b60-edfa-4b45-a461-9be4cd813e8a · outbound

This paper cites InstructionBench: An Instructional Video Understanding Benchmark.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos InstructionBench: An Instructional Video Understanding Benchmark

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.424835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.424835Z digest=sha256:e56d092b3b2c11cd8cf7096a92109a639f42a3e0b36e6791ff659b15094e7c9c

Observation a6e1dc5a-0739-4643-a327-6e6be8c8e448 · outbound

This paper cites Qwen2.5 Technical Report.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos Qwen2.5 Technical Report

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.434938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.434938Z digest=sha256:a61a5c8a4ee0db12fc5f51b03534f66f3d2094a1c80aeec37b6d65b72bae5263

Observation 8f43ed20-4d5f-404e-a453-983c4777fad8 · outbound

This paper cites LLaVA-Video: Video Instruction Tuning With Synthetic Data.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos LLaVA-Video: Video Instruction Tuning With Synthetic Data

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.441895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.441895Z digest=sha256:4cf644b909b1a28fab845f9d3c537ef94264e39dcdba661a852500d6e60fedb8

Observation e5af482a-0791-40c7-92f8-9b122181d297 · outbound

This paper cites LLaVAR: Enhanced Visual Instruction Tuning for Text-Rich Image Understanding.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos LLaVAR: Enhanced Visual Instruction Tuning for Text-Rich Image Understanding

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.448064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.448064Z digest=sha256:4b8b50d12c4b71ee6b035bb294d97ab35082d62890c55068b7e949bad807242c

Observation bae59f16-d4d4-4009-9745-26ea5d6be32e · outbound

This paper cites MLVU: Benchmarking Multi-task Long Video Understanding.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos MLVU: Benchmarking Multi-task Long Video Understanding

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.453302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.453302Z digest=sha256:4f4bf2a6287f854f428cda4a5b9bcc2df856d275522e384e90fe3ee99ed9cfd1

Observation a9fde74e-668b-4f48-bf5f-efb072057b4b · outbound

This paper cites From Machine Learning to Robotics: Challenges and Opportunities for Embodied Intelligence.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos From Machine Learning to Robotics: Challenges and Opportunities for Embodied Intelligence

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.403330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.403330Z digest=sha256:3970835bcfec10b0a57de494427c6db5b0f34719ea20c2d384f949b72c752805

Observation 54096e98-9ef3-4da3-b58e-4a5a492f900e · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos LLaVA-OneVision: Easy Visual Task Transfer

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.370377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.370377Z digest=sha256:6d1649bfd933cfce35b25e0b45c3f72c34df48389c802d52e3759e733f1ff8a1

Observation 2b62ba60-92a1-440b-9ad4-4855146de43e · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.382712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.382712Z digest=sha256:8f2dab8524e1d3b8b8c394b428d8f3461b0b64f93fd95c7bba40dede400d942e

Observation ad492195-bd08-400d-a93c-a65b26054333 · outbound

This paper cites TemporalBench: Benchmarking Fine-grained Temporal Understanding for Multimodal Video Models.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos TemporalBench: Benchmarking Fine-grained Temporal Understanding for Multimodal Video Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.287285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.287285Z digest=sha256:06243435ecf939a45a18e6d130c3a398d159e8d50d645dbc5557e84700e38f60

Observation 804b90cb-38a1-42fd-a2cb-48c5c8aa7450 · outbound

This paper cites Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.279011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.279011Z digest=sha256:53c8112c24e719e455d96966ce2ca1f33efee0d7fc9f23a4e87ebf1213c097fb

Pith citing papers

No inbound Pith citation observations are available.