Pith. sign in

Paper Citation Record · LEDGER

FunQA: Towards Surprising Video Comprehension

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2306.14899.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.14899 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:28:39.566822Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T06:39:37.649762Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1533d825-7d5b-4c47-aca5-84e873861d36 · inbound

Otter: A Multi-Modal Model with In-Context Instruction Tuning cites this paper.

Otter: A Multi-Modal Model with In-Context Instruction Tuning FunQA: Towards Surprising Video Comprehension

Reference 89

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:43:47.881383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T02:43:47.775691Z digest=sha256:cf79d8d171b5a0d059b355c9a96e8e4906a8154acbae0a0bc2446094bae48fb9

Observation b724392e-2dc2-40f1-99df-59cc4c16f0bd · inbound

MVBench: A Comprehensive Multi-modal Video Understanding Benchmark cites this paper.

MVBench: A Comprehensive Multi-modal Video Understanding Benchmark FunQA: Towards Surprising Video Comprehension

Reference 87

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:22:35.095848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-17T20:22:34.954228Z digest=sha256:988e6bb72bf5715c8c96bbfa3d0df3d98a94e237681782c307928c5d43f36cb9

Observation 7f389b7f-1475-498c-b2df-b63e8207e84d · inbound

MLVU: Benchmarking Multi-task Long Video Understanding cites this paper.

MLVU: Benchmarking Multi-task Long Video Understanding FunQA: Towards Surprising Video Comprehension

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:55:26.473171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-14T19:55:26.333923Z digest=sha256:016c74d1c4049f5484f3332b11c0ab33089cf7e7541ecb5750d135d8f36deb4a

Observation b6a748ff-101d-4466-8d90-0f444d2f7b56 · inbound

Black Swan: Abductive and Defeasible Video Reasoning in Unpredictable Events cites this paper.

Black Swan: Abductive and Defeasible Video Reasoning in Unpredictable Events FunQA: Towards Surprising Video Comprehension

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:39.566822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:39.566822Z digest=sha256:fc6c8bb79c676c88ecedb28c4ee4cd0ac8da0d68e35aea777b3417a31f6723d5

Observation ba970cdd-60a5-4ae7-a513-b40f9f1bef12 · inbound

Exploring What Why and How: A Multifaceted Benchmark for Causation Understanding of Video Anomaly cites this paper.

Exploring What Why and How: A Multifaceted Benchmark for Causation Understanding of Video Anomaly FunQA: Towards Surprising Video Comprehension

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T19:09:32.656000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:09:32.656000Z digest=sha256:7ea93ffa5f37cd4bf84450d859a7898a19c0d3f975eec9fa65627fc35d4aa85f

Observation 94ab8cb5-71cb-45b7-b2f0-abab6cb0a966 · inbound

How Vision-Language Tasks Benefit from Large Pre-trained Models: A Survey cites this paper.

How Vision-Language Tasks Benefit from Large Pre-trained Models: A Survey FunQA: Towards Surprising Video Comprehension

Reference 135

Resolution
unresolved
no resolver link, observed 2026-08-11T18:11:54.804455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T18:11:54.804455Z digest=sha256:72181587b8dea859b84fc1a171905de3a0699740820587218bfd2700a8f93534

Observation cebcaa12-1d22-47e2-9e6b-74f8d28985bb · inbound

Movie2Story: A framework for understanding videos and telling stories in the form of novel text cites this paper.

Movie2Story: A framework for understanding videos and telling stories in the form of novel text FunQA: Towards Surprising Video Comprehension

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T11:48:07.228234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:48:07.228234Z digest=sha256:0c431b3e6872f7952a85c2eeb3a9adddb4a712b2ca5a5533b4e8eedbe0c79557

Observation 32dc8df7-cfb2-4296-8d38-38d6549b07f6 · inbound

AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs cites this paper.

AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs FunQA: Towards Surprising Video Comprehension

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-10T22:19:53.559199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:19:53.559199Z digest=sha256:e33e305618fd7baf0c3ec7ae35e7816151dcf5968307488ad3be148ec5c2d451

Observation be4cf59e-3032-4c5d-bc8b-fb4d9ffe5508 · inbound

Towards Video Thinking Test: A Holistic Benchmark for Advanced Video Reasoning and Understanding cites this paper.

Towards Video Thinking Test: A Holistic Benchmark for Advanced Video Reasoning and Understanding FunQA: Towards Surprising Video Comprehension

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T15:48:35.898141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:48:35.898141Z digest=sha256:cd8ad98512a25bf651c3123eacca5474e5cb98d7cbb2666841df7d85ccde8916

Observation 8bd7043d-772c-4c8f-9a92-23e72e684a2c · inbound

HPP: Hierarchical Programmatic Probing for Long Video Understanding by Decoupling Perception and Reasoning cites this paper.

HPP: Hierarchical Programmatic Probing for Long Video Understanding by Decoupling Perception and Reasoning FunQA: Towards Surprising Video Comprehension

Reference 167

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T06:39:37.650987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-26T14:19:53.450263Z digest=sha256:24048f0564f039810526e997950595210c1b97be59f4b32e29321992e19ddca4