Pith. sign in

Paper Citation Record · LEDGER

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

As of 18 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 13 inbound Pith citation observations for arXiv:2507.09313.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.09313 v2

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:03:14.058499Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T08:39:01.053662Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T01:19:20.288097Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c6d6f3e2-9cf3-46ce-b728-a333e7d61e68 · outbound

This paper cites Qwen2.5-VL Technical Report.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Qwen2.5-VL Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:13.969005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:13.969005Z digest=sha256:05049254f2d18ebfad70e455e9fd0df61e6d6edb9a03bc4e6efb9c0caa5fbe2d

Observation e849daa4-bf69-475c-8e19-353675a6c28c · outbound

This paper cites TemporalBench: Benchmarking Fine-grained Temporal Understanding for Multimodal Video Models.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models TemporalBench: Benchmarking Fine-grained Temporal Understanding for Multimodal Video Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:13.972207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:13.972207Z digest=sha256:0d157846f2e8aef424e66a5b02cb4e1e894a1f61ad7fe431b5af4c0852f5c978

Observation 10285eaf-f2e5-487f-b655-23d9189d1e48 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.449302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T18:03:13.975552Z digest=sha256:7ac661dbbb7f62ce9ccea69493c4a70d0657e8e3322aaba7bf057a062a5f334f

Observation 486bec5f-f7ab-45ac-8921-1a765c1009d4 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:13.978291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:13.978291Z digest=sha256:1c1a39fa036d25d839aec8e3570876b345934f5ba3dcf2c7c61f0191eed82608

Observation 9eac1249-9df0-4a71-9598-5a9417cdf906 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:13.980956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:13.980956Z digest=sha256:137b6ea6de8552b1295106c30ea98a114ecc487a6e82963a0b128326f56b7c44

Observation 2b04f4a6-1cdd-49c6-a73d-34ebafa03d10 · outbound

This paper cites MMBench-Video: A Long-Form Multi-Shot Benchmark for Holistic Video Understanding.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models MMBench-Video: A Long-Form Multi-Shot Benchmark for Holistic Video Understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:13.983556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:13.983556Z digest=sha256:0f84d8f9777a327d13d02985c6975e540c682218994ca0e1c8b48af9a94609be

Observation dd5322e8-be16-40f6-9692-512e4e38a101 · outbound

This paper cites Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:13.986518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:13.986518Z digest=sha256:74d4a1c818d3dc30ec28ddf5d8c1660fd363d8707326d96df4e10f177081f786

Observation 1ba1b039-4af8-4801-a916-db5fd5cb4f99 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.437624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T18:03:13.989094Z digest=sha256:b23c66afed1cc241f44f32e0c8b87317943cadc53a8dfa7c03828ecf65b78830

Observation 5c973cca-26f8-45ae-802c-4afd7c5f51b5 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.430735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T18:03:13.991479Z digest=sha256:7b04384e7c6f7aa840accca179944523c3f05e93f84340be7a3971f577a8bae0

Observation 3a6bf101-c3d0-4f4c-9281-3e7231bac870 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.423704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T18:03:13.993760Z digest=sha256:21f4fe16ce014357ba4c225bd24a7c6065a1d4f9e2a8add5692d0fc582b42a5a

Observation 143a51d1-6530-49b9-8546-705b6ec35f4d · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models LLaVA-OneVision: Easy Visual Task Transfer

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:13.996017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:13.996017Z digest=sha256:5cdb02c97eba2570a6a5333d73641601ec0ed10954a0c7b58274a1bf840af007

Observation 0f0d8dc1-58fc-44ab-9453-12b949975ae9 · outbound

This paper cites SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:13.998784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:13.998784Z digest=sha256:923a41ba8215967f01695116bfeaa84df4c21aed857aab4b52997371f615293f

Observation 52c34cc5-e012-4452-8163-c7f287325711 · outbound

This paper cites SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.001294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.001294Z digest=sha256:3ed5cc01437c90d1a767521c7dfe77f16bffda73d761db2020498ab541f670f0

Observation 00ce54f3-6064-41f4-8251-322351b7cbac · outbound

This paper cites LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.003554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.003554Z digest=sha256:fdded0613cde09efa1c3b3f954e15b063a4457c97542caacff3f9d4fa433cd04

Observation 22fe2f4a-7f46-4fed-bf37-23f2ebc16f34 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.416481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T18:03:14.005895Z digest=sha256:f662ba26914afd2d9cd4847b1ef1d791ef294d6bfca502380e15e86b97fae27b

Observation 6822cf65-13fb-425d-801a-e6569e8fef1e · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.408974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T18:03:14.008278Z digest=sha256:22d75b2166079e126cb9dd58dea9f683828481028be57aada249bde037e6e4c4

Observation 467ce196-84d0-45fb-be40-1309c5607a5a · outbound

This paper cites OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.010398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.010398Z digest=sha256:bca1de6bf4d3b548a98befd3ddbf381a6743a846adc2706b0774effbfc5bd4ee

Observation ed36c7b1-e5cb-491c-b3c1-ecc56dfc9c8c · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.012841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.012841Z digest=sha256:2fd3cc4f4f49a3dfafe4a00091120379c57ced3f3da9c9bce0c1610e3752187a

Observation a49e3718-f95e-4026-a5f5-0ac0de340ec5 · outbound

This paper cites StreamingBench: Assessing the Gap for MLLMs to Achieve Streaming Video Understanding.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models StreamingBench: Assessing the Gap for MLLMs to Achieve Streaming Video Understanding

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.015160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.015160Z digest=sha256:1878d3b5a75b6f3aa55a81ff1eb30190fbf528ea6af2fb8c6997d30bea7c0cea

Observation 420a5d6b-82a0-4e9c-97f4-e00fa50e180c · outbound

This paper cites E.T. Bench: Towards Open-Ended Event-Level Video-Language Understanding.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models E.T. Bench: Towards Open-Ended Event-Level Video-Language Understanding

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.017689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.017689Z digest=sha256:0f226a6e325fcb5119853df13604ed9e2a47a16c429253104a5f631437c094a9

Observation ebc798c4-f520-4f9e-8b3f-6eef9486475d · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.401107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T18:03:14.020500Z digest=sha256:3b4b047097b8eb9c41bf878bac580ed4798b7203bc0694c11ac88e6de90da177

Observation 7b0e7426-b2f9-414d-8ade-fed24f1340f5 · outbound

This paper cites Dispider: Enabling Video LLMs with Active Real-Time Interaction via Disentangled Perception, Decision, and Reaction.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Dispider: Enabling Video LLMs with Active Real-Time Interaction via Disentangled Perception, Decision, and Reaction

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.023334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.023334Z digest=sha256:8fef3eec72e46e5e94fe5da8f632acbdf5c8b94006e22063ae51546a97999156

Observation 176d99bb-a4b2-4dbd-b5ed-83a21c18dae9 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.393918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T18:03:14.025661Z digest=sha256:f46a8aa1e291c557eb6b3a8b176fa161500799017efa89561a0124a1e1576ea4

Observation f07fedcc-806e-4850-a9eb-4833e927b7d5 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.386759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T18:03:14.028068Z digest=sha256:26e8f8965451a9ab302c6ecb4da2a57f31003570ad454366142af20f09b0ace1

Observation 5a549eee-007e-4704-830e-b56c9a39e8ce · outbound

This paper cites Lawrence Zitnick, and Devi Parikh.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Lawrence Zitnick, and Devi Parikh

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:03:14.379613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T18:03:14.030365Z digest=sha256:c0271479d6566a8c92bad96d41f070a4994fdc1c41f30d479c7caf000c195d0d

Observation 998eb457-fc7c-4e94-9cfc-dd622733a284 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.032751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.032751Z digest=sha256:2bdd86ac82613bd7f1fa33cc33d57abdc0b9fdc52991d8f4df3ce866653dc0dc

Observation 8eb0e784-8b2e-4aff-8245-0d2dbdd6e43e · outbound

This paper cites OmniMMI: A Comprehensive Multi-modal Interaction Benchmark in Streaming Video Contexts.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models OmniMMI: A Comprehensive Multi-modal Interaction Benchmark in Streaming Video Contexts

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.035067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.035067Z digest=sha256:6dee96816d3a1357998bfbdcc1fe029a548d38d822c58f0d6b50a1b09c5f0e0c

Observation 85291699-b20d-458e-8586-54e298b76647 · outbound

This paper cites Qwen2.5-Omni Technical Report.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Qwen2.5-Omni Technical Report

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.037742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.037742Z digest=sha256:61d8240c4d4a845cc542f20fb03adcc7fbc7dcf3d20e590e8badf350512e6518

Observation 46f8794c-4d6e-4420-8c39-7195e583ed0c · outbound

This paper cites TimeChat-Online: 80% Visual Tokens are Naturally Redundant in Streaming Videos.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models TimeChat-Online: 80% Visual Tokens are Naturally Redundant in Streaming Videos

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.040372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.040372Z digest=sha256:9809fa4daa94f0b9cde16b66c1f981f7d06c7ac22477bf842acc24ada5b13c40

Observation 20623d6c-9e16-4f52-a6bc-4b75886e1e2c · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.371311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T18:03:14.042972Z digest=sha256:aa5bfc8a20e1ba9566db9d87d8c39a23ba6f9580e7e2081251c060ebc4834fcf

Observation 209fa54e-22a5-428c-89cd-1c8afee2b76c · outbound

This paper cites InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.045371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.045371Z digest=sha256:dc5a1f17f55ab3b5f356ceabe5eb2ad430079be25618ec17e68142cb670324e7

Observation 009ac0f9-9498-40c2-b9cd-b9a93ba48836 · outbound

This paper cites Long Context Transfer from Language to Vision.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Long Context Transfer from Language to Vision

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.047880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.047880Z digest=sha256:d1afc3279c15376dbcb719d81d6f9919c1fd68a64276cb6aa037100ddd8f75a1

Observation 7a6aeb78-0594-46dc-9dd5-0891c5923a75 · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models BERTScore: Evaluating Text Generation with BERT

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.052687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.052687Z digest=sha256:7448475d1ca6270c6f2b8d72a45fd1e17ec526db9b3d459840f1a94d9863a5ea

Observation 8fd3de2e-9751-44be-9eab-992a8c16393c · outbound

This paper cites online" 'onlinestring :=.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models online" 'onlinestring :=

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.055616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.055616Z digest=sha256:e05d46236794ded25d31380d8de1c2d4b759f04841fe8e69e6d19eb6a5c7967b

Observation 6e4d065e-de94-4dc6-b3b2-278454a4024a · outbound

This paper cites write newline.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models write newline

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.058499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.058499Z digest=sha256:12b032e9778a8fb5056277d455888b945f97b91fa8d627b14cd1671fe4b6045b

Pith citing papers

Observation 5f6fd4f2-22b5-45a3-80e1-e540f88c8c65 · inbound

Beyond Reactivity: Measuring Proactive Problem Solving in LLM Agents cites this paper.

Beyond Reactivity: Measuring Proactive Problem Solving in LLM Agents ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T08:39:01.053662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:39:01.053662Z digest=sha256:6840e0b220eab12b435682f25487c4fc81678ad1296b270367a7b9ae9a3a88f1

Observation 4027df28-ad4c-4b0f-917b-4e48460d6b75 · inbound

Proact-VL: A Proactive VideoLLM for Real-Time AI Companions cites this paper.

Proact-VL: A Proactive VideoLLM for Real-Time AI Companions ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 336

Resolution
malformed identifier
no resolver link, observed 2026-08-02T19:09:09.440908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:09:09.440908Z digest=sha256:7ffd332853c5fb4295e89199bd2edc636924fe54492f825c03da8d933d1288f1

Observation 4572f0d0-b18f-45ee-804b-a0cc2491ab37 · inbound

From Reactive to Proactive: Assessing the Proactivity of Voice Agents via ProVoice-Bench cites this paper.

From Reactive to Proactive: Assessing the Proactivity of Voice Agents via ProVoice-Bench ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.033091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T11:44:12.373082Z digest=sha256:56aee121ad47ba1118d92b60e6994a02335914c7ee64ffa6ead14b8b0048c308

Observation 6c92c69d-1533-4d11-8cc6-ae801125e38f · inbound

Don't Pause! Every prediction matters in a streaming video cites this paper.

Don't Pause! Every prediction matters in a streaming video ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:41:18.097956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-08T04:32:01.379605Z digest=sha256:bedfa3bd047ad9985f3e4fd72400da46cfbc8478d7986ac7ead80d637ab396ec

Observation 068e4c34-be0c-4134-84b3-3b821d3b358c · inbound

StreamPro: From Reactive Perception to Proactive Decision-Making in Streaming Video cites this paper.

StreamPro: From Reactive Perception to Proactive Decision-Making in Streaming Video ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:33:50.748812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-20T23:32:38.332828Z digest=sha256:bb503c39a5b7285d326259587c74f4f4cd6d7166f8961dc251516e67fe0355ff

Observation 41aead78-12f1-4da2-9363-84e7b519d662 · inbound

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction cites this paper.

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-20T13:38:19.157208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-20T13:36:44.071188Z digest=sha256:f37feb11124ee991efe191d542401a03aa8de3f62689a13eb469431b13ff433f

Observation ca766167-9a2e-46cc-bc82-858fed26fffc · inbound

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction cites this paper.

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-04T01:19:20.291251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-04T01:11:42.073993Z digest=sha256:41133a7258a3a900d1e3bd992e9a56a107ef407eeb1b9420269b9c1743b0f676

Observation fb22717f-9da6-4ae1-acb9-ffaba75f432d · inbound

An Efficient Streaming Video Understanding Framework with Agentic Control cites this paper.

An Efficient Streaming Video Understanding Framework with Agentic Control ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T11:33:14.470074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-20T11:30:22.151045Z digest=sha256:c790afd85fcaf46a71f6ad1dca64d9dcb3c0cb65d2ee352219f434c5017196fb

Observation 3e967234-cadc-4cee-95cf-5fbac879f8a2 · inbound

IPIBench: Evaluating Interactive Proactive Intelligence of MLLMs under Continuous Streams cites this paper.

IPIBench: Evaluating Interactive Proactive Intelligence of MLLMs under Continuous Streams ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:33:51.065824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T18:24:57.881644Z digest=sha256:115399894a67e7f73f48218082a04c1fc77bd80439265d825e77f49ecf0f267a

Observation c4af456a-a752-4d00-9ce7-3a177dc71ce5 · inbound

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding cites this paper.

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:46:18.934183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T15:06:22.102725Z digest=sha256:0c1f82296a77b231395139ff36ec52c7ffbc7eeb2fa78a7661d011fdeb548e8a

Observation 0dd45640-7b1f-485c-b538-17545bed0fe1 · inbound

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding cites this paper.

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T10:44:36.467381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T10:42:37.401221Z digest=sha256:db63e3023e4e08684bff33dab5f36a3d9c5325c35059334d318dc71a19835ccf

Observation 209f9f59-9164-42e5-8a37-7f88f5dc3007 · inbound

GuideMe: Multi-Domain Task Guidance and Intervention in Streaming Video cites this paper.

GuideMe: Multi-Domain Task Guidance and Intervention in Streaming Video ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-12T05:36:40.278384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:36:40.278384Z digest=sha256:c462e9ef22d9d99cd5c3520d4a1e4a0ef009af2db1a9a7e68443f44b5144f841

Observation 7523c331-160b-4a39-b74a-911205ccb356 · inbound

VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding cites this paper.

VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-02T00:44:48.909159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:44:48.909159Z digest=sha256:628a5de1aaf5c7abf4aeb55c60e001761a67cc81225a16463f45aa311c7e1330