Pith. sign in

Paper Citation Record · LEDGER

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

As of 18 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 13 inbound Pith citation observations for arXiv:2507.09313.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.09313 v2

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:03:14.058499Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T08:39:01.053662Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T01:19:20.288097Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c6d6f3e2-9cf3-46ce-b728-a333e7d61e68 · outbound

This paper cites Qwen2.5-VL Technical Report.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Qwen2.5-VL Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:13.969005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:13.969005Z digest=sha256:05049254f2d18ebfad70e455e9fd0df61e6d6edb9a03bc4e6efb9c0caa5fbe2d

Observation e849daa4-bf69-475c-8e19-353675a6c28c · outbound

This paper cites TemporalBench: Benchmarking Fine-grained Temporal Understanding for Multimodal Video Models.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models TemporalBench: Benchmarking Fine-grained Temporal Understanding for Multimodal Video Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:13.972207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:13.972207Z digest=sha256:0d157846f2e8aef424e66a5b02cb4e1e894a1f61ad7fe431b5af4c0852f5c978

Observation 10285eaf-f2e5-487f-b655-23d9189d1e48 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.449302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T18:03:13.975552Z digest=sha256:427e11011b77a2221ad2dfb028fc82561f7bb4463bad0d2e08af231e6fdcea4d

Observation 486bec5f-f7ab-45ac-8921-1a765c1009d4 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:13.978291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:13.978291Z digest=sha256:1c1a39fa036d25d839aec8e3570876b345934f5ba3dcf2c7c61f0191eed82608

Observation 9eac1249-9df0-4a71-9598-5a9417cdf906 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:13.980956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:13.980956Z digest=sha256:137b6ea6de8552b1295106c30ea98a114ecc487a6e82963a0b128326f56b7c44

Observation 2b04f4a6-1cdd-49c6-a73d-34ebafa03d10 · outbound

This paper cites MMBench-Video: A Long-Form Multi-Shot Benchmark for Holistic Video Understanding.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models MMBench-Video: A Long-Form Multi-Shot Benchmark for Holistic Video Understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:13.983556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:13.983556Z digest=sha256:0f84d8f9777a327d13d02985c6975e540c682218994ca0e1c8b48af9a94609be

Observation dd5322e8-be16-40f6-9692-512e4e38a101 · outbound

This paper cites Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:13.986518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:13.986518Z digest=sha256:74d4a1c818d3dc30ec28ddf5d8c1660fd363d8707326d96df4e10f177081f786

Observation 1ba1b039-4af8-4801-a916-db5fd5cb4f99 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.437624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T18:03:13.989094Z digest=sha256:a27f0c21a3f07cbedd97e9ff5014038daf766990b9ad98cbf141343d84e12b75

Observation 5c973cca-26f8-45ae-802c-4afd7c5f51b5 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.430735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T18:03:13.991479Z digest=sha256:04d33296e917faae464e18be71d8c0a20027a266e52af8a7678c494aaae8ae1a

Observation 3a6bf101-c3d0-4f4c-9281-3e7231bac870 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.423704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T18:03:13.993760Z digest=sha256:15a95af5bada5f43d9981448350583cdef9083dab08443a32f3a7c56924fe711

Observation 143a51d1-6530-49b9-8546-705b6ec35f4d · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models LLaVA-OneVision: Easy Visual Task Transfer

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:13.996017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:13.996017Z digest=sha256:a0a2798964bd7670cb4f9f20b7839ff22590409689e32645c734f86e400acf7d

Observation 0f0d8dc1-58fc-44ab-9453-12b949975ae9 · outbound

This paper cites SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:13.998784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:13.998784Z digest=sha256:923a41ba8215967f01695116bfeaa84df4c21aed857aab4b52997371f615293f

Observation 52c34cc5-e012-4452-8163-c7f287325711 · outbound

This paper cites SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.001294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.001294Z digest=sha256:3ed5cc01437c90d1a767521c7dfe77f16bffda73d761db2020498ab541f670f0

Observation 00ce54f3-6064-41f4-8251-322351b7cbac · outbound

This paper cites LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.003554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.003554Z digest=sha256:fdded0613cde09efa1c3b3f954e15b063a4457c97542caacff3f9d4fa433cd04

Observation 22fe2f4a-7f46-4fed-bf37-23f2ebc16f34 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.416481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T18:03:14.005895Z digest=sha256:961fa2c96ea7cddaf34e6ca4f5d8f0cbfd299d708c1ffc778cc0883e6746f5db

Observation 6822cf65-13fb-425d-801a-e6569e8fef1e · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.408974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T18:03:14.008278Z digest=sha256:611758c62d36b61db1c2b83c9282cf21a42ea632a8ecc06586996880a7d53449

Observation 467ce196-84d0-45fb-be40-1309c5607a5a · outbound

This paper cites OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.010398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.010398Z digest=sha256:bca1de6bf4d3b548a98befd3ddbf381a6743a846adc2706b0774effbfc5bd4ee

Observation ed36c7b1-e5cb-491c-b3c1-ecc56dfc9c8c · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.012841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.012841Z digest=sha256:2fd3cc4f4f49a3dfafe4a00091120379c57ced3f3da9c9bce0c1610e3752187a

Observation a49e3718-f95e-4026-a5f5-0ac0de340ec5 · outbound

This paper cites StreamingBench: Assessing the Gap for MLLMs to Achieve Streaming Video Understanding.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models StreamingBench: Assessing the Gap for MLLMs to Achieve Streaming Video Understanding

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.015160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.015160Z digest=sha256:1878d3b5a75b6f3aa55a81ff1eb30190fbf528ea6af2fb8c6997d30bea7c0cea

Observation 420a5d6b-82a0-4e9c-97f4-e00fa50e180c · outbound

This paper cites E.T. Bench: Towards Open-Ended Event-Level Video-Language Understanding.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models E.T. Bench: Towards Open-Ended Event-Level Video-Language Understanding

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.017689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.017689Z digest=sha256:0f226a6e325fcb5119853df13604ed9e2a47a16c429253104a5f631437c094a9

Observation ebc798c4-f520-4f9e-8b3f-6eef9486475d · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.401107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T18:03:14.020500Z digest=sha256:46043436b37be5675b50e9262ddfd31ec6d466775cf23fe90236523c369bd37e

Observation 7b0e7426-b2f9-414d-8ade-fed24f1340f5 · outbound

This paper cites Dispider: Enabling Video LLMs with Active Real-Time Interaction via Disentangled Perception, Decision, and Reaction.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Dispider: Enabling Video LLMs with Active Real-Time Interaction via Disentangled Perception, Decision, and Reaction

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.023334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.023334Z digest=sha256:8fef3eec72e46e5e94fe5da8f632acbdf5c8b94006e22063ae51546a97999156

Observation 176d99bb-a4b2-4dbd-b5ed-83a21c18dae9 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.393918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T18:03:14.025661Z digest=sha256:cc1386e814f02c5360da94e0d1d6596004d3cb2128cdfed8492658dd4547f047

Observation f07fedcc-806e-4850-a9eb-4833e927b7d5 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.386759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T18:03:14.028068Z digest=sha256:9eb2723d4f4d3422f0357eb445048afc4e433f0040d92f51d9bfa7dd6843a355

Observation 5a549eee-007e-4704-830e-b56c9a39e8ce · outbound

This paper cites Lawrence Zitnick, and Devi Parikh.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Lawrence Zitnick, and Devi Parikh

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:03:14.379613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T18:03:14.030365Z digest=sha256:355425f040b8f6a0839c2e57d82a0606996a0ca4f1dc06a5a2169cc78b3f9d6b

Observation 998eb457-fc7c-4e94-9cfc-dd622733a284 · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.032751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.032751Z digest=sha256:2bdd86ac82613bd7f1fa33cc33d57abdc0b9fdc52991d8f4df3ce866653dc0dc

Observation 8eb0e784-8b2e-4aff-8245-0d2dbdd6e43e · outbound

This paper cites OmniMMI: A Comprehensive Multi-modal Interaction Benchmark in Streaming Video Contexts.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models OmniMMI: A Comprehensive Multi-modal Interaction Benchmark in Streaming Video Contexts

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.035067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.035067Z digest=sha256:6dee96816d3a1357998bfbdcc1fe029a548d38d822c58f0d6b50a1b09c5f0e0c

Observation 85291699-b20d-458e-8586-54e298b76647 · outbound

This paper cites Qwen2.5-Omni Technical Report.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Qwen2.5-Omni Technical Report

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.037742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.037742Z digest=sha256:61d8240c4d4a845cc542f20fb03adcc7fbc7dcf3d20e590e8badf350512e6518

Observation 46f8794c-4d6e-4420-8c39-7195e583ed0c · outbound

This paper cites TimeChat-Online: 80% Visual Tokens are Naturally Redundant in Streaming Videos.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models TimeChat-Online: 80% Visual Tokens are Naturally Redundant in Streaming Videos

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.040372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.040372Z digest=sha256:9809fa4daa94f0b9cde16b66c1f981f7d06c7ac22477bf842acc24ada5b13c40

Observation 20623d6c-9e16-4f52-a6bc-4b75886e1e2c · outbound

This paper cites an unresolved cited work.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:03:14.371311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T18:03:14.042972Z digest=sha256:33ce7168f53d41ee59930c973d63b511dff8b1901d1cb5966841f088fd3ea781

Observation 209fa54e-22a5-428c-89cd-1c8afee2b76c · outbound

This paper cites InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.045371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.045371Z digest=sha256:dc5a1f17f55ab3b5f356ceabe5eb2ad430079be25618ec17e68142cb670324e7

Observation 009ac0f9-9498-40c2-b9cd-b9a93ba48836 · outbound

This paper cites Long Context Transfer from Language to Vision.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models Long Context Transfer from Language to Vision

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.047880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.047880Z digest=sha256:d1afc3279c15376dbcb719d81d6f9919c1fd68a64276cb6aa037100ddd8f75a1

Observation 7a6aeb78-0594-46dc-9dd5-0891c5923a75 · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models BERTScore: Evaluating Text Generation with BERT

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.052687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.052687Z digest=sha256:7448475d1ca6270c6f2b8d72a45fd1e17ec526db9b3d459840f1a94d9863a5ea

Observation 8fd3de2e-9751-44be-9eab-992a8c16393c · outbound

This paper cites online" 'onlinestring :=.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models online" 'onlinestring :=

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.055616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.055616Z digest=sha256:e05d46236794ded25d31380d8de1c2d4b759f04841fe8e69e6d19eb6a5c7967b

Observation 6e4d065e-de94-4dc6-b3b2-278454a4024a · outbound

This paper cites write newline.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models write newline

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.058499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.058499Z digest=sha256:12b032e9778a8fb5056277d455888b945f97b91fa8d627b14cd1671fe4b6045b

Pith citing papers

Observation 5f6fd4f2-22b5-45a3-80e1-e540f88c8c65 · inbound

Beyond Reactivity: Measuring Proactive Problem Solving in LLM Agents cites this paper.

Beyond Reactivity: Measuring Proactive Problem Solving in LLM Agents ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T08:39:01.053662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:39:01.053662Z digest=sha256:6840e0b220eab12b435682f25487c4fc81678ad1296b270367a7b9ae9a3a88f1

Observation 4027df28-ad4c-4b0f-917b-4e48460d6b75 · inbound

Proact-VL: A Proactive VideoLLM for Real-Time AI Companions cites this paper.

Proact-VL: A Proactive VideoLLM for Real-Time AI Companions ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 336

Resolution
malformed identifier
no resolver link, observed 2026-08-02T19:09:09.440908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:09:09.440908Z digest=sha256:43ba25840d2e5d48c4c0897d9407beb834c4915e27546d79591ad46e4523e76d

Observation 4572f0d0-b18f-45ee-804b-a0cc2491ab37 · inbound

From Reactive to Proactive: Assessing the Proactivity of Voice Agents via ProVoice-Bench cites this paper.

From Reactive to Proactive: Assessing the Proactivity of Voice Agents via ProVoice-Bench ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.033091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T11:44:12.373082Z digest=sha256:0ac1dcd5bcaf171dcd909327ba378795aed5c5e0eb4a1164399064ec92987a3a

Observation 6c92c69d-1533-4d11-8cc6-ae801125e38f · inbound

Don't Pause! Every prediction matters in a streaming video cites this paper.

Don't Pause! Every prediction matters in a streaming video ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:41:18.097956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-08T04:32:01.379605Z digest=sha256:1fd40d4c56e6c89c3a579a0031ff86b2f34dd5db9119bde6bae47264d468b2d0

Observation 068e4c34-be0c-4134-84b3-3b821d3b358c · inbound

StreamPro: From Reactive Perception to Proactive Decision-Making in Streaming Video cites this paper.

StreamPro: From Reactive Perception to Proactive Decision-Making in Streaming Video ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:33:50.748812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-20T23:32:38.332828Z digest=sha256:c31f896665c54902c2dae8a8447491891ecb1a7b74c6499bc518f0156fea3688

Observation 41aead78-12f1-4da2-9363-84e7b519d662 · inbound

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction cites this paper.

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-20T13:38:19.157208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-20T13:36:44.071188Z digest=sha256:d758c1c90072e29d479f1f2d9defc4c81d574a442b1822d3548519db2f921fba

Observation ca766167-9a2e-46cc-bc82-858fed26fffc · inbound

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction cites this paper.

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-04T01:19:20.291251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-04T01:11:42.073993Z digest=sha256:4f4f5eb2490ad5b8a2920b9a495f3f2a4fb07b5729ba2be902636fb27fcde5f3

Observation fb22717f-9da6-4ae1-acb9-ffaba75f432d · inbound

An Efficient Streaming Video Understanding Framework with Agentic Control cites this paper.

An Efficient Streaming Video Understanding Framework with Agentic Control ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T11:33:14.470074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-20T11:30:22.151045Z digest=sha256:9a50eabcc2cc6a4deac3ee5004f92a0b6e96c358fbf5caa93043485bbe944c99

Observation 3e967234-cadc-4cee-95cf-5fbac879f8a2 · inbound

IPIBench: Evaluating Interactive Proactive Intelligence of MLLMs under Continuous Streams cites this paper.

IPIBench: Evaluating Interactive Proactive Intelligence of MLLMs under Continuous Streams ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:33:51.065824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-29T18:24:57.881644Z digest=sha256:c14393f246926894f031c76803f6eeed51efcc1e56dfad6ba2fd1ab70b5ecb31

Observation c4af456a-a752-4d00-9ce7-3a177dc71ce5 · inbound

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding cites this paper.

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:46:18.934183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-28T15:06:22.102725Z digest=sha256:9748236e80144519a089bc497a5ec97d6690c6f2a86393830d59a08308f2add4

Observation 0dd45640-7b1f-485c-b538-17545bed0fe1 · inbound

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding cites this paper.

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T10:44:36.467381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-30T10:42:37.401221Z digest=sha256:aae692ba4cff0913337a79de752540f26c3b990dfa389dfce961f5268b70e097

Observation 209f9f59-9164-42e5-8a37-7f88f5dc3007 · inbound

GuideMe: Multi-Domain Task Guidance and Intervention in Streaming Video cites this paper.

GuideMe: Multi-Domain Task Guidance and Intervention in Streaming Video ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-12T05:36:40.278384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:36:40.278384Z digest=sha256:c462e9ef22d9d99cd5c3520d4a1e4a0ef009af2db1a9a7e68443f44b5144f841

Observation 7523c331-160b-4a39-b74a-911205ccb356 · inbound

VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding cites this paper.

VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-02T00:44:48.909159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:44:48.909159Z digest=sha256:628a5de1aaf5c7abf4aeb55c60e001761a67cc81225a16463f45aa311c7e1330