Pith. sign in

Paper Citation Record · LEDGER

OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2501.05510.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.05510 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T20:06:35.153650Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:19:29.950974Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ff4f5f3d-1bd8-4344-9e74-500a973541fc · inbound

VideoRoPE: What Makes for Good Video Rotary Position Embedding? cites this paper.

VideoRoPE: What Makes for Good Video Rotary Position Embedding? OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T20:06:35.153650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T20:06:35.153650Z digest=sha256:8b69a92900c556b43063e8331f84cc259fcef30000e73f18dacd3428ed764579

Observation 6cccf783-ee74-4d1e-b9d0-eac6fdb1ab84 · inbound

Seed1.5-VL Technical Report cites this paper.

Seed1.5-VL Technical Report OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:26:05.803350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-11T05:26:04.960844Z digest=sha256:63f54963217779df1f62f0d2feabca47bb85562e2f4594af143913af439659a6

Observation 467ce196-84d0-45fb-be40-1309c5607a5a · inbound

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models cites this paper.

ProactiveVideoQA: A Comprehensive Benchmark Evaluating Proactive Interactions in Video Large Language Models OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.010398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.010398Z digest=sha256:2c3bb6fc81a2893b1fbacc3d13d78ae39d5103ee7773a9b67d6313986d120a4e

Observation 19be690f-f556-425b-bdba-a60ae34096f3 · inbound

StreamingVLM: Real-Time Understanding for Infinite Video Streams cites this paper.

StreamingVLM: Real-Time Understanding for Infinite Video Streams OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T11:51:33.410043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-17T11:51:33.345812Z digest=sha256:1c8a90fce26e8172d4e1fa8f987568899aacf1d8940813b008f9556082244af8

Observation e7bf43dc-d2b6-43ad-b09e-2d068f17ed6b · inbound

Can Multi-Modal LLMs Provide Live Step-by-Step Task Guidance? cites this paper.

Can Multi-Modal LLMs Provide Live Step-by-Step Task Guidance? OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:39:06.137048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-17T05:36:09.208754Z digest=sha256:8b67a00d2c66d219812a782112a7581c40721486d63a676f3dd9289dd4e911ad

Observation 6c1d7f5c-e6bb-4cd9-8748-5f0d6ce6cb62 · inbound

Streaming Video Instruction Tuning cites this paper.

Streaming Video Instruction Tuning OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:48:21.828285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T19:44:11.032898Z digest=sha256:96a0d906be0d9cc216b9afddd478d70b81a06f891b0da6e8c8153ebe0597ebed

Observation 64639c42-d365-455b-84b7-b22b794da0bd · inbound

HERMES: KV Cache as Hierarchical Memory for Efficient Streaming Video Understanding cites this paper.

HERMES: KV Cache as Hierarchical Memory for Efficient Streaming Video Understanding OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:57:53.869588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:04.564442Z digest=sha256:bbf5b21ca2f49052e0eb9e3940154b5c553a3280800050d1214ba95c85e27d59

Observation ea116109-53b1-4fa7-906c-7669049d2d05 · inbound

EasyVideoR1: Easier RL for Video Understanding cites this paper.

EasyVideoR1: Easier RL for Video Understanding OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:47:12.633041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T07:41:27.231098Z digest=sha256:89725905782eca3ecf13b3b005581a13ef9c76f68f8d39bf51de6b3635e9c461

Observation 26a2a049-a79e-4c0f-82dd-dfb1e09714d9 · inbound

MTT-Bench: Predicting Social Dominance in Mice via Multimodal Large Language Models cites this paper.

MTT-Bench: Predicting Social Dominance in Mice via Multimodal Large Language Models OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:26:09.355931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T09:14:03.068865Z digest=sha256:458457838c30434b9cd8a9f4416ea02e038a3cae5368062c5ce00d53e8a7a566

Observation 8a59e2fa-817c-45b3-851b-5e630d61dcf6 · inbound

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction cites this paper.

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-20T13:38:19.135527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-20T13:36:44.071188Z digest=sha256:3b1b932a33d399c9e290d4eb31e45c1a7fafdc9288377929cbc92df193737bba

Observation 5588ce10-5083-4545-92a4-fb0d72ffe780 · inbound

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction cites this paper.

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-04T01:19:20.330777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-04T01:11:42.073993Z digest=sha256:6e78ac8ae1bf0269a5983d0cb76770130fc5d84bac3352e3563cc5d6273231d7

Observation 26f134f2-fede-4f1b-bdb9-13fd41323578 · inbound

Don't Pause: Streaming Video-Language Synchrony for Online Video Understanding cites this paper.

Don't Pause: Streaming Video-Language Synchrony for Online Video Understanding OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 102

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:07:12.787861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-27T22:11:01.690237Z digest=sha256:84919f8f7191564f5859e1140fc6b0eb83ef44c8300ce8ec95ef0ed6425fec21

Observation 0f9ca443-8094-4aa5-aabb-c23d93a03b40 · inbound

MOSS-Video-Preview: Toward Real-Time Video Understanding via Cross-Attention cites this paper.

MOSS-Video-Preview: Toward Real-Time Video Understanding via Cross-Attention OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:26:17.968523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-28T15:22:31.310003Z digest=sha256:7bb4cfcbae53ec4e742e38b0580dd557ab1a338f72bdd031522daf3490ae6a6c

Observation f3f73c38-d09e-4657-bb3f-1b8532e31bb6 · inbound

Streaming Interventions: Can Video Large Language Models Correct Mistakes as They Occur? cites this paper.

Streaming Interventions: Can Video Large Language Models Correct Mistakes as They Occur? OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:57:30.637473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-27T16:52:22.811857Z digest=sha256:0e4770a24766883c2bda08d1e855452a8de51c03f5cca813f521fbead3b330c0

Observation 1f0a41bd-83de-4108-83c6-2735fe6e7321 · inbound

LiveStarPro: Proactive Streaming Video Understanding with Hierarchical Memory for Long-Horizon Streams cites this paper.

LiveStarPro: Proactive Streaming Video Understanding with Hierarchical Memory for Long-Horizon Streams OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:38:56.213701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-27T01:12:46.295455Z digest=sha256:b7d5945796bbde8afa10047c7ac37a2282c1a4c1acf9625aa5f887f0e755c343

Observation 693ee081-db65-464a-b9b1-9cb3d0fbaeb2 · inbound

ViCoStream: Streaming VideoLLMs Can Run Beyond 100 FPS with Stage-Wise Coordinated Inference cites this paper.

ViCoStream: Streaming VideoLLMs Can Run Beyond 100 FPS with Stage-Wise Coordinated Inference OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:19:29.952804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-26T18:17:53.013043Z digest=sha256:5d802dcaf2761decc2ed2bf1e7841013cb4230bd010e86f9cb277ef95135ae51

Observation 11656f91-4a1a-48e8-a1a7-f3f01c71e4ae · inbound

MedStreamBench: A Time-Aware Benchmark for Streaming and Proactive Medical Video Understanding cites this paper.

MedStreamBench: A Time-Aware Benchmark for Streaming and Proactive Medical Video Understanding OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:38:39.536434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-07-03T16:37:09.666491Z digest=sha256:3622867c701149d3dddbd9fbb189b57989fedff617d57d7479931e2ca03deaec

Observation 204053fd-c1cc-4cfc-85d2-5c5e45f60e3b · inbound

VIABench: A Comprehensive Video Benchmark Collected from Blind Individuals for Visual Impairment Assistance cites this paper.

VIABench: A Comprehensive Video Benchmark Collected from Blind Individuals for Visual Impairment Assistance OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T01:31:31.984732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:31:31.984732Z digest=sha256:110bcc1c82ca9e40b0b2f1e7347849ed9340d5f648c25a69bc8e29c17edd1833

Observation 7775ae64-939a-4fd1-b6b3-28f414a6c3b1 · inbound

ObjectStream: Latent Objects as Memory Anchors for Streaming Video Understanding cites this paper.

ObjectStream: Latent Objects as Memory Anchors for Streaming Video Understanding OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T03:21:44.755135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T03:21:44.755135Z digest=sha256:1a339d0cd621fa32aab9606b34b5e787d1f66b3fe8ac8419559194f85492d368