Pith. sign in

Paper Citation Record · LEDGER

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design

As of 15 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 6 inbound Pith citation observations for arXiv:2505.16175.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.16175 v2

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:09:02.863981Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-13T21:37:55.887477Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact2
  • verified fuzzy13
  • unresolved33
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation ad220a81-1908-4937-95dd-ef66c6e778ae · outbound

This paper cites GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:57.325344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:57.325344Z digest=sha256:8a7ffa217d621cf6c3dac986f07ab9b2da3dbf9c296a3b8b1d28d0108d351388

Observation 3f980e2e-5d23-4ff3-b0f0-3c5d3cccaf3e · outbound

This paper cites Deep Architectures for Content Moderation and Movie Content Rating.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Deep Architectures for Content Moderation and Movie Content Rating

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:09:04.034398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:08:57.466754Z digest=sha256:c0603914149b84c902390931d3ad7aa8a7c64a34573bd56b7a221df87467c87f

Observation 2ac4bdae-3bda-4181-858f-70268bb873ac · outbound

This paper cites Amazon ec2 p5 instances.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Amazon ec2 p5 instances

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:07.635821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:08:57.597010Z digest=sha256:d5ae9572bad14bc14940af82232fde73178912d5bd04631cb9aab4446062ca24

Observation 252502dc-737e-48bd-ad87-fce68d6b572d · outbound

This paper cites Qwen2.5-vl technical report,.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Qwen2.5-vl technical report,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:57.755911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:57.755911Z digest=sha256:c87578add770b8ac2b86501447d19dab87a8bd320fdebb330c3dde55a8b6efb3

Observation fbf6c014-a83f-47b5-8413-cfae70d5ef7b · outbound

This paper cites YouTube: hours of video uploaded every minute 2022| Statista — statista.com.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design YouTube: hours of video uploaded every minute 2022| Statista — statista.com

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:07.331028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:08:58.001673Z digest=sha256:7a3b9453e19606d0bba632f7658f7751a1247e8c3d9a3b2f78dd76e7d3272662

Observation 8d9b09a8-591c-47ad-9595-2972c42d2b4b · outbound

This paper cites An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision-language models.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision-language models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:07.158292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:08:58.182960Z digest=sha256:a152f3696123a075f319ae91ba615a196aefb70d7ee16fb71e476275d9e7bbe0

Observation cfd80a0d-2535-4e0e-8341-55ac8772c364 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:58.336854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:58.336854Z digest=sha256:452672d34e6bc9eb71ee4a85124529845357ad8bd5d86ce85609d549b07557ad

Observation 811c1184-7641-4290-a1b1-cf8ad31f6424 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:58.643779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:58.643779Z digest=sha256:2454726a9120d041d2f583873278a85865debc78278968b8ec33fb4aae2df2c5

Observation 64bf8cb9-74cb-4ed2-8d39-8b08676803cb · outbound

This paper cites A simple and effective l_2 norm-based strategy for KV cache compression.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design A simple and effective l_2 norm-based strategy for KV cache compression

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:58.734457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:58.734457Z digest=sha256:d4543c152949a2473ceff04e095dc0ec559a2b8e0dea846b71f8d4ee0edd7d23

Observation 9063a399-1985-458e-8132-c815e7c1f289 · outbound

This paper cites an unresolved cited work.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:09:06.915683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:08:58.825571Z digest=sha256:945161a46994cb89aecb0ee0f2b1f498e6b782a501dd4c643daebbf894a3419f

Observation c6b01153-5754-4bcb-9386-c324953cad71 · outbound

This paper cites Global Expansion of AI Surveillance.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Global Expansion of AI Surveillance

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:06.612281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:08:58.938549Z digest=sha256:6366c89c70925caadbcea1e92ae9c8c33ab0be310aaefe2095fe8d0a47146f30

Observation aefb289d-6cb3-4faa-9153-0fb45712fa7d · outbound

This paper cites Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:59.036704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:59.036704Z digest=sha256:ef9205550b0fafdaf30f54f2b130a6115fdfec619dad0b28c72174dd12fb93d6

Observation bd9548cd-1a58-47a9-a7f1-e3b680b73237 · outbound

This paper cites Gpu machine types.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Gpu machine types

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:06.360188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:08:59.124942Z digest=sha256:7c6ef0567260a50ba4c325e31fc6ae2162871ed71b5113694043c2a56038fa17

Observation 2de940d4-e8e9-4719-8d0e-5928ab4f91f4 · outbound

This paper cites Attention score is not all you need for token importance indicator in kv cache reduction: Value also matters.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Attention score is not all you need for token importance indicator in kv cache reduction: Value also matters

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:06.109502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:08:59.275533Z digest=sha256:af46a46e0636f65f1af28c89359e1518a6f69b361809298a497f0d2b80506219

Observation b54852ed-8063-4b14-8351-72249f017395 · outbound

This paper cites The video codec landscape in 2020.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design The video codec landscape in 2020

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:05.835354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:08:59.378650Z digest=sha256:77fa9d779db5289b1af64ada1782af6a63018f7ed4e530652ae11b59b2a279ef

Observation c2d29906-50b1-4075-b8a9-94a8ac06066b · outbound

This paper cites Mpeg-4 overview.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Mpeg-4 overview

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:05.600116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:08:59.495223Z digest=sha256:348f69d8d986de8a0c580fa3244a65c67a11f098bca6b71784b26866c23d8e39

Observation db9fcf24-695a-4aec-a574-7ae3ef652bb9 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Gonzalez, Hao Zhang, and Ion Stoica

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:59.606163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:59.606163Z digest=sha256:fcfd7a3e679edfca3b3474a2eb2e121c1ba00d8156ac462580858515d63efb90

Observation 12fb61f6-4b5d-4952-aa51-67b730de0eb5 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design LLaVA-OneVision: Easy Visual Task Transfer

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:59.692862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:59.692862Z digest=sha256:ceb808d8056d7e17a75ba716768e376fcc05857fe1ef97d24fdc9bf2cab3c84a

Observation ac245947-e8e5-4ba4-898e-2918d830290a · outbound

This paper cites Video-llava: Learning united visual representation by alignment before projection.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Video-llava: Learning united visual representation by alignment before projection

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:05.331274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:08:59.745158Z digest=sha256:9aa68f0ae74243eadf9179e86fa30e1ad34555dc8a0673bbcc46555134fe7ef1

Observation 71d3d810-404c-4111-99a7-9e411e6639ed · outbound

This paper cites an unresolved cited work.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:59.809043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:59.809043Z digest=sha256:d4f447caf85a2d75851da8a7ceec72e232b4dc2893e725c81b4f904bb5319bdf

Observation 9552e308-723a-4e74-bdbc-aa9c6648000a · outbound

This paper cites Torchvision: Pytorch’s computer vision library.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Torchvision: Pytorch’s computer vision library

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:59.880737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:59.880737Z digest=sha256:5b312784863a8d9e5c4a7809525ea159594b64b024b1b40a4743b8eb2a6ce2ad

Observation 96f232b4-0979-4d45-8381-e1451d40afc5 · outbound

This paper cites Nd-h100-v5 sizes series.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Nd-h100-v5 sizes series

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:05.109963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:08:59.924660Z digest=sha256:2cb15c31f6980c724333085c44cd7e3afed62237042eeb352a4c269f1b4dc44b

Observation 2658dd93-e59c-4028-a90c-33833fc8716c · outbound

This paper cites Slowfocus: Enhancing fine-grained temporal understanding in video llm.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Slowfocus: Enhancing fine-grained temporal understanding in video llm

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:04.842954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:09:00.015450Z digest=sha256:6d21de84d1608dba11f3fac69d55281ebfd8606078f11d2ae7585a4964512053

Observation 4e2ee009-0eea-42d5-80d2-b1e18af85000 · outbound

This paper cites Cosmos World Foundation Model Platform for Physical AI.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Cosmos World Foundation Model Platform for Physical AI

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:00.103049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:00.103049Z digest=sha256:864862225d83f57c0a3a131b5fcb87f4b404b2fbafcbe42ac0106853ab4471a6

Observation f0ec15a1-ae3c-48d0-8bb1-3800a70dee46 · outbound

This paper cites torchcodec.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design torchcodec

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:04.570573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:09:00.206310Z digest=sha256:8de8fc0caab536499533cf3fefe031b1da54bfa654b31ed0a74f9a94ec68363d

Observation 2dcecacd-925a-46e2-b249-3da4fd109665 · outbound

This paper cites Llava-prumerge: Adaptive token reduction for efficient large multimodal models.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Llava-prumerge: Adaptive token reduction for efficient large multimodal models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:00.322271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:00.322271Z digest=sha256:cdc27e86095020d3fdc1190488f6ef45a7bffda50e286d43a2ff276e40e8c4d8

Observation 89736361-1637-42ea-a8a8-008f86d131d4 · outbound

This paper cites GLU Variants Improve Transformer.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design GLU Variants Improve Transformer

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:00.413151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:00.413151Z digest=sha256:1bccd31a51f630c37e5e90955b383861dab9cbc5a866ee2f8bfe8b771084442c

Observation bf034485-8057-420d-aab4-16d8ef8b1483 · outbound

This paper cites Reducing traffic wastage in video streaming via bandwidth-efficient bitrate adaptation.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Reducing traffic wastage in video streaming via bandwidth-efficient bitrate adaptation

Reference 30

Resolution
malformed identifier
no resolver link, observed 2026-08-07T15:09:00.505815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:00.505815Z digest=sha256:138add9260371c6002f75d8da6a61cbba7050e29d9ebca9fd57fce4ffffffd33

Observation 2271bb92-5821-4c86-808c-57a53eba606e · outbound

This paper cites Sullivan, Jens-Rainer Ohm, Woo-Jin Han, and Thomas Wiegand.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Sullivan, Jens-Rainer Ohm, Woo-Jin Han, and Thomas Wiegand

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:00.626260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:00.626260Z digest=sha256:99cdc0ad746b607ca6ba6a37137dfecc24416b5813fb2abccd33c2b6b90fa290

Observation 7cbbf95d-23e6-4de7-9713-305bd8a0bb45 · outbound

This paper cites Converting video formats with ffmpeg.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Converting video formats with ffmpeg

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:00.747842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:00.747842Z digest=sha256:12987ddacd99a43631c48ca3e0ee86c7941b4ef5425800119f71ee318aba0f01

Observation d9eb447b-a3df-4d50-88d7-8f26f853b5eb · outbound

This paper cites Oliphant, Matt Haberland, Tyler Reddy, David Cournapeau, Evgeni Burovski, Pearu Peterson, Warren Weckesser, Jonathan Bright, Stéfan J.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Oliphant, Matt Haberland, Tyler Reddy, David Cournapeau, Evgeni Burovski, Pearu Peterson, Warren Weckesser, Jonathan Bright, Stéfan J

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:00.846963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:00.846963Z digest=sha256:e8f8a47c0f987e7a7844acd8cb0fdc97fe0ab7b1240c12c8556c2c83570eb492

Observation c5a491e5-87fc-457d-b189-90a230f2a774 · outbound

This paper cites LVBench: An Extreme Long Video Understanding Benchmark.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design LVBench: An Extreme Long Video Understanding Benchmark

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:00.920595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:00.920595Z digest=sha256:8c70a2685229b2f5d5889d5a9411555cee735a0303340392ef08ca9b5f4dc8e9

Observation 3b4f82a9-424b-4afe-922d-3d11b78ce6a3 · outbound

This paper cites InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:01.009435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:01.009435Z digest=sha256:a1f0c68435d86804e7ee1a15adbe83df3f041cd7278ac3f4b463910f85f6a1c2

Observation 556af35a-f87c-47f4-9e94-d83197cf4798 · outbound

This paper cites Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:01.081340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:01.081340Z digest=sha256:4a38053e3fdaade084a1f10d50bbc979e98806db55ac29600af29ab4db93adc4

Observation 8d4d7a48-0c57-4724-95c2-1949778989c6 · outbound

This paper cites LongVLM: Efficient Long Video Understanding via Large Language Models.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design LongVLM: Efficient Long Video Understanding via Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:01.176007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:01.176007Z digest=sha256:d06976f27bd82f53ddf17eb831573e547149cf19fd01d0b355e64c8b74f060a2

Observation 21f4183b-859e-434e-81fc-e3e0ce269c87 · outbound

This paper cites Wiegand, G.J.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Wiegand, G.J

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:01.308068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:01.308068Z digest=sha256:87ca42a7a4c085cdf40b7e452480c214076ebb317c2871d3a2427a65c05c1bff

Observation ddaeca31-db4e-410a-88f6-5ddb7030a003 · outbound

This paper cites an unresolved cited work.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:01.428834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:01.428834Z digest=sha256:30cc7d1a2b2f3b65bf0a61bd4b2960e4df5b7dfd68a5ff9f155f4fabfdfb374a

Observation 12eb9dac-3327-471b-a1bd-e057e52bebb7 · outbound

This paper cites LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:01.557498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:01.557498Z digest=sha256:6c768a3b6588275f07264a19e52090fc1e63e80ac00e450c7c8a7088444d7173

Observation d9797340-f40b-47eb-b2ac-873e1785ae1f · outbound

This paper cites Efficient Streaming Language Models with Attention Sinks.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Efficient Streaming Language Models with Attention Sinks

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:01.693596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:01.693596Z digest=sha256:34b6f20f4989f3e7d570e07606432004cd3fdd7ba3e04a40c34ea19fe52e7483

Observation d845edd9-fd66-4d33-9543-b7d481dec836 · outbound

This paper cites PyramidDrop: Accelerating Your Large Vision-Language Models via Pyramid Visual Redundancy Reduction.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design PyramidDrop: Accelerating Your Large Vision-Language Models via Pyramid Visual Redundancy Reduction

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:01.834128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:01.834128Z digest=sha256:e560a268f96e8006bff6340cd1e07d3659ffaaea5d571bd9384c83067d320547

Observation b2be132b-8cf0-4626-8968-7c65d98a184e · outbound

This paper cites Towards surveillance video-and-language understanding: New dataset, baselines, and challenges,.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Towards surveillance video-and-language understanding: New dataset, baselines, and challenges,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:04.292742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:09:01.923406Z digest=sha256:3f1d1c7042fe3adeab71d4bfcfd9a19992f3ae287cdfd41d546e39e5ecf3fe73

Observation a1a68158-024f-4428-8b46-b41501f1b8a0 · outbound

This paper cites LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:02.207599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:02.207599Z digest=sha256:72ced4feaeacf3ea74ebb70964027f9960e46442568c76b617e64fb39281b571

Observation 0446f2e1-6718-42ef-9695-d2069951736d · outbound

This paper cites H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:02.326672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:02.326672Z digest=sha256:4fd79619943a143c60bd26ade9d23680ba8e33b2c1c1fcb6bec8a9e281075764

Observation 6c627c0b-e94d-41a4-9e07-decad9278543 · outbound

This paper cites MLVU: Benchmarking Multi-task Long Video Understanding.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design MLVU: Benchmarking Multi-task Long Video Understanding

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:02.433510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:02.433510Z digest=sha256:9aff43008c1bc3712c3d5d5b5cd1862b71a28dc7298e25598bc0c41300c6b48a

Observation aa11f0f8-8e95-4899-95da-55d7c9a06ec2 · outbound

This paper cites A Survey on Generative AI and LLM for Video Generation, Understanding, and Streaming.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design A Survey on Generative AI and LLM for Video Generation, Understanding, and Streaming

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:02.584217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:02.584217Z digest=sha256:6cc1aaee9b15e74f5cb7d1a8dc3bef583c6320f243ac6362e18a37f499812477

Observation c3b24cbf-3b98-4fd6-9d57-9e382a28dd35 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:02.863981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:02.863981Z digest=sha256:1d30b45880284c9240527fe0184c57b46249ad254b4e265a72f2a9d1078ae71f

Observation b6fc6b63-aef3-41a6-933a-504c6ef83247 · outbound

This paper cites Towards Surveillance Video-and-Language Understanding: New Dataset, Baselines, and Challenges.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Towards Surveillance Video-and-Language Understanding: New Dataset, Baselines, and Challenges

Reference 2023

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:09:03.187420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:09:02.065778Z digest=sha256:a68b3ae6759fc499a38319e5de3c44ceea847537656fb4ac6b25c4a9b315bafd

Observation ec2dfef2-143e-46ba-a653-f8d36c27625e · outbound

This paper cites VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:58.536816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:58.536816Z digest=sha256:07534246f0bee0c58fa1737210311b03f88d16c9d7bb7bcf495453834ea425de

Observation cf2bdaea-b699-480c-8978-b137f78de9a6 · outbound

This paper cites Qwen2.5-VL Technical Report.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Qwen2.5-VL Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:57.859255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:57.859255Z digest=sha256:34ff2091542063700e79bd14b95b632ee7fd065fb2e98c0df90d8fd19ff36dbc

Pith citing papers

Observation c3255245-49d4-4b80-a211-f65760303b0c · inbound

CoVR-R:Reason-Aware Composed Video Retrieval cites this paper.

CoVR-R:Reason-Aware Composed Video Retrieval QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-13T21:37:55.887477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T21:37:55.887477Z digest=sha256:c505be2d427a90e518562b5916a678e5cb419182f3224adea63f658144959fcf

Observation f928d3bc-10f3-4e00-a3ea-9b161ef66450 · inbound

cuRAMSES: Scalable AMR Optimizations for Large-Scale Cosmological Simulations cites this paper.

cuRAMSES: Scalable AMR Optimizations for Large-Scale Cosmological Simulations QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design

Reference 63

Resolution
unresolved
no resolver link, observed 2026-07-13T09:09:32.932051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T09:09:32.932051Z digest=sha256:6e5bd7ec1713a0c9fcb2ecce34c7805e23dd71fb668fccbee9e3d1ff80c6c9c7

Observation 8bdc8817-a71d-4a51-8316-bb6c919749e2 · inbound

CodecSight: Leveraging Video Codec Signals for Efficient Streaming VLM Inference cites this paper.

CodecSight: Leveraging Video Codec Signals for Efficient Streaming VLM Inference QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:00:52.799503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T18:45:26.555395Z digest=sha256:c97d4d13c995f8d980fd4e434b41e9a61b9b3f7767174c0a773f967da9cfd260

Observation d625753e-86c1-4920-8ef6-d5fcf85243e5 · inbound

Mosaic: Cross-Modal Clustering for Efficient Video Understanding cites this paper.

Mosaic: Cross-Modal Clustering for Efficient Video Understanding QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-10T16:10:34.374605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T16:07:36.133404Z digest=sha256:0ebaed78164d68541d4fab984c00bea12fab94d66db352647b7359ab47cddaea

Observation 3df7cf18-54bf-490a-bce5-0e112bc67269 · inbound

VLMaxxing through FrameMogging Training-Free Anti-Recomputation for Video Vision-Language Models cites this paper.

VLMaxxing through FrameMogging Training-Free Anti-Recomputation for Video Vision-Language Models QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:11:13.760186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-08T01:30:15.463051Z digest=sha256:41446ba21f3a45aca504d5fcb7900a834d476fc62b266ff1310e8b8045f8a664

Observation 42def0b6-30ec-4fda-a44a-e3722280bbdc · inbound

Watch, Remember, Reason: Human-View Video Understanding with MLLMs cites this paper.

Watch, Remember, Reason: Human-View Video Understanding with MLLMs QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design

Reference 173

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:27:15.530520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T22:00:28.350003Z digest=sha256:66ca07ed2b06c85ffd53cddeac51a6c905beb101ba2dfc4f7e570f68aee22368