Pith. sign in

Paper Citation Record · LEDGER

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design

As of 8 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 6 inbound Pith citation observations for arXiv:2505.16175.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.16175 v2

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:09:02.863981Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-13T21:37:55.887477Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact2
  • verified fuzzy13
  • unresolved33
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation ad220a81-1908-4937-95dd-ef66c6e778ae · outbound

This paper cites GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:57.325344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:57.325344Z digest=sha256:043677a6a237920613321b24fb49b8b064e1cafc28963e571bcf8f0fb223f42f

Observation 3f980e2e-5d23-4ff3-b0f0-3c5d3cccaf3e · outbound

This paper cites Deep Architectures for Content Moderation and Movie Content Rating.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Deep Architectures for Content Moderation and Movie Content Rating

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:09:04.034398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:08:57.466754Z digest=sha256:d6cb692df05b30e57e92a8586ad68ea2cb99be1152288bc04fb1337a1ca179b7

Observation 2ac4bdae-3bda-4181-858f-70268bb873ac · outbound

This paper cites Amazon ec2 p5 instances.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Amazon ec2 p5 instances

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:07.635821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:08:57.597010Z digest=sha256:5f0e4771bd8b7af91dfe62e8851a842c6e5175a2a66f69d381cc69d5c94501ef

Observation 252502dc-737e-48bd-ad87-fce68d6b572d · outbound

This paper cites Qwen2.5-vl technical report,.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Qwen2.5-vl technical report,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:57.755911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:57.755911Z digest=sha256:8ba286c52ff062170c4611353776994ec58b26a13fd1b01f7b5cb7aa62f270ec

Observation fbf6c014-a83f-47b5-8413-cfae70d5ef7b · outbound

This paper cites YouTube: hours of video uploaded every minute 2022| Statista — statista.com.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design YouTube: hours of video uploaded every minute 2022| Statista — statista.com

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:07.331028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:08:58.001673Z digest=sha256:82e36b934f5a211263181a286f85b350c3da52d0c18cec608b822088b18f79ec

Observation 8d9b09a8-591c-47ad-9595-2972c42d2b4b · outbound

This paper cites An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision-language models.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision-language models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:07.158292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:08:58.182960Z digest=sha256:616d7814be26009c2b188b896a519dce65c0da75afc19eb00bd60bb0948313e9

Observation cfd80a0d-2535-4e0e-8341-55ac8772c364 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:58.336854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:58.336854Z digest=sha256:d8fd894591df47c5153f50973fb2bcbf4284a79befd83c9f3589b5fc06ac3097

Observation 811c1184-7641-4290-a1b1-cf8ad31f6424 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:58.643779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:58.643779Z digest=sha256:456c1e8dd8b39c33a87f9a1e9766d8dd510ee1772bcae57ebff044de2b642fdf

Observation 64bf8cb9-74cb-4ed2-8d39-8b08676803cb · outbound

This paper cites A simple and effective l_2 norm-based strategy for KV cache compression.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design A simple and effective l_2 norm-based strategy for KV cache compression

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:58.734457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:58.734457Z digest=sha256:f408844e4e8cc732238709e220f06830d5773f5957050c06aa6eb03a46d2ca8b

Observation 9063a399-1985-458e-8132-c815e7c1f289 · outbound

This paper cites an unresolved cited work.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:09:06.915683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:08:58.825571Z digest=sha256:68187a4b331973f3ae59d085b84d0c670d1ea46610d1fb3e0893551f3f5e9e5f

Observation c6b01153-5754-4bcb-9386-c324953cad71 · outbound

This paper cites Global Expansion of AI Surveillance.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Global Expansion of AI Surveillance

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:06.612281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:08:58.938549Z digest=sha256:f7d54faccaa801533cec720e19eec3bf49d9a070b70a72f5e4229ce34dfc3ada

Observation aefb289d-6cb3-4faa-9153-0fb45712fa7d · outbound

This paper cites Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:59.036704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:59.036704Z digest=sha256:16145aeb8477582fe15a0a63d4e55b8844ed3fa3f838298386178cdb0dbfbdc8

Observation bd9548cd-1a58-47a9-a7f1-e3b680b73237 · outbound

This paper cites Gpu machine types.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Gpu machine types

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:06.360188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:08:59.124942Z digest=sha256:7634d785d5be3a36d2f19a8ac870e3718b012632e916095e35dbac43cd2e162c

Observation 2de940d4-e8e9-4719-8d0e-5928ab4f91f4 · outbound

This paper cites Attention score is not all you need for token importance indicator in kv cache reduction: Value also matters.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Attention score is not all you need for token importance indicator in kv cache reduction: Value also matters

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:06.109502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:08:59.275533Z digest=sha256:f58c786d6740dcbf08e1e85d5e4ca9b2a0d1346a9b9bef93bb47b4616b169452

Observation b54852ed-8063-4b14-8351-72249f017395 · outbound

This paper cites The video codec landscape in 2020.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design The video codec landscape in 2020

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:05.835354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:08:59.378650Z digest=sha256:45a5de81641a07baf880c54b430a3f32445ad8d98fe238236744caeeedf21720

Observation c2d29906-50b1-4075-b8a9-94a8ac06066b · outbound

This paper cites Mpeg-4 overview.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Mpeg-4 overview

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:05.600116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:08:59.495223Z digest=sha256:34bb2ddb33052a4eddfb674125b46405fe3bad2538485133a7d43740859011e5

Observation db9fcf24-695a-4aec-a574-7ae3ef652bb9 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Gonzalez, Hao Zhang, and Ion Stoica

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:59.606163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:59.606163Z digest=sha256:4f6b834faad8bd23b25abd604a2915f12e9b4600e6a42ed36b7bd2a72335a40d

Observation 12fb61f6-4b5d-4952-aa51-67b730de0eb5 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design LLaVA-OneVision: Easy Visual Task Transfer

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:59.692862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:59.692862Z digest=sha256:423d061256194844a9494ff9866234d6e3243105c3768c3740e8c5587bd1705a

Observation ac245947-e8e5-4ba4-898e-2918d830290a · outbound

This paper cites Video-llava: Learning united visual representation by alignment before projection.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Video-llava: Learning united visual representation by alignment before projection

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:05.331274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:08:59.745158Z digest=sha256:f9a2ea72d4e7e140b5d7045232211c9521a2b29ec3db892424aeff4f0223cc69

Observation 71d3d810-404c-4111-99a7-9e411e6639ed · outbound

This paper cites an unresolved cited work.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:59.809043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:59.809043Z digest=sha256:11a13fdf6a088ef2c1545d892d9f870a6e883fa3aa45c78e0115870b74b37db1

Observation 9552e308-723a-4e74-bdbc-aa9c6648000a · outbound

This paper cites Torchvision: Pytorch’s computer vision library.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Torchvision: Pytorch’s computer vision library

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:59.880737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:59.880737Z digest=sha256:fa6aa29c401802d9b0a302bb78f97ac0af40c0e3986b95bc40abc0872aae23fa

Observation 96f232b4-0979-4d45-8381-e1451d40afc5 · outbound

This paper cites Nd-h100-v5 sizes series.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Nd-h100-v5 sizes series

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:05.109963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:08:59.924660Z digest=sha256:6cd8d6248161a0e2962e219c8a7a6d8e329a75208fee2dbb021760b5a1618ab7

Observation 2658dd93-e59c-4028-a90c-33833fc8716c · outbound

This paper cites Slowfocus: Enhancing fine-grained temporal understanding in video llm.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Slowfocus: Enhancing fine-grained temporal understanding in video llm

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:04.842954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:09:00.015450Z digest=sha256:5ca77afc32cdd5981faddf0988f7a094ab523cd9dd35ff27923969d4bac6ea7d

Observation 4e2ee009-0eea-42d5-80d2-b1e18af85000 · outbound

This paper cites Cosmos World Foundation Model Platform for Physical AI.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Cosmos World Foundation Model Platform for Physical AI

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:00.103049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:00.103049Z digest=sha256:c1d376038f2067b1a5d94a32bd2e4e521dc1df85ed56ef273d53a34da495f8fc

Observation f0ec15a1-ae3c-48d0-8bb1-3800a70dee46 · outbound

This paper cites torchcodec.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design torchcodec

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:04.570573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:09:00.206310Z digest=sha256:3d9b277b8815b9a2efb202700785cc58e507f0463769be46376d4e6c2703a513

Observation 2dcecacd-925a-46e2-b249-3da4fd109665 · outbound

This paper cites Llava-prumerge: Adaptive token reduction for efficient large multimodal models.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Llava-prumerge: Adaptive token reduction for efficient large multimodal models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:00.322271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:00.322271Z digest=sha256:4c0d15fd1bb4495534572e60e0f7d2723ffcaf4355c949d45d5e82192f1562b8

Observation 89736361-1637-42ea-a8a8-008f86d131d4 · outbound

This paper cites GLU Variants Improve Transformer.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design GLU Variants Improve Transformer

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:00.413151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:00.413151Z digest=sha256:bfaba87ed3387dc0a8ec85db102fd5f9f10be9ce7787553edcb66d6555eadc9c

Observation bf034485-8057-420d-aab4-16d8ef8b1483 · outbound

This paper cites Reducing traffic wastage in video streaming via bandwidth-efficient bitrate adaptation.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Reducing traffic wastage in video streaming via bandwidth-efficient bitrate adaptation

Reference 30

Resolution
malformed identifier
no resolver link, observed 2026-08-07T15:09:00.505815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:00.505815Z digest=sha256:b1b778bd616e1b72f9f39c0978248cb581ba25f5d6f221609c7ac6f25b920f0d

Observation 2271bb92-5821-4c86-808c-57a53eba606e · outbound

This paper cites Sullivan, Jens-Rainer Ohm, Woo-Jin Han, and Thomas Wiegand.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Sullivan, Jens-Rainer Ohm, Woo-Jin Han, and Thomas Wiegand

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:00.626260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:00.626260Z digest=sha256:c80ffb6084ae96c52a1075491d85b704b01eb04553bdd44e623ced8493499e50

Observation 7cbbf95d-23e6-4de7-9713-305bd8a0bb45 · outbound

This paper cites Converting video formats with ffmpeg.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Converting video formats with ffmpeg

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:00.747842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:00.747842Z digest=sha256:769419fd36948e35dd9d6e60d5736a36e9fcdcc3ee0bf989225b9394e33bbad5

Observation d9eb447b-a3df-4d50-88d7-8f26f853b5eb · outbound

This paper cites Oliphant, Matt Haberland, Tyler Reddy, David Cournapeau, Evgeni Burovski, Pearu Peterson, Warren Weckesser, Jonathan Bright, Stéfan J.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Oliphant, Matt Haberland, Tyler Reddy, David Cournapeau, Evgeni Burovski, Pearu Peterson, Warren Weckesser, Jonathan Bright, Stéfan J

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:00.846963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:00.846963Z digest=sha256:b31b5d0ebd119a2c72259ad52c68f85649faf6d65a736a3c7533baae8dc351d1

Observation c5a491e5-87fc-457d-b189-90a230f2a774 · outbound

This paper cites LVBench: An Extreme Long Video Understanding Benchmark.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design LVBench: An Extreme Long Video Understanding Benchmark

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:00.920595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:00.920595Z digest=sha256:9b18d9d72168cf213114163ff529e0ec7469e8001195c9c8277c14ef07e43299

Observation 3b4f82a9-424b-4afe-922d-3d11b78ce6a3 · outbound

This paper cites InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:01.009435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:01.009435Z digest=sha256:12f7365eedd2aaae2b1ffed31e07cfa0ebf0dd516d35b42cc9d44261b1d28898

Observation 556af35a-f87c-47f4-9e94-d83197cf4798 · outbound

This paper cites Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:01.081340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:01.081340Z digest=sha256:7d7bee58b5ff79cd5e2f58b011f929fda6fc6f282cedd5acab8d45aba53bf7f4

Observation 8d4d7a48-0c57-4724-95c2-1949778989c6 · outbound

This paper cites LongVLM: Efficient Long Video Understanding via Large Language Models.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design LongVLM: Efficient Long Video Understanding via Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:01.176007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:01.176007Z digest=sha256:5e137979bebc9c07d94bbdb83c3acb9f76b17b3779d1512ef5fd3d0575ae0830

Observation 21f4183b-859e-434e-81fc-e3e0ce269c87 · outbound

This paper cites Wiegand, G.J.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Wiegand, G.J

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:01.308068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:01.308068Z digest=sha256:4b86e8e8ff87cd47ea0187835f28928559b00850f9e3ac19d5bd3509d3f826ed

Observation ddaeca31-db4e-410a-88f6-5ddb7030a003 · outbound

This paper cites an unresolved cited work.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:01.428834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:01.428834Z digest=sha256:3a287180f634ddc2ba251ee3e0a67d51cf691f181b732397967a1ef4aa4697fa

Observation 12eb9dac-3327-471b-a1bd-e057e52bebb7 · outbound

This paper cites LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:01.557498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:01.557498Z digest=sha256:645be1a57bb5149172b6c77d53d3232ecd45bfd440fe4ad5c3e32ec1c37fa46e

Observation d9797340-f40b-47eb-b2ac-873e1785ae1f · outbound

This paper cites Efficient Streaming Language Models with Attention Sinks.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Efficient Streaming Language Models with Attention Sinks

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:01.693596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:01.693596Z digest=sha256:64bb8bf47b514cfd44dee202b506fff7af4973f1bf4db61607c3be6441d8c058

Observation d845edd9-fd66-4d33-9543-b7d481dec836 · outbound

This paper cites PyramidDrop: Accelerating Your Large Vision-Language Models via Pyramid Visual Redundancy Reduction.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design PyramidDrop: Accelerating Your Large Vision-Language Models via Pyramid Visual Redundancy Reduction

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:01.834128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:01.834128Z digest=sha256:e766f38d004cc3e7835921e1a855367786f3cfe24b2d9ce0394134f476af904a

Observation b2be132b-8cf0-4626-8968-7c65d98a184e · outbound

This paper cites Towards surveillance video-and-language understanding: New dataset, baselines, and challenges,.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Towards surveillance video-and-language understanding: New dataset, baselines, and challenges,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:09:04.292742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:09:01.923406Z digest=sha256:996bbdc442832a5d0422f87b06639b2f0d71bdc44481f2481450ba37b7c453ba

Observation a1a68158-024f-4428-8b46-b41501f1b8a0 · outbound

This paper cites LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:02.207599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:02.207599Z digest=sha256:490773169414c1df4b72429f9d53d13d572cb959d967f17cee7e8d8f83a796ce

Observation 0446f2e1-6718-42ef-9695-d2069951736d · outbound

This paper cites H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:02.326672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:02.326672Z digest=sha256:1a2d990bd43b7b44e3fdd23f89beb8a9fbb1848e5445afa30a5eea587899e8ce

Observation 6c627c0b-e94d-41a4-9e07-decad9278543 · outbound

This paper cites MLVU: Benchmarking Multi-task Long Video Understanding.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design MLVU: Benchmarking Multi-task Long Video Understanding

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:02.433510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:02.433510Z digest=sha256:b58e4bd33a615de19481952546beff0c67d73fec042e997076ad3fa341b95677

Observation aa11f0f8-8e95-4899-95da-55d7c9a06ec2 · outbound

This paper cites A Survey on Generative AI and LLM for Video Generation, Understanding, and Streaming.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design A Survey on Generative AI and LLM for Video Generation, Understanding, and Streaming

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:02.584217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:02.584217Z digest=sha256:15d9ebc3b0f207c483a81e96a4e23c9e749a7052bcd472c9b1b7de37d7703e92

Observation c3b24cbf-3b98-4fd6-9d57-9e382a28dd35 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:02.863981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:02.863981Z digest=sha256:48cca55fbfec3f728ecc59ab4f3248ab77b5d21df74cc6907d29bd10e681d968

Observation b6fc6b63-aef3-41a6-933a-504c6ef83247 · outbound

This paper cites Towards Surveillance Video-and-Language Understanding: New Dataset, Baselines, and Challenges.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Towards Surveillance Video-and-Language Understanding: New Dataset, Baselines, and Challenges

Reference 2023

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:09:03.187420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:09:02.065778Z digest=sha256:6a9d28cdf91359246198b116a7001b052cd5de8963d6c90d1a3df482baa5f187

Observation ec2dfef2-143e-46ba-a653-f8d36c27625e · outbound

This paper cites VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:58.536816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:58.536816Z digest=sha256:c7ea8e1687a2f0686d06b564dc63d673f2b85a45e8378419c74bb3f964611ab2

Observation cf2bdaea-b699-480c-8978-b137f78de9a6 · outbound

This paper cites Qwen2.5-VL Technical Report.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Qwen2.5-VL Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:57.859255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:57.859255Z digest=sha256:1a954f9c091b511cce445f469ddb565fbb6e5284b94192895beb70bbd053d2d4

Pith citing papers

Observation c3255245-49d4-4b80-a211-f65760303b0c · inbound

CoVR-R:Reason-Aware Composed Video Retrieval cites this paper.

CoVR-R:Reason-Aware Composed Video Retrieval QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-13T21:37:55.887477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T21:37:55.887477Z digest=sha256:955f53b197af59b8458fcf29e27f909fdf2dec6acab095b5e36d6f41ef9274bf

Observation f928d3bc-10f3-4e00-a3ea-9b161ef66450 · inbound

cuRAMSES: Scalable AMR Optimizations for Large-Scale Cosmological Simulations cites this paper.

cuRAMSES: Scalable AMR Optimizations for Large-Scale Cosmological Simulations QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design

Reference 63

Resolution
unresolved
no resolver link, observed 2026-07-13T09:09:32.932051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T09:09:32.932051Z digest=sha256:658e3521164081805ca3373a8ee380405996293e4fdeeaa0c8462da650a63088

Observation 8bdc8817-a71d-4a51-8316-bb6c919749e2 · inbound

CodecSight: Leveraging Video Codec Signals for Efficient Streaming VLM Inference cites this paper.

CodecSight: Leveraging Video Codec Signals for Efficient Streaming VLM Inference QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:00:52.799503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:45:26.555395Z digest=sha256:b06949b9110fedb391e3dba3fb0d3073775ea3c24fbcf97f8bb7b42b8ad89b0e

Observation d625753e-86c1-4920-8ef6-d5fcf85243e5 · inbound

Mosaic: Cross-Modal Clustering for Efficient Video Understanding cites this paper.

Mosaic: Cross-Modal Clustering for Efficient Video Understanding QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-10T16:10:34.374605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T16:07:36.133404Z digest=sha256:cb4b50cbe7eb7b0b9bc9cdf9cfcc489a37d79f1042128c86cdf5f62e3f053cb0

Observation 3df7cf18-54bf-490a-bce5-0e112bc67269 · inbound

VLMaxxing through FrameMogging Training-Free Anti-Recomputation for Video Vision-Language Models cites this paper.

VLMaxxing through FrameMogging Training-Free Anti-Recomputation for Video Vision-Language Models QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:11:13.760186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T01:30:15.463051Z digest=sha256:5c02584d51bfd2abaae1f984c2dc9f2bcbc0bfc6995afdcc7232147ded93e246

Observation 42def0b6-30ec-4fda-a44a-e3722280bbdc · inbound

Watch, Remember, Reason: Human-View Video Understanding with MLLMs cites this paper.

Watch, Remember, Reason: Human-View Video Understanding with MLLMs QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design

Reference 173

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:27:15.530520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T22:00:28.350003Z digest=sha256:365b4b36ec54299909b3c6c036977b2378f09956cbe29285c09a2cf9c18f3dbb