Pith. sign in

Paper Citation Record · LEDGER

Free Video-LLM: Prompt-guided Visual Perception for Efficient Training-free Video LLMs

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2410.10441.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.10441 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T05:44:05.841307Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T04:27:36.972367Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cd6bc861-253c-4f62-a883-eb7f56059d91 · inbound

Sparrow: Data-Efficient Video-LLM with Text-to-Image Augmentation cites this paper.

Sparrow: Data-Efficient Video-LLM with Text-to-Image Augmentation Free Video-LLM: Prompt-guided Visual Perception for Efficient Training-free Video LLMs

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:05.841307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:05.841307Z digest=sha256:c4c344531198e55bb0b14ef07cf854bb1e2d4fd9c31fa3ad7871ca0d045737d4

Observation 35fde9f8-37f7-4c12-a864-5052672b7b74 · inbound

p-MoD: Building Mixture-of-Depths MLLMs via Progressive Ratio Decay cites this paper.

p-MoD: Building Mixture-of-Depths MLLMs via Progressive Ratio Decay Free Video-LLM: Prompt-guided Visual Perception for Efficient Training-free Video LLMs

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T21:31:44.361989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:31:44.361989Z digest=sha256:a3905bb1a46ffc61c04d9087867e650f60f810b54b7ef43eee3ebcaafbed8569

Observation 949f35d5-cce9-4119-ab27-3a397753b64e · inbound

Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos cites this paper.

Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos Free Video-LLM: Prompt-guided Visual Perception for Efficient Training-free Video LLMs

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:39:22.560732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-16T11:39:22.340737Z digest=sha256:8ef3971714cad072be00bbb1132a89e8ffad6674f74898da2e12408a84d3b3b7

Observation 9848067d-40d2-44bd-a3a3-aac8c40b6c7f · inbound

Mixed-R1: Unified Reward Perspective For Reasoning Capability in Multimodal Large Language Models cites this paper.

Mixed-R1: Unified Reward Perspective For Reasoning Capability in Multimodal Large Language Models Free Video-LLM: Prompt-guided Visual Perception for Efficient Training-free Video LLMs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:37:44.473998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:37:44.473998Z digest=sha256:0f2031064e5ab9f2d0891736b79bec0f0563d8f2be6346a061ec97e882f528a5

Observation 52664408-8232-4b51-b2a6-1494f9cc6290 · inbound

CyberV: Cybernetics for Test-time Scaling in Video Understanding cites this paper.

CyberV: Cybernetics for Test-time Scaling in Video Understanding Free Video-LLM: Prompt-guided Visual Perception for Efficient Training-free Video LLMs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:26:47.096138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:26:47.096138Z digest=sha256:e219519b762aecd63552132717c196033d670397d9ccb6e53b597753062d424f

Observation 25103cae-a729-4d1b-8808-c80a0cd947d2 · inbound

Dense360: Dense Understanding from Omnidirectional Panoramas cites this paper.

Dense360: Dense Understanding from Omnidirectional Panoramas Free Video-LLM: Prompt-guided Visual Perception for Efficient Training-free Video LLMs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:23:48.496547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:23:48.496547Z digest=sha256:0ff1055ade5e8e2fd7535cf69c62b747c82b502fae4b338637fb5f2e6d943af1

Observation 8045657a-32af-4ee3-8888-4b699826fde3 · inbound

DenseWorld-1M: Towards Detailed Dense Grounded Caption in the Real World cites this paper.

DenseWorld-1M: Towards Detailed Dense Grounded Caption in the Real World Free Video-LLM: Prompt-guided Visual Perception for Efficient Training-free Video LLMs

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T21:27:39.419795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:27:39.419795Z digest=sha256:7452b04073abf5ebc9c46dccb70f4cf27256bdd9d2559be4c4315e04554bf178

Observation d2107171-335e-420e-bf51-ec1c4782e98b · inbound

A Glimpse to Compress: Dynamic Visual Token Pruning for Large Vision-Language Models cites this paper.

A Glimpse to Compress: Dynamic Visual Token Pruning for Large Vision-Language Models Free Video-LLM: Prompt-guided Visual Perception for Efficient Training-free Video LLMs

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T05:37:35.440786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T05:37:35.440786Z digest=sha256:120aa13fdbcb525d5ff875d85cdc9a8061f27d212eb0cb6247a109b7b9e97bb3

Observation 9556da37-baf6-4c23-ab30-eaeffddf2b41 · inbound

Kwai Keye-VL-2.0 Technical Report cites this paper.

Kwai Keye-VL-2.0 Technical Report Free Video-LLM: Prompt-guided Visual Perception for Efficient Training-free Video LLMs

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T04:27:36.974043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-27T13:53:10.352603Z digest=sha256:f068276760739016b4bbfb65a4c9465ba925f7caac774a6deb92449e81b07078