Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T16:58:52.999563Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 4 inbound Pith citation observations for arXiv:2508.18634.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T16:58:52.999563Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-03T16:37:06.384435Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T16:38:39.639869Z
18 of 18 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b97a79e9-fe98-433b-b9cf-4ebefee7d1a0 · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward InternLM2 Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3298519f-8116-47e2-8a3c-fb51cd469b27 · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward Video-R1: Reinforcing Video Reasoning in MLLMs
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2b2e4fd-e6ba-4b96-a960-1b1af6077640 · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward HunyuanVideo: A Systematic Framework For Large Video Generative Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c663f4be-611a-43b8-91d6-94d3891572d7 · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f791870f-4b2e-47f3-afa1-70dbaf1f3a1a · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward VILA: On Pre-training for Visual Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43a16375-456b-421b-87b2-f4b40d938276 · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward Visual-RFT: Visual Reinforcement Fine-Tuning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c65558d-f0b6-46a8-9c1d-539e9579eef0 · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward VideoCap-R1: Enhancing MLLMs for Video Captioning via Structured Thinking
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f02bf7b2-069e-42b9-8cfe-efc42760be66 · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward Proximal Policy Optimization Algorithms
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27db9b83-f703-45f2-9d58-ff3b4941cbcb · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37757e5e-2466-411a-9b05-09d7f95e4a25 · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward Tarsier: Recipes for Training and Evaluating Large Video Description Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd7e4223-1481-439d-ac3b-8115644aff46 · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward CaReBench: A Fine-Grained Benchmark for Video Captioning and Retrieval
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec9cfe10-3809-44bc-b74a-418a1a71f429 · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward Tarsier2: Advancing Large Vision-Language Models from Detailed Video Description to Comprehensive Video Understanding
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74ee5bd9-610d-4e0f-9cb9-a8505077d0b0 · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward LLaVA-Video: Video Instruction Tuning With Synthetic Data
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b7af4c0-29d8-451d-9805-9121ce189f10 · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward LLaVA-OneVision: Easy Visual Task Transfer
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc6828fc-2087-4544-b2c8-cc2177230532 · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward VideoCoCa: Video-Text Modeling with Zero-Shot Transfer from Contrastive Captioners
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9da50164-5600-40e3-a7df-fe7fd3211cd2 · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward VideoChat: Chat-Centric Video Understanding
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 375f232f-a76c-4df4-a412-9bb4755dab52 · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward In SIGGRAPH Asia 2024 Conference Papers, 1–11
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3879f5be-25da-48dd-9d08-45726cd1717b · outbound
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward Qwen2.5-VL Technical Report
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 642faeba-3ec0-4052-a2a4-953c14d80d2d · inbound
Building a Precise Video Language with Human-AI Oversight OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e40fc9d6-b49a-477b-bb01-c6ebda7129c6 · inbound
VCap: Hypergeometric Rewards for Weak-to-Strong Visual Captioning OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a29d4349-76e2-4025-9e38-e08a8d8b8536 · inbound
Watch, Remember, Reason: Human-View Video Understanding with MLLMs OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 36344ea0-3cc8-4847-9cb4-71aec4a2f91c · inbound
Temporal and Cross-Modal Alignment for Enhanced Audiovisual Video Captioning OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.