Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2304.14407.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T20:06:35.185621Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-23T22:53:33.136794Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 19efc1b1-4ac3-49c7-892a-bae99dec71b5 · inbound
What to Say and When to Say it: Live Fitness Coaching as a Testbed for Situated Interaction ChatVideo: A Tracklet-centric Multimodal and Versatile Video Understanding System
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ec2a373d-f810-4828-a02f-1d37a7f2e28d · inbound
VideoRoPE: What Makes for Good Video Rotary Position Embedding? ChatVideo: A Tracklet-centric Multimodal and Versatile Video Understanding System
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd22ef6a-5f86-4416-ad29-b5ba31469af1 · inbound
CoS: Chain-of-Shot Prompting for Long Video Understanding ChatVideo: A Tracklet-centric Multimodal and Versatile Video Understanding System
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06d262f9-5160-4354-bb14-4ce73c8dbb54 · inbound
Video-CoT: A Comprehensive Dataset for Spatiotemporal Understanding of Videos Based on Chain-of-Thought ChatVideo: A Tracklet-centric Multimodal and Versatile Video Understanding System
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0dd44e4c-e53b-44f3-be17-484d8fd9ff9b · inbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering ChatVideo: A Tracklet-centric Multimodal and Versatile Video Understanding System
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac6a474e-5933-453f-9972-1f4d49aeb7e2 · inbound
Empowering Multimodal LLMs with External Tools: A Comprehensive Survey ChatVideo: A Tracklet-centric Multimodal and Versatile Video Understanding System
Reference 269
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9e2d0ce-b931-4d3f-84e6-2435b6d1b4d1 · inbound
SurgLLM: A Versatile Large Multimodal Model with Spatial Focus and Temporal Awareness for Surgical Video Understanding ChatVideo: A Tracklet-centric Multimodal and Versatile Video Understanding System
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5763ab29-2d37-442c-86c7-bfe7cc13504d · inbound
Token Reduction via Local and Global Contexts Optimization for Efficient Video Large Language Models ChatVideo: A Tracklet-centric Multimodal and Versatile Video Understanding System
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1f8f433f-4812-4bc4-8563-c4b538c5550d · inbound
From Verbatim to Gist: Distilling Pyramidal Multimodal Memory via Semantic Information Bottleneck for Long-Horizon Video Agents ChatVideo: A Tracklet-centric Multimodal and Versatile Video Understanding System
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 46f6a05a-b353-4bc2-8862-e30fd90d321a · inbound
Semantic-Aware Adaptive Visual Memory for Streaming Video Understanding ChatVideo: A Tracklet-centric Multimodal and Versatile Video Understanding System
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.