Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2312.08870.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:36:40.834090Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-23T06:02:37.425276Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation eca306c0-c0e9-4e29-9fb6-332af2e57044 · inbound
PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning Vista-LLaMA: Reducing Hallucination in Video Language Models via Equal Distance to Visual Tokens
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a8563319-ad34-476c-bb1e-81520d738250 · inbound
VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs Vista-LLaMA: Reducing Hallucination in Video Language Models via Equal Distance to Visual Tokens
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 60027b7f-0915-4a79-b308-14f27afd7e3d · inbound
LLaVA-Octopus: Unlocking Instruction-Driven Adaptive Projector Fusion for Video Understanding Vista-LLaMA: Reducing Hallucination in Video Language Models via Equal Distance to Visual Tokens
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e16ea5c9-ee3c-424d-bcab-23f5c894bfdb · inbound
FaVChat: Hierarchical Prompt-Query Guided Facial Video Understanding with Data-Efficient GRPO Vista-LLaMA: Reducing Hallucination in Video Language Models via Equal Distance to Visual Tokens
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8e190d31-4252-4533-9a0a-c3cbdab840a8 · inbound
Is Extending Modality The Right Path Towards Omni-Modality? Vista-LLaMA: Reducing Hallucination in Video Language Models via Equal Distance to Visual Tokens
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73b8ef74-1160-4320-a0b4-1d300b3dc3c7 · inbound
Mitigating Behavioral Hallucination in Multimodal Large Language Models for Sequential Images Vista-LLaMA: Reducing Hallucination in Video Language Models via Equal Distance to Visual Tokens
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3da03465-d944-41e8-853e-75afe820e43f · inbound
IPFormer-VideoLLM: Enhancing Multi-modal Video Understanding for Multi-shot Scenes Vista-LLaMA: Reducing Hallucination in Video Language Models via Equal Distance to Visual Tokens
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fef4cecf-0169-4367-8e2e-009733dc94b6 · inbound
AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding Vista-LLaMA: Reducing Hallucination in Video Language Models via Equal Distance to Visual Tokens
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22f7bdd3-7318-42a1-9806-3829e0bcfd82 · inbound
DynImg: Key Frames with Visual Prompts are Good Representation for Multi-Modal Video Understanding Vista-LLaMA: Reducing Hallucination in Video Language Models via Equal Distance to Visual Tokens
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b63e8b36-84b4-41f8-8858-b028fb630932 · inbound
Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey Vista-LLaMA: Reducing Hallucination in Video Language Models via Equal Distance to Visual Tokens
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.