Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2401.06071.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:32:52.287195Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T15:09:55.080073Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 4644ef79-c8e4-4276-aac2-14c94566a4e0 · inbound
DisTime: Distribution-based Time Representation for Video Large Language Models GroundingGPT:Language Enhanced Multi-modal Grounding Model
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71987732-166b-49b7-b8f0-00bddd142d99 · inbound
AV-Reasoner: Improving and Benchmarking Clue-Grounded Audio-Visual Counting for MLLMs GroundingGPT:Language Enhanced Multi-modal Grounding Model
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ca8b1aa-90ba-405f-be7c-d9c6e6b4a20a · inbound
Improving the Reasoning of Multi-Image Grounding in MLLMs via Reinforcement Learning GroundingGPT:Language Enhanced Multi-modal Grounding Model
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7b59254b-d862-46ce-b065-c05fb0b57cd9 · inbound
Grounded Gesture Generation: Language, Motion, and Space GroundingGPT:Language Enhanced Multi-modal Grounding Model
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9871138c-18d8-4817-bdbf-69dc38409ff6 · inbound
ReMeREC: Relation-aware and Multi-entity Referring Expression Comprehension GroundingGPT:Language Enhanced Multi-modal Grounding Model
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6149dfaf-44e7-4ebe-91ac-a649344baf6b · inbound
IntentVCNet: Bridging Spatio-Temporal Gaps for Intention-Oriented Controllable Video Captioning GroundingGPT:Language Enhanced Multi-modal Grounding Model
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a3c96a9-3833-424e-9c10-bbe1ecb46b2d · inbound
KnowDR-REC: A Benchmark for Referring Expression Comprehension with Real-World Knowledge GroundingGPT:Language Enhanced Multi-modal Grounding Model
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61d11a63-b771-45c9-946d-e869eb9d780e · inbound
Video Understanding by Design: How Datasets Shape Video Models GroundingGPT:Language Enhanced Multi-modal Grounding Model
Reference 249
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b8c1b47-2992-4ede-99ee-0329898dec86 · inbound
Encode Once, Decode Never: Reusing Audio LM Internals for Efficient Temporal Localization GroundingGPT:Language Enhanced Multi-modal Grounding Model
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5191aed-f75e-4845-afc1-82b81ca4610d · inbound
From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models GroundingGPT:Language Enhanced Multi-modal Grounding Model
Reference 157
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8b22c4a7-72f1-463e-92b9-91554591db82 · inbound
DART: Difficulty-Adaptive Routing for Zero-Shot Video Temporal Grounding GroundingGPT:Language Enhanced Multi-modal Grounding Model
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5aa2aea9-8e4c-4430-bfd1-916b284a0fdc · inbound
Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Calibrated GRPO GroundingGPT:Language Enhanced Multi-modal Grounding Model
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.