Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2304.04227.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T18:39:28.378789Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-20T13:18:18.384182Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 170d1152-6a5f-484d-816b-7688094b65bc · inbound
MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models Video ChatCaptioner: Towards Enriched Spatiotemporal Descriptions
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ea6f3787-07f1-460f-93d8-7f411c2c5db1 · inbound
MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning Video ChatCaptioner: Towards Enriched Spatiotemporal Descriptions
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7df9b4b3-0afd-40bb-b2d7-644e30574c97 · inbound
NoteIt: A System Converting Instructional Videos to Interactable Notes Through Multimodal Video Understanding Video ChatCaptioner: Towards Enriched Spatiotemporal Descriptions
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8af754fa-db72-4002-893f-3063f07f9f40 · inbound
Soap2Soap: Long Cinematic Video Remaking via Multi-Agent Collaboration Video ChatCaptioner: Towards Enriched Spatiotemporal Descriptions
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.