Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2407.08303.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T11:41:44.453339Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T00:17:29.025254Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 6f81342e-bd97-4abe-baba-5a01192baf65 · inbound
Emu3: Next-Token Prediction is All You Need DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a8f2966c-ba4d-41d2-9f6c-40cda4fe35e9 · inbound
Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b65a97ec-b51e-406b-b27d-24f6728f9732 · inbound
DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1c76186e-249a-4591-8610-c7f7bf3feff7 · inbound
VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 43b09f0f-cfdd-4c8f-9c43-1143b7357f62 · inbound
COCONut-PanCap: Joint Panoptic Segmentation and Grounded Captions for Fine-Grained Understanding and Generation DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5b52559-e7e7-42ee-8c3b-0be5be797755 · inbound
EVEv2: Improved Baselines for Encoder-Free Vision-Language Models DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f93e43ac-949c-4386-9124-a7ab07fdf843 · inbound
Show-o2: Improved Native Unified Multimodal Models DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e3aa9530-bfdd-4844-9640-dd64e7f01dc9 · inbound
GenRecal: Generation after Recalibration from Large to Small Vision-Language Models DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea64e559-00bb-49fa-9315-8e5b1445afbb · inbound
OmniGen2: Towards Instruction-Aligned Multimodal Generation DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 75fc37fc-9c21-4d22-829f-a0592d761a66 · inbound
Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 736aeb7c-7f49-471f-88a5-e1ac314cf19b · inbound
CapRL++: Unified Reinforcement Learning with Verifiable Rewards for Dense Image and Video Captioning DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.