Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 25 inbound Pith citation observations for arXiv:2403.15378.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T23:40:21.599948Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T16:29:57.890361Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 03a6e17e-2b29-4780-b90a-e008cf09c589 · inbound
InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 171
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e14f8811-1c17-4665-b344-4424d20d57a6 · inbound
E5-V: Universal Embeddings with Multimodal Large Language Models Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6cb0ee77-9ec0-4717-8e75-9b406f2599c7 · inbound
Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b51070d-d63d-4fc0-976a-77c714bd43a5 · inbound
VideoRoPE: What Makes for Good Video Rotary Position Embedding? Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d6f4e5d-57fc-41b3-9b1f-fd9d40e586a1 · inbound
Distill CLIP (DCLIP): Enhancing Image-Text Retrieval via Cross-Modal Transformer Distillation Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 766675df-dbfd-4115-975d-536f1a31855b · inbound
Multiverse Through Deepfakes: The MultiFakeVerse Dataset of Person-Centric Visual and Conceptual Manipulations Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a9871a1-57b7-40e4-bd16-eda9ed779b15 · inbound
ANT: Adaptive Neural Temporal-Aware Text-to-Motion Model Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38f709d8-3a45-402a-a049-1a51bc49f3f1 · inbound
GenEscape: Hierarchical Multi-Agent Generation of Escape Room Puzzles Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d91d90a1-acd1-4d8e-ad0a-ea4de23d1568 · inbound
FIX-CLIP: Dual-Branch Hierarchical Contrastive Learning via Synthetic Captions for Better Understanding of Long Text Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fecc64b-1d59-4c7c-9aa8-12589220b431 · inbound
Domain-Enhanced Dual-Branch Model for Efficient and Interpretable Accident Anticipation Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4160fdf5-77db-4682-8f0b-cf6b232dfb42 · inbound
PoemTale Diffusion: Minimising Information Loss in Poem to Image Generation with Multi-Stage Prompt Refinement Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4c08ade-a215-4f17-8092-44de9aca2019 · inbound
Zero-Effort Image-to-Music Generation: An Interpretable RAG-based VLM Approach Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 21865d75-2172-4a81-abe7-a9c3a11ce13c · inbound
LFS: Learnable Frame Selector for Event-Aware and Temporally Diverse Video Captioning Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b9a0f02a-d2d4-4809-99f8-99d96cc9c187 · inbound
AnyStyle: Single-Pass Multimodal Stylization for 3D Gaussian Splatting Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b837b95-8915-4b84-aca9-6f6f0cc107e9 · inbound
VideoNet: A Large-Scale Dataset for Domain-Specific Action Recognition Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 84f41b62-ca6e-407d-8f7b-90a806f33b21 · inbound
Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 62d63b4f-850f-4b2b-8183-983e4e591e26 · inbound
Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 873a7819-028c-47e2-a0c4-a30081ab958e · inbound
Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 65d9bc3f-5700-4c3f-847a-0ce1a24e5f20 · inbound
Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ac12bc4f-551e-49c0-9619-29fecc202c2d · inbound
Offline Semantic Guidance for Efficient Vision-Language-Action Policy Distillation Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 29d9322a-d926-4c8b-a69c-81a353b8bb01 · inbound
HPP: Hierarchical Programmatic Probing for Long Video Understanding by Decoupling Perception and Reasoning Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 103
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ab0c928c-5f9b-456d-95c7-e6a77702e2f0 · inbound
TuringViT: Making SOTA Vision Transformers Accessible to All Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f8971470-a521-43ab-9a07-60b75aa612e5 · inbound
TuringViT: Making SOTA Vision Transformers Accessible to All Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 016f124c-dda2-4a77-b01e-e8e688f0ee17 · inbound
JuZhou 1.0 Technical Report: The First Edge-Native Text-to-Image Foundation Model Trained Entirely on China-Developed AI Accelerators Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c9eee4a0-f03e-48e6-b829-770ae20f3b26 · inbound
JuZhou 1.0 Technical Report: The First Edge-Native Text-to-Image Foundation Model Trained Entirely on China-Developed AI Accelerators Long-CLIP: Unlocking the Long-Text Capability of CLIP
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.