Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T20:16:34.534847Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2608.04302.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T20:16:34.534847Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
44 of 44 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 90e525e9-c911-4328-8b70-a47174a6dd66 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e682979-f303-4655-b98c-24f72a54a037 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f2bfdb1-5068-4b29-be5b-870da079cc2d · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Chen and William B
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a7d1e4bd-17d0-4489-930c-6456b1bfe081 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2970c550-7d5a-40d3-af60-fceb9b8a81b4 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1582c9de-4f47-4d4a-8780-68c4ea6c95a6 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a3f21455-f719-443b-b591-f809e8fcb211 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 753642d1-50af-464c-bab4-1dbbf9911e8d · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4e8b7751-3cab-4e9c-99a7-fcdc69b33bf6 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8a56f043-1a1d-424c-b117-8fd3bc998e03 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models KaLM-Embedding: Superior Training Data Brings A Stronger Embedding Model
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62aedf90-4082-49ee-a8cf-ed397a9707f1 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 019bf947-3c2d-4f9a-a848-4d6e6bd4dfc0 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d71e7e95-8a7d-4cc2-b0a1-5d27b15878c0 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models NV-Embed: Improved Techniques for Training LLMs as Generalist Embedding Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3831fa8c-bf59-40f4-8171-172052cb4d29 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7bad7fbf-c2d1-4e61-9f9f-4fdd4728a3ee · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models VideoChat: Chat-Centric Video Understanding
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a617dda4-f763-4c75-84bd-f36af7609cbb · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models MVBench: A Comprehensive Multi-modal Video Understanding Benchmark
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2dd8dab-8c8d-41b1-a819-aae22ea875c6 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Towards General Text Embeddings with Multi-stage Contrastive Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 889c4ed5-d568-4b95-9361-65fcc4896386 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b467bf61-ccd9-4622-bc0d-5a605a726cc7 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1a17fd2-de85-469d-80e4-623e86192b12 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models MTEB: Massive Text Embedding Benchmark
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 507ddfc6-018c-4dad-a926-47cd89f360e5 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8e850048-92e3-4996-93b0-6887b42a6d0d · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b69efeb4-e38f-409a-9cd4-4175dc29d9cc · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a96b9b52-63ca-4bae-a291-4e8dc2f7bad5 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8c17c894-8ce3-49cd-8f47-08d5f941c46d · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 771885ca-b742-45d0-a498-9af0ca2f7ee1 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 73fa8f2e-070a-432e-8175-c93a6b3a08b2 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Learning Video Representations using Contrastive Bidirectional Transformer
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f6d5388-38ec-4119-a6ee-04a42b624e94 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fa63ce4-1d95-4a3b-9e06-517f3a94d643 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d70a72fd-307e-41a1-a376-8e986d791e13 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models G-VEval: A Versatile Metric for Evaluating Image and Video Captions Using GPT-4o
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 689f95e0-4559-4c78-b7d5-3a11add0247e · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Lawrence Zitnick, and Devi Parikh
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2b055535-333b-46f3-842e-8ed682e17685 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e54ea5e1-88ee-4922-9c2b-4b77e1e5fece · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d072fc63-4471-401d-bf6e-e8000cacb45b · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b7356d2-bb7e-4cfa-a6dd-1917d7cbd8d4 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8f2f3108-e4f0-452b-b364-3f7e77b4f820 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e000a909-aa83-4292-8add-ad6365a077d3 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce66fc99-86aa-49a4-b4f6-8ba6306c0928 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 021e247e-457f-43cd-9895-70703f97fa2f · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4ba473f-b6a0-47c2-ae54-75b5cb21f295 · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Weinberger, and Yoav Artzi
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3bb8d43-6a44-48f6-97ce-44bf046ab13d · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models The Three Amigos
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation befd7d9a-3877-4b66-8cf0-7bd50b3d465f · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models VideoBERT: A Joint Model for Video and Language Representation Learning
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c6cc8b1-86bc-4485-b3ff-4f98c9eb1f1a · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models InInternational Con- ference on Learning Representations
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2377a7b3-aeab-4418-b237-7c6ff0290edb · outbound
CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models Unresolved cited work
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
No inbound Pith citation observations are available.