Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T23:19:46.019472Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 2 inbound Pith citation observations for arXiv:2505.04946.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T23:19:46.019472Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:10:59.330611Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T00:07:28.151343Z
23 of 23 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ff374fb2-fcd8-4be1-90ab-cd42505ff5aa · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models Text-to-image diffusion models cannot count, and prompt refinement cannot help
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb76a354-c470-4734-9513-319da0526a52 · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models High-Order Matching for One-Step Shortcut Diffusion Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59f134ff-d6f4-42a4-89db-921c29027afc · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models Provable Failure of Language Models in Learning Majority Boolean Logic via Gradient Descent
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 631b5329-b302-454a-bf18-1270a1c42b94 · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56dee2c6-9eb8-416d-a47b-4aef9609a5cc · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models Can You Count to Nine? A Human Evaluation Benchmark for Counting Limits in Modern Text-to-Video Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ff221b9-cba9-4a39-abba-aad082528da1 · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models LTX-Video: Realtime Video Latent Diffusion
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1636a2b-8414-46e6-8da9-bf50daf6f16f · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models Wordart designer: User- driven artistic typography synthesis using large language models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b8b4c785-cf5f-408e-860d-7161fff3e218 · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models VBench++: Comprehensive and Versatile Benchmark Suite for Video Generative Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4929530-d18c-4c7c-a04e-2b7c74f51a4b · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models On the expressive power of modern hopfield networks
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b76c554-49a7-44b0-8173-f4ead2ec32f0 · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models Text-Animator: Controllable Visual Text Video Generation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3ce2210d-bba0-4c29-af49-5cb70a629055 · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models On the Computational Capability of Graph Neural Networks: A Circuit Complexity Bound Perspective
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1af7c76-f89a-49cd-bc2a-5480e0df4f2d · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models PhyBench: A Physical Commonsense Benchmark for Evaluating Text-to-Image Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b84d62b2-cfd6-4861-a432-255d234cea72 · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models BizGen: Advancing Article-level Visual Text Rendering for Infographics Generation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 047b4ec9-ce8a-4a93-b413-0b4091a44877 · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ad058c4-f0c3-4871-bb06-c41493397b9a · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models Dolfin: Diffusion Layout Transformers without Autoencoder
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec9180d6-b003-482b-b97d-6d1d5d894570 · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models Alignab: Pareto- optimal energy alignment for designing nature-like antibodies
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a42ca0a-b428-4bfb-988a-b485017e4892 · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models Typedance: Creating semantic typographic logos from image through personalized generation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 23400afb-78ac-427b-927c-309011a3091a · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c3cb00f-e799-4bde-b641-efcc3d4ea1fb · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models The dragon’s weakness?
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation df0ca87a-ac19-4919-863d-fc1e00ae2568 · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models HOFAR: High-Order Augmentation of Flow Autoregressive Transformers
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3be70a02-f180-4c9e-b8e9-322690d25b05 · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models Videophy: Eval- uating physical commonsense for video generation
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 30b3f5d9-030c-4d11-bd7f-8c9de14228f9 · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models Force Matching with Relativistic Constraints: A Physics-Inspired Approach to Stable and Efficient Generative Modeling
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 215aacc4-35ff-45d3-b38b-e8e278159f07 · outbound
T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b4ff729-256f-4dc6-98e9-7ba0a5ad3299 · inbound
Only Large Weights (And Not Skip Connections) Can Prevent the Perils of Rank Collapse T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 352aabec-db72-4833-8c91-ef0e4d5d7bd4 · inbound
CineDance: Towards Next-Generation Multi-Shot Long-Form Cinematic Audio-Video Generation T2VTextBench: A Human Evaluation Benchmark for Textual Control in Video Generation Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.