Pith. sign in

Paper Citation Record · LEDGER

TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2406.08656.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.08656 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T16:22:44.163417Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:49:31.003751Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 61dc7b8c-49a0-4425-808b-f3f33c9bedb5 · inbound

VBench++: Comprehensive and Versatile Benchmark Suite for Video Generative Models cites this paper.

VBench++: Comprehensive and Versatile Benchmark Suite for Video Generative Models TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 116

Resolution
unresolved
no resolver link, observed 2026-08-12T16:22:44.163417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:22:44.163417Z digest=sha256:77bb8e8d89e3d0a56c290ee800e5c027562330148039125fb1676a82f37b9230

Observation 281c7401-6ffb-457a-b9d7-48afa78b3874 · inbound

Neuro-Symbolic Evaluation of Text-to-Video Models using Formal Verification cites this paper.

Neuro-Symbolic Evaluation of Text-to-Video Models using Formal Verification TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T14:26:21.916622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:26:21.916622Z digest=sha256:43616d71c865ee1628c11108064148fb71656bfaa3dad02bc6b44dad3b9903a3

Observation 91ef17fe-07c7-4694-8578-06be7ef9ddfa · inbound

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models cites this paper.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.583658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.583658Z digest=sha256:1c892721cd9ff8560796b431af6e38e990f36807cc55746a41f790cd217b7e75

Observation c1346412-4a35-4d28-8834-171cfd93aad5 · inbound

Is Your World Simulator a Good Story Presenter? A Consecutive Events-Based Benchmark for Future Long Video Generation cites this paper.

Is Your World Simulator a Good Story Presenter? A Consecutive Events-Based Benchmark for Future Long Video Generation TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T13:15:28.331739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:15:28.331739Z digest=sha256:b0a2a1a630f6028881baddc6ee48bcf1ec2b924f28e291f02716f56bfabe6dfb

Observation 491fb305-9892-4bae-a1d2-045effcab649 · inbound

BlobGEN-Vid: Compositional Text-to-Video Generation with Blob Video Representations cites this paper.

BlobGEN-Vid: Compositional Text-to-Video Generation with Blob Video Representations TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:46.702000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:46.702000Z digest=sha256:d7afa56df5b559cf9c989007b0d527c1ff0b5b2ea65a694fbe18c1df8d935b27

Observation 24e2b302-5f59-4efa-9066-1467aaa50ebf · inbound

We'll Fix it in Post: Improving Text-to-Video Generation with Neuro-Symbolic Feedback cites this paper.

We'll Fix it in Post: Improving Text-to-Video Generation with Neuro-Symbolic Feedback TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-22T18:56:58.222912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-22T18:56:39.735334Z digest=sha256:6b47e48e3c727d91212b57389d691f279f5097cc0e39bd2de80b437e3f2e0d77

Observation f9d3a925-7cf2-41e0-8c54-ce28d147afb5 · inbound

ARGUS: Hallucination and Omission Evaluation in Video-LLMs cites this paper.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.629628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.629628Z digest=sha256:260c797859526709ad3a2c9b04fa6952bcf8a268c4f452dfb6993165d5a049c3

Observation ab23c3cf-1d1e-496d-9759-c525e4513816 · inbound

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations cites this paper.

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:46.458920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:46.458920Z digest=sha256:fee21ce957888c6ecc9a207c884e73d17f0ca6335833642ba51faca430fb29e9

Observation f2308e59-9a08-42ce-a73a-64f2db7c1330 · inbound

T2VWorldBench: A Benchmark for Evaluating World Knowledge in Text-to-Video Generation cites this paper.

T2VWorldBench: A Benchmark for Evaluating World Knowledge in Text-to-Video Generation TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:26.566236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:26.566236Z digest=sha256:f71a307db8da8bf0aaade3d68c2834a535fe82d957e9202f08119d3318d589f8

Observation 070e9f70-c0f7-4fc6-813a-ad34bd17ea0e · inbound

NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding cites this paper.

NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T13:01:58.045344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:01:58.045344Z digest=sha256:65307f0e279b9d7a9e2d548e4dca58b702766fbf7c6d8b738a6d7b3e7caeabf3

Observation f7706508-171a-4b27-b1a5-33634d400074 · inbound

NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding cites this paper.

NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T06:36:33.811054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:36:33.811054Z digest=sha256:49da23e530c2a81f263ac39daafb35137bd38f61de2e038dd712c877141895b3

Observation eed2fe64-e570-4751-8fc1-4a7c37f37f1a · inbound

RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation cites this paper.

RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T03:03:37.455762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T03:02:26.084185Z digest=sha256:1ef159503a50baeca8a8f6dda6b41769e3e931d1a8120aff65aba836e412c80e

Observation 319912a3-7652-4639-ac05-94c56e90beeb · inbound

RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation cites this paper.

RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:19:50.198674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-15T06:15:32.881140Z digest=sha256:3c34acf6b5ba5a4c418655ef6e3ab1e2a38064da71d49e2e408011e8426cb6ad

Observation d9c6c93f-4d63-40ca-a3a5-deb230cecd16 · inbound

CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models cites this paper.

CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:40:23.130772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-25T04:39:22.400458Z digest=sha256:fb0b831cb59b1ef9e56be5cfe0a0044bc191662867850a9a508d257f577a1cdc

Observation 42c44632-6e7c-4385-a01e-cb5d6b8075c6 · inbound

LongAV-Compass: Towards Unified Evaluation of Minute-Scale Audio-Visual Generation Across T2AV, I2AV, and V2AV cites this paper.

LongAV-Compass: Towards Unified Evaluation of Minute-Scale Audio-Visual Generation Across T2AV, I2AV, and V2AV TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:00.793852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T22:52:38.330851Z digest=sha256:7b69bd3ea0400e385d032e35f5ea8f784680daa9e4cf9c81f3ba9c6e3a131494

Observation b2015940-614c-4655-a5ad-580580669802 · inbound

Towards Interactive Video World Modeling: Frontiers, Challenges, Benchmarks, and Future Trends cites this paper.

Towards Interactive Video World Modeling: Frontiers, Challenges, Benchmarks, and Future Trends TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 134

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T21:06:14.094042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-28T17:29:18.513507Z digest=sha256:208b9f95421d6a29de0c3b14fef51018f72191d719a30f5724e490f2e340035c

Observation 01349f55-1a4f-44c8-86dd-85674ac5176b · inbound

Can Image Models Imagine Time? ImageTime: A Novel Benchmark for Probing Visual World Modeling Through Spatiotemporal Consistency cites this paper.

Can Image Models Imagine Time? ImageTime: A Novel Benchmark for Probing Visual World Modeling Through Spatiotemporal Consistency TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-03T04:57:38.101059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-27T13:34:25.079037Z digest=sha256:52cc7358bd237c6551b279587d1634697f88dffc904f8b204589ffeb3f7add4a

Observation 1287a9a5-e700-4982-99b2-cbe2f21ee39a · inbound

Current World Models Lack a Persistent State Core cites this paper.

Current World Models Lack a Persistent State Core TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:49:31.006734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-26T17:33:41.461245Z digest=sha256:cb62cb8daa6535892e8109560a8c1107af25c765b36e27dbd5f38445a4b95038

Observation e92bca83-aa2c-4ec9-8b0f-80b4951b5ec4 · inbound

From World Action Models to Embodied Brains: A Roadmap for Open-World Physical Intelligence cites this paper.

From World Action Models to Embodied Brains: A Roadmap for Open-World Physical Intelligence TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 115

Resolution
unresolved
no resolver link, observed 2026-07-14T03:51:24.547781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T03:51:24.547781Z digest=sha256:0d0d9ad2b55f4df0d6fa81515b74716bd21c00a396b7dd5f970b81f089537922

Observation 763678e9-f1d8-4bc6-81bc-4e197037d136 · inbound

VGIF-Score: Interpretable and Diagnostic Evaluation of Spatio-Temporal Instruction Following in Video Generation cites this paper.

VGIF-Score: Interpretable and Diagnostic Evaluation of Spatio-Temporal Instruction Following in Video Generation TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T04:57:34.021666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:57:34.021666Z digest=sha256:efeccc6cce19fad52a9c81cbafaa1addc347dc079ae53aa00324149b80cca25c