Pith. sign in

Paper Citation Record · LEDGER

Compositional Video Synthesis by Temporal Object-Centric Learning

As of 9 August 2026, this Paper Citation Record lists 73 of 73 outbound references and 0 inbound Pith citation observations for arXiv:2507.20855.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.20855 v1

Coverage vector

measured 73 of 73 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:16:44.217929Z

measured 73 of 73 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

73 of 73 outbound references displayed

  • verified exact0
  • verified fuzzy66
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4a53892a-e5ca-4463-a595-7d52443aef09 · outbound

This paper cites Slamp: Stochastic latent appearance and motion pre- diction.

Compositional Video Synthesis by Temporal Object-Centric Learning Slamp: Stochastic latent appearance and motion pre- diction

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.808932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:42.664655Z digest=sha256:40446fa726df6f3d5eb086be46397b6555f346200eda50de5b3756c07cd3287e

Observation eb8ca687-a7ee-4de6-87da-0115022032ca · outbound

This paper cites Stretchbev: Stretching future instance prediction spatially and temporally.

Compositional Video Synthesis by Temporal Object-Centric Learning Stretchbev: Stretching future instance prediction spatially and temporally

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.800561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:42.726015Z digest=sha256:15b6fe0f6037ad9021d7727d600a21fdf8fd9e9f065854b7e8c79fb21ec26ea3

Observation 1d62f899-9ec3-4a13-b283-b9a7c11f328d · outbound

This paper cites Slot-guided adaptation of pre-trained diffusion models for object-centric learning and compositional generation.

Compositional Video Synthesis by Temporal Object-Centric Learning Slot-guided adaptation of pre-trained diffusion models for object-centric learning and compositional generation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.792481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:42.800382Z digest=sha256:6a9491b88d0dcdb25b9d4b63952560eac189ea217552ced92fe16d5c941c1529

Observation b98a97f7-cbe1-49a6-a34e-55b60c85acfd · outbound

This paper cites Self- supervised Object-centric Learning for Videos.

Compositional Video Synthesis by Temporal Object-Centric Learning Self- supervised Object-centric Learning for Videos

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.784142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:42.882950Z digest=sha256:9f37786a2dc6c299aaf77583ba226efcc814fd21c2e2b1378e64d5fca82926db

Observation fa7374ae-00d6-49ec-bd27-1016364b0d8b · outbound

This paper cites Systematic generalization: What is required and can it be learned? In Proc.

Compositional Video Synthesis by Temporal Object-Centric Learning Systematic generalization: What is required and can it be learned? In Proc

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.776126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:42.996989Z digest=sha256:5976d83a6e37481df8be20b83789ef912145969e9bb5cf29f367de917b4b7312

Observation 95eb09a6-88ba-4d1d-b393-f1c94fef3206 · outbound

This paper cites Object discovery from motion- guided tokens.

Compositional Video Synthesis by Temporal Object-Centric Learning Object discovery from motion- guided tokens

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.768077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:43.070604Z digest=sha256:222e20f11dd0678c4072c740daf9c1609f627d9ecded4bdd3a5d5e8ec82aea90

Observation 7533f941-d8d2-41d5-915f-96c2da1f9353 · outbound

This paper cites Lumiere: A space-time diffusion model for video generation.

Compositional Video Synthesis by Temporal Object-Centric Learning Lumiere: A space-time diffusion model for video generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T13:16:43.142294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:16:43.142294Z digest=sha256:14afbc3ff84de37800d6be9e8f3da9ab8c754c16fe73923d93d5a74ad621c580

Observation 904ad9e4-e0ac-4776-8d6f-6d8f52199e08 · outbound

This paper cites Invariant slot attention: Object discovery with slot-centric reference frames.

Compositional Video Synthesis by Temporal Object-Centric Learning Invariant slot attention: Object discovery with slot-centric reference frames

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.754950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:43.213561Z digest=sha256:da1d5a0838b4c52a482d9713db8600188777b6fb3c08737a1808826a0df07301

Observation 564176c5-8881-49e7-bf47-90de864fcdac · outbound

This paper cites Align your latents: High-resolution video synthesis with latent diffusion models.

Compositional Video Synthesis by Temporal Object-Centric Learning Align your latents: High-resolution video synthesis with latent diffusion models

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.747508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:43.244544Z digest=sha256:12b38669368a1f5ff328e92dc096a738552c6d01c829212e1a2ae061916ef701

Observation 7a2ab687-3873-4bdf-a200-01ed3cfdfe65 · outbound

This paper cites Emerging properties in self-supervised vision transformers.

Compositional Video Synthesis by Temporal Object-Centric Learning Emerging properties in self-supervised vision transformers

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.740093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:43.317792Z digest=sha256:7d81e474aeeb1f65dc7802de80f99bf0387d6712bcc68cdc9c45691feb10b54c

Observation 707f5e7d-a837-45ad-96ec-57af4f0a52be · outbound

This paper cites Pixart-α: Fast training of diffusion trans- former for photorealistic text-to-image synthesis.

Compositional Video Synthesis by Temporal Object-Centric Learning Pixart-α: Fast training of diffusion trans- former for photorealistic text-to-image synthesis

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.732572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:43.429828Z digest=sha256:882c79bb83609b3751a88c240666afaff257c95119fc1b1030755700d82f28dd

Observation 59be87bd-7ec8-4363-b3bc-ad85fe3548b2 · outbound

This paper cites Vision transformers need registers.

Compositional Video Synthesis by Temporal Object-Centric Learning Vision transformers need registers

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.725128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:43.547864Z digest=sha256:5e5666c8bd8ed43bdc62811bd4eb2bcc55b5a2bb16c50f5c0895413160fdb00e

Observation 9d9630db-6022-4695-9b67-8a9adce56d21 · outbound

This paper cites Diffusion models beat GANs on image synthesis.

Compositional Video Synthesis by Temporal Object-Centric Learning Diffusion models beat GANs on image synthesis

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.717663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:43.664202Z digest=sha256:c74435c7f96214353ca5c3670939981e2c94890a33a188c6e3b6a5aaef2911f0

Observation 8c2f57a6-5277-4fe7-93d7-3b620ce15ed6 · outbound

This paper cites Betrayed by attention: A simple yet ef- fective approach for self-supervised video object segmenta- tion.

Compositional Video Synthesis by Temporal Object-Centric Learning Betrayed by attention: A simple yet ef- fective approach for self-supervised video object segmenta- tion

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.710132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:43.782017Z digest=sha256:2c527d42376793e24c07eaae8f6405ecfd56f498826ac27fd88a34c17511f4a3

Observation 74e803c0-7c62-43c1-b6b8-4245ad24607d · outbound

This paper cites SAVi++: Towards end-to-end object-centric learning from real-world videos.

Compositional Video Synthesis by Temporal Object-Centric Learning SAVi++: Towards end-to-end object-centric learning from real-world videos

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.702365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:43.921894Z digest=sha256:f6ea12b89cfb807a198c1c33f7684f6fc2005d633ccd0ff2a7662aaaeadd562c

Observation 5f54637c-c83b-4f9a-901e-cc3a1f26adbf · outbound

This paper cites Attend, infer, re- peat: Fast scene understanding with generative models.

Compositional Video Synthesis by Temporal Object-Centric Learning Attend, infer, re- peat: Fast scene understanding with generative models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.694512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:43.928307Z digest=sha256:325ddb1e9fe4142a93d1ac67226d08ce284a30fee90d314c826bab1f6ba6e8eb

Observation 29f9374a-5e7b-447c-a0ca-b79e2474428c · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

Compositional Video Synthesis by Temporal Object-Centric Learning Scaling rectified flow transformers for high-resolution image synthesis

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.687199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.034897Z digest=sha256:574af958a17dc69956ae0e26b5cdfe49cfb87b6a51e28f331ff89074e00f7eff

Observation e848e0dd-2fde-4e9c-9db3-7ee48f205314 · outbound

This paper cites The PASCAL visual object classes (VOC) challenge.

Compositional Video Synthesis by Temporal Object-Centric Learning The PASCAL visual object classes (VOC) challenge

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.679727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.050887Z digest=sha256:33bfd9ef554c9e8619248d160ae067548610fd56af8846fa2a35adfce89a8df7

Observation 2c509ebe-7e2b-4dbc-ad6d-22c102cae5ff · outbound

This paper cites Connectionism and cognitive architecture: A critical analysis.

Compositional Video Synthesis by Temporal Object-Centric Learning Connectionism and cognitive architecture: A critical analysis

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.672196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.057618Z digest=sha256:3592f2afadc3d563df6423ecec02a552f40508ea1909d336c51185f17a3de16b

Observation 534e61bc-8442-41ee-9984-7d04bdbb1073 · outbound

This paper cites Understanding the diffi- culty of training deep feedforward neural networks.

Compositional Video Synthesis by Temporal Object-Centric Learning Understanding the diffi- culty of training deep feedforward neural networks

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.664636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.085924Z digest=sha256:369bb9bc39d1577ddf983c66f9d2a61b85aebe04ceedfde238e7edec41cb6907

Observation 77e30ea2-d3d8-4895-a2cf-b9afd15ebbfc · outbound

This paper cites Multi-object representation learning with iterative variational inference.

Compositional Video Synthesis by Temporal Object-Centric Learning Multi-object representation learning with iterative variational inference

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.656961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.088511Z digest=sha256:b7be03e3da258957d8ab6bf5b44aa4201f856de83561eacddf0e205967a06c99

Observation 58c9c3fa-53e4-411f-a45c-a0f2267e8589 · outbound

This paper cites On the Binding Problem in Artificial Neural Networks.

Compositional Video Synthesis by Temporal Object-Centric Learning On the Binding Problem in Artificial Neural Networks

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T13:16:44.091002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:16:44.091002Z digest=sha256:832b14057c7da437d4e0421f11821c1a0cb0dadef8810db9c7127c90fd053574

Observation 1477d4d0-274e-45d9-b34f-952f1459db35 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equi- librium.

Compositional Video Synthesis by Temporal Object-Centric Learning Gans trained by a two time-scale update rule converge to a local nash equi- librium

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.648983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.093688Z digest=sha256:cc6bb348973d2b6447165a1fcd12dd317a5eabc164cad2de2702a469bddcc395

Observation 7d369d06-fc9f-4b2d-a042-6b0553d869a1 · outbound

This paper cites Denoising diffu- sion probabilistic models.

Compositional Video Synthesis by Temporal Object-Centric Learning Denoising diffu- sion probabilistic models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.641276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.096080Z digest=sha256:c37d199b288729d68f69332f2ca546dc7ccd339b359c506e9b015ebf59632b60

Observation 8f00cd9d-d67b-4eee-9010-f96cbbebff78 · outbound

This paper cites Imagen Video: High Definition Video Generation with Diffusion Models.

Compositional Video Synthesis by Temporal Object-Centric Learning Imagen Video: High Definition Video Generation with Diffusion Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T13:16:44.098463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:16:44.098463Z digest=sha256:ed3355f72c3477714c1c72dff50df54eb29ae15ea2ab037cf8f25aaa833a01c8

Observation 0b8f085c-cbb8-4239-b402-210cb3196108 · outbound

This paper cites Video dif- fusion models.

Compositional Video Synthesis by Temporal Object-Centric Learning Video dif- fusion models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.633260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.101111Z digest=sha256:119c72f89a9f720811d1005935283c8c60c5f51aeaca9c603ea9ba3e90794a23

Observation e8569ae3-5caf-4512-a0f6-1878af20ae95 · outbound

This paper cites Object-centric slot diffusion.

Compositional Video Synthesis by Temporal Object-Centric Learning Object-centric slot diffusion

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.625355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.103608Z digest=sha256:4cbbb6deee8efd0647fb87689230833c6cad07b62011b9a52d759b7a3cc129d4

Observation 27f4d3f3-0c29-4eb7-b51f-4772803cad9c · outbound

This paper cites CLEVR: A diagnostic dataset for compositional language and elementary visual reasoning.

Compositional Video Synthesis by Temporal Object-Centric Learning CLEVR: A diagnostic dataset for compositional language and elementary visual reasoning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.617467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.106047Z digest=sha256:54e1f14033ae4f6ecdc164e4e91d5d26147b9cccfa4d5b827f66d0d55df8f1f9

Observation 77265028-906b-4e4e-aba8-e68f208e99f6 · outbound

This paper cites ClevrTex: A Texture-Rich Benchmark for Unsupervised Multi-Object Segmentation.

Compositional Video Synthesis by Temporal Object-Centric Learning ClevrTex: A Texture-Rich Benchmark for Unsupervised Multi-Object Segmentation

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.609002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.108500Z digest=sha256:cdeb70a4f6cca668067736ec1d243ac3957d5472bda6a744a62f9dfaefc6dc4e

Observation 1ed7a1e7-71f1-4bc6-815d-02ac1b9c81a6 · outbound

This paper cites Con- ditional object-centric learning from video.

Compositional Video Synthesis by Temporal Object-Centric Learning Con- ditional object-centric learning from video

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.600628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.111133Z digest=sha256:af404b0370a7f33b1f314ead440e9300eea3855ae46397c6e049db3d2fbb1708

Observation 30f913bd-9f30-4e1b-a0dc-deaa3d39e6d0 · outbound

This paper cites Sequential attend, infer, repeat: Generative mod- elling of moving objects.

Compositional Video Synthesis by Temporal Object-Centric Learning Sequential attend, infer, repeat: Generative mod- elling of moving objects

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.592327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.113684Z digest=sha256:580b8cccbb5fffe157cebd53c7c1cfa2bc958e793d0381593f3178f6d8c330db

Observation 17a89202-5cc8-4743-b9af-57f04207f6ea · outbound

This paper cites Structured object-aware physics prediction for video modeling and planning.

Compositional Video Synthesis by Temporal Object-Centric Learning Structured object-aware physics prediction for video modeling and planning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.584228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.116118Z digest=sha256:eda3386ebea3ff5fc5e4b4df596ac140754998c6bba1626be8076dd7e4df489a

Observation e6fb7c66-938a-4e30-912d-14fc1de3b904 · outbound

This paper cites Hierarchical compact clustering attention (coca) for unsupervised object-centric learning.

Compositional Video Synthesis by Temporal Object-Centric Learning Hierarchical compact clustering attention (coca) for unsupervised object-centric learning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.575661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.118397Z digest=sha256:4254c0e68e0ea3c1e719fafce54bf016c295a4151eec085af6a45b84a2c3353a

Observation 5ddaa5c3-53a9-47ec-b5bc-1618ee2143fb · outbound

This paper cites an unresolved cited work.

Compositional Video Synthesis by Temporal Object-Centric Learning Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T13:16:44.120650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:16:44.120650Z digest=sha256:9b7c425e457967534c46a349f92a5d3cb67036818ed64ebcbb01647cf384fc1a

Observation 834353d7-8e1f-4680-a4b7-7da091f144c5 · outbound

This paper cites Building machines that learn and think like people.

Compositional Video Synthesis by Temporal Object-Centric Learning Building machines that learn and think like people

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.562155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.123075Z digest=sha256:b8f1c4de01504e374527df4344558504cfa16de513d3d569dfc2d60454c54c2b

Observation 0d7f7806-0acb-42db-a5c5-f5a61b987d54 · outbound

This paper cites an unresolved cited work.

Compositional Video Synthesis by Temporal Object-Centric Learning Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:16:44.554659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.125394Z digest=sha256:99944bf31d9471223e00509b62b6aacd221df7ed2dfa2f827ddca782ff739e5d

Observation eba8958f-3f05-401b-a4ab-d739601f684b · outbound

This paper cites Microsoft COCO: Common objects in context.

Compositional Video Synthesis by Temporal Object-Centric Learning Microsoft COCO: Common objects in context

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.547077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.127855Z digest=sha256:3d97f4ceda8a8a03277abe27f49a1f8fbdd25670e23642e196e1b844426a5cb3

Observation 3545ae88-31fd-481a-8f73-0a4d3a4493e6 · outbound

This paper cites Improving generative imagination in object-centric world models.

Compositional Video Synthesis by Temporal Object-Centric Learning Improving generative imagination in object-centric world models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.539180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.130351Z digest=sha256:8d000168ab1d3c87b713f086ab629006c41ffee858ed12743e3e9d182ce06030

Observation 0e3c82ae-6228-47e3-9f6e-e060c820ef0a · outbound

This paper cites Object- centric learning with slot attention.

Compositional Video Synthesis by Temporal Object-Centric Learning Object- centric learning with slot attention

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.531327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.132668Z digest=sha256:8cfa66eb89a827417348a3c33abc56d509f758a50e033de6f23aa2e5ffe5250a

Observation 32da4fa8-bedc-423f-8fe7-39cb16b2ee4e · outbound

This paper cites Decoupled weight decay regularization.

Compositional Video Synthesis by Temporal Object-Centric Learning Decoupled weight decay regularization

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.523772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.135327Z digest=sha256:f97bc4d188c964865f9a3b0cfd5827a4a490a19af9d6749acd74277125787c05

Observation 5adc5170-5997-46c5-9570-514199a6fa17 · outbound

This paper cites Temporally consistent object-centric learning by contrasting slots.

Compositional Video Synthesis by Temporal Object-Centric Learning Temporally consistent object-centric learning by contrasting slots

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.515469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.138069Z digest=sha256:c35c02300436f749b6b2b4f7082cf4f54454f18f9d6a77c49e3f3b651f5fb4ed

Observation 2fc8e8f7-388a-47d0-a9ff-cdd64dffacbf · outbound

This paper cites T2i-adapter: Learning adapters to dig out more controllable ability for text-to- image diffusion models.

Compositional Video Synthesis by Temporal Object-Centric Learning T2i-adapter: Learning adapters to dig out more controllable ability for text-to- image diffusion models

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.507184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.140512Z digest=sha256:355ddbd49ac49aef8695a0e27a8494c18cbe9fd2ae2d9a9d411f4fcb874855f6

Observation 00dd651f-33cd-41ae-b79d-05ca2e1c01a3 · outbound

This paper cites Segmentation of moving objects by long term video analysis.

Compositional Video Synthesis by Temporal Object-Centric Learning Segmentation of moving objects by long term video analysis

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.499643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.142834Z digest=sha256:e7d082e5edb101d671a63ab791bd7e66973aa394075fe3467046210bd3e37c80

Observation 7f7c8fda-9ce4-40af-8397-092a650ea45d · outbound

This paper cites Dinov2: Learning robust visual features without supervision.

Compositional Video Synthesis by Temporal Object-Centric Learning Dinov2: Learning robust visual features without supervision

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.491763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.145544Z digest=sha256:82527f1e8424d393450147f4a29005add5703c8611f80f66997d34f0fa6b0e6d

Observation 24f67be7-53c4-4945-9708-08e717ecb150 · outbound

This paper cites A benchmark dataset and evaluation methodology for video object segmentation.

Compositional Video Synthesis by Temporal Object-Centric Learning A benchmark dataset and evaluation methodology for video object segmentation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.484295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.148072Z digest=sha256:f298c1e916a01fa1a80d1e93bb4a62f25120a05e549dd40ede6b840d7ab8a507

Observation a7e07070-f2f3-4527-867b-36a12b29eaf6 · outbound

This paper cites Sdxl: Improving latent diffusion models for high-resolution image synthesis.

Compositional Video Synthesis by Temporal Object-Centric Learning Sdxl: Improving latent diffusion models for high-resolution image synthesis

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.476532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.150674Z digest=sha256:51774793c81a7c0b7a9968f4d373a687931cd0fd447a03230365c9ee2b6e973d

Observation 19ade172-0b41-4ab7-91ad-6c6fd7a0acf7 · outbound

This paper cites Rethinking image-to-video adaptation: An object-centric perspective.

Compositional Video Synthesis by Temporal Object-Centric Learning Rethinking image-to-video adaptation: An object-centric perspective

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.468916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.153056Z digest=sha256:854fe1cfab2549a8a658ecc2e8a28a9d7825f98b8ca5dc49403b9b3d440496f2

Observation 701887ea-ba26-49eb-8e74-852cdb8f9523 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Compositional Video Synthesis by Temporal Object-Centric Learning Learning transferable visual models from natural language supervision

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.461242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.155324Z digest=sha256:6274212d06237c6daaf4d6a1b2da2b843d1c51143b8dccca2ccf448cc59debad

Observation 9b7f2095-64c9-4937-bfd2-2b70c1e364d3 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Compositional Video Synthesis by Temporal Object-Centric Learning Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T13:16:44.157831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:16:44.157831Z digest=sha256:f99d19511bdb5ff2a28d4907f624e106d621f6167bed24a04b545cfe44cdd9d7

Observation 21d5838a-89ce-4f63-b000-97a33195b8a6 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Compositional Video Synthesis by Temporal Object-Centric Learning High-resolution image synthesis with latent diffusion models

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.453298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.160495Z digest=sha256:52d7ad0666434f9f42d0e3934dbf56a409a03b699b808facbb770e7e0ddd847a

Observation 73df8951-2164-4f06-ac6b-f976ac6425ea · outbound

This paper cites Photorealistic text-to-image JOURNAL OF LATEX CLASS FILES, VOL.

Compositional Video Synthesis by Temporal Object-Centric Learning Photorealistic text-to-image JOURNAL OF LATEX CLASS FILES, VOL

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.444819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.162837Z digest=sha256:b9a14ab3cb55e62d49b1d71bc494703c002830a31b4ff7710189e40fce10f8e5

Observation c56b3bf3-90b6-4d05-b79c-9b3d626a4569 · outbound

This paper cites Toward causal representation learning.

Compositional Video Synthesis by Temporal Object-Centric Learning Toward causal representation learning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.437460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.165116Z digest=sha256:1ee67d25a3515ae104e1bd022fcda9bfaccb1917519320d45bdeb042f6287fc6

Observation 79c76354-ae9d-4e40-8cb6-4c0501799db6 · outbound

This paper cites Bridging the gap to real-world object-centric learning.

Compositional Video Synthesis by Temporal Object-Centric Learning Bridging the gap to real-world object-centric learning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.430034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.167520Z digest=sha256:0274e556cb2da4d113d2ca18ce7e1ecaac45fdd15d12dc30311426af51b69e90

Observation 3ab10f8a-824f-47d9-91f1-a7c1d2ab0205 · outbound

This paper cites Make-a-video: Text-to-video generation without text-video data.

Compositional Video Synthesis by Temporal Object-Centric Learning Make-a-video: Text-to-video generation without text-video data

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.422188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.169814Z digest=sha256:d48ab3df4a8909fb7a27becca291d1c84e4e6c928e43659a8c607ba2eeab98ee

Observation 8a7ca348-9b49-46d5-8aa7-70c0098f49b6 · outbound

This paper cites Illiterate dall- e learns to compose.

Compositional Video Synthesis by Temporal Object-Centric Learning Illiterate dall- e learns to compose

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.414080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.172171Z digest=sha256:fb67d93c5f0e41b19842351124c9c7e344b2dfe8365cd995239392aefd98c122

Observation 9d97a6e9-7422-43a5-8a2c-5a424caf13f7 · outbound

This paper cites Simple unsu- pervised object-centric learning for complex and natural- istic videos.

Compositional Video Synthesis by Temporal Object-Centric Learning Simple unsu- pervised object-centric learning for complex and natural- istic videos

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.406073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.175236Z digest=sha256:fb31afdb904b121722ef4548dd7e1262cc790fec30fdcb6ed5e33f0f690b84cd

Observation 425ccc73-ca74-4252-bcb6-d1e1c1ad69e4 · outbound

This paper cites Guided latent slot diffusion for object-centric learning.

Compositional Video Synthesis by Temporal Object-Centric Learning Guided latent slot diffusion for object-centric learning

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.398611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.177916Z digest=sha256:62449818b53268d604eb8685cd861736c7b3bc4e009757d2dcca673b59ab820d

Observation 7ac35c2b-f2d0-487e-bede-3d8421ad2798 · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

Compositional Video Synthesis by Temporal Object-Centric Learning Deep unsupervised learning using nonequilibrium thermodynamics

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.390287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.180311Z digest=sha256:7f623ee7e988e6dfb9bddd477ed0a9a7da404b980a4b7d70a377df124cb53ef3

Observation bd9ce06d-2791-4be4-8d9d-68151011e5e9 · outbound

This paper cites Core knowl- edge.

Compositional Video Synthesis by Temporal Object-Centric Learning Core knowl- edge

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.382246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.182654Z digest=sha256:13037403ce9465c74c4f0a9052d84279090387b261f104044b3767334b0a0b74

Observation b0fbe04d-6cdf-42a2-a350-706fcfb4e2c0 · outbound

This paper cites Mind games: Game engines as an architecture for intuitive physics.

Compositional Video Synthesis by Temporal Object-Centric Learning Mind games: Game engines as an architecture for intuitive physics

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.375066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.185445Z digest=sha256:ed00e259180f705699c24469f07a5e5548c0d7081250ae6108e6e3b9e3c2f355

Observation af1b4f84-5a23-4632-aedb-2b80edaab6a8 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

Compositional Video Synthesis by Temporal Object-Centric Learning Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T13:16:44.188237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:16:44.188237Z digest=sha256:b2b9ec6006495c9bf59bfdadddf2e68a06b2834aafef44feb87aa06bac41db9d

Observation 823ef0a0-2908-4923-874a-54a70b1d1b13 · outbound

This paper cites Attention is all you need.

Compositional Video Synthesis by Temporal Object-Centric Learning Attention is all you need

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.367744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.191030Z digest=sha256:ae0cac51df8a46da241bfbcbde2f2580a2d66ddb1bad72ce056532b21dcd6107

Observation 47df23b6-9d0d-4d93-b779-f07af27600c2 · outbound

This paper cites Phenaki: Variable length video generation from open domain textual descrip- tions.

Compositional Video Synthesis by Temporal Object-Centric Learning Phenaki: Variable length video generation from open domain textual descrip- tions

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.360071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.193315Z digest=sha256:af0e9e0fe36b388fdc5026fe3d731d0bab2df218c54f22a135a28ad06ccce10b

Observation df72c4e7-c5a8-41bf-b7c3-552be28f8323 · outbound

This paper cites Videocomposer: Compositional video syn- thesis with motion controllability.

Compositional Video Synthesis by Temporal Object-Centric Learning Videocomposer: Compositional video syn- thesis with motion controllability

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.352046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.195673Z digest=sha256:234582969766969b164436778e0900d99487b8ad9352c77846c85d42d2256f26

Observation edcc823e-9d61-4eba-8783-04ba52c8c529 · outbound

This paper cites Bovik, Hamid R.

Compositional Video Synthesis by Temporal Object-Centric Learning Bovik, Hamid R

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.344281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.198465Z digest=sha256:554f97ea4c0809e1f0d82c1eaa6811544be9bcfde12d69e0d1c380b4c579cc96

Observation ee08bb91-16ae-485c-939d-fdc14c5731c6 · outbound

This paper cites Tune-a-video: One-shot tun- ing of image diffusion models for text-to-video generation.

Compositional Video Synthesis by Temporal Object-Centric Learning Tune-a-video: One-shot tun- ing of image diffusion models for text-to-video generation

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.336566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.200811Z digest=sha256:71131fb1d3c68bd6eaa2aebc1f6d89011855be2ae15252ea9cfe576483c562c3

Observation 430bb969-8803-40e6-b8c4-8ed33041241e · outbound

This paper cites SlotFormer: Unsupervised visual dynamics simulation with object-centric models.

Compositional Video Synthesis by Temporal Object-Centric Learning SlotFormer: Unsupervised visual dynamics simulation with object-centric models

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.328676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.203102Z digest=sha256:7b8e8aff89ec51823619dcb4b0d43bb332a9a14775f254fbcdafa64d1cd079ad

Observation e0d99ffa-f4c9-4e52-b5e7-f53cfec49e3b · outbound

This paper cites Slotdiffusion: Object-centric generative model- ing with diffusion models.

Compositional Video Synthesis by Temporal Object-Centric Learning Slotdiffusion: Object-centric generative model- ing with diffusion models

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.319819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.205454Z digest=sha256:67084117f6922540ced560e806afa90fe01bc591628bec6af14ef05073a62afa

Observation 049bc11e-77fc-4b40-946f-fe0b3b0f08ef · outbound

This paper cites Segment- ing moving objects via an object-centric layered represen- tation.

Compositional Video Synthesis by Temporal Object-Centric Learning Segment- ing moving objects via an object-centric layered represen- tation

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.311397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.208326Z digest=sha256:9c7a1a7e5f6b3cbcd645ff6995b1107f3e3ae9bdfdbf3b0a7355e4bd421afd7a

Observation ab3234e6-fe68-4713-a9ce-81bc16a5decc · outbound

This paper cites Youtube-vos: A large-scale video object segmentation benchmark.

Compositional Video Synthesis by Temporal Object-Centric Learning Youtube-vos: A large-scale video object segmentation benchmark

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.303372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.210743Z digest=sha256:38d80844ea7550ba6ac66f08b2d67e12975fe531cbbd09112b167a4836c1eb9c

Observation 94a6aae8-6199-45ce-943e-2c72a445de2b · outbound

This paper cites Video instance segmentation.

Compositional Video Synthesis by Temporal Object-Centric Learning Video instance segmentation

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.295284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.213086Z digest=sha256:8400fcfaa886153a653609ea77b3e5ee5a20a0a1689076435bf284bc6485b65d

Observation 53ca6954-638a-4aff-90f2-0f387802af24 · outbound

This paper cites Efros, Eli Shechtman, and Oliver Wang.

Compositional Video Synthesis by Temporal Object-Centric Learning Efros, Eli Shechtman, and Oliver Wang

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.287579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.215584Z digest=sha256:c90f71d60577c5313654f9f7d511b4ce7f30925bf2f55c8cf8c7c2b09482fb6b

Observation cdae38e8-f316-438b-ab0c-21b77609e9d0 · outbound

This paper cites Controlvideo: Training-free controllable text-to-video generation.

Compositional Video Synthesis by Temporal Object-Centric Learning Controlvideo: Training-free controllable text-to-video generation

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.278578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T13:16:44.217929Z digest=sha256:a0976a88c2e125fdd11d8369478bca332749f10f8626383de38cbb9a4645b778

Pith citing papers

No inbound Pith citation observations are available.