Pith. sign in

Paper Citation Record · LEDGER

Towards Precise Scaling Laws for Video Diffusion Transformers

As of 13 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2411.17470.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.17470 v2

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:58:57.353357Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact0
  • verified fuzzy15
  • unresolved48
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6959664b-9adb-4de5-be0e-a573a00b1a82 · outbound

This paper cites DeepSeek LLM: Scaling Open-Source Language Models with Longtermism.

Towards Precise Scaling Laws for Video Diffusion Transformers DeepSeek LLM: Scaling Open-Source Language Models with Longtermism

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.075789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.075789Z digest=sha256:02f432a2fe7969b16df13d53382a97e2d6a5cf3069eefb3352dec0a385337bf5

Observation 121e8797-2caf-4cb6-9959-b24eda427fd7 · outbound

This paper cites u-$\mu$P: The Unit-Scaled Maximal Update Parametrization.

Towards Precise Scaling Laws for Video Diffusion Transformers u-$\mu$P: The Unit-Scaled Maximal Update Parametrization

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.081120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.081120Z digest=sha256:fe226622aa3fbea0a43ff2e43fd4e262a029e55df8ad110218a4b105f32d3fba

Observation 523ba06a-a3aa-4b23-ab8f-f07675bb14a4 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Towards Precise Scaling Laws for Video Diffusion Transformers Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.085344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.085344Z digest=sha256:db30ec5535184532aff5007eba51be24b0efb6cac1ae2208fec73d4194ac8a36

Observation 86828b5b-13e7-4606-aa7e-e8d188465b3d · outbound

This paper cites Align your latents: High-resolution video synthesis with la- tent diffusion models.

Towards Precise Scaling Laws for Video Diffusion Transformers Align your latents: High-resolution video synthesis with la- tent diffusion models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:58:58.247754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:58:57.089294Z digest=sha256:388e38745ce828f0b580b8190baa5e05516c66eb8eb31af7a46523804e0a6577

Observation 23ae8ba5-13e2-4e27-8e80-3375ba641bb3 · outbound

This paper cites Video generation models as world simulators, 2024.

Towards Precise Scaling Laws for Video Diffusion Transformers Video generation models as world simulators, 2024

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.093407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.093407Z digest=sha256:7b308f69f3d17ebdcd6c1674974e9b6074ffc98035aee61ce28468fa24775998

Observation 2e5c425c-659c-4459-bf58-d83a745250ba · outbound

This paper cites Videocrafter2: 9 Overcoming data limitations for high-quality video diffu- sion models.

Towards Precise Scaling Laws for Video Diffusion Transformers Videocrafter2: 9 Overcoming data limitations for high-quality video diffu- sion models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:58:58.227466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:58:57.097072Z digest=sha256:54a630ab5af834249e233ea25d73d50605fc235063f3364b47fecac1cd82668a

Observation 5f67f53a-93e8-4dcb-a53a-c8f44139d549 · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

Towards Precise Scaling Laws for Video Diffusion Transformers PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.104946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.104946Z digest=sha256:610e4fff161f65cc5437b0729ba2fecbe1883a2cf02457d791fc6fc946b7b5a3

Observation 27030f13-9a31-4096-90a1-2e1f9b33cc0a · outbound

This paper cites Panda-70m: Captioning 70m videos with multiple cross-modality teachers.

Towards Precise Scaling Laws for Video Diffusion Transformers Panda-70m: Captioning 70m videos with multiple cross-modality teachers

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:58:58.214310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:58:57.108850Z digest=sha256:aa3182dea6da2edf2d63a74d3d9348bd8cd4770fee9987bdc04d201645f8b22c

Observation 239e98e6-c1dc-4962-b088-5800c70fbf7c · outbound

This paper cites Seine: Short-to-long video diffu- sion model for generative transition and prediction.

Towards Precise Scaling Laws for Video Diffusion Transformers Seine: Short-to-long video diffu- sion model for generative transition and prediction

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:58:58.201523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:58:57.112801Z digest=sha256:7279954c2c3333f605a4a70b17e02ce1f8da62beafe5c4bc94728be8a647dae3

Observation 48708aa3-a572-46ae-8c7a-4fb831e2474c · outbound

This paper cites Adversarial Video Generation on Complex Datasets.

Towards Precise Scaling Laws for Video Diffusion Transformers Adversarial Video Generation on Complex Datasets

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.116735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.116735Z digest=sha256:551c7493881555b667b4add34a8d7ee96da1496bfff7b05e5e4afc8974f34280

Observation 4ba04073-6539-46c9-bf6f-3e2d34f5ec09 · outbound

This paper cites The Llama 3 Herd of Models.

Towards Precise Scaling Laws for Video Diffusion Transformers The Llama 3 Herd of Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.120912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.120912Z digest=sha256:96a9d1428a8d8413d6bf42dae3c2878c982539982f3e3095783fb3337d0ded3f

Observation fcb6e01f-93f8-46e6-a087-fff832362720 · outbound

This paper cites Scaling recti- fied flow transformers for high-resolution image synthesis.

Towards Precise Scaling Laws for Video Diffusion Transformers Scaling recti- fied flow transformers for high-resolution image synthesis

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.125086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.125086Z digest=sha256:d2d2d2ed8fba80fb03d61239b05e328eefdd47ba58db2b4e66d6c9c0a977f2ca

Observation 9dd5c5e8-c743-4fee-a509-75eef8c6e838 · outbound

This paper cites Lumina-T2X: Transforming Text into Any Modality, Resolution, and Duration via Flow-based Large Diffusion Transformers.

Towards Precise Scaling Laws for Video Diffusion Transformers Lumina-T2X: Transforming Text into Any Modality, Resolution, and Duration via Flow-based Large Diffusion Transformers

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.129116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.129116Z digest=sha256:03358287c036ede65d35132218ff2db229badc9ee375dd79db04cb941b340b97

Observation f0625430-8b40-479d-b21c-e00cb918bcaa · outbound

This paper cites Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning.

Towards Precise Scaling Laws for Video Diffusion Transformers Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.137762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.137762Z digest=sha256:9aee1accf9d9750d083cffab9062471e8814fa5a2f242b6e0b3305d01ae2f037

Observation 0441af0c-0cc6-41de-a59e-fed8017eb81b · outbound

This paper cites Accurate, Large Minibatch SGD: Training ImageNet in 1 Hour.

Towards Precise Scaling Laws for Video Diffusion Transformers Accurate, Large Minibatch SGD: Training ImageNet in 1 Hour

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.141192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.141192Z digest=sha256:02f9edc8748e7e25bb7f52256a79b339483caa0a8ed687912824f5f1ef31953f

Observation 69a9fec3-811e-4398-8b6d-48bba4422a5c · outbound

This paper cites Latent Video Diffusion Models for High-Fidelity Long Video Generation.

Towards Precise Scaling Laws for Video Diffusion Transformers Latent Video Diffusion Models for High-Fidelity Long Video Generation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.145574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.145574Z digest=sha256:522188d57f4fcdb6d5cbe17e47af04fef406e14acc03e8ab8a896de00f1ce8fb

Observation 073df9a6-4ee6-463a-8f99-ba741dea8c5e · outbound

This paper cites Scaling Laws for Autoregressive Generative Modeling.

Towards Precise Scaling Laws for Video Diffusion Transformers Scaling Laws for Autoregressive Generative Modeling

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.153432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.153432Z digest=sha256:990cb993e54f983f3c45d4a2c48f948c698b3b632284b78161ad832805fbb80f

Observation ac41a400-3674-404e-90dc-b7815959f048 · outbound

This paper cites StreamingT2V: Consistent, Dynamic, and Extendable Long Video Generation from Text.

Towards Precise Scaling Laws for Video Diffusion Transformers StreamingT2V: Consistent, Dynamic, and Extendable Long Video Generation from Text

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.156983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.156983Z digest=sha256:a10fc1161bdc2f1e2348ba7b1c956468a860162bc8b2ec3ba7dcc817ecc33b61

Observation 1e98f59e-12d8-4299-8fb0-3b28d16be15e · outbound

This paper cites Local Lipschitz Bounds of Deep Neural Networks.

Towards Precise Scaling Laws for Video Diffusion Transformers Local Lipschitz Bounds of Deep Neural Networks

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.160989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.160989Z digest=sha256:46636970eaf84014cca7b34f8f04987c4950cfd95b734d168ceb57402f8c590c

Observation b81af0fb-8aa1-4a35-a82c-f5b818c30df7 · outbound

This paper cites Deep Learning Scaling is Predictable, Empirically.

Towards Precise Scaling Laws for Video Diffusion Transformers Deep Learning Scaling is Predictable, Empirically

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.164893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.164893Z digest=sha256:c2fe04e1d9272c1355bfd5cd93c9bf9e91c1b3b4b8b6ed0f4fdb14fc4419f2d8

Observation 05201b75-7e64-42ba-a163-e90c5599d204 · outbound

This paper cites Denoising dif- fusion probabilistic models.

Towards Precise Scaling Laws for Video Diffusion Transformers Denoising dif- fusion probabilistic models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.169522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.169522Z digest=sha256:0debd5c251e26a5df01ebfab26008411b6977469c5d7c1022a39acc3ae54a63f

Observation 34aa8a0c-fe31-45a6-a434-57e4226866e0 · outbound

This paper cites Imagen Video: High Definition Video Generation with Diffusion Models.

Towards Precise Scaling Laws for Video Diffusion Transformers Imagen Video: High Definition Video Generation with Diffusion Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.173429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.173429Z digest=sha256:0962f11a3a55fd657b3119ed2062a5782992370ca9d6aedb904bf4acf220428c

Observation a399adbe-96a3-4e3c-a374-be1b23908883 · outbound

This paper cites Training Compute-Optimal Large Language Models.

Towards Precise Scaling Laws for Video Diffusion Transformers Training Compute-Optimal Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.177934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.177934Z digest=sha256:80ff08e01432551fe835c2e6fba1e8c212dc6ae02ede553fba89cfe7098444aa

Observation ac75d370-f0fd-461a-9257-8a71e4f11fd1 · outbound

This paper cites CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers.

Towards Precise Scaling Laws for Video Diffusion Transformers CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.182926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.182926Z digest=sha256:d302a8103b3a85f465764460ed1edcd2038f1132ea8c7dc1b2b6f2a72335ce26

Observation 5674c9ff-7cf3-4108-9623-f4809b357269 · outbound

This paper cites DepthCrafter: Generating Consistent Long Depth Sequences for Open-world Videos.

Towards Precise Scaling Laws for Video Diffusion Transformers DepthCrafter: Generating Consistent Long Depth Sequences for Open-world Videos

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.187234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.187234Z digest=sha256:8fcd92b918dfd970f796321b349110f933adb7cd3ae4c8d1e9ad97420abe95b9

Observation 5f808a54-7347-4b75-96ba-73401d1c1386 · outbound

This paper cites Scaling Laws for Neural Language Models.

Towards Precise Scaling Laws for Video Diffusion Transformers Scaling Laws for Neural Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.191259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.191259Z digest=sha256:9a5c71d35c46248646c6361b8025c107fb74a7a20df3e850f677be3e1a5b1ad8

Observation 89143203-6888-476f-b378-3a6c0eaa9a69 · outbound

This paper cites Klingai, 2024.

Towards Precise Scaling Laws for Video Diffusion Transformers Klingai, 2024

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:58:58.171563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:58:57.195142Z digest=sha256:779c150c5b39abfc35cf31986e7d1317d89420f6e52eb1fcf119f089fa1ecb38

Observation e703bb33-e5cd-493a-b05e-823776b64579 · outbound

This paper cites VideoPoet: A Large Language Model for Zero-Shot Video Generation.

Towards Precise Scaling Laws for Video Diffusion Transformers VideoPoet: A Large Language Model for Zero-Shot Video Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.198691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.198691Z digest=sha256:cf6373a46ac5e03eeb69bd2fb343a422923d39eeb28cfae99840c21b095d5e2e

Observation 290835e3-5c72-4468-a4f5-ed8240a28293 · outbound

This paper cites On the scalability of diffusion-based text-to-image generation.

Towards Precise Scaling Laws for Video Diffusion Transformers On the scalability of diffusion-based text-to-image generation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:58:58.159348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:58:57.202621Z digest=sha256:f04e2c69638ebfa29445c8c8b4b15bd8e6686fcf0f41d9a7097142cbc59171dc

Observation 8487dc0d-9166-45f7-a414-4bf0d7732189 · outbound

This paper cites Scaling laws for diffusion transformers.

Towards Precise Scaling Laws for Video Diffusion Transformers Scaling laws for diffusion transformers

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.206368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.206368Z digest=sha256:3ff87740621e3d6f1887d20be33637bb7c001329248af20eb69f85ea99ecc212

Observation 5f009bd5-a877-460a-884d-dbbf479fee0f · outbound

This paper cites MarDini: Masked Autoregressive Diffusion for Video Generation at Scale.

Towards Precise Scaling Laws for Video Diffusion Transformers MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.210336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.210336Z digest=sha256:9ddbbdf23098a88b93efa5b8a8b9541110a6e8c865a12475745e0caa09bb366d

Observation 98c15704-7791-411b-ab39-520ae39c2635 · outbound

This paper cites VDT: General-purpose Video Diffusion Transformers via Mask Modeling.

Towards Precise Scaling Laws for Video Diffusion Transformers VDT: General-purpose Video Diffusion Transformers via Mask Modeling

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.214702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.214702Z digest=sha256:8729f30fc7f13d5e6c338b2d9e9e52ccb2f3174fafd99f0042cea654ce550b45

Observation fe7e8a26-a7e7-4441-b964-838b303adaf5 · outbound

This paper cites SiT: Exploring Flow and Diffusion-based Generative Models with Scalable Interpolant Transformers.

Towards Precise Scaling Laws for Video Diffusion Transformers SiT: Exploring Flow and Diffusion-based Generative Models with Scalable Interpolant Transformers

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.218694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.218694Z digest=sha256:6279923d771c048ac2b379dc4d9838d1a080c8b9b6b39579548c47c23995a823

Observation e2e0e4c3-8574-464b-b192-af4da32c6d1a · outbound

This paper cites Latte: Latent Diffusion Transformer for Video Generation.

Towards Precise Scaling Laws for Video Diffusion Transformers Latte: Latent Diffusion Transformer for Video Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.222584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.222584Z digest=sha256:a6dfd2dfa0a149e0be94f918baef5ad2cb04377847617cb5604f3c0c99ef3fde

Observation 427014aa-0976-4f90-9497-e97d0f9c1a8a · outbound

This paper cites An Empirical Model of Large-Batch Training.

Towards Precise Scaling Laws for Video Diffusion Transformers An Empirical Model of Large-Batch Training

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.226574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.226574Z digest=sha256:34103af22d3b03d1a79a4bb0ce9984e02603ef6d8e917b465dff324cddb0624b

Observation 6a5edc9d-e900-49f9-b244-471ee9fdbdb8 · outbound

This paper cites Bigger is not Always Better: Scaling Properties of Latent Diffusion Models.

Towards Precise Scaling Laws for Video Diffusion Transformers Bigger is not Always Better: Scaling Properties of Latent Diffusion Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.230612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.230612Z digest=sha256:748f15bfaa9efc5fb913145f6144a42e2441ce66e16e67bf9cb604f0695a96fc

Observation 21b465df-66bc-4b3f-bef9-0cd4dae2f1e8 · outbound

This paper cites Sora, 2024.

Towards Precise Scaling Laws for Video Diffusion Transformers Sora, 2024

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:58:58.146475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:58:57.234566Z digest=sha256:a7f0672c93bcfc7a180c34985f276c8b7e2990f70aeb767303a7af54194fbff0

Observation d17b79ed-2c06-4b43-8b34-a9140e7474e3 · outbound

This paper cites Scalable diffusion models with transformers.

Towards Precise Scaling Laws for Video Diffusion Transformers Scalable diffusion models with transformers

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.241828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.241828Z digest=sha256:ac9de0f986e631f1c50bff4413cca3806fb1eefda40cdb015b5812b09c796cb2

Observation 1bb1d455-877d-4e43-aee1-aec96878df42 · outbound

This paper cites Movie Gen: A Cast of Media Foundation Models.

Towards Precise Scaling Laws for Video Diffusion Transformers Movie Gen: A Cast of Media Foundation Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.245236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.245236Z digest=sha256:8826388880c8c38a11adefcdec39834c3137b43838fb57140bdd75e8d7eeaeb7

Observation 43828e3a-0e2e-47ae-a80a-29bdf1ab1509 · outbound

This paper cites Tempo- ral generative adversarial nets with singular value clipping.

Towards Precise Scaling Laws for Video Diffusion Transformers Tempo- ral generative adversarial nets with singular value clipping

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:58:58.126785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:58:57.249522Z digest=sha256:51c926ac051f50427e62e7031e458d470d63e28c502afab848be97060becfbdd

Observation 8dae8a06-b1a9-44d4-b608-45a4c9204f79 · outbound

This paper cites Fast High-Resolution Image Synthesis with Latent Adversarial Diffusion Distillation.

Towards Precise Scaling Laws for Video Diffusion Transformers Fast High-Resolution Image Synthesis with Latent Adversarial Diffusion Distillation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.253471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.253471Z digest=sha256:83aa407c9fe611d8975204f8b2f80b0abc9444dd95959456a810b6df85a7ba28

Observation d2dfee5a-d7fc-4ba3-8374-c14028255dba · outbound

This paper cites Measuring the effects of data parallelism on neural network training.

Towards Precise Scaling Laws for Video Diffusion Transformers Measuring the effects of data parallelism on neural network training

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:58:58.114943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:58:57.257816Z digest=sha256:4c32772e3f496a1c2645e32307257cdaed7b70017d9688ccb9348f40215f2054

Observation fe91a8cf-b0f1-47ee-b514-f21d73011611 · outbound

This paper cites Power Scheduler: A Batch Size and Token Number Agnostic Learning Rate Scheduler.

Towards Precise Scaling Laws for Video Diffusion Transformers Power Scheduler: A Batch Size and Token Number Agnostic Learning Rate Scheduler

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.261939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.261939Z digest=sha256:d64707fb29c668d864f77a0466ba296067a115f498b8c29db8e71a336052f45c

Observation b03d27a5-8d73-4733-b022-ab8807ce6ae3 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

Towards Precise Scaling Laws for Video Diffusion Transformers Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.265902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.265902Z digest=sha256:e255c0a62b06503288e51a431e546b11f8ebb96389ca5f651a36ac28d54aecc3

Observation 3e98d66c-2eef-48e9-a91d-d09074420dd7 · outbound

This paper cites Don't Decay the Learning Rate, Increase the Batch Size.

Towards Precise Scaling Laws for Video Diffusion Transformers Don't Decay the Learning Rate, Increase the Batch Size

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.270766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.270766Z digest=sha256:08b8ef7dd05de722a1d2d71f8576516e99d089e3a1f8c0164832591b0c678aa2

Observation f51b78f6-d8af-438b-afaa-050ceb8d494f · outbound

This paper cites Denoising Diffusion Implicit Models.

Towards Precise Scaling Laws for Video Diffusion Transformers Denoising Diffusion Implicit Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.275848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.275848Z digest=sha256:ee3c1d3bb495e08108315264321f57f6a81565a52c2c4ff0e1b5bb7bacc51d12

Observation 681fb9ba-e4f2-4b34-8b25-418904736e6a · outbound

This paper cites Video-Infinity: Distributed Long Video Generation.

Towards Precise Scaling Laws for Video Diffusion Transformers Video-Infinity: Distributed Long Video Generation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.281092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.281092Z digest=sha256:18cc8c195c984a7254abe2759becf6738080bb546a3034c24af564dd237af58c

Observation 55ce4d28-063a-442a-b43a-e8bf89e8976e · outbound

This paper cites Mocogan: Decomposing motion and content for video generation.

Towards Precise Scaling Laws for Video Diffusion Transformers Mocogan: Decomposing motion and content for video generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.286579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.286579Z digest=sha256:32e2adf0f84439730bef086c47ac974d99089d02990b06e3c7b7c3be4d692289

Observation 94561091-ef8d-4d61-aba3-5d5ec211b40f · outbound

This paper cites Phenaki: Variable length video generation from open domain textual descriptions.

Towards Precise Scaling Laws for Video Diffusion Transformers Phenaki: Variable length video generation from open domain textual descriptions

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:58:58.096353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:58:57.290684Z digest=sha256:dd7392e32984fbfc93c4dd85ee0e8b711d02d3ab0978e4e209e4f824d19de52e

Observation d62a95e3-e629-4f28-8429-ffcb2c61a67c · outbound

This paper cites Generating videos with scene dynamics.

Towards Precise Scaling Laws for Video Diffusion Transformers Generating videos with scene dynamics

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:58:58.083579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:58:57.294695Z digest=sha256:e466ce25d1e832cdc30deabc2c12dac05525fc59a8525b8d2c98f7ce5a8f597e

Observation 0f79c6ea-effe-4c7b-9f36-98830d650975 · outbound

This paper cites Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising.

Towards Precise Scaling Laws for Video Diffusion Transformers Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.299416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.299416Z digest=sha256:161d6e28f43d6215efa636dccb2fa1deccd80bfc2b280fc1cdd098b660105b3a

Observation 4eeb25bb-cdaf-4e4d-9070-769665fcd4cf · outbound

This paper cites Adapting to smoothness: A more universal algorithm for on- line convex optimization.

Towards Precise Scaling Laws for Video Diffusion Transformers Adapting to smoothness: A more universal algorithm for on- line convex optimization

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:58:58.070711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:58:57.303450Z digest=sha256:75a72b034caedbc83dcebcf18b2f85d4015f5fd9744edbcb23bd1308540d34ff

Observation a63ae0b9-8093-4009-b810-b9cf60a0ed87 · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

Towards Precise Scaling Laws for Video Diffusion Transformers Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.307392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.307392Z digest=sha256:5975d7b3333418e61c6f0e541c5a09849d40f8e108e8143f1022aff9a0d74069

Observation ada8b994-365b-48ae-b22e-69089c4a60e2 · outbound

This paper cites VideoGPT: Video Generation using VQ-VAE and Transformers.

Towards Precise Scaling Laws for Video Diffusion Transformers VideoGPT: Video Generation using VQ-VAE and Transformers

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.312178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.312178Z digest=sha256:939d075f4008df448117df7c12fc2d7144584505127dc26ad560baf77b9049a6

Observation a3381a06-ed1b-4b35-9dc9-e8900b1ff2a3 · outbound

This paper cites Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer.

Towards Precise Scaling Laws for Video Diffusion Transformers Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.316711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.316711Z digest=sha256:0a120f81e56b7c1ab59d4476966a202709753002e661d32fcb1b0b4940592047

Observation efdaa03c-4afe-486d-b26e-1de95e965cc3 · outbound

This paper cites Rerender a video: Zero-shot text-guided video-to-video translation.

Towards Precise Scaling Laws for Video Diffusion Transformers Rerender a video: Zero-shot text-guided video-to-video translation

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:58:58.055536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:58:57.322749Z digest=sha256:6da87010e8e71548822ef7894bb6ea9e557992023a40f71aed0313abf82e406f

Observation e0f695d4-8db7-4e38-90d8-76b3c2e43858 · outbound

This paper cites Space-time diffusion features for zero-shot text-driven motion transfer.

Towards Precise Scaling Laws for Video Diffusion Transformers Space-time diffusion features for zero-shot text-driven motion transfer

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:58:58.043210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:58:57.328218Z digest=sha256:bbe493f1b8361b8decac2306c8749a0355659abb1aa8143cde1b418305e39187

Observation ac178b8d-662a-4fbb-a554-74a39c0478d9 · outbound

This paper cites NUWA-XL: Diffusion over Diffusion for eXtremely Long Video Generation.

Towards Precise Scaling Laws for Video Diffusion Transformers NUWA-XL: Diffusion over Diffusion for eXtremely Long Video Generation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.332537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.332537Z digest=sha256:b5196aee5feec2676f045af336bc51e31094fa1378761b6052f4500b38a12433

Observation 16c40c50-f5a8-44e9-8276-668a4013d80a · outbound

This paper cites Generating Videos with Dynamics-aware Implicit Generative Adversarial Networks.

Towards Precise Scaling Laws for Video Diffusion Transformers Generating Videos with Dynamics-aware Implicit Generative Adversarial Networks

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.337071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.337071Z digest=sha256:24170479f6b4261b0af747c77332516a3538f7ab04c2cb1a059b64479c885f1f

Observation 2b883efa-dfc3-4297-a227-f26424a6a543 · outbound

This paper cites Tora: Trajectory-oriented Diffusion Transformer for Video Generation.

Towards Precise Scaling Laws for Video Diffusion Transformers Tora: Trajectory-oriented Diffusion Transformer for Video Generation

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.341054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.341054Z digest=sha256:0f8e2c9cc85e5d7c60eac9f043062049e825626ec7e23c44ad78d1aac3e85b7d

Observation 911ded9e-c48a-461a-9aed-eb8f0a5676cd · outbound

This paper cites Moviedreamer: Hier- archical generation for coherent long visual sequence.

Towards Precise Scaling Laws for Video Diffusion Transformers Moviedreamer: Hier- archical generation for coherent long visual sequence

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.345097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.345097Z digest=sha256:6d02ff78a8209f721e6374aa1e6383665b5d5f0901a4d4fd8898cf407c4044cc

Observation c661acd3-f623-4a74-8240-08e034d3f65b · outbound

This paper cites Open-sora: Democratizing efficient video production for all, 2024.

Towards Precise Scaling Laws for Video Diffusion Transformers Open-sora: Democratizing efficient video production for all, 2024

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:58:58.030147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:58:57.349217Z digest=sha256:93f4ab77d1afd0ddd02edea0aabcf7b1386d73451941003dcbc31128abd1a557

Observation a764416c-a39e-4bc0-b39f-f87f97179d32 · outbound

This paper cites StoryDiffusion: Consistent Self-Attention for Long-Range Image and Video Generation.

Towards Precise Scaling Laws for Video Diffusion Transformers StoryDiffusion: Consistent Self-Attention for Long-Range Image and Video Generation

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.353357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.353357Z digest=sha256:3539786a30470d7c4edfcf1978dcfb815b9d9a6b315ed8c230cea237080b504c

Pith citing papers

No inbound Pith citation observations are available.