Pith. sign in

Paper Citation Record · LEDGER

Understanding Attention Mechanism in Video Diffusion Models

As of 18 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 4 inbound Pith citation observations for arXiv:2504.12027.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.12027 v2

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:40:31.484674Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:57:01.325386Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T00:39:35.661863Z

Reference resolution

50 of 50 outbound references displayed

  • verified exact0
  • verified fuzzy25
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 28184173-d3e3-4b75-8c3b-ce7f221821d5 · outbound

This paper cites Self-Rectifying Diffusion Sampling with Perturbed-Attention Guidance.

Understanding Attention Mechanism in Video Diffusion Models Self-Rectifying Diffusion Sampling with Perturbed-Attention Guidance

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.084822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.084822Z digest=sha256:46e7fc5a057c8b8740c58ece7dba314502b1119b52154ddff1aa51be1ab679fd

Observation 5d3f72be-4427-4338-a82d-eb16a8037929 · outbound

This paper cites Masactrl: Tuning-free mu- tual self-attention control for consistent image synthesis and editing.

Understanding Attention Mechanism in Video Diffusion Models Masactrl: Tuning-free mu- tual self-attention control for consistent image synthesis and editing

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.782458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.092317Z digest=sha256:63f522fdd19bac7e4b63d96f74529f46c81f7422557bec48ef2fd9a8b823ccc0

Observation 04fc034c-9831-44a4-acb6-cb2d038170fa · outbound

This paper cites Huang, and Niloy J.

Understanding Attention Mechanism in Video Diffusion Models Huang, and Niloy J

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.763310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.100292Z digest=sha256:170bab0a8af68d226e1e12d4c9de2ca0202a613f66fce5e8d90fc9f245b14ac3

Observation b4e85b25-60e0-4ad7-85b4-aefd6026532e · outbound

This paper cites Videocrafter2: Overcoming data limitations for high-quality video diffusion models, 2024.

Understanding Attention Mechanism in Video Diffusion Models Videocrafter2: Overcoming data limitations for high-quality video diffusion models, 2024

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.744161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.107586Z digest=sha256:41e6a478aa23747f09f678a6e3a4146c8bf78c7b6d7273b266b2b3699803ae13

Observation eacc89e2-7e6d-43f4-b4d9-49bce61cc9b7 · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

Understanding Attention Mechanism in Video Diffusion Models PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.115922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.115922Z digest=sha256:328dfd1b273bfcc0c256b245f0594ad60a7223ba4887846960c7cf479991fadf

Observation 1e0866b4-d85a-42b6-9893-81c5a1aaf99f · outbound

This paper cites Slicedit: Zero- shot video editing with text-to-image diffusion models using spatio-temporal slices.

Understanding Attention Mechanism in Video Diffusion Models Slicedit: Zero- shot video editing with text-to-image diffusion models using spatio-temporal slices

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.726248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.123810Z digest=sha256:6f33a75286a157b96f04435b2ca26842849ce5c5e46433087c667d70f2ad9c05

Observation b1492cd3-7c55-4fe6-bec8-d5cfe02febf0 · outbound

This paper cites FLATTEN: optical FLow-guided ATTENtion for consistent text-to-video editing.

Understanding Attention Mechanism in Video Diffusion Models FLATTEN: optical FLow-guided ATTENtion for consistent text-to-video editing

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.131828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.131828Z digest=sha256:9915d91409bb1fd6b606f5f7531d3a1e234f10dac15d516beb57d8b9763c6275

Observation 99cc7e0c-fa27-4c63-9f4b-fdb46b409465 · outbound

This paper cites Stylegan-nada: Clip- guided domain adaptation of image generators.

Understanding Attention Mechanism in Video Diffusion Models Stylegan-nada: Clip- guided domain adaptation of image generators

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.706857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.137998Z digest=sha256:e9ab9665aaa931a7c61f3db1e375b9c22c8c337e5808226893bd590a375f5e6c

Observation 36a7cf5c-89f1-4f08-adc5-cce126937d41 · outbound

This paper cites Efros, and Jacob Steinhardt.

Understanding Attention Mechanism in Video Diffusion Models Efros, and Jacob Steinhardt

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.679355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.146213Z digest=sha256:d664b704d43a706d5fd11632d71b207e4eb9fc05c27815af4ec7b26177b3b87b

Observation 0324039f-efae-47e2-9911-9c0fd7b0dfba · outbound

This paper cites TokenFlow: Consistent Diffusion Features for Consistent Video Editing.

Understanding Attention Mechanism in Video Diffusion Models TokenFlow: Consistent Diffusion Features for Consistent Video Editing

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.157854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.157854Z digest=sha256:355e3b4fa4c7e84127c228f1b95b4b7ed4f6f6800bbb2bbf01481588bdf9ca7f

Observation 2aaebf9c-ef78-4e9c-a11a-11f1ed0937ae · outbound

This paper cites Entropy and information theory.

Understanding Attention Mechanism in Video Diffusion Models Entropy and information theory

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.652649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.167634Z digest=sha256:769bfb24ffac6964d6d371a9ec134a193133e0b60488255e5da9fc5930c32194

Observation f8743714-09d6-41d9-b96c-403e5de9c1be · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

Understanding Attention Mechanism in Video Diffusion Models AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.175567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.175567Z digest=sha256:3d3d2d68c681cc1e8ed654bb404c9cb8adf742cfb2da7ef01a1dac3c0ac600c6

Observation 4c01a08e-169b-45cd-81d2-fb133e03c532 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

Understanding Attention Mechanism in Video Diffusion Models Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.183276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.183276Z digest=sha256:ed91d601ba80127b4669a16f5ff2c2e8b04bf10bb50157d6e971d512efb8ecd3

Observation d6e8add7-66bf-41ef-a166-44d0ad7e0dff · outbound

This paper cites Classifier-Free Diffusion Guidance.

Understanding Attention Mechanism in Video Diffusion Models Classifier-Free Diffusion Guidance

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.191844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.191844Z digest=sha256:25278c2ed67c825b5726c20f2832e19bcfbc10f6272e52d096b01a88ae3ba5a5

Observation e90fb4c2-3a3c-4e1f-a070-2c74823400a9 · outbound

This paper cites Denoising diffu- sion probabilistic models.

Understanding Attention Mechanism in Video Diffusion Models Denoising diffu- sion probabilistic models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.199397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.199397Z digest=sha256:0ebfc29f4db49cec2cee9ed8087df181b740dcf9fcc70c5be2724d37294759dc

Observation d740040b-6acb-4f78-9194-619546a1a3a4 · outbound

This paper cites CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers.

Understanding Attention Mechanism in Video Diffusion Models CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.206444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.206444Z digest=sha256:9478c3fff5335560a72db450acbb0e8ce1327e318275657abae92c54d2a3f839

Observation 1540c488-1d25-455b-a95e-3b139abf59bb · outbound

This paper cites VideoControlNet: A Motion-Guided Video-to-Video Translation Framework by Using Diffusion Model with ControlNet.

Understanding Attention Mechanism in Video Diffusion Models VideoControlNet: A Motion-Guided Video-to-Video Translation Framework by Using Diffusion Model with ControlNet

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.216256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.216256Z digest=sha256:3eb73dc17b09e6fe79bc85a538ca157d572c2d92b59bd0735d1cb78fb70e42ab

Observation 1ea0d7ce-5ed1-423b-9969-7bc34301cca0 · outbound

This paper cites Vbench: Comprehensive benchmark suite for video generative models.

Understanding Attention Mechanism in Video Diffusion Models Vbench: Comprehensive benchmark suite for video generative models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.607531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.224344Z digest=sha256:a2037f6a0a4d6bf9f0e1dc28ee75ab9c5f57bff5c6b2c0874453303d3bc10004

Observation b9f640ff-4ffe-48ac-ad78-990de883dacc · outbound

This paper cites Pnp inversion: Boosting diffusion-based editing with 3 lines of code.

Understanding Attention Mechanism in Video Diffusion Models Pnp inversion: Boosting diffusion-based editing with 3 lines of code

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.581567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.232005Z digest=sha256:28dd3481fd33e17da716f4dee2eefe7db823258fdcf3c1b3079cab42ff736c38

Observation cf3d74dd-480d-47de-b6fe-5430babe5956 · outbound

This paper cites AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks.

Understanding Attention Mechanism in Video Diffusion Models AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.245290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.245290Z digest=sha256:ba8a89f68c6144aaedfd9ea0b6a3af2e484e8c3df07bc3c0e1aaf84e1f30456d

Observation 6f2313c0-1aff-4298-8efd-b447edf1bff2 · outbound

This paper cites Shannon entropy: a rigorous notion at the crossroads between probability, information theory, dynami- cal systems and statistical physics.

Understanding Attention Mechanism in Video Diffusion Models Shannon entropy: a rigorous notion at the crossroads between probability, information theory, dynami- cal systems and statistical physics

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.558297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.256478Z digest=sha256:ebaae8136a2f03b343ba58e0a42b9adbebc0eccc03844bcd627ac1836e2551eb

Observation dc9f41c5-f5f1-47ab-b673-6f68df039d55 · outbound

This paper cites Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding.

Understanding Attention Mechanism in Video Diffusion Models Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.264007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.264007Z digest=sha256:5f5aaa23060e7716f3af4232ac66d9bc781f96cf20ad1b8ff1c7ab80cbf7e2ea

Observation 789b0293-39b2-4d68-a181-ff02139600df · outbound

This paper cites Flowvid: Taming imperfect optical flows for consistent video-to-video synthe- sis.

Understanding Attention Mechanism in Video Diffusion Models Flowvid: Taming imperfect optical flows for consistent video-to-video synthe- sis

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.525401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.269752Z digest=sha256:182c635ecc48c6a5cc515c3837c4d3036b77e68df14f372a76abb67537efd3bc

Observation cec16c4d-3e28-4857-983a-6ebf3fd2661c · outbound

This paper cites Animatediff-lightning: Cross- model diffusion distillation, 2024.

Understanding Attention Mechanism in Video Diffusion Models Animatediff-lightning: Cross- model diffusion distillation, 2024

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.504733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.275421Z digest=sha256:2e23147199c17f1ee78f989686da2d9b35c640b7e5136c2ff60f4bff8ca9831f

Observation b40ae706-80d3-4cee-88a5-d9c4769158ae · outbound

This paper cites AnimateDiff-Lightning: Cross-Model Diffusion Distillation.

Understanding Attention Mechanism in Video Diffusion Models AnimateDiff-Lightning: Cross-Model Diffusion Distillation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.283276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.283276Z digest=sha256:651a136fbee13e92603415868468f3109e89ad200eacd968d46aa3856abf7042

Observation 1d10e7c1-baec-4af9-b417-382d3d416689 · outbound

This paper cites Towards understanding cross and self-attention in stable diffusion for text-guided image editing.

Understanding Attention Mechanism in Video Diffusion Models Towards understanding cross and self-attention in stable diffusion for text-guided image editing

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.484181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.291775Z digest=sha256:f5e5e89e60a6b787b11e6e654ea1c14c18140f0a9e651285a453a775d5e25722

Observation d0100158-66be-42fd-8689-0668975c6369 · outbound

This paper cites Video-p2p: Video editing with cross-attention control.

Understanding Attention Mechanism in Video Diffusion Models Video-p2p: Video editing with cross-attention control

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.464746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.298618Z digest=sha256:d1445336bd9c286c51c2268631ea9cba736d97bc4ab5d0f5fdfabd72a32d473c

Observation 28af9a09-e01c-4d5a-9faa-6dd7e6b5c10e · outbound

This paper cites sora, 2024.

Understanding Attention Mechanism in Video Diffusion Models sora, 2024

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.444011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.307320Z digest=sha256:3947f9d68987e3a3934410d65d970bacd9de31ce094a0ac26950274ec0958896

Observation b13e67c8-5381-4740-b0c9-55a4b46ba2c5 · outbound

This paper cites Scalable diffusion models with transformers.

Understanding Attention Mechanism in Video Diffusion Models Scalable diffusion models with transformers

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.314453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.314453Z digest=sha256:78893de74db983e96d2b2d776a9e910a5ff9731cf24fedd27988ce373bf70daf

Observation f61be9e5-8588-4538-ba22-660cb9280207 · outbound

This paper cites The 2017 DAVIS Challenge on Video Object Segmentation.

Understanding Attention Mechanism in Video Diffusion Models The 2017 DAVIS Challenge on Video Object Segmentation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.319871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.319871Z digest=sha256:9d23963c4c577e1fbfe960297b9a51c2538da3be03253aa768f7c6706a179768

Observation 646aff54-ba56-4cac-b55f-cb435c56813b · outbound

This paper cites Fatezero: Fusing attentions for zero-shot text-based video editing.

Understanding Attention Mechanism in Video Diffusion Models Fatezero: Fusing attentions for zero-shot text-based video editing

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.409320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.326729Z digest=sha256:be43a23317eb220ff18d5a174245a027d8ea8d27bcfe9a0c7bb0036f149d77cb

Observation 315a05b8-8db4-45f9-b69e-a6b5a1985461 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Understanding Attention Mechanism in Video Diffusion Models Learning transferable visual models from natural language supervi- sion

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.386584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.333131Z digest=sha256:813f3362d1bca3a5bc9ba2f472b34c026411fefaae1f5c90f46392dcb67f72da

Observation 8ce765c9-dd34-442b-87e3-612363d75a10 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Understanding Attention Mechanism in Video Diffusion Models High-resolution image synthesis with latent diffusion models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.339181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.339181Z digest=sha256:2046c0633330d6eea8b4c84532fa6e454da56c2473ed59f6635c919f408e77e9

Observation 21cab84d-4957-4f36-a071-21b01ce5ea4f · outbound

This paper cites an unresolved cited work.

Understanding Attention Mechanism in Video Diffusion Models Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-16T12:40:32.347559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.346281Z digest=sha256:95bdfc0072e1ba4829d055f789089f7c17f3baf7b399bb420382340bab3d48a6

Observation 34096daa-462a-4624-87ac-1f2ca2473c7d · outbound

This paper cites Laion-5b: An open large-scale dataset for training next gen- eration image-text models.

Understanding Attention Mechanism in Video Diffusion Models Laion-5b: An open large-scale dataset for training next gen- eration image-text models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.328141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.354674Z digest=sha256:9da250ff8bf85981675aae62db7b1480e9c058afd752de4e6f53404643a6c9cf

Observation d1bb81df-cc30-4667-873c-7f51698414f3 · outbound

This paper cites Freeu: Free lunch in diffusion u-net.

Understanding Attention Mechanism in Video Diffusion Models Freeu: Free lunch in diffusion u-net

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.296437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.365288Z digest=sha256:7c426a3d10125f82a720b283045d198adeb75035c6d541e2733d09aa64264aa6

Observation 2f25519f-e36f-461c-9192-b2156a7f4577 · outbound

This paper cites Denoising diffusion implicit models.

Understanding Attention Mechanism in Video Diffusion Models Denoising diffusion implicit models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.271804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.373217Z digest=sha256:5927da09865b0bea6e06358de933dce0d645a6340820c5d840a28a0e029d9f53

Observation a2806c3a-d8b5-4abd-bf1f-3015bbb17737 · outbound

This paper cites Plug-and-play diffusion features for text-driven image-to- image translation.

Understanding Attention Mechanism in Video Diffusion Models Plug-and-play diffusion features for text-driven image-to- image translation

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.244343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.383603Z digest=sha256:a6752bff6eb3a18456e9976641a77543e75f2402a82026ce54eea6e3e563c749

Observation e4623132-61b5-41e4-a80e-2d57c3223313 · outbound

This paper cites LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models.

Understanding Attention Mechanism in Video Diffusion Models LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.391194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.391194Z digest=sha256:02c135ca69d7ab751232ec163fcd56ca0481295997c1a9cd8cc972183ceb17eb

Observation 9bf943fe-85fc-4199-a5a0-248332b78032 · outbound

This paper cites Image quality assessment: from error visibility to structural similarity.

Understanding Attention Mechanism in Video Diffusion Models Image quality assessment: from error visibility to structural similarity

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.401282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.401282Z digest=sha256:de8cd801dd9d49f430bfa9ec26d55792a780d9357460e96981924e144181eb45

Observation 7247f747-e053-49b3-bf45-3528eea1ca3e · outbound

This paper cites Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation.

Understanding Attention Mechanism in Video Diffusion Models Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.208345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.407483Z digest=sha256:d1f0cfd8489cecf4cd20d3bd8a811882717d9d958885b18985600a6276440c98

Observation bc73c21b-723b-4b54-afea-89d2fd853a82 · outbound

This paper cites CVPR 2023 Text Guided Video Editing Competition.

Understanding Attention Mechanism in Video Diffusion Models CVPR 2023 Text Guided Video Editing Competition

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.422477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.422477Z digest=sha256:bc6025427498231ba8990159b3e84f15fed9cf5b7cde4cf718e0a4df947321ac

Observation 1e1a0d22-40af-4bc5-85ca-17a9a3f792ad · outbound

This paper cites Easyanimate: A high-performance long video generation method based on transformer architecture.

Understanding Attention Mechanism in Video Diffusion Models Easyanimate: A high-performance long video generation method based on transformer architecture

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.434122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.434122Z digest=sha256:d51c4655940318d3def7d32c656a10fbfc2b9ffcf43df8a3ce744f36944cd26f

Observation 41c8cb1a-7afa-4c3c-a385-f4386f676aa3 · outbound

This paper cites Rerender a video: Zero-shot text-guided video-to-video trans- lation.

Understanding Attention Mechanism in Video Diffusion Models Rerender a video: Zero-shot text-guided video-to-video trans- lation

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.177894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.442518Z digest=sha256:d3aa723193685cc79a8878929af68b99b429ee3350bb7f77d6b2a468cc7bf3bf

Observation e0be45d1-9e91-4d9e-b66c-4c69658507ba · outbound

This paper cites Fresco: Spatial-temporal correspondence for zero-shot video translation.

Understanding Attention Mechanism in Video Diffusion Models Fresco: Spatial-temporal correspondence for zero-shot video translation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.153477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.449282Z digest=sha256:2b81186682f19c44ed318e21c49cd11799dc14a355243eb812e96fc86f943f43

Observation 2cd3b3e0-693f-4fdf-b178-14fd3de352b8 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

Understanding Attention Mechanism in Video Diffusion Models CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.456358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.456358Z digest=sha256:453ad353223f89fe30d7baff969fad02faf9d60e142705ffadfc5fbc98096d03

Observation b8a5c7dd-e42f-47f1-986c-0cdf94360be4 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

Understanding Attention Mechanism in Video Diffusion Models The unreasonable effectiveness of deep features as a perceptual metric

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.464055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.464055Z digest=sha256:63cf168f44f9a026164e792882332a479a3a0666eb955497f141fcb87f97601b

Observation fbac40cc-c878-40c5-800f-2e7b0d975166 · outbound

This paper cites ControlVideo: Training-free Controllable Text-to-Video Generation.

Understanding Attention Mechanism in Video Diffusion Models ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:31.473106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:31.473106Z digest=sha256:54cdf6d729464df5e7699a5a75e5a63aae60d7efcb1907c7fb5d0e82b0f22631

Observation 5bc4f304-4ac9-4a23-afb1-0f99d901b187 · outbound

This paper cites Implementation Details Our experimental setup is described as follows: For video generation using AnimateDiff,1 the model is initialized with a ‘torch.float16‘ data type.

Understanding Attention Mechanism in Video Diffusion Models Implementation Details Our experimental setup is described as follows: For video generation using AnimateDiff,1 the model is initialized with a ‘torch.float16‘ data type

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:32.106980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.478852Z digest=sha256:4e6c9a1c3ff5acfd8950b6101e4ece5ff9b9160188a30ada601e3bce69ff6ea5

Observation 67e21b67-a049-4573-ad0d-0f3e6b2e12cc · outbound

This paper cites an unresolved cited work.

Understanding Attention Mechanism in Video Diffusion Models Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-16T12:40:32.075030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:40:31.484674Z digest=sha256:fb5ab77fcbef3637537a97bbf3b3f2949dd322e81739b2a5f5e8d6b78cd11217

Pith citing papers

Observation 9b77efe5-856d-4308-be80-e0fd4af3d99b · inbound

Zero-to-Hero: Zero-Shot Initialization Empowering Reference-Based Video Appearance Editing cites this paper.

Zero-to-Hero: Zero-Shot Initialization Empowering Reference-Based Video Appearance Editing Understanding Attention Mechanism in Video Diffusion Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:57:01.325386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:57:01.325386Z digest=sha256:6d83f553a651a1b7ce5839f5310f4895060d5ce3b2bfe3503b955d3e55beb1ed

Observation f54742b7-ba7b-4154-8d07-93cca00622ba · inbound

Attention of a Kiss: Exploring Attention Maps in Video Diffusion for XAIxArts cites this paper.

Attention of a Kiss: Exploring Attention Maps in Video Diffusion for XAIxArts Understanding Attention Mechanism in Video Diffusion Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T13:27:41.689046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:27:41.689046Z digest=sha256:21051e7510be7baa0ece89e6a65afe6118d1d5fcee5e0abc3cc3ac7eb7881bdf

Observation 716819a8-bbee-4039-babd-ae4268c12ee8 · inbound

CLEAR: Context-Aware Learning with End-to-End Mask-Free Inference for Adaptive Video Subtitle Removal cites this paper.

CLEAR: Context-Aware Learning with End-to-End Mask-Free Inference for Adaptive Video Subtitle Removal Understanding Attention Mechanism in Video Diffusion Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:39:35.663851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T00:39:15.790131Z digest=sha256:51a10a63c4ad709fdc63afa25d3750958321c46f47796a64124355e9301c3039

Observation 1eb62197-462c-4f29-b6cd-c815ac47e8c0 · inbound

Controlling Motion Transfer in Diffusion Transformers via Attention Heads cites this paper.

Controlling Motion Transfer in Diffusion Transformers via Attention Heads Understanding Attention Mechanism in Video Diffusion Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-14T07:11:45.493916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T07:11:45.493916Z digest=sha256:68d7a167a2befc6ff88ec684cab18d80ebc00fe7e5644d82d3570ca7f7878fd8