Pith. sign in

Paper Citation Record · LEDGER

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models

As of 14 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 5 inbound Pith citation observations for arXiv:2506.09229.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.09229 v2

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:59:00.878129Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-02T23:56:39.865433Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T11:08:03.362718Z

Reference resolution

52 of 52 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved42
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 32cabdd5-345d-416c-ad84-abcc019dd647 · outbound

This paper cites Cosmos World Foundation Model Platform for Physical AI.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:54.075933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:54.075933Z digest=sha256:0252b7e3b3faf21d414a20af54ff20025f72f875c76b2c369d1a1a6c12176ce0

Observation b9e20daa-daee-4be6-8256-fd27d91260e7 · outbound

This paper cites Baker, I.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Baker, I

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:04.547903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:54.147170Z digest=sha256:3e38c11f9693138aefe591b5644dcd2d75bd8b36fa280ceff4fb2a00889c8215

Observation 484b1eae-df14-458f-bfc8-16c603cb49cb · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:59:04.393193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:54.240845Z digest=sha256:be664313daa5251acf463a30872108bc9744a2eee70fde970eff0cdb03bf8e0c

Observation 212d2a9a-5448-41fa-9c4e-23250c39ebe9 · outbound

This paper cites Scenes dataset.https://huggingface.co/datasets/bigdata-pw/scenes, 2025.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Scenes dataset.https://huggingface.co/datasets/bigdata-pw/scenes, 2025

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:04.254765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:54.331646Z digest=sha256:5a329a5243f9f50a34be91521287b5ccdb7a39ac2bfb0979cc9a24472ecb6b59

Observation 53e1fde2-3f87-4667-beeb-46ecd2477af3 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:54.413395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:54.413395Z digest=sha256:21286fe695a6965a3886f16007e6e45dd4746aa8b72b26c9e73a4a38b87c18ff

Observation 82630d41-0eee-4c6e-ad39-e8e85365d390 · outbound

This paper cites Blattmann, R.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Blattmann, R

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:54.475871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:54.475871Z digest=sha256:7579a1515ff130b008a1ef1376fe11236a2033a2c317cc1fc34e36c799b21383

Observation acfd326d-117e-4a67-bc07-936ccbe3078c · outbound

This paper cites Brown, B.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Brown, B

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:54.548379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:54.548379Z digest=sha256:f3a1cc07d9f8c51636013fe0e9520e1ab25038fb9d7ffe04de01975758dad5ed

Observation 82b0d012-c855-47e5-9950-e5cbf78187ef · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:54.642936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:54.642936Z digest=sha256:975c3f1493f9db543b3c490a49a4c4e85bda62a6ddd847b9fb86b4a227b0e951

Observation 42931469-6608-4aba-a780-144f405df876 · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:54.722164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:54.722164Z digest=sha256:c4503c806b178650c7a8672fbd37e4d803f2090de42d36f364f984800795ff40

Observation d2f99442-5670-492f-b528-85186922546f · outbound

This paper cites Cakeify smol dataset.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Cakeify smol dataset

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:04.080717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:54.781019Z digest=sha256:cfe9cfd82e15c60f87ba9d5ba6938c9ce5b229b367524885f948a02a060da7c5

Observation 649f0bf5-6d9b-4a95-b46a-9115e20ed2d0 · outbound

This paper cites Crush smol dataset.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Crush smol dataset

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:03.922273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:54.858367Z digest=sha256:9db92ab7f25cdf7ab7cf4d5bb3856a6869eff4de9449dc414062aabd8c09de6a

Observation dc1f7ff7-30dc-4055-be0f-c60ea87df8aa · outbound

This paper cites Squish pika dataset.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Squish pika dataset

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:03.773661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:54.963160Z digest=sha256:7c7539d92527ea5c059496b7f7427553cfb2cf62415c6c410958521930fbba65

Observation b4f1e179-26a3-4a52-807d-ae4c96412b42 · outbound

This paper cites Long Context Tuning for Video Generation.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Long Context Tuning for Video Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:55.068782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:55.068782Z digest=sha256:09c87dccc6c4b67f7485da2edcafdced3cd965c1bf622514239e6bba25c36523

Observation 23d133e8-920f-4f41-ad68-41bfcba13ccb · outbound

This paper cites Don't Stop Pretraining: Adapt Language Models to Domains and Tasks.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Don't Stop Pretraining: Adapt Language Models to Domains and Tasks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:55.232129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:55.232129Z digest=sha256:64e41239404c40968774b595f62ec7cf0da069475d9d0c30dd436bd64f03eb8e

Observation 77280d0b-4a9e-40bd-8acc-b87ddb79c0cb · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:59:03.640450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:55.457490Z digest=sha256:6a72aef15aadf729ad94fbaa47010bc6487a2a1ecba2cad9f4c273285d602e26

Observation 50a983f2-0a97-4467-8978-967f0bc78d18 · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:55.691200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:55.691200Z digest=sha256:8576895fe98a862358e33ed123c0f69e66aec3ae7493e3eaa087ec3aaa8a44ff

Observation f76c8697-c3b6-44c8-8c19-bafeb7f3a54f · outbound

This paper cites CogVLM2: Visual Language Models for Image and Video Understanding.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models CogVLM2: Visual Language Models for Image and Video Understanding

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:55.911607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:55.911607Z digest=sha256:b3fbe696be1541fbbf031e7c72bbcf819154f01fc0ff6358d08e9ccb1a12b99b

Observation 90c07177-cd3a-4101-8641-b532ae18ea0a · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:56.113423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:56.113423Z digest=sha256:ffa30f79dfc5ddcca6e2e04a96aab0d123b0209bd2ed34d5f39c8a54b0a15246

Observation 814a947a-5024-480f-bbbc-f78e0c64b73d · outbound

This paper cites How to Efficiently Adapt Large Segmentation Model(SAM) to Medical Images.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models How to Efficiently Adapt Large Segmentation Model(SAM) to Medical Images

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:56.274007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:56.274007Z digest=sha256:b8154e247166081006e1465715ec3c42e4cd58859489bef35e30eee183daa552

Observation a2e96d83-8c85-4b3d-8c5a-4c4b26f0f309 · outbound

This paper cites Huang, Y.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Huang, Y

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:56.406571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:56.406571Z digest=sha256:64ac89fb80f2e07d68b73aa27fc81984e2ac163c20da25353844b3b5b497ae79

Observation c4de849f-3f50-4f31-b512-d293228bc8fe · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:59:03.445152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:56.625548Z digest=sha256:e117b0c7cfcae9e40360eb176976b7ac7215a20ed51a35c31759227a8f9eead2

Observation cb381b1f-6502-4a4a-b228-d47deeba0cd3 · outbound

This paper cites Kerbl, G.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Kerbl, G

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:56.769059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:56.769059Z digest=sha256:7a0bdce251fb3197b35eea19765d04e01193037804891d44ce125a01eef5e170

Observation 4cd1b8cd-4a87-4f6c-953d-79834911d265 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:56.895532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:56.895532Z digest=sha256:2d58dcbc8723ea73f830c18fb4f1760593b334d96bf6e78677fae480706e0939

Observation 650cbaaa-7fc6-4ec5-bf33-a3e031ff69b2 · outbound

This paper cites Kornblith, M.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Kornblith, M

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:56.946497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:56.946497Z digest=sha256:45d396fce1940e0f82701f7ecdd4589e9a0f598e0bdd244002d1a9fa8dabeb4f

Observation d88d6dd3-420c-467a-bca7-fe1699492d03 · outbound

This paper cites Wonderland: Navigating 3D Scenes from a Single Image.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Wonderland: Navigating 3D Scenes from a Single Image

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:57.092872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:57.092872Z digest=sha256:bad7f9d8759ea8188b84539930614a8463d3ee795c0933f969011b3124f283fa

Observation f9999628-37dc-41a8-858d-b6df74cf6716 · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:57.226813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:57.226813Z digest=sha256:2c0ef05df9522aa4d6a5eb1fbd2c3dcb0e08e276af2b1752bf84cd53b0b98ea8

Observation 948327a2-94cf-4e92-91a0-12147e52608b · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:59:03.303275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:57.378218Z digest=sha256:48dff08cd55d5eb475aa2a3a7098b16af9a9b2a80596e6e876ecdffe948b57e0

Observation 79a89b01-c357-4afc-90d1-a9c17c84c86d · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:57.624050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:57.624050Z digest=sha256:4be01ef30627b635fdff2c46c31c469f2d7f31ddd0fb9a664f6564c4a161a483

Observation 2861068b-d4e5-4de7-83a3-17defcf5ec11 · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:59:03.158274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:57.806559Z digest=sha256:b596833509964b9fea0c28f4fa2971d33b3caafc59e1360d0e74ba0db5584952

Observation 34e8cc7b-84cc-4e46-9428-ba9a762804f3 · outbound

This paper cites LoRA3D: Low-Rank Self-Calibration of 3D Geometric Foundation Models.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models LoRA3D: Low-Rank Self-Calibration of 3D Geometric Foundation Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:57.941066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:57.941066Z digest=sha256:0a04d42a97895acf2b76ddfb70d65e7688f91d6217627b85b80409801fc71ecb

Observation 2a21b18b-2e71-46bf-ac81-e24fc942ab43 · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:59:03.023638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:58.092897Z digest=sha256:e6bd4d8766abf9495fd0a68a7f17b966c72dbef90c3ca15e95a3f8e909a49d81

Observation f47a902a-a281-45d0-9412-ef8ec2499fa4 · outbound

This paper cites Mahendran and A.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Mahendran and A

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:02.888084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:58.190029Z digest=sha256:47e1be1f5298780423064f977a60ed1298d1847b287459e8752226135b5a6622

Observation be8bc8ef-9ca7-4ffd-a473-154bc18b4658 · outbound

This paper cites Oquab, T.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Oquab, T

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:02.761707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:58.326772Z digest=sha256:51af35b7305ec9cbdaaef4a2b9848ce5b5119d55d6392d04672913627b0a4a54

Observation 527b7160-ebef-46e0-b097-c51c3210dd7c · outbound

This paper cites Peebles and S.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Peebles and S

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:58.423770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:58.423770Z digest=sha256:c9d036782e084af92a0bc7246ba03fbdc6aef427654f839bf1c7ad31d882711c

Observation 28803efb-76e8-4332-a6d4-d2337ab15f35 · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:58.550560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:58.550560Z digest=sha256:225290aca544526f68e91606a04cf5fb4cfeb3642727e11961444f6ad79e73e1

Observation 2b114bc8-d6bf-4dd7-84b6-09c9a796cff2 · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:59:02.606531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:58.647520Z digest=sha256:7be177c9e3922f7fd275e6ed39650897b53443311b8c3bd23699eee9304086f2

Observation ca972199-9970-43bc-a138-abe2203e1904 · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:59:02.433471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:58.788656Z digest=sha256:2150ccb99963bbd0abe1c803455febb30c203e56631d39840b177451d948112a

Observation fa87bace-6e62-4c44-9dba-cd9769ae9f8d · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:59:02.314036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:58.925893Z digest=sha256:033495c70f49f574a1675401828da49119e4b996295f1faac0d242e03d8ea82f

Observation 7a83ecd2-f632-46bb-b42b-6c2076c64569 · outbound

This paper cites Sohl-Dickstein, E.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Sohl-Dickstein, E

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:58.999429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:58.999429Z digest=sha256:50bf38b0486fddc3aa548968ef22457eac83afcfd7b5b0ce929ee6c19107f011

Observation c9a00406-d5f2-4675-bcb9-ec4c402ed411 · outbound

This paper cites Denoising Diffusion Implicit Models.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Denoising Diffusion Implicit Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:59.134015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:59.134015Z digest=sha256:dbf9d6bba9fbf18d328002d50c039c07d354ca3a0d770139a50380126923548c

Observation 20bfeda4-0466-4e40-8923-eeb37a2c6abc · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Score-Based Generative Modeling through Stochastic Differential Equations

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:59.265853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:59.265853Z digest=sha256:7850933c0743667d6ef9b8fde74ce242fc35a2a3d2fb527f756bca1c51d0d15c

Observation 9de21364-8d91-4d7e-b1b0-be0ec841ca9b · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Wan: Open and Advanced Large-Scale Video Generative Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:59.392779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:59.392779Z digest=sha256:7af4d225e91a97ed29df073f073a6cca56606d4dadc312e751b7da4cb0e32a45

Observation 7a6c2743-04ff-4c9a-b596-8b86fc65adfb · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:59:02.142489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:59.552589Z digest=sha256:db30be9df6d6f8a25e79b692b7af62953147a43445ed41f887e2be33fa7d30d2

Observation 68824346-05e2-4fcd-8d76-8c9e55b44672 · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:59:02.012157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:59.675013Z digest=sha256:8dce4ab3f7b60975a452749cc7854be08dbeec463b352d1d9f69d61ee4128848

Observation 1fce435d-6b28-43ea-9bc9-27d8f14af809 · outbound

This paper cites Disney video generation dataset.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Disney video generation dataset

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:01.900089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:59.780924Z digest=sha256:184bd9ed2af10f0226a60515b02faf70aad78cddaef0b801fe097484aa397c03

Observation 0e4e0981-9d19-4072-b796-9033393f27df · outbound

This paper cites Tom and jerry video generation dataset.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Tom and jerry video generation dataset

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:01.770396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:58:59.950113Z digest=sha256:9145fcd29c3088ff29732bc6752808fdd84ea38b1ec20f8069df8d17993c1dc7

Observation 5b90e42e-be20-4fd7-9174-67555a3f6b0d · outbound

This paper cites Xiang, H.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Xiang, H

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:01.644663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:59:00.081997Z digest=sha256:5029fede4a3a2720540d5d459a3fd952c3e03a15d2748fecab96260825088b21

Observation 584051e3-5bcf-4549-af5c-fa5966d48ab5 · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:59:01.537258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:59:00.309419Z digest=sha256:8d6315bf6f9ecf16c7518b2573a892b80f1fc017837159126401b368acf4fbbc

Observation 4265131c-7609-47fe-860d-62cc9724cbc4 · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:59:01.429395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:59:00.471626Z digest=sha256:aeba5f18d4b8736a5d581e6866a7ac551ef63254cf9603bdaf48e7fd7aa0d766

Observation bf9ad171-3aaa-49af-939f-bd57ef2b24f4 · outbound

This paper cites Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:00.595140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:59:00.595140Z digest=sha256:f5e5f964855e8aeced6cd347e3169bb8cada4b7b03111ce99bed12aeac03c64e

Observation 1e6db635-0be2-4ab5-9fc0-9b3f8bd28ef0 · outbound

This paper cites an unresolved cited work.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:59:01.316779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T04:59:00.736880Z digest=sha256:23c89e977afb3d035be162d4094a3740e35662e22b4ec59b90e4fbbd32459475

Observation d3923465-b429-4008-909c-96c613a3f77f · outbound

This paper cites Open-Sora: Democratizing Efficient Video Production for All.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Open-Sora: Democratizing Efficient Video Production for All

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:00.878129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:59:00.878129Z digest=sha256:dec81dda64f0dec98a45d97b966b3c0b89c06823abe90050421cf5beaed74a93

Pith citing papers

Observation 0c581420-517e-46eb-bb5e-b46dd599550e · inbound

Representations Before Pixels: Semantics-Guided Hierarchical Video Prediction cites this paper.

Representations Before Pixels: Semantics-Guided Hierarchical Video Prediction Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:45:58.546740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T16:30:53.578491Z digest=sha256:e50a139561876ef5da461aaf8e86d08f149fa9a56d52908be246453adfbf33b6

Observation 9b9a431c-cd03-4a26-a05d-a30ccbe19679 · inbound

Divide and Conquer: Decoupled Representation Alignment for Multimodal World Models cites this paper.

Divide and Conquer: Decoupled Representation Alignment for Multimodal World Models Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:46:02.839422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T15:51:12.604102Z digest=sha256:a1d74124d92d151540eccb5da2a31031ce9a88eac2ca4189703e60be78bc6e39

Observation e8bcbbe9-8751-4278-abd6-76fd8212487c · inbound

Divide and Conquer: Decoupled Representation Alignment for Multimodal World Models cites this paper.

Divide and Conquer: Decoupled Representation Alignment for Multimodal World Models Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-02T23:57:27.892610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-02T23:56:39.865433Z digest=sha256:577727d4dfb9e66db6281fe0d1923314237071c4759c0a0d1baecb78f21300cb

Observation 01c8bdf5-8ded-4a4d-b33e-bcc1b905f302 · inbound

Tempered Self-Similarity Alignment for Physically Plausible Video Generation cites this paper.

Tempered Self-Similarity Alignment for Physically Plausible Video Generation Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:44:38.388278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-30T11:39:06.597513Z digest=sha256:4b8cdcfa24c67fb3da55d7758cf44b0bceb08b14de55ce7c93bc52aee7e42754

Observation a4091bff-f3fe-4c4d-8b34-ebe2ff424477 · inbound

Making Foresight Actionable: Repurposing Representation Alignment in World Action Models cites this paper.

Making Foresight Actionable: Repurposing Representation Alignment in World Action Models Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-07-03T11:08:03.363913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-27T09:40:54.963759Z digest=sha256:7543d341f2ca1d238ad4cff5e86f642830dcd0af4d058956776f7c9fe71ff237