Pith. sign in

Paper Citation Record · LEDGER

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation

As of 8 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2607.18789.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.18789 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T14:24:46.123867Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 693973c7-37fe-447c-abe1-2a7cd9eabb6e · outbound

This paper cites VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.067822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.067822Z digest=sha256:1e6f3862fc3fea7e6317393121c1580e9a26d9935de0bb7ffbdeea9bf4d50cfd

Observation 0d35a70b-b969-4c19-8217-0940589ede0c · outbound

This paper cites The Pile: An 800GB Dataset of Diverse Text for Language Modeling.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation The Pile: An 800GB Dataset of Diverse Text for Language Modeling

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.074247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.074247Z digest=sha256:baa3c41412b132754fcf3f78a1ff323a7f1ebf9d0efd55b03656dcefd7f8586e

Observation 8a13afa6-1a26-45c5-acc6-b8c4bcb271d9 · outbound

This paper cites Imagen Video: High Definition Video Generation with Diffusion Models.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Imagen Video: High Definition Video Generation with Diffusion Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.080717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.080717Z digest=sha256:f277c856c9c71c4b604977c5cca18a8826299a68d858eb60e82e7abf0a4c14fe

Observation 413adb3b-7dc5-4ed3-b8d5-8b34deb0ba8c · outbound

This paper cites Latte: Latent Diffusion Transformer for Video Generation.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Latte: Latent Diffusion Transformer for Video Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.092789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.092789Z digest=sha256:a31153724018dda29760f99ee5d242b1c50727b466ef6faea1a3bd917e3fdc15

Observation c9b6c558-2227-4405-a9c4-9c87dd2f600e · outbound

This paper cites Scaling Data-Constrained Language Models.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Scaling Data-Constrained Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.095609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.095609Z digest=sha256:cc9994043a00498a46be157d7514e0ff95c4ddadea8aab08d2835b52404e9a46

Observation 58b82d0c-4db1-4c0e-a33d-613e99a7c9d2 · outbound

This paper cites Movie Gen: A Cast of Media Foundation Models.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Movie Gen: A Cast of Media Foundation Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.098413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.098413Z digest=sha256:ce9c7737f7147b8836ff1af407dfbc3eaf01d0195e2b070a4177a46a1181beb4

Observation ae690f68-872e-40e2-b250-fb8819585b03 · outbound

This paper cites The layout bet, June 2026.https://reve.com.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation The layout bet, June 2026.https://reve.com

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.101322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.101322Z digest=sha256:b16e5c555cd4bd356096e52984b6434e270b442843442abc9c2ec02dd99f390a

Observation dc336a12-9133-4c2f-ac8d-cf08f5064aef · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.104418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.104418Z digest=sha256:992bea2bae63367f3035434cd5dab0460b4fa014c20bfedeeaa02ae267e95141

Observation 2d70f25d-88dd-472f-a9bd-ff865892816d · outbound

This paper cites Kimi-VL Technical Report.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Kimi-VL Technical Report

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.107195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.107195Z digest=sha256:c1f635f468a1ca3f6834be06accca18df7a2ef146d2eeb86d0a101c8e1721100

Observation 2c439224-b816-4c81-a181-933ffe2909de · outbound

This paper cites InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.109969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.109969Z digest=sha256:6c5d9ecbcff4575a97c4214e253ff3949abd55feb65c21dbac75fbf7ed4a1fc7

Observation 74041b2e-5c19-4a61-af00-d706787a76e5 · outbound

This paper cites Qwen3 Technical Report.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Qwen3 Technical Report

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.112606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.112606Z digest=sha256:3486ec2535bdb2797fdd110d95d4df8e77ed556abf027bf2352850055e451f33

Observation 769ef145-26bf-4f10-98fb-44334ddf91d5 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.115337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.115337Z digest=sha256:025c13769c45333e0fdea21a72cc687445489ea78da3c2af04f680fe38e69bc3

Observation 22de8f3c-08d4-4808-a226-b81939f715aa · outbound

This paper cites Data Mixing Laws: Optimizing Data Mixtures by Predicting Language Modeling Performance.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Data Mixing Laws: Optimizing Data Mixtures by Predicting Language Modeling Performance

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.118299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.118299Z digest=sha256:ba6dcd064a10d3c06a98c5f995c60cf98813c30051ee6183cc2f7687f0593787

Observation b3bf8dc0-9717-4763-8671-a51cf65f0121 · outbound

This paper cites Open-Sora: Democratizing Efficient Video Production for All.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Open-Sora: Democratizing Efficient Video Production for All

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.121154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.121154Z digest=sha256:1457b5a0d7e810df25da945163a32f66b41267fe68e99b2a64cd99b30ee3ab00

Observation 3d381a34-1eec-44a7-ba60-4cb63def9d77 · outbound

This paper cites 4 lists all attributes used in the Moving Alphabet dataset and their possible values.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation 4 lists all attributes used in the Moving Alphabet dataset and their possible values

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.123867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.123867Z digest=sha256:b405cb0ebe53c7043f67c77ff705b66d5d30b355891ff7b0ca6671c34427c075

Observation b2c9eb7a-83df-48e9-b52f-0adfc932bb45 · outbound

This paper cites Scaling Laws for Neural Language Models.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Scaling Laws for Neural Language Models

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.086843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.086843Z digest=sha256:ddc763195e8aa078d46793e2d0f91a1e6068deafa12ec8312e9098d6326cbece

Observation b6b3cb75-426b-4279-ab51-ca0b52e2215d · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.089947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.089947Z digest=sha256:7deda05cd44bc2800ff74f18ce4c9b8d07c701e14f7a4b97d762187ed4be0859

Observation ccea6b25-c376-4495-8069-b4f379bc0dea · outbound

This paper cites Classifier-Free Diffusion Guidance.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Classifier-Free Diffusion Guidance

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.077417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.077417Z digest=sha256:02836a642b491aa6c013febab5da653aeb7bb10d7c1a88784030d845a26169fa

Observation 73280caa-25ab-44dd-b35d-df6f459a6509 · outbound

This paper cites T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.083996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.083996Z digest=sha256:225bd00d29adb40efe9121d66b76253b88ee10300df39beae4feb3b9f04963d2

Observation fecae1f6-d53c-4734-abb8-b07da78f8706 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.064539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.064539Z digest=sha256:1fed3a4e3e519f403e13c4f20c1ef63246a1aa8d22f9bfb53d974fb93b8cd98d

Observation 1607a6e6-7344-473d-bc6e-1917f964994a · outbound

This paper cites Lumiere: A Space-Time Diffusion Model for Video Generation.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Lumiere: A Space-Time Diffusion Model for Video Generation

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.060545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.060545Z digest=sha256:d470e228d591c088149475e490ed39f9c8aeadf4126e867e335cc8b453bfc900

Observation f06cfb1f-d129-48dc-bc79-00a4a6f15663 · outbound

This paper cites TinyStories: How Small Can Language Models Be and Still Speak Coherent English?.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation TinyStories: How Small Can Language Models Be and Still Speak Coherent English?

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.071187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.071187Z digest=sha256:6f90b21fe219026fa7eaafb51de139635bcae20c3cab015b1d222528570d3a6a

Pith citing papers

No inbound Pith citation observations are available.