Pith. sign in

Paper Citation Record · LEDGER

Vidu S1: A Real-Time Interactive Video Generation Model

As of 18 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 2 inbound Pith citation observations for arXiv:2607.03118.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.03118 v2

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T08:56:31.626723Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T04:52:49.585526Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T04:52:50.130247Z

Reference resolution

61 of 61 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved60
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3e363ebb-bca1-4d0b-a364-4ccab1a3f003 · outbound

This paper cites Video generation models as world simulators.

Vidu S1: A Real-Time Interactive Video Generation Model Video generation models as world simulators

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.463504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.463504Z digest=sha256:f2cb05d7076dd06686f88facfdfa782af0b32c78b7cd0938bc21150959729216

Observation 3aeb8d28-103a-4275-8e24-ceb48a86870c · outbound

This paper cites Veo: A text-to-video generation system.

Vidu S1: A Real-Time Interactive Video Generation Model Veo: A text-to-video generation system

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.467458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.467458Z digest=sha256:34eeb0513d518c9b19ef7a873068e480dcb6d8c7b9a6fff3b6301d664056af84

Observation cab3bf8a-6dcd-4df4-9eef-844027cd223d · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Vidu S1: A Real-Time Interactive Video Generation Model Wan: Open and Advanced Large-Scale Video Generative Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.470703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.470703Z digest=sha256:7e79e1dfdac21d1a700251e170cdd4f9c98fa54e0649532c24c347e6b41acfb0

Observation d3c08625-a7aa-4c0a-be1f-03b1b8be0acf · outbound

This paper cites Seedance 2.0: Advancing Video Generation for World Complexity.

Vidu S1: A Real-Time Interactive Video Generation Model Seedance 2.0: Advancing Video Generation for World Complexity

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.474238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.474238Z digest=sha256:b8c2968ae99d7cef2de20d6a9a8370b844bc1b6f123478587650650a04a9f076

Observation ecccb5ee-80b0-48a0-a499-e5cf55ef0736 · outbound

This paper cites Diffusion forcing: Next-token prediction meets full-sequence diffusion.

Vidu S1: A Real-Time Interactive Video Generation Model Diffusion forcing: Next-token prediction meets full-sequence diffusion

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.477580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.477580Z digest=sha256:2e8de16c3e61cd73dd20c589cbca634490b91f5894dac0f78ab452d4709dc2d8

Observation b6abb58f-af3c-4794-80ac-d74398402727 · outbound

This paper cites Fifo-diffusion: Generating infinite videos from text without training.

Vidu S1: A Real-Time Interactive Video Generation Model Fifo-diffusion: Generating infinite videos from text without training

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.480537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.480537Z digest=sha256:27daaccaf0b88aa1f9a88c0c24bc37cf907b2539c5c754046c91732b0ef962cd

Observation ff47484f-91e5-47ee-ace7-21a04e51327f · outbound

This paper cites Streamingt2v: Consistent, dynamic, and extendable long video generation from text.

Vidu S1: A Real-Time Interactive Video Generation Model Streamingt2v: Consistent, dynamic, and extendable long video generation from text

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.483719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.483719Z digest=sha256:36bba18eae439e6eaae0260de3dfede68f16c48acd69ae128b7da9cab5955670

Observation 581655c9-ad42-4b91-8ac9-1c790b4fdbe3 · outbound

This paper cites History-Guided Video Diffusion.

Vidu S1: A Real-Time Interactive Video Generation Model History-Guided Video Diffusion

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.486576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.486576Z digest=sha256:79fec77fc4032048e78d4d3432c5f05cf49c3b20eabbe2c94a60dc1289f178cb

Observation c4b0d531-7a5c-4974-a016-ad469f72aa68 · outbound

This paper cites Rolling Forcing: Autoregressive Long Video Diffusion in Real Time.

Vidu S1: A Real-Time Interactive Video Generation Model Rolling Forcing: Autoregressive Long Video Diffusion in Real Time

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.489618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.489618Z digest=sha256:f29bf6d599f2747b9d3f082f0a0249480c6660ea1665235aea67dca48cb6d5c1

Observation 0c05cea9-6547-4cc4-ac7f-d63ab2c9fc8e · outbound

This paper cites Ar-diffusion: Asynchronous video generation with auto-regressive diffusion.

Vidu S1: A Real-Time Interactive Video Generation Model Ar-diffusion: Asynchronous video generation with auto-regressive diffusion

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.492622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.492622Z digest=sha256:b055a7ecd8b26e15ab959b2049d771274b15a8ffb3b956b664efe28d716a6159

Observation b2c57829-ae31-4cc4-ba59-fdf2bc436448 · outbound

This paper cites Progressive autoregressive video diffusion models.

Vidu S1: A Real-Time Interactive Video Generation Model Progressive autoregressive video diffusion models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.495691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.495691Z digest=sha256:562ca7abfeae28dd76e0271c2e0b91952d860ceeccb7f2003b295ceb51ebfbf7

Observation 5d57c7ad-8073-405f-8070-3b48af4375bd · outbound

This paper cites Streamdiffusionv2: A streaming system for dynamic and interactive video generation.

Vidu S1: A Real-Time Interactive Video Generation Model Streamdiffusionv2: A streaming system for dynamic and interactive video generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.498709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.498709Z digest=sha256:84dc04d530a89316434aa96314d00eea76cb93e5e8425420d0146e5c19c4f2f7

Observation c4a4b9e2-570e-40f6-b829-1a6faac2be07 · outbound

This paper cites From slow bidirectional to fast autoregressive video diffusion models.

Vidu S1: A Real-Time Interactive Video Generation Model From slow bidirectional to fast autoregressive video diffusion models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.501590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.501590Z digest=sha256:e61d3eb14a12df1f268a4c51f745f50b159893f390dad5912386ed5f8c853525

Observation 067a425a-2ee4-429b-8fc8-d83be035a332 · outbound

This paper cites Self forcing: Bridging the train-test gap in autoregressive video diffusion.

Vidu S1: A Real-Time Interactive Video Generation Model Self forcing: Bridging the train-test gap in autoregressive video diffusion

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.504457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.504457Z digest=sha256:7e92703159ec013ed32a75b5b1b187efcec0e60942d35c8f66b0635520c6d22d

Observation 3248f336-7196-4c2b-a7a5-984225c59fa5 · outbound

This paper cites Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation.

Vidu S1: A Real-Time Interactive Video Generation Model Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.507401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.507401Z digest=sha256:805267c40756656cac6fb7b8183c8a03fe6516459767404f8063b12673f420ee

Observation 348b9833-f2b8-4e1b-a326-4bcb07919cf1 · outbound

This paper cites Causal Forcing++: Scalable Few-Step Autoregressive Diffusion Distillation for Real-Time Interactive Video Generation.

Vidu S1: A Real-Time Interactive Video Generation Model Causal Forcing++: Scalable Few-Step Autoregressive Diffusion Distillation for Real-Time Interactive Video Generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.510408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.510408Z digest=sha256:216e76ae716d3d0ad07f70b0d90ef3b6ea6137c2c6066b529e1694f89fd0badd

Observation bc26c285-1f33-44d8-af4e-02ca1c9744fc · outbound

This paper cites LongLive: Real-time Interactive Long Video Generation.

Vidu S1: A Real-Time Interactive Video Generation Model LongLive: Real-time Interactive Long Video Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.513461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.513461Z digest=sha256:e17bde09632a406128bfcf17ab265ec70c9fcaa6c0414309a6531d2f22bd69ae

Observation a82db0fd-a0c7-46b3-b959-ff2f213dfb0c · outbound

This paper cites MAGI-1: Autoregressive Video Generation at Scale.

Vidu S1: A Real-Time Interactive Video Generation Model MAGI-1: Autoregressive Video Generation at Scale

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.516388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.516388Z digest=sha256:3b1c24c410fc3fa1d518f9810635aa8c10c3d88cca592340008d74f18c101eac

Observation f7ef0785-6436-4487-a2d4-6e82a188ba99 · outbound

This paper cites SkyReels-V2: Infinite-length Film Generative Model.

Vidu S1: A Real-Time Interactive Video Generation Model SkyReels-V2: Infinite-length Film Generative Model

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.519369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.519369Z digest=sha256:d891b47869cae270a8b1bb786b3de1c3f39b1e47b636d0160ca97e80e207a01d

Observation 2218d08e-c25b-4132-993c-bcb47a358fec · outbound

This paper cites Packing input frame context in next-frame prediction models for video generation.

Vidu S1: A Real-Time Interactive Video Generation Model Packing input frame context in next-frame prediction models for video generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.522348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.522348Z digest=sha256:fda06a4e1e03020c2cb2c69b58d6c0c72fec64715c19e65d4238cea6a412b44a

Observation f9925a27-de13-409c-9838-6f113ba0d756 · outbound

This paper cites Wan-S2V: Audio-Driven Cinematic Video Generation.

Vidu S1: A Real-Time Interactive Video Generation Model Wan-S2V: Audio-Driven Cinematic Video Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.525012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.525012Z digest=sha256:40244d23d9443853e7ad30a65a611d6c96ceefa631a637843777069213ab0718

Observation 2bd93fa1-c131-48d6-95ee-b48d84c7e2b8 · outbound

This paper cites Turbodiffusion: Accelerating video diffusion models by 100-200 times.

Vidu S1: A Real-Time Interactive Video Generation Model Turbodiffusion: Accelerating video diffusion models by 100-200 times

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.527895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.527895Z digest=sha256:789f72ee547141dd1d3ad4e04efbf7e4087ad8527aee0eb9018239967475579e

Observation 91a812e8-d2b7-461d-ba52-e2324aa14959 · outbound

This paper cites TurboServe: Serving Streaming Video Generation Efficiently and Economically.

Vidu S1: A Real-Time Interactive Video Generation Model TurboServe: Serving Streaming Video Generation Efficiently and Economically

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.530403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.530403Z digest=sha256:aa2922db2a13d28218565c1030b47cb68119e136985ba9f0ec72822aadbe0774

Observation 80dde387-c348-401a-89c8-e55c6bc8a88a · outbound

This paper cites Qwen3-Omni Technical Report.

Vidu S1: A Real-Time Interactive Video Generation Model Qwen3-Omni Technical Report

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.533158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.533158Z digest=sha256:014c6bbc9d18c3001ac0fba0f781825ce27e7ea6801b5e98afde15477d53fd6f

Observation c5e46572-31eb-402d-8272-29e4bf9de353 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Vidu S1: A Real-Time Interactive Video Generation Model Gemini: A Family of Highly Capable Multimodal Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.535865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.535865Z digest=sha256:0caf4a4a60a0e719de929b9dd70e8a5b35a700eca125dbb5446f5570553d0371

Observation 90abec4b-4c0b-4ca5-830a-35c121473a6d · outbound

This paper cites Ca2-vdm: Efficient autoregres- sive video diffusion model with causal generation and cache sharing, 2025.

Vidu S1: A Real-Time Interactive Video Generation Model Ca2-vdm: Efficient autoregres- sive video diffusion model with causal generation and cache sharing, 2025

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.538797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.538797Z digest=sha256:6463b59f56d16edf3a138e7c0d73e6f1601b4b522c3db573ef92143e0b065125

Observation 844b100c-1434-4ba1-be53-fb5578ab5542 · outbound

This paper cites Pyramidal flow matching for efficient video generative modeling.

Vidu S1: A Real-Time Interactive Video Generation Model Pyramidal flow matching for efficient video generative modeling

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.541421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.541421Z digest=sha256:628321e0e7540fa9ba7f067ed5ac9392261c0b51c6e424d75b06dc4d7ac6f8df

Observation f53fc15a-4014-4084-926d-6038ff82afdd · outbound

This paper cites One-step diffusion with distribution matching distillation.

Vidu S1: A Real-Time Interactive Video Generation Model One-step diffusion with distribution matching distillation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.543850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.543850Z digest=sha256:97cfa0d60649d0c19e13cb05544ff806eff474c4b496facc92550b4360a7f554

Observation e483012a-357b-4ffd-a522-b4d0a5dd4aa3 · outbound

This paper cites Phased consistency models.

Vidu S1: A Real-Time Interactive Video Generation Model Phased consistency models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.546327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.546327Z digest=sha256:76835fa3d2554de4e75bacff96d1c3d6fe70b5aee5c98c829724baa524490c74

Observation 32310bdd-18a0-4d66-8659-30afe28ceb70 · outbound

This paper cites Efficient streaming language models with attention sinks.

Vidu S1: A Real-Time Interactive Video Generation Model Efficient streaming language models with attention sinks

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.548773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.548773Z digest=sha256:37e858d6761477ab1d202bbeba0b18e0194d1dc433950de7213f20298886b845

Observation a3526f36-03ed-4c48-8697-14a7388cfa1f · outbound

This paper cites Infinity-rope: Action-controllable infinite video generation emerges from autoregressive self-rollout.

Vidu S1: A Real-Time Interactive Video Generation Model Infinity-rope: Action-controllable infinite video generation emerges from autoregressive self-rollout

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.551331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.551331Z digest=sha256:4dbcd6e5bd0bc67cc040006b16e1778f8511fd07c3aefc805d70572064376289

Observation 20b37558-ec6d-4ac8-856c-26e2e753c268 · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.

Vidu S1: A Real-Time Interactive Video Generation Model Roformer: Enhanced transformer with rotary position embedding

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.553831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.553831Z digest=sha256:0c1f492592d345187a4ab3b63d12253f21982e26511d262157677ca181e50596

Observation d3d86871-2b0a-4c2d-b6c6-f19a4c8bd6ab · outbound

This paper cites Deep forcing: Training-free long video generation with deep sink and participative compression.

Vidu S1: A Real-Time Interactive Video Generation Model Deep forcing: Training-free long video generation with deep sink and participative compression

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.556217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.556217Z digest=sha256:3ed688bbeff7dad00bdd3cdebe44de99735ae45b3c4e0c61d8655140be2185e9

Observation 3a63176c-696c-445e-ba39-027af13592af · outbound

This paper cites Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion.

Vidu S1: A Real-Time Interactive Video Generation Model Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.558808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.558808Z digest=sha256:16edd8fff99a5e2120594e5df5cbb35a4a427eb97c6262d851a896907559ed5a

Observation fc8ac0a1-ce5a-4ddc-a3f6-eeeeb1e2d13f · outbound

This paper cites Grounded Forcing: Bridging Time-Independent Semantics and Proximal Dynamics in Autoregressive Video Synthesis.

Vidu S1: A Real-Time Interactive Video Generation Model Grounded Forcing: Bridging Time-Independent Semantics and Proximal Dynamics in Autoregressive Video Synthesis

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.561574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.561574Z digest=sha256:cdcd6022cd187d450360134d963f5548de389c518ef0d7d821f005281943c571

Observation f8ad4ad4-25dd-4bc8-9e75-e6d9a8d17dd9 · outbound

This paper cites Memrope: Training-free infinite video generation via evolving memory tokens.

Vidu S1: A Real-Time Interactive Video Generation Model Memrope: Training-free infinite video generation via evolving memory tokens

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.564278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.564278Z digest=sha256:de0b82024133499de0212c0a73395bcc5c4e7bb247ebb0c2388cfa320dd58123

Observation 40026c4a-3386-4c42-9dce-a3c30f010f97 · outbound

This paper cites Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length.

Vidu S1: A Real-Time Interactive Video Generation Model Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.566881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.566881Z digest=sha256:9e84d7df1a6cebcb70ca6ad8278117b4c1e2fde8a06241af11206fe7e691b8e2

Observation 9b39dc43-2355-46b4-a07e-7ba9e4f8bba1 · outbound

This paper cites Causal-rCM: A Unified Teacher-Forcing and Self-Forcing Open Recipe for Autoregressive Diffusion Distillation in Streaming Video Generation and Interactive World Models.

Vidu S1: A Real-Time Interactive Video Generation Model Causal-rCM: A Unified Teacher-Forcing and Self-Forcing Open Recipe for Autoregressive Diffusion Distillation in Streaming Video Generation and Interactive World Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.569666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.569666Z digest=sha256:89b152ce18a944a252d81d0ac05a16532d6b8545bec3ed20df89973821ec71b1

Observation 4d6cf2d9-c8ad-4f51-90cd-55ef574df26e · outbound

This paper cites LPM 1.0: Video-based Character Performance Model.

Vidu S1: A Real-Time Interactive Video Generation Model LPM 1.0: Video-based Character Performance Model

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.572316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.572316Z digest=sha256:077b223efca8b6d27ba685772949211282f1fa570a15d6abc1094b34316b820d

Observation a5fbe3e1-45f0-46e9-9495-94e6005945bc · outbound

This paper cites Efficient attention methods: Hardware-efficient, sparse, compact, and linear attention.

Vidu S1: A Real-Time Interactive Video Generation Model Efficient attention methods: Hardware-efficient, sparse, compact, and linear attention

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.575254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.575254Z digest=sha256:ff0f98a7f0836b5950a2ed16341ab5074dfe342acd33ec9b0e93ec3ea56ac445

Observation aea88014-b7a7-4c7a-b921-e7557d3be766 · outbound

This paper cites Sageattention: Accurate 8-bit attention for plug-and-play inference acceleration.

Vidu S1: A Real-Time Interactive Video Generation Model Sageattention: Accurate 8-bit attention for plug-and-play inference acceleration

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.577869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.577869Z digest=sha256:271f7a41e6a21be9123798647bfd8ab1c6f6eac9045f3f0b52c5fda8e5724efe

Observation 89b775ef-041d-4a9c-b2e2-24aa4c7dd394 · outbound

This paper cites Sageattention2: Efficient attention with thorough outlier smoothing and per-thread int4 quantization.

Vidu S1: A Real-Time Interactive Video Generation Model Sageattention2: Efficient attention with thorough outlier smoothing and per-thread int4 quantization

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.580166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.580166Z digest=sha256:79eb7e2d175ce117cc0b7bf9cb0647e95c3bed007b2a46aef098b25c089693d9

Observation c63f6e16-8a37-4f22-ad1a-41449ec5fce3 · outbound

This paper cites SageAttention2++: A More Efficient Implementation of SageAttention2.

Vidu S1: A Real-Time Interactive Video Generation Model SageAttention2++: A More Efficient Implementation of SageAttention2

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.582638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.582638Z digest=sha256:7846b20dee7ec7876be92348e7e61545284614afd80ab66aaaf7e55a3170b797

Observation 12bb74d1-4e59-4bca-9baa-a160ba739091 · outbound

This paper cites Sageattention3: Microscaling fp4 attention for inference and an exploration of 8-bit training.

Vidu S1: A Real-Time Interactive Video Generation Model Sageattention3: Microscaling fp4 attention for inference and an exploration of 8-bit training

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.585274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.585274Z digest=sha256:01ccfc85d89dedb7a1329d1b9c9daf250a97b8a5a5229888fa92299d6e2464b0

Observation 89b897d3-425f-466e-9771-2ae98351a798 · outbound

This paper cites Sagebwd: A trainable low-bit attention.

Vidu S1: A Real-Time Interactive Video Generation Model Sagebwd: A trainable low-bit attention

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.587582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.587582Z digest=sha256:e75f1ba64165ece44ea7637d14ad83d9319e3b9e5d16b26cc32c859eca0db1e7

Observation e2343944-edae-4aab-9f22-0a80487b7aca · outbound

This paper cites Spargeattention: Accurate and training-free sparse attention accelerating any model inference.

Vidu S1: A Real-Time Interactive Video Generation Model Spargeattention: Accurate and training-free sparse attention accelerating any model inference

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.589966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.589966Z digest=sha256:79f31fd6248009336e3184dd73c0d6baa1199ed31c7c67595936d8719573f702

Observation 585f5161-1ea3-4582-bcb9-015ff28b0062 · outbound

This paper cites Spargeattention2: Trainable sparse attention via hybrid top-k+ top-p masking and distillation fine-tuning.

Vidu S1: A Real-Time Interactive Video Generation Model Spargeattention2: Trainable sparse attention via hybrid top-k+ top-p masking and distillation fine-tuning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.592261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.592261Z digest=sha256:896602dc2141601127ddfeeb8fc3dddfbd83723b2d18a73dae9df2ac566028ee

Observation 26d5b5f4-4764-4b0e-a758-e474bc62ee61 · outbound

This paper cites Gonzalez, Jun Zhu, and Jianfei Chen.

Vidu S1: A Real-Time Interactive Video Generation Model Gonzalez, Jun Zhu, and Jianfei Chen

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.594594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.594594Z digest=sha256:8b3e2c091dec6ac58889b796ae395998bb396c1d0bdc146f42aef5eb5a82e1b3

Observation e94ba322-1a22-4e46-817b-4fe618e2438c · outbound

This paper cites Sla2: Sparse-linear attention with learnable routing and qat.

Vidu S1: A Real-Time Interactive Video Generation Model Sla2: Sparse-linear attention with learnable routing and qat

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.596847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.596847Z digest=sha256:42789a3d8a68246082361b1dbb81e751e119a54922ba6a58d4753dec6a3fbf76

Observation 880d7b0c-8379-4b62-ba22-fee057930146 · outbound

This paper cites DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models.

Vidu S1: A Real-Time Interactive Video Generation Model DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.599291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.599291Z digest=sha256:9eef94e53d5b25b9a63f2605d0fde7871fc8d7bb5bd2fc63a4188c9c3c6bf671

Observation ec607be8-1c1d-4669-bcc2-c93d41decb29 · outbound

This paper cites Flow-guided one-shot talking face generation with a high-resolution audio-visual dataset.

Vidu S1: A Real-Time Interactive Video Generation Model Flow-guided one-shot talking face generation with a high-resolution audio-visual dataset

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.601964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.601964Z digest=sha256:fc2afad0dc3d03aa4c4c5a04f2f57cc95e7f375c9fa2732c63b45f843d45e763

Observation b0ef7be3-eaa4-4688-8e47-b50b0d1cae7e · outbound

This paper cites Heygen ai video avatar.

Vidu S1: A Real-Time Interactive Video Generation Model Heygen ai video avatar

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.604304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.604304Z digest=sha256:3a8302a46e3da2c019297a782b230af015abd914617b3a051ee190abb549c734

Observation a1cfbc37-775c-4c06-b731-a8d685c99c37 · outbound

This paper cites Lemonslice studio: Create talking and singing ai avatar videos.

Vidu S1: A Real-Time Interactive Video Generation Model Lemonslice studio: Create talking and singing ai avatar videos

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.606615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.606615Z digest=sha256:9b23a9141323c78fd4df53529d5de2d3bdbed220109c76e4e8b5beff3a507956

Observation 21712d40-31e3-45b6-a168-1421154120c9 · outbound

This paper cites Klingavatar 2.0 technical report, 2025.

Vidu S1: A Real-Time Interactive Video Generation Model Klingavatar 2.0 technical report, 2025

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.611583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.611583Z digest=sha256:26b6daa60afb122bc1cda4e6549161d8899f0f6494fbc545211f0ef5bd224e41

Observation 0d4e942e-947d-46e0-95e1-092e8fcd4b2d · outbound

This paper cites OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation.

Vidu S1: A Real-Time Interactive Video Generation Model OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.614075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.614075Z digest=sha256:a06a2ef4863014ae1d43059b7ef94f2bd6c2aca719e8a01530c7126972e3337f

Observation 8f59d1de-8a7f-4be0-ae84-aaea78b0a6da · outbound

This paper cites Hallo3: Highly dynamic and realistic portrait image animation with video diffusion transformer.

Vidu S1: A Real-Time Interactive Video Generation Model Hallo3: Highly dynamic and realistic portrait image animation with video diffusion transformer

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.616682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.616682Z digest=sha256:da1d54f74601c13620d76f76123e913d1cd4c16b362f45fc80ba98ec47d92290

Observation 62419ab1-da7b-4530-bc99-fb8a4f8b7e7d · outbound

This paper cites StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation.

Vidu S1: A Real-Time Interactive Video Generation Model StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.619017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.619017Z digest=sha256:3f321d58081a9782ff63878ce69f88a2fc965a98258c03fbe513f32a26fe053e

Observation cf1e41b7-bd69-4204-bff0-6aa2dc81e491 · outbound

This paper cites Arcface: Additive angular margin loss for deep face recognition.

Vidu S1: A Real-Time Interactive Video Generation Model Arcface: Additive angular margin loss for deep face recognition

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.621637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.621637Z digest=sha256:5343a117aaece50ba4eb4ea5dc5c433f1cf829c86338a9e535cc06b05fc2c676

Observation e8267e36-cfb7-4f07-ad67-2fe614302a9a · outbound

This paper cites A lip sync expert is all you need for speech to lip generation in the wild.

Vidu S1: A Real-Time Interactive Video Generation Model A lip sync expert is all you need for speech to lip generation in the wild

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.624150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.624150Z digest=sha256:0191a67bfb45249858382828de7fa15fc16e693e1202d7bad9bd1f414f10e9d7

Observation 0a2c29d2-6d73-401b-937d-778037eea38a · outbound

This paper cites Exploring video quality assessment on user generated contents from aesthetic and technical perspectives.

Vidu S1: A Real-Time Interactive Video Generation Model Exploring video quality assessment on user generated contents from aesthetic and technical perspectives

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.626723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.626723Z digest=sha256:a6abce4e1fc0b48b5639c0e23e2cdc43f51c039ce5896927c249db2fb77cd44f

Observation dea6cf42-a1ea-4d65-850d-f2713d97858c · outbound

This paper cites an unresolved cited work.

Vidu S1: A Real-Time Interactive Video Generation Model Unresolved cited work

Reference 2026

Resolution
parse uncertain
no resolver link, observed 2026-08-02T08:56:31.609050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.609050Z digest=sha256:1b526e27dd7c2442adb0baabaaf720c6691b7251c41268721849bd7d213c19e1

Pith citing papers

Observation 07857c41-bc51-4ef3-80d4-f446320f17d1 · inbound

Visko Orbis 1.0: A Live Model for Real-Time Interactive Long Video Generation cites this paper.

Visko Orbis 1.0: A Live Model for Real-Time Interactive Long Video Generation Vidu S1: A Real-Time Interactive Video Generation Model

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-30T23:47:57.729629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T23:47:57.729629Z digest=sha256:8a3d041c731a5c47f29580825c8b193c4159bd2bab49a8f845b47fdf75edc8eb

Observation 8c4e5496-cf18-4387-9ff8-e09321f7374c · inbound

JoyAI-Video-Edit: Real-Time Open-Ended Video Editing with Autoregressive Diffusion cites this paper.

JoyAI-Video-Edit: Real-Time Open-Ended Video Editing with Autoregressive Diffusion Vidu S1: A Real-Time Interactive Video Generation Model

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-05T04:52:50.219179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T04:52:49.585526Z digest=sha256:6b23b06c5fe3d28d0972c0ee1f3bf90155ca1821d7fedd2790f1055d48e0651c