Pith. sign in

Paper Citation Record · LEDGER

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices

As of 9 August 2026, this Paper Citation Record lists 86 of 86 outbound references and 4 inbound Pith citation observations for arXiv:2502.04363.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.04363 v2

Coverage vector

measured 86 of 86 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T10:46:29.637028Z

measured 90 of 90 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T08:21:44.482775Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T14:44:59.784059Z

Reference resolution

86 of 86 outbound references displayed

  • verified exact2
  • verified fuzzy33
  • unresolved51
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5da9d11d-6aab-484b-847c-515295b1fcb9 · outbound

This paper cites iphone 15 pro—technical specifications, 2023.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices iphone 15 pro—technical specifications, 2023

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.365358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.365358Z digest=sha256:90f4d8681eac68a675d87b3b39e260c33493a7199846adaf80209dfe1a86510a

Observation 5b08b0ec-586f-4a16-bce6-156c0db86aa4 · outbound

This paper cites Swift, 2024.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Swift, 2024

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.369925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.369925Z digest=sha256:df4ce3afb02b9ceb36041dcf99a1000cb7d10841c8bc579d86b10122f68fe532

Observation 65e576f8-e1ce-4dbe-876d-ff6a956d5a1a · outbound

This paper cites A discussion on euler method: A review.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices A discussion on euler method: A review

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.373458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.373458Z digest=sha256:fc78b0f962b231abaa017aeb5a55bd3b266361010ed5f01af700690946060400

Observation e888b8db-7500-4ae4-a049-d8c98beee61e · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.376984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.376984Z digest=sha256:465668a72c73bb9d30ecb30a2e781f9cf04f11dc2c003f68004a1ca7fc35bafa

Observation c049aab2-8e65-474c-8cdc-4634c85792e7 · outbound

This paper cites Align your latents: High-resolution video synthesis with la- tent diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Align your latents: High-resolution video synthesis with la- tent diffusion models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.381180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.381180Z digest=sha256:3c19867266a2e1c3d5b470e8f095868df6252a640b132e8026b93c929fcb4e48

Observation a3d877ef-dded-4ebd-b451-eb989955aa70 · outbound

This paper cites Token Merging: Your ViT But Faster.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Token Merging: Your ViT But Faster

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.384729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.384729Z digest=sha256:9b58ebf44fb8b4697bbdf141060dd28093eff86735a4affa1e84cc6f68f5767b

Observation d5265d9a-2cf8-42fd-acb9-d7572fc2fb97 · outbound

This paper cites Token merging: Your ViT but faster.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Token merging: Your ViT but faster

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.389029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.389029Z digest=sha256:9e584c73d6da2ac432a5294b50ab71d69028c2e5f9d14f298587cf11e35738b4

Observation da626f79-ee59-4c72-ba06-fd2c0ccfb6d5 · outbound

This paper cites EdgeFusion: On-Device Text-to-Image Generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices EdgeFusion: On-Device Text-to-Image Generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.392456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.392456Z digest=sha256:a4ce62550cacea7a48d31bf6ce317d598f0120d4337614d4fdaf52444a54771c

Observation b4405442-c3de-4264-b25e-9f6c9170d8b0 · outbound

This paper cites Tempme: Towards the explain- ability of temporal graph neural networks via motif discov- ery.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Tempme: Towards the explain- ability of temporal graph neural networks via motif discov- ery

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.396621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.396621Z digest=sha256:2068989e1c551650c98f9655972e872021be567c158c7a1374b64ddd4548c24d

Observation 206bfa80-3ccf-4ac1-9169-6c60e56dfd6e · outbound

This paper cites Neural ordinary differential equa- tions.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Neural ordinary differential equa- tions

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.621189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.401233Z digest=sha256:e71b66896ec31dd21bf0f148506bd22a7b13bfa1bb9ce01c331440ab61c2a054

Observation 8bac707c-32be-4615-8705-715a5b0d3b89 · outbound

This paper cites Panda-70M: Captioning 70M Videos with Multiple Cross-Modality Teachers.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Panda-70M: Captioning 70M Videos with Multiple Cross-Modality Teachers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.404444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.404444Z digest=sha256:f7b97a2f7e32172bd72461712daa1b85f3b92aff96bcccd14b25454adc572e28

Observation d05b8f2a-ff8b-428f-938b-c4e34653b043 · outbound

This paper cites Speed is all you need: On-device acceleration of large diffu- sion models via gpu-aware optimizations.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Speed is all you need: On-device acceleration of large diffu- sion models via gpu-aware optimizations

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.611589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.407840Z digest=sha256:ad290e8714eedd2a92c57959a90bc903f13aad61f5865de1ab83741958a7db9d

Observation b2fbcf9f-e4d9-4d0c-b419-632d18016632 · outbound

This paper cites Squeezing Large-Scale Diffusion Models for Mobile.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Squeezing Large-Scale Diffusion Models for Mobile

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-09T10:46:30.194008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.411129Z digest=sha256:9b346fbdabae34415ca2ab181a443089c24a5761eec6108dba4c548995c39d56

Observation e98c8e93-0e81-4338-8830-f5418dc0e2d1 · outbound

This paper cites Diffusion models beat gans on image synthesis.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Diffusion models beat gans on image synthesis

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.414863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.414863Z digest=sha256:5ae4a32d68ab525009429c4d7e0643a0653355e6951ee540cac24d87bea9cbfd

Observation 082b9bcf-43e1-467e-b6f1-980fc64182da · outbound

This paper cites Tutorial on Variational Autoencoders.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Tutorial on Variational Autoencoders

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.418025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.418025Z digest=sha256:07fa872a5706a8ab47a202ef1240f5a2e2d1e73d77d2d7df3486612cd2ed8e99

Observation 09da7e6f-0232-475c-bc7f-9b0d6f4241ac · outbound

This paper cites Scaling recti- fied flow transformers for high-resolution image synthesis.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Scaling recti- fied flow transformers for high-resolution image synthesis

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.421885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.421885Z digest=sha256:9c5e648b6d243148af4508ee2ac795487a541654d9be54581ecd6f6be6e02839

Observation 3ee9dc0f-dd24-4f34-b040-c1ed32ef953f · outbound

This paper cites Efficient vision trans- former via token merger.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Efficient vision trans- former via token merger

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.591324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.425244Z digest=sha256:697655e2e23a12e546d4736308d1192605f7118867f2a4ae51374367bd71787a

Observation a1df505b-b859-420f-929f-e541a7540628 · outbound

This paper cites Efficient Time Series Processing for Transformers and State-Space Models through Token Merging.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Efficient Time Series Processing for Transformers and State-Space Models through Token Merging

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.428727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.428727Z digest=sha256:3dbabc300f0f175c39310026f809e8fd7577d93fc455073a407a368842a3bc83

Observation 78da2a77-51d5-448e-a97a-564a67492963 · outbound

This paper cites Knowledge distillation: A survey.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Knowledge distillation: A survey

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.583382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.431950Z digest=sha256:174e859f32491143f7c0ae8f558d102d5fc88cf9b6d5414b8bf3f0d5ac60b18a

Observation a48c0b2f-fd1d-43fe-9243-333fa2e2c433 · outbound

This paper cites Gray and David L.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Gray and David L

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.575665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.434964Z digest=sha256:6258922e044209490f92e16807c6eb34fc453af9dbc598ddf45f72aa01594c18

Observation ff597029-a652-4459-be09-7b3730f90384 · outbound

This paper cites Flexible diffusion modeling of long videos.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Flexible diffusion modeling of long videos

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.568339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.438121Z digest=sha256:18b5098a520952deaf3b178d7c85c989aced5fef7f0a9f383b96d0037b428287

Observation fc0e1528-0f4f-4dfa-8a90-f3494fdd3f40 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Distilling the Knowledge in a Neural Network

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.440873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.440873Z digest=sha256:9238391e1cedffca5c234241da815c0313fbea139afe1033e8ef42e38424370a

Observation 8764cf74-0058-4d4f-8b1f-286b5eeb4633 · outbound

This paper cites Denoising dif- fusion probabilistic models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Denoising dif- fusion probabilistic models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.444744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.444744Z digest=sha256:1a669476fc33654cade3001d9a4521ea79a6b17d101c844b51651da92571c5fc

Observation 2e2f7bd8-f1e1-4dfc-9fd9-06eed829b9b7 · outbound

This paper cites Imagen Video: High Definition Video Generation with Diffusion Models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Imagen Video: High Definition Video Generation with Diffusion Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.447958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.447958Z digest=sha256:cd180539b0fd91ee5a96fd01e082044f52abcc7070a2697b10f5e22b909dfbc5

Observation 1233f79a-4018-4e93-831b-e5ef2d1e072a · outbound

This paper cites Cascaded diffu- sion models for high fidelity image generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Cascaded diffu- sion models for high fidelity image generation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.554150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.451148Z digest=sha256:eec535480c91b143331e997015be5c588c8ff5acfc35519fdf3722b8b1706140

Observation 21cd6704-090f-419d-913f-c7584f4e3d58 · outbound

This paper cites Video dif- fusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Video dif- fusion models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.454778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.454778Z digest=sha256:dcfdf6c8affc0b731d7d0fee3ec1caf7d86e46d036c28048b221c6acbf7d2fc8

Observation bb64e608-68c8-4f73-8a1e-8552655ca4cb · outbound

This paper cites Toward controlled generation of text.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Toward controlled generation of text

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.539287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.458128Z digest=sha256:4c8ebfb145f6cd34adf15e0d2dcd5ac39365fa0cb17e0fbfe08dcf6db67ae8fc

Observation 26e2f92f-9178-4318-b5ee-54de7701daad · outbound

This paper cites Vbench: Comprehensive bench- mark suite for video generative models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Vbench: Comprehensive bench- mark suite for video generative models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.530247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.460899Z digest=sha256:b031004be10714a07f325325791c3bd13fb6b44f040aa7c916fd7f70fd053691

Observation 2e21fd85-4f33-494d-8fa3-236bda7c6376 · outbound

This paper cites Pyramidal flow matching for efficient video generative modeling.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Pyramidal flow matching for efficient video generative modeling

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.463824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.463824Z digest=sha256:099f3b833bb17ea207145638be7c4d11a317e639b37de602135e7c6036c985fa

Observation bed72ca8-75b0-4ad6-a0c1-8dbc7d2dde3d · outbound

This paper cites Imagic: Text-based real image editing with diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Imagic: Text-based real image editing with diffusion models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.467231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.467231Z digest=sha256:4c139092e3471bbdd06c77d19ebf746e1045383d573a980f35a07efb2d687788

Observation bd918bd0-689b-47fd-9b43-b6cd1ef0a494 · outbound

This paper cites Text2video-zero: Text- to-image diffusion models are zero-shot video generators.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Text2video-zero: Text- to-image diffusion models are zero-shot video generators

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.515093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.470388Z digest=sha256:45db9121e64cbf60b1f425aff59b14ec54d87a4f230807dec877654e4b21902d

Observation 3771e4df-b0c1-476c-9726-0a91c3d84054 · outbound

This paper cites an unresolved cited work.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:46:30.505452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.473312Z digest=sha256:29e5e83fc4ed789686c53aefd06a12faa0b376fe2be92bb6a54c32a29429a445

Observation 3ba24b8c-a1c6-4dee-8573-c2c4688b98f9 · outbound

This paper cites xformers: A modular and hackable trans- former modelling library.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices xformers: A modular and hackable trans- former modelling library

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.476228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.476228Z digest=sha256:fd48982f38893b43a25df34d9b6a565ccaa7f6a94d42ac63dc6b3d20367ad728

Observation 696a61ec-b843-4404-ace7-91a6a286ff04 · outbound

This paper cites Vidtome: Video token merging for zero-shot video editing.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Vidtome: Video token merging for zero-shot video editing

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.490227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.479554Z digest=sha256:236499b8aac81f6b6dd587adaa690284bd24819e8bb158c043c968fe236b1912

Observation fd85b3f3-3459-495d-8fe5-69f607bb5947 · outbound

This paper cites Snap- fusion: Text-to-image diffusion model on mobile devices within two seconds.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Snap- fusion: Text-to-image diffusion model on mobile devices within two seconds

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.480295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.482337Z digest=sha256:334c94c263481b9e9fdedd5d27b4ba2feb2f47cfbc8fc2902b8fdfddc60b3c7f

Observation 4e656760-e6ab-46ee-b9ad-1163943b6ca1 · outbound

This paper cites AnimateDiff-Lightning: Cross-Model Diffusion Distillation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices AnimateDiff-Lightning: Cross-Model Diffusion Distillation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.485547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.485547Z digest=sha256:bf718190584fcce6e32a397a87a963997d963a1a8b2145744e0e6b87c9fae864

Observation 8e9088e1-93ae-4668-bcc5-07bec1ac74ee · outbound

This paper cites Generative adversarial networks for image and video synthesis: Algorithms and applications.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Generative adversarial networks for image and video synthesis: Algorithms and applications

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.470396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.488642Z digest=sha256:8a67d1206af4eedf5a2ec3872cbef15811a09a788358c51d4475e322707274c1

Observation 586056c2-deff-44d2-8da3-fa18a7d3801e · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.491631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.491631Z digest=sha256:6e610bc44d3a36a09bdddaaae1afbb84de36045a8625707e66bf8a689911eb71

Observation 1ef678db-8144-48e5-a6f3-f87dfa795533 · outbound

This paper cites Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.494514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.494514Z digest=sha256:69656ac946f37d69dc81ce4bb19906abf9670f906de4e776183b28a300ee0b13

Observation d32a0d21-28df-417a-ab70-bd08cd9ad0cf · outbound

This paper cites DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.497660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.497660Z digest=sha256:45b26f056f3acab7a1eae274bc685eaf45b290c87faa291fd67db19d4b57d63b

Observation 39e8164d-8b45-4f5a-89ed-9685c831a4cb · outbound

This paper cites Snap video: Scaled spatiotemporal transformers for text-to-video synthesis.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Snap video: Scaled spatiotemporal transformers for text-to-video synthesis

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.460738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.501071Z digest=sha256:d67186aca2ef7c56db57e77dc9df3edfc48bcbde309fb3328d0258e9123b3e5d

Observation 2602ef1d-0efb-4b3f-b0d7-7e312d1c4747 · outbound

This paper cites Dreamix: Video Diffusion Models are General Video Editors.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Dreamix: Video Diffusion Models are General Video Editors

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.504193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.504193Z digest=sha256:fa1495b967fe7a133c806ede310264588bff0f182a45388f783ea0039a7a657d

Observation 07263c9f-0af4-4d22-96eb-3b29230b11ce · outbound

This paper cites A review on the attention mechanism of deep learning.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices A review on the attention mechanism of deep learning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.451211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.507341Z digest=sha256:5ca020c703baa1c35e6e5daf44f42556d3fdd1510aa668950f4af77223aee62d

Observation ef2d38b3-4c91-455e-96cb-e45467715c29 · outbound

This paper cites Generative models for video analysis and 3D range data applications.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Generative models for video analysis and 3D range data applications

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.441817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.510095Z digest=sha256:827d0d468636a031a8aff048a16d98495544c1dc1a3b7b02a968bf9cd497b266

Observation 44b0ef8b-c9a2-4588-9887-dacc24c32669 · outbound

This paper cites Pytorch: An im- perative style, high-performance deep learning library.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Pytorch: An im- perative style, high-performance deep learning library

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.512914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.512914Z digest=sha256:bba20905ac3ae2f6280ba34f990d4579b0cbda1f6e24dd79e7ff7af4e7c6b27e

Observation c8932e36-323f-4ccd-880b-5dc008f09e21 · outbound

This paper cites Scalable diffusion models with transformers.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Scalable diffusion models with transformers

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.515489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.515489Z digest=sha256:fc943e1c3fc815b0fd533b3048e9581b8b2251e45b4029f6570c742b6530688d

Observation 9db81d19-5ed9-4496-b2ab-bcc34e9519d0 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.518458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.518458Z digest=sha256:aaa1f27399bddf7c9af2b2afde261561b472cb0017cda85ff95840d73c624497

Observation c85b33ee-5cda-4f47-8d0f-312674fd8187 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Exploring the limits of transfer learning with a unified text-to-text transformer

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.422866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.521518Z digest=sha256:968f1cf9146d3e2da555c6d1643a290802a63ab1cf1f9f927dbe04a929957f43

Observation 3039f780-abeb-461e-94ba-7d4d0770f774 · outbound

This paper cites Pruning algorithms-a survey.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Pruning algorithms-a survey

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.414999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.524299Z digest=sha256:2d8031472aa4fba1664c2c3484bf9b08515f0fcb4003c519197937db501f942f

Observation 270a6090-4cf2-4f80-bab7-08636b7ae85b · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices High-resolution image synthesis with latent diffusion models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.527356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.527356Z digest=sha256:88dfb86a234b6bf54af77b4af1eec9bc4b584cf0237358a41734d5649c52936f

Observation 0b270534-f790-4bbf-8cfa-2ab72efd0f3d · outbound

This paper cites U- net: Convolutional networks for biomedical image segmen- tation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices U- net: Convolutional networks for biomedical image segmen- tation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.402361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.530078Z digest=sha256:872dc1527b0e3f2223fd330d12b5cfceb5984efdef988327e9db8cedf194df0f

Observation ead3ae0b-4280-4114-a803-5351e5bf52f4 · outbound

This paper cites ¨Uber die numerische aufl¨osung von differential- gleichungen.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices ¨Uber die numerische aufl¨osung von differential- gleichungen

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.394529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.532752Z digest=sha256:4f6e81f42dc3f448e4aab85a9323af80452cd00c0e3a6418adfcc505f7480d85

Observation ce006bea-9d8d-4f78-9619-3e29b5e88b1c · outbound

This paper cites Palette: Image-to-image diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Palette: Image-to-image diffusion models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.535718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.535718Z digest=sha256:3eaff66e249c3d9b6b57d006347e8077b08abb7c8e25c8124a60cb2d53ae1a4d

Observation 10783bd1-71cd-420a-a34f-ab04eb992168 · outbound

This paper cites Introduction to apple ml tools.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Introduction to apple ml tools

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.380252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.538372Z digest=sha256:b37e6fdeff122b4f9d78ff639526179c85595231312cb1fb001a0a5b46fc5fe2

Observation 93457ce0-f9d8-44ca-be70-9ee14336f7f0 · outbound

This paper cites Sin- gan: Learning a generative model from a single natural im- age.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Sin- gan: Learning a generative model from a single natural im- age

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.371044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.541123Z digest=sha256:e3cd23019d929c19e419e1e4d24949301ae9d1afedd21852eaf4478e0d83b75e

Observation 04f5b16c-d2f1-4cb1-8064-e7250f50cd06 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.544214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.544214Z digest=sha256:43ee45eff3bf8add17aaeb928d4916591fc20adecb680771a50326215f34be5b

Observation feeebf05-e61a-42bf-a061-22ca690883e4 · outbound

This paper cites Video edit- ing via factorized diffusion distillation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Video edit- ing via factorized diffusion distillation

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.361436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.547315Z digest=sha256:6ac77229288afde3d08282608c250d47ce6f0e06e129c17a709132f4e6448776

Observation 9b8fe24f-ce86-4ddb-bc7c-ca3ac3a8181a · outbound

This paper cites Denoising Diffusion Implicit Models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Denoising Diffusion Implicit Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.550642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.550642Z digest=sha256:c445c73da6cd4c0b28d15d237656d5cc95cb3e35ada479643850794adbba951e

Observation f2162860-5301-493d-b63d-ae706d0add3d · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Roformer: Enhanced transformer with rotary position embedding

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.553654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.553654Z digest=sha256:447e5dfe59948a00b82a378997c8d893c63554c14166560531c18c3b4ffbc3d0

Observation df0ca600-10e6-44fd-a97f-c0206f804237 · outbound

This paper cites A survey of multi- modal deep generative models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices A survey of multi- modal deep generative models

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.347655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.556817Z digest=sha256:365cc92c02d9d76c9b63e746a6fa97c5d41a00920c5d8ec426859986ca8d0d4f

Observation 67f6d166-4f8b-4b80-aed4-917be129f93e · outbound

This paper cites VidGen-1M: A Large-Scale Dataset for Text-to-video Generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.559618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.559618Z digest=sha256:00d865bd0e76e4903c197e7d2a7c57973abe0b29bf6a76591e15b526cbd31d59

Observation 61403ab1-9933-43aa-8083-c041c56facff · outbound

This paper cites Qvd: Post-training quantization for video diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Qvd: Post-training quantization for video diffusion models

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.338317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.562630Z digest=sha256:9d0ed161beefeee81bfa799f29ac2035604cc28b8abcbdf70ba079b043a51ba7

Observation f615afc3-0795-42a7-a66b-9d8d0080fe72 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.565679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.565679Z digest=sha256:99275fcac854cd0cc5560ea6b953ea81c7d1470358449c2b628bddf81e392a6e

Observation 67196b35-7a68-4182-a219-44d2f8e63072 · outbound

This paper cites Mobileone: An im- proved one millisecond mobile backbone.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Mobileone: An im- proved one millisecond mobile backbone

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.329357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.569420Z digest=sha256:1f674696b688511c58076d74af97e252999711e2fa022ea66d81f839cc9f6d4f

Observation b43bfd82-e448-4158-b1c2-76e8f47eafba · outbound

This paper cites Mcvd-masked conditional video diffusion for prediction, generation, and interpolation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Mcvd-masked conditional video diffusion for prediction, generation, and interpolation

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.572440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.572440Z digest=sha256:a294512f841f028ca60391c70a737c4c541db69a0ebf9721cd280a76556e9ee1

Observation e6309612-bcbf-4f10-ba22-3aa9b716c1ad · outbound

This paper cites Animatelcm: Computation-efficient personalized style video generation without personalized video data.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Animatelcm: Computation-efficient personalized style video generation without personalized video data

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.314893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.575531Z digest=sha256:4d0bddfb75f69aaaf2fcc715e091bc9f27c0aefd4aad6fbea8146f71adccff2e

Observation df7b0ce5-2024-4fdb-a622-8d598d2fde69 · outbound

This paper cites Transformers: State-of-the-art natural language processing.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Transformers: State-of-the-art natural language processing

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.299223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.581493Z digest=sha256:9245eb1fa66da96ed5465440d4690eba652b3487bead138d29dcbb500b6241a2

Observation 6423c304-6464-4871-8a85-ddd87977ee88 · outbound

This paper cites Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.584244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.584244Z digest=sha256:3a42043c456679e6b3272f003a4783af8bc10e69d53c16b0958b3db9dd2299ae

Observation 1b161b5f-bf0a-4b58-be06-fd4762d0e4f2 · outbound

This paper cites Individual Content and Motion Dynamics Preserved Pruning for Video Diffusion Models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Individual Content and Motion Dynamics Preserved Pruning for Video Diffusion Models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.587315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.587315Z digest=sha256:a55211e85b173b37237bafd4ca7e30392016b09439a533538c3dfeaaedc9ba17

Observation aaa0229f-04b8-48cb-804a-699ef936120f · outbound

This paper cites SnapGen-V: Generating a Five-Second Video within Five Seconds on a Mobile Device.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices SnapGen-V: Generating a Five-Second Video within Five Seconds on a Mobile Device

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-08-09T10:46:29.899800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.590722Z digest=sha256:3c1d424e0ce3c868226e237e73423843af0815f88c43476b3bda960d9e3caf28

Observation 2557feb2-562d-42f1-9143-2911dd6c48d5 · outbound

This paper cites Mobile Video Diffusion.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Mobile Video Diffusion

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.593427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.593427Z digest=sha256:f15b0a4e467bb4469a40f60f6efb52b2abd023741a23e070d7b750336d8e09ba

Observation ffc90300-e8d1-4b71-a906-a569e90c518f · outbound

This paper cites Stat: Spatial-temporal attention mechanism for video cap- tioning.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Stat: Spatial-temporal attention mechanism for video cap- tioning

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.284615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.596470Z digest=sha256:22504400ec224ac0007c86406b18d74529a2cb3c9f17d548fec30434c086e949

Observation d234fb92-3657-4860-b74a-e40336f19908 · outbound

This paper cites Diffusion models: A comprehensive survey of methods and applications.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Diffusion models: A comprehensive survey of methods and applications

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.276159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.599249Z digest=sha256:14ae09b1d5fab0640b1d00b8589f3526e1803854fd4470c1de6a5075d59ae8d6

Observation 05eeb88a-777c-48c2-9c35-3bd3b4e96fcf · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.602176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.602176Z digest=sha256:c050f91ed082c148355f776448188d9bf99e306d201b0d3d7b94395ef367e138

Observation 3a406da9-43db-4b6b-b7f1-9cb8a42e4270 · outbound

This paper cites Schedule On the Fly: Diffusion Time Prediction for Faster and Better Image Generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Schedule On the Fly: Diffusion Time Prediction for Faster and Better Image Generation

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.605280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.605280Z digest=sha256:f75480ce9419e8c0f8ee0298e9aaea47381d1277f81d4588d1f05d406b7ae083

Observation ebd32267-17fe-4703-b90e-a0b28354014c · outbound

This paper cites Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.608228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.608228Z digest=sha256:d1b052b91785072cdff265673a1fc4c8325ad11df5ec75f77166666bc9a2717d

Observation 72439b63-bdc9-42bb-bf5d-ac11f153215b · outbound

This paper cites Text-to-image Diffusion Models in Generative AI: A Survey.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Text-to-image Diffusion Models in Generative AI: A Survey

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.611120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.611120Z digest=sha256:f9433791660ca5c34ad141422ad48d122491f18e3058e01ddede547ebefd58fb

Observation 450f723c-a591-439f-95a7-4a682fd8655c · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Adding conditional control to text-to-image diffusion models

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.614652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.614652Z digest=sha256:93ac9631c2ee698d954f0550fe0557b3ae634f17ff7f8831cc10fe2472f800fc

Observation 65158356-8f47-4256-b959-f77e707502df · outbound

This paper cites A survey on personalized content synthesis with diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices A survey on personalized content synthesis with diffusion models

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.617436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.617436Z digest=sha256:dc498fe1db10b6539ff3338d3b4d624bd3f00415874fdd9aeda4b4aa50c935f4

Observation 7cdbc0f6-9940-4f43-a386-28a5475f1d57 · outbound

This paper cites ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.620727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.620727Z digest=sha256:0899889ea462fbcac79d5b29df9e0263e2edef7233c5049d836c7f62baf3065e

Observation 05ce1640-800c-4717-97f1-40bc758d01ea · outbound

This paper cites MobileDiffusion: Instant Text-to-Image Generation on Mobile Devices.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices MobileDiffusion: Instant Text-to-Image Generation on Mobile Devices

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.624096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.624096Z digest=sha256:3d2376dba0722bd976eb5176f2c9d706460179e5a066b846588f9c0083521527

Observation cf07997b-b53d-4c9a-8985-0f8c7fabc346 · outbound

This paper cites Dpm- solver-v3: Improved diffusion ode solver with empirical model statistics.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Dpm- solver-v3: Improved diffusion ode solver with empirical model statistics

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.262870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.627415Z digest=sha256:fca64a2421d2217626a422e09116528d125eb1be486ac2f6e822a97cd32af4a8

Observation 3b8a40ab-2be7-4bfd-9aca-9995e0813d28 · outbound

This paper cites Open-sora: Democratizing efficient video production for all, 2024.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Open-sora: Democratizing efficient video production for all, 2024

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.253872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.630373Z digest=sha256:f7f9db3e06c2d634e71e43d15852d76dc9a7d7d19391e91468e66f95a84089fa

Observation c2c03a01-b1e0-4aca-aa7c-47f814db441b · outbound

This paper cites Slimflow: Training smaller one-step diffusion models with rectified flow.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Slimflow: Training smaller one-step diffusion models with rectified flow

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.244253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.633290Z digest=sha256:8d8a98bf7541c99282cd095971e8dd59cbfe0cf2d255315eb399171e490d53ba

Observation 54e43f0d-4266-4184-85f0-f8e662045ba0 · outbound

This paper cites an unresolved cited work.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Unresolved cited work

Reference 86

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:46:30.233552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.637028Z digest=sha256:44c866a3279578f2d68ccb8831e1f52a27f7bbf7d1656b4c86b6e8ec968bdf76

Observation 3c6a34cc-29a6-495e-80a3-9366d3de6440 · outbound

This paper cites an unresolved cited work.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Unresolved cited work

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.578325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.578325Z digest=sha256:4ff9970ddb433b33b073a5ba1d41240990f44d11f1d70ccaf395833219a4d43e

Pith citing papers

Observation 47c6b7c4-9155-4791-8fcb-745aa841879f · inbound

Video Generation Models as World Models: Efficient Paradigms, Architectures and Algorithms cites this paper.

Video Generation Models as World Models: Efficient Paradigms, Architectures and Algorithms On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices

Reference 155

Resolution
verified exact
arxiv_id, observed 2026-05-14T01:38:36.426016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T01:35:14.878069Z digest=sha256:688992eda3805361a680bcfb08706ca0a99913d69b0beb94efd54dd7343a5a6f

Observation 2f294d63-0116-4b9e-bba8-7afd5bf3364b · inbound

Efficient Video Diffusion Models: Advancements and Challenges cites this paper.

Efficient Video Diffusion Models: Advancements and Challenges On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:03:26.202726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T08:28:29.706249Z digest=sha256:f504bd1fecc6486b273f3b800b583ca7ca85a315da2fd75ed454ce4dbf650f77

Observation d90978ab-05e8-4c89-aa73-051ddd806df8 · inbound

MobileWan: Closing the Quality Gap for Mobile Video Diffusion cites this paper.

MobileWan: Closing the Quality Gap for Mobile Video Diffusion On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-07-08T14:44:59.785369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-08T14:37:46.957265Z digest=sha256:ea623f5fbd5c5a9736d39805e611a670d474ce94544e2dbdad8cd875e7faa40d

Observation 28edac6e-cc20-40fd-9a52-b3ee690f6fbd · inbound

MobileWan: Closing the Quality Gap for Mobile Video Diffusion cites this paper.

MobileWan: Closing the Quality Gap for Mobile Video Diffusion On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T08:21:44.482775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:21:44.482775Z digest=sha256:582e0bb261c2dab592e45fa4a4f7ba4919ab655eac1b19bc10339dca4eac5417