Pith. sign in

Paper Citation Record · LEDGER

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices

As of 13 August 2026, this Paper Citation Record lists 86 of 86 outbound references and 4 inbound Pith citation observations for arXiv:2502.04363.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.04363 v2

Coverage vector

measured 86 of 86 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T10:46:29.637028Z

measured 90 of 90 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T08:21:44.482775Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T14:44:59.784059Z

Reference resolution

86 of 86 outbound references displayed

  • verified exact2
  • verified fuzzy33
  • unresolved51
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5da9d11d-6aab-484b-847c-515295b1fcb9 · outbound

This paper cites iphone 15 pro—technical specifications, 2023.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices iphone 15 pro—technical specifications, 2023

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.365358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.365358Z digest=sha256:097226462037463215537ec133ba874d8ef20661b413cbcb7ecb4241a8e36894

Observation 5b08b0ec-586f-4a16-bce6-156c0db86aa4 · outbound

This paper cites Swift, 2024.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Swift, 2024

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.369925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.369925Z digest=sha256:af0cce42aba2de59449dfb183f86a0087893c6b7649e2af5e18f579afcf2bb35

Observation 65e576f8-e1ce-4dbe-876d-ff6a956d5a1a · outbound

This paper cites A discussion on euler method: A review.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices A discussion on euler method: A review

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.373458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.373458Z digest=sha256:5047a178bafe04060b14b772369d3ecfa3c94d32e868482fadb9f141ece9e67d

Observation e888b8db-7500-4ae4-a049-d8c98beee61e · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.376984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.376984Z digest=sha256:c97f152cff055fc1c4efcc06f22d80b26f41604cae4c2884a4b9a953b392b6e7

Observation c049aab2-8e65-474c-8cdc-4634c85792e7 · outbound

This paper cites Align your latents: High-resolution video synthesis with la- tent diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Align your latents: High-resolution video synthesis with la- tent diffusion models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.381180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.381180Z digest=sha256:20033258e346249d95fa89177e80a8ccdc6c08ce020c94ddee2ad34d16cc1042

Observation a3d877ef-dded-4ebd-b451-eb989955aa70 · outbound

This paper cites Token Merging: Your ViT But Faster.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Token Merging: Your ViT But Faster

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.384729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.384729Z digest=sha256:5a7aabcc3831e94e1c49adb2862ba1ac053003ae564fd2cd0af6e131615ad690

Observation d5265d9a-2cf8-42fd-acb9-d7572fc2fb97 · outbound

This paper cites Token merging: Your ViT but faster.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Token merging: Your ViT but faster

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.389029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.389029Z digest=sha256:71bdf3a0fc4e7fd46adaea7042b46b4627e512ad03bd43db201f271d757c5d43

Observation da626f79-ee59-4c72-ba06-fd2c0ccfb6d5 · outbound

This paper cites EdgeFusion: On-Device Text-to-Image Generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices EdgeFusion: On-Device Text-to-Image Generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.392456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.392456Z digest=sha256:b0481fb6c7833727e520fb7733a03365718311cbeb14a5e44c6b195d061083a2

Observation b4405442-c3de-4264-b25e-9f6c9170d8b0 · outbound

This paper cites Tempme: Towards the explain- ability of temporal graph neural networks via motif discov- ery.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Tempme: Towards the explain- ability of temporal graph neural networks via motif discov- ery

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.396621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.396621Z digest=sha256:707514957f3b266cfcb3d677f7e696a729ea86c46aca29654efa0b5f4a9836da

Observation 206bfa80-3ccf-4ac1-9169-6c60e56dfd6e · outbound

This paper cites Neural ordinary differential equa- tions.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Neural ordinary differential equa- tions

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.621189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.401233Z digest=sha256:fd970847af5efb6dd80c19708c79549c9b3f7bc36c6651bb158488c5f471bdb2

Observation 8bac707c-32be-4615-8705-715a5b0d3b89 · outbound

This paper cites Panda-70M: Captioning 70M Videos with Multiple Cross-Modality Teachers.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Panda-70M: Captioning 70M Videos with Multiple Cross-Modality Teachers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.404444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.404444Z digest=sha256:8f889e9a6f8ddde20e00efe82954f837eb2bbd3ab8630eb44721efccd19d179a

Observation d05b8f2a-ff8b-428f-938b-c4e34653b043 · outbound

This paper cites Speed is all you need: On-device acceleration of large diffu- sion models via gpu-aware optimizations.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Speed is all you need: On-device acceleration of large diffu- sion models via gpu-aware optimizations

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.611589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.407840Z digest=sha256:75ea57d6ff5b75efb61cc3c42c58e525385739c46ff71daa38fe8b120c69199f

Observation b2fbcf9f-e4d9-4d0c-b419-632d18016632 · outbound

This paper cites Squeezing Large-Scale Diffusion Models for Mobile.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Squeezing Large-Scale Diffusion Models for Mobile

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-09T10:46:30.194008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.411129Z digest=sha256:38fc90e94b6237f0a07922697e489c44f15bd0ce36b5e78151a6d97ad160de10

Observation e98c8e93-0e81-4338-8830-f5418dc0e2d1 · outbound

This paper cites Diffusion models beat gans on image synthesis.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Diffusion models beat gans on image synthesis

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.414863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.414863Z digest=sha256:afa9dab4c650ceebb75f2cba6ae0c6550380761f4529592d6597bb90dd37ce3b

Observation 082b9bcf-43e1-467e-b6f1-980fc64182da · outbound

This paper cites Tutorial on Variational Autoencoders.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Tutorial on Variational Autoencoders

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.418025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.418025Z digest=sha256:3080cb7f17b962887332b167ca6c05d3328d45c1a2cca9b66d02747d6eda12d3

Observation 09da7e6f-0232-475c-bc7f-9b0d6f4241ac · outbound

This paper cites Scaling recti- fied flow transformers for high-resolution image synthesis.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Scaling recti- fied flow transformers for high-resolution image synthesis

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.421885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.421885Z digest=sha256:588a01b4029727d12382a571eae5df3c2128320f2a3cb2bdddb3b7d95bbc5947

Observation 3ee9dc0f-dd24-4f34-b040-c1ed32ef953f · outbound

This paper cites Efficient vision trans- former via token merger.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Efficient vision trans- former via token merger

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.591324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.425244Z digest=sha256:93d0bd602fc5138c6cdc2a35ff1bb6b8d8b45ee7b04054253a1f6ac04803e368

Observation a1df505b-b859-420f-929f-e541a7540628 · outbound

This paper cites Efficient Time Series Processing for Transformers and State-Space Models through Token Merging.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Efficient Time Series Processing for Transformers and State-Space Models through Token Merging

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.428727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.428727Z digest=sha256:f0f7864f724e62c614d7bfe5cc774e8f73736f4941fc4c8870b097db7fb690eb

Observation 78da2a77-51d5-448e-a97a-564a67492963 · outbound

This paper cites Knowledge distillation: A survey.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Knowledge distillation: A survey

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.583382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.431950Z digest=sha256:68d6f16f555ce9392e33b068f128440fa839bf7e12ee15f2efcae66730efff0a

Observation a48c0b2f-fd1d-43fe-9243-333fa2e2c433 · outbound

This paper cites Gray and David L.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Gray and David L

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.575665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.434964Z digest=sha256:4497c54946db32821a4d570d854750319b7c72d6cb8865f7623cac405c89a155

Observation ff597029-a652-4459-be09-7b3730f90384 · outbound

This paper cites Flexible diffusion modeling of long videos.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Flexible diffusion modeling of long videos

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.568339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.438121Z digest=sha256:5a2d457bc2728da71952a8db882728f667d2ebb54b26a6866c65f8726b706e11

Observation fc0e1528-0f4f-4dfa-8a90-f3494fdd3f40 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Distilling the Knowledge in a Neural Network

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.440873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.440873Z digest=sha256:ef34f7d48ce9593aef535cf6d492de62a1a2c3d0b6147fe1f7bb56ff3b5a48e4

Observation 8764cf74-0058-4d4f-8b1f-286b5eeb4633 · outbound

This paper cites Denoising dif- fusion probabilistic models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Denoising dif- fusion probabilistic models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.444744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.444744Z digest=sha256:24c9c0d22a696bc77240937b6320e15f3ba8906d1c14cc2e2732e43ac1039294

Observation 2e2f7bd8-f1e1-4dfc-9fd9-06eed829b9b7 · outbound

This paper cites Imagen Video: High Definition Video Generation with Diffusion Models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Imagen Video: High Definition Video Generation with Diffusion Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.447958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.447958Z digest=sha256:961d84a2b53f8ae072f3af4f22001aa3d8078fae9aaeec164e56099a1c530328

Observation 1233f79a-4018-4e93-831b-e5ef2d1e072a · outbound

This paper cites Cascaded diffu- sion models for high fidelity image generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Cascaded diffu- sion models for high fidelity image generation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.554150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.451148Z digest=sha256:0a228fda3650a568e991342439cb583e19641769478eecdddb0400457f68c237

Observation 21cd6704-090f-419d-913f-c7584f4e3d58 · outbound

This paper cites Video dif- fusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Video dif- fusion models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.454778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.454778Z digest=sha256:13c76be285ae5dc1cbf640d8ea47f1b336567131b40ec0903ca5d9fa83fbec28

Observation bb64e608-68c8-4f73-8a1e-8552655ca4cb · outbound

This paper cites Toward controlled generation of text.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Toward controlled generation of text

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.539287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.458128Z digest=sha256:0d363e20806df36141312d28a1a804b519ebd73f797dc2240b5941cb3f768772

Observation 26e2f92f-9178-4318-b5ee-54de7701daad · outbound

This paper cites Vbench: Comprehensive bench- mark suite for video generative models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Vbench: Comprehensive bench- mark suite for video generative models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.530247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.460899Z digest=sha256:4e76b7db643daea95b4e2f2559b58d2413e7c43997e2fb8314347d70e89e0459

Observation 2e21fd85-4f33-494d-8fa3-236bda7c6376 · outbound

This paper cites Pyramidal flow matching for efficient video generative modeling.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Pyramidal flow matching for efficient video generative modeling

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.463824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.463824Z digest=sha256:7cbd407636fe9581c654d1143b08fee800ed9655ce731602ee445dbe8e078a81

Observation bed72ca8-75b0-4ad6-a0c1-8dbc7d2dde3d · outbound

This paper cites Imagic: Text-based real image editing with diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Imagic: Text-based real image editing with diffusion models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.467231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.467231Z digest=sha256:80bc356d36ddb93a19e62bed857b8b0df7dbc4982adfe685ac49112ddf2cc951

Observation bd918bd0-689b-47fd-9b43-b6cd1ef0a494 · outbound

This paper cites Text2video-zero: Text- to-image diffusion models are zero-shot video generators.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Text2video-zero: Text- to-image diffusion models are zero-shot video generators

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.515093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.470388Z digest=sha256:e46c64bd21d915d53608754fb1ba7f4ce3ac191d71df1bc685c475e9ac0d45b5

Observation 3771e4df-b0c1-476c-9726-0a91c3d84054 · outbound

This paper cites an unresolved cited work.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:46:30.505452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.473312Z digest=sha256:38d9f9aba1c1f982d22dd08d2bc6d4b0a7e640b62705d24f045897b3397e780d

Observation 3ba24b8c-a1c6-4dee-8573-c2c4688b98f9 · outbound

This paper cites xformers: A modular and hackable trans- former modelling library.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices xformers: A modular and hackable trans- former modelling library

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.476228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.476228Z digest=sha256:b5746ab5cf2ee5245d40508eba732eeb8d55a6fca8d1ca244ae51bcb765a14e8

Observation 696a61ec-b843-4404-ace7-91a6a286ff04 · outbound

This paper cites Vidtome: Video token merging for zero-shot video editing.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Vidtome: Video token merging for zero-shot video editing

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.490227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.479554Z digest=sha256:e14e2671f4b07999ecc5c016bd43a66ffd9c2041ccf9df31238dc0e1f41ad080

Observation fd85b3f3-3459-495d-8fe5-69f607bb5947 · outbound

This paper cites Snap- fusion: Text-to-image diffusion model on mobile devices within two seconds.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Snap- fusion: Text-to-image diffusion model on mobile devices within two seconds

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.480295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.482337Z digest=sha256:fee8767100025cafe40304b376f12cb93d477ed879eeac86853730e00306357a

Observation 4e656760-e6ab-46ee-b9ad-1163943b6ca1 · outbound

This paper cites AnimateDiff-Lightning: Cross-Model Diffusion Distillation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices AnimateDiff-Lightning: Cross-Model Diffusion Distillation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.485547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.485547Z digest=sha256:042bdd2997093b2d264a4f0634bc092742931a9d5dbe0fcd247b2dd0c64a0760

Observation 8e9088e1-93ae-4668-bcc5-07bec1ac74ee · outbound

This paper cites Generative adversarial networks for image and video synthesis: Algorithms and applications.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Generative adversarial networks for image and video synthesis: Algorithms and applications

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.470396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.488642Z digest=sha256:1ed4883f245e4ff1f18cf1350c0ac7ff3af068ce2dd8e8fa7ec0284edc06988a

Observation 586056c2-deff-44d2-8da3-fa18a7d3801e · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.491631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.491631Z digest=sha256:9f0b7f7acc24f5c9c91e3e70905489d4b880ad01e0769ccddd7ff0874a4011bd

Observation 1ef678db-8144-48e5-a6f3-f87dfa795533 · outbound

This paper cites Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.494514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.494514Z digest=sha256:3f76a765a49dc5df20339bc62f060e09b9ddaf4e4e21f35ba689009dcc81099c

Observation d32a0d21-28df-417a-ab70-bd08cd9ad0cf · outbound

This paper cites DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.497660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.497660Z digest=sha256:8b19dad9e759205460abd0fbd6aae9684e3137ff00d02c923c905b2763019be5

Observation 39e8164d-8b45-4f5a-89ed-9685c831a4cb · outbound

This paper cites Snap video: Scaled spatiotemporal transformers for text-to-video synthesis.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Snap video: Scaled spatiotemporal transformers for text-to-video synthesis

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.460738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.501071Z digest=sha256:8462cad4eea52b7819d2222e8a01653ac78022ee4dddb429b7aca47d91381128

Observation 2602ef1d-0efb-4b3f-b0d7-7e312d1c4747 · outbound

This paper cites Dreamix: Video Diffusion Models are General Video Editors.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Dreamix: Video Diffusion Models are General Video Editors

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.504193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.504193Z digest=sha256:85e80b6154be7b180ef0f55384697591bc3d92c81b02f72fade3229c742abfa1

Observation 07263c9f-0af4-4d22-96eb-3b29230b11ce · outbound

This paper cites A review on the attention mechanism of deep learning.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices A review on the attention mechanism of deep learning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.451211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.507341Z digest=sha256:618c14b7ffb448e2fec27cdfe3c7437d9a7df5c45f05a786276efd73dbc56799

Observation ef2d38b3-4c91-455e-96cb-e45467715c29 · outbound

This paper cites Generative models for video analysis and 3D range data applications.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Generative models for video analysis and 3D range data applications

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.441817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.510095Z digest=sha256:2eff15c5b4634bdbd4f24cf6fa96d0446cc354f1768c28f8ec5c6fb187c1bd98

Observation 44b0ef8b-c9a2-4588-9887-dacc24c32669 · outbound

This paper cites Pytorch: An im- perative style, high-performance deep learning library.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Pytorch: An im- perative style, high-performance deep learning library

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.512914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.512914Z digest=sha256:566262d716ace905c10bf639b536381a430866e4f640a0617e7e83030644017d

Observation c8932e36-323f-4ccd-880b-5dc008f09e21 · outbound

This paper cites Scalable diffusion models with transformers.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Scalable diffusion models with transformers

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.515489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.515489Z digest=sha256:2a0fad6b3086afd6255d3e608d319f39e1bb448c1eabb15711d4afc2e161958c

Observation 9db81d19-5ed9-4496-b2ab-bcc34e9519d0 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.518458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.518458Z digest=sha256:8b0bb38aa3bafdd67d9daefc0661d9cde028f9ea2998cdef42796c6f4bc87b2e

Observation c85b33ee-5cda-4f47-8d0f-312674fd8187 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Exploring the limits of transfer learning with a unified text-to-text transformer

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.422866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.521518Z digest=sha256:0cbc58c19ff80ddc3c1cc819e6f390ad509a38497ea9026b4677ca72c88a7721

Observation 3039f780-abeb-461e-94ba-7d4d0770f774 · outbound

This paper cites Pruning algorithms-a survey.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Pruning algorithms-a survey

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.414999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.524299Z digest=sha256:490d1f850ddfd1079a8489731ab83a89f8b475e0601688fb50a345893f432362

Observation 270a6090-4cf2-4f80-bab7-08636b7ae85b · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices High-resolution image synthesis with latent diffusion models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.527356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.527356Z digest=sha256:2408a3930f3b38fc8cdf2fd3fb280e7b08a23b6dd5cd081cfdf7a0372a266255

Observation 0b270534-f790-4bbf-8cfa-2ab72efd0f3d · outbound

This paper cites U- net: Convolutional networks for biomedical image segmen- tation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices U- net: Convolutional networks for biomedical image segmen- tation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.402361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.530078Z digest=sha256:c47dd13fde219b5208bfea91cf9be28dd6fcb02fabc7084d49c17377cc5a563a

Observation ead3ae0b-4280-4114-a803-5351e5bf52f4 · outbound

This paper cites ¨Uber die numerische aufl¨osung von differential- gleichungen.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices ¨Uber die numerische aufl¨osung von differential- gleichungen

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.394529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.532752Z digest=sha256:19e37cf56f28c80a6d1949babd5fb611a8ed86b10062577d55e98b526a5f9745

Observation ce006bea-9d8d-4f78-9619-3e29b5e88b1c · outbound

This paper cites Palette: Image-to-image diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Palette: Image-to-image diffusion models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.535718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.535718Z digest=sha256:2dfc16d80ca427694d86fabcd5b62042af5a791927ea0cbbbabb7e80f8631d04

Observation 10783bd1-71cd-420a-a34f-ab04eb992168 · outbound

This paper cites Introduction to apple ml tools.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Introduction to apple ml tools

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.380252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.538372Z digest=sha256:445c52724ceb9c65060d9a9dd1939f41052044742854e6c6d14b53e6a1a1874b

Observation 93457ce0-f9d8-44ca-be70-9ee14336f7f0 · outbound

This paper cites Sin- gan: Learning a generative model from a single natural im- age.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Sin- gan: Learning a generative model from a single natural im- age

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.371044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.541123Z digest=sha256:f756c1076aaa60517eeb80547b3a8443bc0c75b3673343ab7647f2b76bf3666e

Observation 04f5b16c-d2f1-4cb1-8064-e7250f50cd06 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.544214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.544214Z digest=sha256:ea5078621f00d0dca06bc401e2967c495a97e24cc2ce19c1687d3c69e445fd96

Observation feeebf05-e61a-42bf-a061-22ca690883e4 · outbound

This paper cites Video edit- ing via factorized diffusion distillation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Video edit- ing via factorized diffusion distillation

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.361436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.547315Z digest=sha256:42978702b9ef5083383fc42116bf37f1e39ed1650b78acaf0d79b05744f461bf

Observation 9b8fe24f-ce86-4ddb-bc7c-ca3ac3a8181a · outbound

This paper cites Denoising Diffusion Implicit Models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Denoising Diffusion Implicit Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.550642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.550642Z digest=sha256:5c3699e2795d8e7c0920e5294d45f584e722990081198824c7c0d30ac3ed42f8

Observation f2162860-5301-493d-b63d-ae706d0add3d · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Roformer: Enhanced transformer with rotary position embedding

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.553654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.553654Z digest=sha256:47dc52ed9c001ab0b3b73a6c7f0eee9d618595068ee8a8d926e85ef29053e39a

Observation df0ca600-10e6-44fd-a97f-c0206f804237 · outbound

This paper cites A survey of multi- modal deep generative models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices A survey of multi- modal deep generative models

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.347655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.556817Z digest=sha256:d338efc8166c03fa08c01851cb0d13885484716461c84a50d80373829bbce9fc

Observation 67f6d166-4f8b-4b80-aed4-917be129f93e · outbound

This paper cites VidGen-1M: A Large-Scale Dataset for Text-to-video Generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.559618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.559618Z digest=sha256:e71c556fcb1dd5f6aed30a7c3699f11d47707044bff41c75ff302ec902556bad

Observation 61403ab1-9933-43aa-8083-c041c56facff · outbound

This paper cites Qvd: Post-training quantization for video diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Qvd: Post-training quantization for video diffusion models

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.338317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.562630Z digest=sha256:28db0ec3622694943876f9f9fb6921b5913647fdadaa1af3287297ba92d596f6

Observation f615afc3-0795-42a7-a66b-9d8d0080fe72 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.565679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.565679Z digest=sha256:762c8596c6f031e12c0d4cc43b3fbc43763d827dbc9894eb470036bc25f1be3f

Observation 67196b35-7a68-4182-a219-44d2f8e63072 · outbound

This paper cites Mobileone: An im- proved one millisecond mobile backbone.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Mobileone: An im- proved one millisecond mobile backbone

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.329357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.569420Z digest=sha256:160551aa78d6fe6caf84decd02685fe1d5f4dfbd6163c7258ec93a7f71282d13

Observation b43bfd82-e448-4158-b1c2-76e8f47eafba · outbound

This paper cites Mcvd-masked conditional video diffusion for prediction, generation, and interpolation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Mcvd-masked conditional video diffusion for prediction, generation, and interpolation

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.572440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.572440Z digest=sha256:a49dccd948d62310ae23cc20fca9871c0543f40a5bfe50f4b3ab25b31d2d939d

Observation e6309612-bcbf-4f10-ba22-3aa9b716c1ad · outbound

This paper cites Animatelcm: Computation-efficient personalized style video generation without personalized video data.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Animatelcm: Computation-efficient personalized style video generation without personalized video data

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.314893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.575531Z digest=sha256:4f20052860de4b250ce5fd9b10f8731372b3f75afdde1f5809cae1b13507fe26

Observation df7b0ce5-2024-4fdb-a622-8d598d2fde69 · outbound

This paper cites Transformers: State-of-the-art natural language processing.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Transformers: State-of-the-art natural language processing

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.299223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.581493Z digest=sha256:aed9e1b5fa2266bd1db3c38b46c5c7f5dad68fad009c6e71c98c7bc0b88c9797

Observation 6423c304-6464-4871-8a85-ddd87977ee88 · outbound

This paper cites Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.584244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.584244Z digest=sha256:bb9c5e5d74025784692c8850e8c10288851a79f15187c95ee8b5570b750b1710

Observation 1b161b5f-bf0a-4b58-be06-fd4762d0e4f2 · outbound

This paper cites Individual Content and Motion Dynamics Preserved Pruning for Video Diffusion Models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Individual Content and Motion Dynamics Preserved Pruning for Video Diffusion Models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.587315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.587315Z digest=sha256:4f395c465b6c7e6061cf70675f350699664390304fec24003bc0672b9f4d63df

Observation aaa0229f-04b8-48cb-804a-699ef936120f · outbound

This paper cites SnapGen-V: Generating a Five-Second Video within Five Seconds on a Mobile Device.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices SnapGen-V: Generating a Five-Second Video within Five Seconds on a Mobile Device

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-08-09T10:46:29.899800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.590722Z digest=sha256:259afe95a7743b99ec8b8b1554397253442215b3bdadf68bf104a5f293c974c2

Observation 2557feb2-562d-42f1-9143-2911dd6c48d5 · outbound

This paper cites Mobile Video Diffusion.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Mobile Video Diffusion

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.593427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.593427Z digest=sha256:eb5f9249f5d7e54026c20f107146910697c69d0c87bb0461c8079d641fedce7e

Observation ffc90300-e8d1-4b71-a906-a569e90c518f · outbound

This paper cites Stat: Spatial-temporal attention mechanism for video cap- tioning.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Stat: Spatial-temporal attention mechanism for video cap- tioning

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.284615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.596470Z digest=sha256:a460f5989a989ac744db745e9160f79e2f72e3a69b25528a7e937e14d6c16c74

Observation d234fb92-3657-4860-b74a-e40336f19908 · outbound

This paper cites Diffusion models: A comprehensive survey of methods and applications.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Diffusion models: A comprehensive survey of methods and applications

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.276159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.599249Z digest=sha256:1db7e25a1c2b65b38f02aeb7447e0431ee0e05d2349a9a9c57d443f10ed0c0b7

Observation 05eeb88a-777c-48c2-9c35-3bd3b4e96fcf · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.602176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.602176Z digest=sha256:7705e0f42bd3e5ac1cadf984164449381a8452a259076c093a0d4ace3eeb5e73

Observation 3a406da9-43db-4b6b-b7f1-9cb8a42e4270 · outbound

This paper cites Schedule On the Fly: Diffusion Time Prediction for Faster and Better Image Generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Schedule On the Fly: Diffusion Time Prediction for Faster and Better Image Generation

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.605280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.605280Z digest=sha256:651da0ae3f6cfb8ac9e621f8a72a054e46fb70e8b1e29cdce1112df4825d743a

Observation ebd32267-17fe-4703-b90e-a0b28354014c · outbound

This paper cites Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.608228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.608228Z digest=sha256:1f438c8319a38fbb66480fa5d777503c7700a76dc1d18d564eb22123c2e3f124

Observation 72439b63-bdc9-42bb-bf5d-ac11f153215b · outbound

This paper cites Text-to-image Diffusion Models in Generative AI: A Survey.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Text-to-image Diffusion Models in Generative AI: A Survey

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.611120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.611120Z digest=sha256:692573df367729642c6096b4c3d40871fa3993f488112df1e2fc63da208a3451

Observation 450f723c-a591-439f-95a7-4a682fd8655c · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Adding conditional control to text-to-image diffusion models

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.614652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.614652Z digest=sha256:0041eb79c69b7c3124565c2a004e7b6b994b7491d29291564849a54924ca84de

Observation 65158356-8f47-4256-b959-f77e707502df · outbound

This paper cites A survey on personalized content synthesis with diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices A survey on personalized content synthesis with diffusion models

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.617436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.617436Z digest=sha256:e36f6abbec19115f600d1b4b588ad3307e29dfc5b92dce769572671f9a79f949

Observation 7cdbc0f6-9940-4f43-a386-28a5475f1d57 · outbound

This paper cites ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.620727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.620727Z digest=sha256:ef7087c8b6f90ce02dfd0b8c71a484c040887ece39223196f1f359f0225832a8

Observation 05ce1640-800c-4717-97f1-40bc758d01ea · outbound

This paper cites MobileDiffusion: Instant Text-to-Image Generation on Mobile Devices.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices MobileDiffusion: Instant Text-to-Image Generation on Mobile Devices

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.624096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.624096Z digest=sha256:6fb59786a3eec11bfcf65e6f69ca812264d388b8d0be657956537f231a716145

Observation cf07997b-b53d-4c9a-8985-0f8c7fabc346 · outbound

This paper cites Dpm- solver-v3: Improved diffusion ode solver with empirical model statistics.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Dpm- solver-v3: Improved diffusion ode solver with empirical model statistics

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.262870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.627415Z digest=sha256:d5788d1841bab855a464e19d05a964961e215317ac15888fd33e6d830c845899

Observation 3b8a40ab-2be7-4bfd-9aca-9995e0813d28 · outbound

This paper cites Open-sora: Democratizing efficient video production for all, 2024.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Open-sora: Democratizing efficient video production for all, 2024

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.253872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.630373Z digest=sha256:3fe520f9a1de8b5db81c9f809ff83684f287ccd5bd19c7ea22e071bd220ccf7d

Observation c2c03a01-b1e0-4aca-aa7c-47f814db441b · outbound

This paper cites Slimflow: Training smaller one-step diffusion models with rectified flow.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Slimflow: Training smaller one-step diffusion models with rectified flow

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.244253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.633290Z digest=sha256:6838d39fa7c24a17a7344cccc9db1138b3b0457c7166244bfb4782cac7f9e2d9

Observation 54e43f0d-4266-4184-85f0-f8e662045ba0 · outbound

This paper cites an unresolved cited work.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Unresolved cited work

Reference 86

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:46:30.233552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T10:46:29.637028Z digest=sha256:2fb6a1de45ce0d69daf1d0fa88657538999dba5963619e927914e4a96dbaa3d2

Observation 3c6a34cc-29a6-495e-80a3-9366d3de6440 · outbound

This paper cites an unresolved cited work.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Unresolved cited work

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.578325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.578325Z digest=sha256:7b64f959c38a82712fd9a8898b207b44a1ff27d497df64d1b32da97efc4bbafe

Pith citing papers

Observation 47c6b7c4-9155-4791-8fcb-745aa841879f · inbound

Video Generation Models as World Models: Efficient Paradigms, Architectures and Algorithms cites this paper.

Video Generation Models as World Models: Efficient Paradigms, Architectures and Algorithms On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices

Reference 155

Resolution
verified exact
arxiv_id, observed 2026-05-14T01:38:36.426016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-14T01:35:14.878069Z digest=sha256:9206de2f4cb9819d177612978a1b5b819196e76261effd20363e59776bb493fb

Observation 2f294d63-0116-4b9e-bba8-7afd5bf3364b · inbound

Efficient Video Diffusion Models: Advancements and Challenges cites this paper.

Efficient Video Diffusion Models: Advancements and Challenges On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:03:26.202726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T08:28:29.706249Z digest=sha256:dfe426776edeeeb86a19b8cac8082533a546d57abab3841cd80446afd83d8f89

Observation d90978ab-05e8-4c89-aa73-051ddd806df8 · inbound

MobileWan: Closing the Quality Gap for Mobile Video Diffusion cites this paper.

MobileWan: Closing the Quality Gap for Mobile Video Diffusion On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-07-08T14:44:59.785369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-08T14:37:46.957265Z digest=sha256:db2bb5815d375b73ef6a58ba18e01be2c4c17cf8c07d16274dc73e9dd0a19cd2

Observation 28edac6e-cc20-40fd-9a52-b3ee690f6fbd · inbound

MobileWan: Closing the Quality Gap for Mobile Video Diffusion cites this paper.

MobileWan: Closing the Quality Gap for Mobile Video Diffusion On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T08:21:44.482775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:21:44.482775Z digest=sha256:2480eaaed70abd0a51635be29445e47ef8503a994f6f29a94f52bcc2a25932ee