Pith. sign in

Paper Citation Record · LEDGER

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices

As of 9 August 2026, this Paper Citation Record lists 86 of 86 outbound references and 4 inbound Pith citation observations for arXiv:2502.04363.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.04363 v2

Coverage vector

measured 86 of 86 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T10:46:29.637028Z

measured 90 of 90 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T08:21:44.482775Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T14:44:59.784059Z

Reference resolution

86 of 86 outbound references displayed

  • verified exact2
  • verified fuzzy33
  • unresolved51
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5da9d11d-6aab-484b-847c-515295b1fcb9 · outbound

This paper cites iphone 15 pro—technical specifications, 2023.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices iphone 15 pro—technical specifications, 2023

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.365358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.365358Z digest=sha256:097226462037463215537ec133ba874d8ef20661b413cbcb7ecb4241a8e36894

Observation 5b08b0ec-586f-4a16-bce6-156c0db86aa4 · outbound

This paper cites Swift, 2024.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Swift, 2024

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.369925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.369925Z digest=sha256:af0cce42aba2de59449dfb183f86a0087893c6b7649e2af5e18f579afcf2bb35

Observation 65e576f8-e1ce-4dbe-876d-ff6a956d5a1a · outbound

This paper cites A discussion on euler method: A review.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices A discussion on euler method: A review

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.373458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.373458Z digest=sha256:5047a178bafe04060b14b772369d3ecfa3c94d32e868482fadb9f141ece9e67d

Observation e888b8db-7500-4ae4-a049-d8c98beee61e · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.376984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.376984Z digest=sha256:c97f152cff055fc1c4efcc06f22d80b26f41604cae4c2884a4b9a953b392b6e7

Observation c049aab2-8e65-474c-8cdc-4634c85792e7 · outbound

This paper cites Align your latents: High-resolution video synthesis with la- tent diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Align your latents: High-resolution video synthesis with la- tent diffusion models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.381180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.381180Z digest=sha256:20033258e346249d95fa89177e80a8ccdc6c08ce020c94ddee2ad34d16cc1042

Observation a3d877ef-dded-4ebd-b451-eb989955aa70 · outbound

This paper cites Token Merging: Your ViT But Faster.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Token Merging: Your ViT But Faster

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.384729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.384729Z digest=sha256:52267df568022d5ab9ec4546cd652a8b55a8e65eb3a9b233962dcc7aebecbf3e

Observation d5265d9a-2cf8-42fd-acb9-d7572fc2fb97 · outbound

This paper cites Token merging: Your ViT but faster.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Token merging: Your ViT but faster

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.389029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.389029Z digest=sha256:71bdf3a0fc4e7fd46adaea7042b46b4627e512ad03bd43db201f271d757c5d43

Observation da626f79-ee59-4c72-ba06-fd2c0ccfb6d5 · outbound

This paper cites EdgeFusion: On-Device Text-to-Image Generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices EdgeFusion: On-Device Text-to-Image Generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.392456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.392456Z digest=sha256:1165a9d0a931451cba9b0611fe2056bd2414100151934afbe0a0982bd4ada355

Observation b4405442-c3de-4264-b25e-9f6c9170d8b0 · outbound

This paper cites Tempme: Towards the explain- ability of temporal graph neural networks via motif discov- ery.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Tempme: Towards the explain- ability of temporal graph neural networks via motif discov- ery

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.396621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.396621Z digest=sha256:707514957f3b266cfcb3d677f7e696a729ea86c46aca29654efa0b5f4a9836da

Observation 206bfa80-3ccf-4ac1-9169-6c60e56dfd6e · outbound

This paper cites Neural ordinary differential equa- tions.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Neural ordinary differential equa- tions

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.621189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.401233Z digest=sha256:6b6ccacbabd928f7afc2fd33a5a94451f2961eb5907c9dc9556fde93278237cb

Observation 8bac707c-32be-4615-8705-715a5b0d3b89 · outbound

This paper cites Panda-70M: Captioning 70M Videos with Multiple Cross-Modality Teachers.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Panda-70M: Captioning 70M Videos with Multiple Cross-Modality Teachers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.404444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.404444Z digest=sha256:376eec2a2d876c2f1bfc2ec52c9821883b5569141970eb8c9be5e9eae7879f0a

Observation d05b8f2a-ff8b-428f-938b-c4e34653b043 · outbound

This paper cites Speed is all you need: On-device acceleration of large diffu- sion models via gpu-aware optimizations.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Speed is all you need: On-device acceleration of large diffu- sion models via gpu-aware optimizations

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.611589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.407840Z digest=sha256:1c08fa44c44cb4f2e3509a12e8c88bfb0e51cb2ecfebf91dd9a7457514c0f20d

Observation b2fbcf9f-e4d9-4d0c-b419-632d18016632 · outbound

This paper cites Squeezing Large-Scale Diffusion Models for Mobile.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Squeezing Large-Scale Diffusion Models for Mobile

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-09T10:46:30.194008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.411129Z digest=sha256:38db434516b89a42c80bf4092a3b2c2cb65bf314f2f55cf55a6facf41610d6e4

Observation e98c8e93-0e81-4338-8830-f5418dc0e2d1 · outbound

This paper cites Diffusion models beat gans on image synthesis.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Diffusion models beat gans on image synthesis

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.414863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.414863Z digest=sha256:afa9dab4c650ceebb75f2cba6ae0c6550380761f4529592d6597bb90dd37ce3b

Observation 082b9bcf-43e1-467e-b6f1-980fc64182da · outbound

This paper cites Tutorial on Variational Autoencoders.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Tutorial on Variational Autoencoders

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.418025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.418025Z digest=sha256:0e4ab8904e2b86f6308bb0ecfe6cd280933b23890d46b70d5872fb21ba061a9a

Observation 09da7e6f-0232-475c-bc7f-9b0d6f4241ac · outbound

This paper cites Scaling recti- fied flow transformers for high-resolution image synthesis.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Scaling recti- fied flow transformers for high-resolution image synthesis

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.421885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.421885Z digest=sha256:588a01b4029727d12382a571eae5df3c2128320f2a3cb2bdddb3b7d95bbc5947

Observation 3ee9dc0f-dd24-4f34-b040-c1ed32ef953f · outbound

This paper cites Efficient vision trans- former via token merger.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Efficient vision trans- former via token merger

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.591324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.425244Z digest=sha256:703cfa416c434d23e45e198628235ce279ab6a60e4c6922052e322c64984f3bf

Observation a1df505b-b859-420f-929f-e541a7540628 · outbound

This paper cites Efficient Time Series Processing for Transformers and State-Space Models through Token Merging.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Efficient Time Series Processing for Transformers and State-Space Models through Token Merging

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.428727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.428727Z digest=sha256:7cb6c93d295450eb7d81ba063fe77a8f584abf7453a61060ab38f5d4b2a8b403

Observation 78da2a77-51d5-448e-a97a-564a67492963 · outbound

This paper cites Knowledge distillation: A survey.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Knowledge distillation: A survey

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.583382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.431950Z digest=sha256:7f84983d44b2d472d3e13c81d3e0e74616d7134808a8f0660047abfed717b3b1

Observation a48c0b2f-fd1d-43fe-9243-333fa2e2c433 · outbound

This paper cites Gray and David L.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Gray and David L

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.575665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.434964Z digest=sha256:1f186eb58231a80766173c1e93a6a5bd0b234729df677e38de012a4a002d743d

Observation ff597029-a652-4459-be09-7b3730f90384 · outbound

This paper cites Flexible diffusion modeling of long videos.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Flexible diffusion modeling of long videos

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.568339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.438121Z digest=sha256:eca4bae8dc1fdec827c242e681d11631f494547e3c355d13fe38eb1a33e7cf4d

Observation fc0e1528-0f4f-4dfa-8a90-f3494fdd3f40 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Distilling the Knowledge in a Neural Network

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.440873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.440873Z digest=sha256:ef34f7d48ce9593aef535cf6d492de62a1a2c3d0b6147fe1f7bb56ff3b5a48e4

Observation 8764cf74-0058-4d4f-8b1f-286b5eeb4633 · outbound

This paper cites Denoising dif- fusion probabilistic models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Denoising dif- fusion probabilistic models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.444744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.444744Z digest=sha256:24c9c0d22a696bc77240937b6320e15f3ba8906d1c14cc2e2732e43ac1039294

Observation 2e2f7bd8-f1e1-4dfc-9fd9-06eed829b9b7 · outbound

This paper cites Imagen Video: High Definition Video Generation with Diffusion Models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Imagen Video: High Definition Video Generation with Diffusion Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.447958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.447958Z digest=sha256:961d84a2b53f8ae072f3af4f22001aa3d8078fae9aaeec164e56099a1c530328

Observation 1233f79a-4018-4e93-831b-e5ef2d1e072a · outbound

This paper cites Cascaded diffu- sion models for high fidelity image generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Cascaded diffu- sion models for high fidelity image generation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.554150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.451148Z digest=sha256:263534264611ea46869d6228252d29e83cff0c9b6e516b6e6f764fd6ee705525

Observation 21cd6704-090f-419d-913f-c7584f4e3d58 · outbound

This paper cites Video dif- fusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Video dif- fusion models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.454778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.454778Z digest=sha256:13c76be285ae5dc1cbf640d8ea47f1b336567131b40ec0903ca5d9fa83fbec28

Observation bb64e608-68c8-4f73-8a1e-8552655ca4cb · outbound

This paper cites Toward controlled generation of text.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Toward controlled generation of text

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.539287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.458128Z digest=sha256:af5e88882fb9f96255fec6f36a2ea2c9d8e93960234a4ceefbde852f746ea75d

Observation 26e2f92f-9178-4318-b5ee-54de7701daad · outbound

This paper cites Vbench: Comprehensive bench- mark suite for video generative models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Vbench: Comprehensive bench- mark suite for video generative models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.530247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.460899Z digest=sha256:8ef80308b2a20c62d66cd4d47d0523abe93a2a30f4130b396f4a162819ef43a1

Observation 2e21fd85-4f33-494d-8fa3-236bda7c6376 · outbound

This paper cites Pyramidal flow matching for efficient video generative modeling.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Pyramidal flow matching for efficient video generative modeling

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.463824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.463824Z digest=sha256:7cbd407636fe9581c654d1143b08fee800ed9655ce731602ee445dbe8e078a81

Observation bed72ca8-75b0-4ad6-a0c1-8dbc7d2dde3d · outbound

This paper cites Imagic: Text-based real image editing with diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Imagic: Text-based real image editing with diffusion models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.467231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.467231Z digest=sha256:80bc356d36ddb93a19e62bed857b8b0df7dbc4982adfe685ac49112ddf2cc951

Observation bd918bd0-689b-47fd-9b43-b6cd1ef0a494 · outbound

This paper cites Text2video-zero: Text- to-image diffusion models are zero-shot video generators.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Text2video-zero: Text- to-image diffusion models are zero-shot video generators

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.515093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.470388Z digest=sha256:e0e5cdf075878ace64f3621c6a768e464a9c80ffeb22625a8dbdc8c809371764

Observation 3771e4df-b0c1-476c-9726-0a91c3d84054 · outbound

This paper cites an unresolved cited work.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:46:30.505452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.473312Z digest=sha256:d42bba4e677934f925f4b54cc80c0b762f4b274838f2d1afd81e0c0845225398

Observation 3ba24b8c-a1c6-4dee-8573-c2c4688b98f9 · outbound

This paper cites xformers: A modular and hackable trans- former modelling library.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices xformers: A modular and hackable trans- former modelling library

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.476228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.476228Z digest=sha256:b5746ab5cf2ee5245d40508eba732eeb8d55a6fca8d1ca244ae51bcb765a14e8

Observation 696a61ec-b843-4404-ace7-91a6a286ff04 · outbound

This paper cites Vidtome: Video token merging for zero-shot video editing.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Vidtome: Video token merging for zero-shot video editing

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.490227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.479554Z digest=sha256:5ebdea8003ccf7e2bd83f6095a2a7717e863c261d93a4612c8b272788505b5ed

Observation fd85b3f3-3459-495d-8fe5-69f607bb5947 · outbound

This paper cites Snap- fusion: Text-to-image diffusion model on mobile devices within two seconds.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Snap- fusion: Text-to-image diffusion model on mobile devices within two seconds

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.480295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.482337Z digest=sha256:2e638c049fe4d5edc1238d1f6d9c4ccef072297b4725392fef860e8dd3071992

Observation 4e656760-e6ab-46ee-b9ad-1163943b6ca1 · outbound

This paper cites AnimateDiff-Lightning: Cross-Model Diffusion Distillation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices AnimateDiff-Lightning: Cross-Model Diffusion Distillation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.485547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.485547Z digest=sha256:9e2e9d314b195120314986d37699fc54218d2394d8906513393cc538992ee922

Observation 8e9088e1-93ae-4668-bcc5-07bec1ac74ee · outbound

This paper cites Generative adversarial networks for image and video synthesis: Algorithms and applications.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Generative adversarial networks for image and video synthesis: Algorithms and applications

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.470396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.488642Z digest=sha256:79f8e1908d42fa43aa24748d55cf8042660b6c42d366975cabc0263aae3f8ad8

Observation 586056c2-deff-44d2-8da3-fa18a7d3801e · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.491631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.491631Z digest=sha256:9f0b7f7acc24f5c9c91e3e70905489d4b880ad01e0769ccddd7ff0874a4011bd

Observation 1ef678db-8144-48e5-a6f3-f87dfa795533 · outbound

This paper cites Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.494514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.494514Z digest=sha256:3f76a765a49dc5df20339bc62f060e09b9ddaf4e4e21f35ba689009dcc81099c

Observation d32a0d21-28df-417a-ab70-bd08cd9ad0cf · outbound

This paper cites DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.497660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.497660Z digest=sha256:987abc506ceb4054923731280471cc2ac7bb7ab1ee9e95be64a76b72604dcf8d

Observation 39e8164d-8b45-4f5a-89ed-9685c831a4cb · outbound

This paper cites Snap video: Scaled spatiotemporal transformers for text-to-video synthesis.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Snap video: Scaled spatiotemporal transformers for text-to-video synthesis

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.460738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.501071Z digest=sha256:5442f95f0d2995f8d76c41b50ec6794dc22ac098c5f58fee6b9e85218c6e8bcb

Observation 2602ef1d-0efb-4b3f-b0d7-7e312d1c4747 · outbound

This paper cites Dreamix: Video Diffusion Models are General Video Editors.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Dreamix: Video Diffusion Models are General Video Editors

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.504193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.504193Z digest=sha256:85e80b6154be7b180ef0f55384697591bc3d92c81b02f72fade3229c742abfa1

Observation 07263c9f-0af4-4d22-96eb-3b29230b11ce · outbound

This paper cites A review on the attention mechanism of deep learning.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices A review on the attention mechanism of deep learning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.451211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.507341Z digest=sha256:ee24448340ea10aa83efd0ffab9060a679b782416c9656810056659e72a23d87

Observation ef2d38b3-4c91-455e-96cb-e45467715c29 · outbound

This paper cites Generative models for video analysis and 3D range data applications.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Generative models for video analysis and 3D range data applications

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.441817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.510095Z digest=sha256:ddbc07b5ad54f97cb3170a2adf68b73a0052881ba7887296dab6185a66d0df94

Observation 44b0ef8b-c9a2-4588-9887-dacc24c32669 · outbound

This paper cites Pytorch: An im- perative style, high-performance deep learning library.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Pytorch: An im- perative style, high-performance deep learning library

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.512914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.512914Z digest=sha256:566262d716ace905c10bf639b536381a430866e4f640a0617e7e83030644017d

Observation c8932e36-323f-4ccd-880b-5dc008f09e21 · outbound

This paper cites Scalable diffusion models with transformers.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Scalable diffusion models with transformers

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.515489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.515489Z digest=sha256:2a0fad6b3086afd6255d3e608d319f39e1bb448c1eabb15711d4afc2e161958c

Observation 9db81d19-5ed9-4496-b2ab-bcc34e9519d0 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.518458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.518458Z digest=sha256:8b0bb38aa3bafdd67d9daefc0661d9cde028f9ea2998cdef42796c6f4bc87b2e

Observation c85b33ee-5cda-4f47-8d0f-312674fd8187 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Exploring the limits of transfer learning with a unified text-to-text transformer

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.422866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.521518Z digest=sha256:bdcaba53023640b58948842d9813ae329657a7a49a4db8ebd06bac2b974d12f1

Observation 3039f780-abeb-461e-94ba-7d4d0770f774 · outbound

This paper cites Pruning algorithms-a survey.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Pruning algorithms-a survey

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.414999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.524299Z digest=sha256:fc3c055df580a97585fd734fc575443b89d6f3bbadec8fd3e5a6cc424ffa8684

Observation 270a6090-4cf2-4f80-bab7-08636b7ae85b · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices High-resolution image synthesis with latent diffusion models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.527356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.527356Z digest=sha256:2408a3930f3b38fc8cdf2fd3fb280e7b08a23b6dd5cd081cfdf7a0372a266255

Observation 0b270534-f790-4bbf-8cfa-2ab72efd0f3d · outbound

This paper cites U- net: Convolutional networks for biomedical image segmen- tation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices U- net: Convolutional networks for biomedical image segmen- tation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.402361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.530078Z digest=sha256:44cebfb4b528f2fbabb69f84d5a2f75e1e733809f4c35bacbc6cc596b318b597

Observation ead3ae0b-4280-4114-a803-5351e5bf52f4 · outbound

This paper cites ¨Uber die numerische aufl¨osung von differential- gleichungen.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices ¨Uber die numerische aufl¨osung von differential- gleichungen

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.394529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.532752Z digest=sha256:1d0755a17efa6725cc8c49bdeb997d80aa03215dc972e6fab2ba52e5346d5343

Observation ce006bea-9d8d-4f78-9619-3e29b5e88b1c · outbound

This paper cites Palette: Image-to-image diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Palette: Image-to-image diffusion models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.535718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.535718Z digest=sha256:2dfc16d80ca427694d86fabcd5b62042af5a791927ea0cbbbabb7e80f8631d04

Observation 10783bd1-71cd-420a-a34f-ab04eb992168 · outbound

This paper cites Introduction to apple ml tools.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Introduction to apple ml tools

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.380252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.538372Z digest=sha256:7cff2294b83666b9baa2268727b3247960b0ad4bde1cdb6ea67c1abe98a6e9a0

Observation 93457ce0-f9d8-44ca-be70-9ee14336f7f0 · outbound

This paper cites Sin- gan: Learning a generative model from a single natural im- age.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Sin- gan: Learning a generative model from a single natural im- age

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.371044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.541123Z digest=sha256:401106ec7bb0037ea1ab2cbfce86924881d1a90d90216eb66a8eb44bce481fac

Observation 04f5b16c-d2f1-4cb1-8064-e7250f50cd06 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.544214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.544214Z digest=sha256:ea5078621f00d0dca06bc401e2967c495a97e24cc2ce19c1687d3c69e445fd96

Observation feeebf05-e61a-42bf-a061-22ca690883e4 · outbound

This paper cites Video edit- ing via factorized diffusion distillation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Video edit- ing via factorized diffusion distillation

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.361436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.547315Z digest=sha256:247278dabdff3a0de0f4591d9967415cb363c2ab0619faa5e1ef1a884ea8448b

Observation 9b8fe24f-ce86-4ddb-bc7c-ca3ac3a8181a · outbound

This paper cites Denoising Diffusion Implicit Models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Denoising Diffusion Implicit Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.550642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.550642Z digest=sha256:b4344781b001fbcf39d562bccea7c90e419d817844a4fc7e0a974db2063c6826

Observation f2162860-5301-493d-b63d-ae706d0add3d · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Roformer: Enhanced transformer with rotary position embedding

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.553654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.553654Z digest=sha256:47dc52ed9c001ab0b3b73a6c7f0eee9d618595068ee8a8d926e85ef29053e39a

Observation df0ca600-10e6-44fd-a97f-c0206f804237 · outbound

This paper cites A survey of multi- modal deep generative models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices A survey of multi- modal deep generative models

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.347655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.556817Z digest=sha256:f18f95611b0c4286e1fe1eb1c9b1c3a26742003658be06167e74ca357f6ff6c6

Observation 67f6d166-4f8b-4b80-aed4-917be129f93e · outbound

This paper cites VidGen-1M: A Large-Scale Dataset for Text-to-video Generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.559618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.559618Z digest=sha256:2524c4d71fd2f54264f28087471ee994d722c1e9fa814713e30d50b2c4281ea3

Observation 61403ab1-9933-43aa-8083-c041c56facff · outbound

This paper cites Qvd: Post-training quantization for video diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Qvd: Post-training quantization for video diffusion models

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.338317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.562630Z digest=sha256:7f0b71d4b22391240373a3296f25f7af22fc04b5555d2e096844f0442ade7793

Observation f615afc3-0795-42a7-a66b-9d8d0080fe72 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.565679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.565679Z digest=sha256:a4c766ca29f37ef845410d1cb34c0c15e3524f3f618dc1d791587efbe8d98287

Observation 67196b35-7a68-4182-a219-44d2f8e63072 · outbound

This paper cites Mobileone: An im- proved one millisecond mobile backbone.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Mobileone: An im- proved one millisecond mobile backbone

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.329357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.569420Z digest=sha256:f93c5c592037d60d76ebc0ad80a6631c0a0ee4648c3f75ef615968434a7f7a10

Observation b43bfd82-e448-4158-b1c2-76e8f47eafba · outbound

This paper cites Mcvd-masked conditional video diffusion for prediction, generation, and interpolation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Mcvd-masked conditional video diffusion for prediction, generation, and interpolation

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.572440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.572440Z digest=sha256:a49dccd948d62310ae23cc20fca9871c0543f40a5bfe50f4b3ab25b31d2d939d

Observation e6309612-bcbf-4f10-ba22-3aa9b716c1ad · outbound

This paper cites Animatelcm: Computation-efficient personalized style video generation without personalized video data.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Animatelcm: Computation-efficient personalized style video generation without personalized video data

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.314893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.575531Z digest=sha256:ddf08ad4a8c6b9566d14e166bb3623bc541dff2a559bde3e1b00ca227d031b07

Observation df7b0ce5-2024-4fdb-a622-8d598d2fde69 · outbound

This paper cites Transformers: State-of-the-art natural language processing.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Transformers: State-of-the-art natural language processing

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.299223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.581493Z digest=sha256:b252de3b9991abaf395c03346ce3f764a4b0aef7eb55825260f4f636f52517ae

Observation 6423c304-6464-4871-8a85-ddd87977ee88 · outbound

This paper cites Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.584244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.584244Z digest=sha256:bb9c5e5d74025784692c8850e8c10288851a79f15187c95ee8b5570b750b1710

Observation 1b161b5f-bf0a-4b58-be06-fd4762d0e4f2 · outbound

This paper cites Individual Content and Motion Dynamics Preserved Pruning for Video Diffusion Models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Individual Content and Motion Dynamics Preserved Pruning for Video Diffusion Models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.587315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.587315Z digest=sha256:4bd2de80fa71b6b6af1d78d8b2318991c8cbdaf324df6086064e2c8c5948627f

Observation aaa0229f-04b8-48cb-804a-699ef936120f · outbound

This paper cites SnapGen-V: Generating a Five-Second Video within Five Seconds on a Mobile Device.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices SnapGen-V: Generating a Five-Second Video within Five Seconds on a Mobile Device

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-08-09T10:46:29.899800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.590722Z digest=sha256:eaaea428425f0f65fced6b15f5d7db7c7f55773c28e02920d5936c5af055ab2a

Observation 2557feb2-562d-42f1-9143-2911dd6c48d5 · outbound

This paper cites Mobile Video Diffusion.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Mobile Video Diffusion

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.593427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.593427Z digest=sha256:950bcf901ced0e5f252e949630daa32ef8a4a838ddc7d8a435ed4d3a00583e54

Observation ffc90300-e8d1-4b71-a906-a569e90c518f · outbound

This paper cites Stat: Spatial-temporal attention mechanism for video cap- tioning.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Stat: Spatial-temporal attention mechanism for video cap- tioning

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.284615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.596470Z digest=sha256:77d7b6c54f8c2f21a858a54e7c5d3c15bacf96ed2bc60a861ed60b319c580b01

Observation d234fb92-3657-4860-b74a-e40336f19908 · outbound

This paper cites Diffusion models: A comprehensive survey of methods and applications.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Diffusion models: A comprehensive survey of methods and applications

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.276159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.599249Z digest=sha256:64f4ffdd5e110482773ace8db4b39c7dc0cf2520669d322b00b7d934e71449ee

Observation 05eeb88a-777c-48c2-9c35-3bd3b4e96fcf · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.602176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.602176Z digest=sha256:7705e0f42bd3e5ac1cadf984164449381a8452a259076c093a0d4ace3eeb5e73

Observation 3a406da9-43db-4b6b-b7f1-9cb8a42e4270 · outbound

This paper cites Schedule On the Fly: Diffusion Time Prediction for Faster and Better Image Generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Schedule On the Fly: Diffusion Time Prediction for Faster and Better Image Generation

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.605280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.605280Z digest=sha256:1bb542a7d31c1a8923396020c1fbca9a4d25ae9272304388628557a5fd4cce6e

Observation ebd32267-17fe-4703-b90e-a0b28354014c · outbound

This paper cites Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.608228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.608228Z digest=sha256:1f438c8319a38fbb66480fa5d777503c7700a76dc1d18d564eb22123c2e3f124

Observation 72439b63-bdc9-42bb-bf5d-ac11f153215b · outbound

This paper cites Text-to-image Diffusion Models in Generative AI: A Survey.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Text-to-image Diffusion Models in Generative AI: A Survey

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.611120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.611120Z digest=sha256:692573df367729642c6096b4c3d40871fa3993f488112df1e2fc63da208a3451

Observation 450f723c-a591-439f-95a7-4a682fd8655c · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Adding conditional control to text-to-image diffusion models

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.614652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.614652Z digest=sha256:0041eb79c69b7c3124565c2a004e7b6b994b7491d29291564849a54924ca84de

Observation 65158356-8f47-4256-b959-f77e707502df · outbound

This paper cites A survey on personalized content synthesis with diffusion models.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices A survey on personalized content synthesis with diffusion models

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.617436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.617436Z digest=sha256:e36f6abbec19115f600d1b4b588ad3307e29dfc5b92dce769572671f9a79f949

Observation 7cdbc0f6-9940-4f43-a386-28a5475f1d57 · outbound

This paper cites ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.620727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.620727Z digest=sha256:16f40f2d07e65f2115f091b05c0b652b17a4f36ee31659f6d2f5e86d22fb9ff5

Observation 05ce1640-800c-4717-97f1-40bc758d01ea · outbound

This paper cites MobileDiffusion: Instant Text-to-Image Generation on Mobile Devices.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices MobileDiffusion: Instant Text-to-Image Generation on Mobile Devices

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.624096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.624096Z digest=sha256:7995f15c3757af94b31165e4e04982727187116d2168286ab3476b483f13851c

Observation cf07997b-b53d-4c9a-8985-0f8c7fabc346 · outbound

This paper cites Dpm- solver-v3: Improved diffusion ode solver with empirical model statistics.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Dpm- solver-v3: Improved diffusion ode solver with empirical model statistics

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.262870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.627415Z digest=sha256:4be7de0f6d5a075753e82b8d4024a7b2d2d561882c444b59f7d50c806c12c684

Observation 3b8a40ab-2be7-4bfd-9aca-9995e0813d28 · outbound

This paper cites Open-sora: Democratizing efficient video production for all, 2024.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Open-sora: Democratizing efficient video production for all, 2024

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.253872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.630373Z digest=sha256:61f3da568667caa1a0a28fdbe1369accecc440a78ca1abe0a294250a8ccb5044

Observation c2c03a01-b1e0-4aca-aa7c-47f814db441b · outbound

This paper cites Slimflow: Training smaller one-step diffusion models with rectified flow.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Slimflow: Training smaller one-step diffusion models with rectified flow

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:46:30.244253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.633290Z digest=sha256:02dd80b6eaeeecdbc7deaec6da953969e9e9ce844757cd678b59c75a68465b29

Observation 54e43f0d-4266-4184-85f0-f8e662045ba0 · outbound

This paper cites an unresolved cited work.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Unresolved cited work

Reference 86

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:46:30.233552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T10:46:29.637028Z digest=sha256:7e89518c8b329bf03e777774d1a8ef25b2c107039af7202ad2329c08ac0197e2

Observation 3c6a34cc-29a6-495e-80a3-9366d3de6440 · outbound

This paper cites an unresolved cited work.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices Unresolved cited work

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.578325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.578325Z digest=sha256:7b64f959c38a82712fd9a8898b207b44a1ff27d497df64d1b32da97efc4bbafe

Pith citing papers

Observation 47c6b7c4-9155-4791-8fcb-745aa841879f · inbound

Video Generation Models as World Models: Efficient Paradigms, Architectures and Algorithms cites this paper.

Video Generation Models as World Models: Efficient Paradigms, Architectures and Algorithms On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices

Reference 155

Resolution
verified exact
arxiv_id, observed 2026-05-14T01:38:36.426016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T01:35:14.878069Z digest=sha256:bd26c05ea5be4e93c7164b08e8d594dcbf4df74ff639834e86552d66ccefac61

Observation 2f294d63-0116-4b9e-bba8-7afd5bf3364b · inbound

Efficient Video Diffusion Models: Advancements and Challenges cites this paper.

Efficient Video Diffusion Models: Advancements and Challenges On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:03:26.202726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T08:28:29.706249Z digest=sha256:5f13647235569e8e6350d116dad287a34389b303227ee42e1419dd22209bf24c

Observation d90978ab-05e8-4c89-aa73-051ddd806df8 · inbound

MobileWan: Closing the Quality Gap for Mobile Video Diffusion cites this paper.

MobileWan: Closing the Quality Gap for Mobile Video Diffusion On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-07-08T14:44:59.785369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-08T14:37:46.957265Z digest=sha256:d5fc3677efb09b5bf5f73f9c968405e7787a6a03115607a01f23f3da5b982319

Observation 28edac6e-cc20-40fd-9a52-b3ee690f6fbd · inbound

MobileWan: Closing the Quality Gap for Mobile Video Diffusion cites this paper.

MobileWan: Closing the Quality Gap for Mobile Video Diffusion On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T08:21:44.482775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:21:44.482775Z digest=sha256:2480eaaed70abd0a51635be29445e47ef8503a994f6f29a94f52bcc2a25932ee