Pith. sign in

Paper Citation Record · LEDGER

Open-Sora Plan: Open-Source Large Video Generation Model

As of 5 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 89 inbound Pith citation observations for arXiv:2412.00131.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.00131 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-23T08:38:27.946746Z

measured 119 of 119 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 89 of 89 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:49:50.891081Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T14:44:59.776032Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact27
  • verified fuzzy2
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6b2b49bb-f279-4efc-b29d-8ca76ff7d42b · outbound

This paper cites MiDaS v3.1 -- A Model Zoo for Robust Monocular Relative Depth Estimation.

Open-Sora Plan: Open-Source Large Video Generation Model MiDaS v3.1 -- A Model Zoo for Robust Monocular Relative Depth Estimation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:42:45.263542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:37db9cb4aad9ba7d3ac048ee3f524c5967837401f9489490e4176027b9a6d8cd

Observation a49ec532-a065-4d3f-af8b-b838a26ca687 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Open-Sora Plan: Open-Source Large Video Generation Model Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-23T08:42:45.258752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:dd6d38ee06cf632b11a0fbad2a8dcd2911aaba43f0265890ac6ba118eed1bd4b

Observation 5efd29ea-e9f0-4d9b-bbb7-c04c01814d08 · outbound

This paper cites PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation.

Open-Sora Plan: Open-Source Large Video Generation Model PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:42:45.242218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:d1e9473fd1feb08656611418eafc1ca960fc3258d7810aa13faaa783a544672f

Observation 8d44dcba-f8c7-4dc3-84b5-2ff666933556 · outbound

This paper cites The Llama 3 Herd of Models.

Open-Sora Plan: Open-Source Large Video Generation Model The Llama 3 Herd of Models

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-23T08:42:45.274366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:581f19b944c147d553f9d750339773f17128c0c014215e2f06d82d12d48d37be

Observation 74ae9d8a-17ad-4595-b1ea-2b11402987ea · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

Open-Sora Plan: Open-Source Large Video Generation Model AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-23T08:42:45.219292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:7fe75077d9ef88e4d4a0be1576b410c14f3d1b12face1394b500b1159fcfbf56

Observation eec5d34f-4858-43c2-9dc5-859c429362d1 · outbound

This paper cites Image quality metrics: Psnr vs.

Open-Sora Plan: Open-Source Large Video Generation Model Image quality metrics: Psnr vs

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T08:42:45.634602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:6729bca7f393856c456de2408ad0ec38b80d355975efa3a9eb753313eb299fa6

Observation 4f87e937-32f7-4444-aa5e-5349f0da9107 · outbound

This paper cites Mistral 7B.

Open-Sora Plan: Open-Source Large Video Generation Model Mistral 7B

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-23T08:42:45.188285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:ada7d08721e3a10404c02273ce6f04c4bea7e4bf5ccfb651ce64e2e2e7496776

Observation 41b9b3fb-433a-4e58-a9b7-1e3ab7cd8412 · outbound

This paper cites Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding.

Open-Sora Plan: Open-Source Large Video Generation Model Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-23T08:42:45.165303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:401c675737ffe29621fa39b42195e59042efd6b9afbde7eee336426fd58ad02b

Observation c1866d45-31ca-4a64-b882-3a571da99dcd · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

Open-Sora Plan: Open-Source Large Video Generation Model Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-23T08:42:45.293672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:f004915e466d50e0b2bf0ace9391ec1dd4a30cf0e6cd4574b71df2c39dc6ca2b

Observation f275a2c5-9834-4991-8383-a77cc9aca797 · outbound

This paper cites MoE-LLaVA: Mixture of Experts for Large Vision-Language Models.

Open-Sora Plan: Open-Source Large Video Generation Model MoE-LLaVA: Mixture of Experts for Large Vision-Language Models

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-23T08:42:45.135889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:b49cc8e2bcc4c2534f6df6438f3bbd0e2dcb3cc6696e2aedf749f8ee841f328f

Observation 84ff7069-da8d-4cd0-a9ed-5cfbf7399167 · outbound

This paper cites Flow Matching for Generative Modeling.

Open-Sora Plan: Open-Source Large Video Generation Model Flow Matching for Generative Modeling

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-23T08:42:45.230701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:1e7462489fb264cde5bb0728dc2d1dac4cb315a7b8af3ba8cdee7911b415a66d

Observation 01114a68-abbe-4329-91d7-036312b2936f · outbound

This paper cites Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models.

Open-Sora Plan: Open-Source Large Video Generation Model Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:42:45.153879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:8063512ad7c2bdfc78f740743338f9b30733aa4203e365d22f44e230699eecd8

Observation 73009a21-2caf-4f69-81db-9eb190b1b2c6 · outbound

This paper cites FiT: Flexible Vision Transformer for Diffusion Model.

Open-Sora Plan: Open-Source Large Video Generation Model FiT: Flexible Vision Transformer for Diffusion Model

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:42:45.269086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:96b13306b1f138713dcffb7eefe7813e68f2a9696ce862dd803da7019e824951

Observation ae45a8e6-3400-4e8d-8a6e-562d2c7483bd · outbound

This paper cites ControlNeXt: Powerful and Efficient Control for Image and Video Generation.

Open-Sora Plan: Open-Source Large Video Generation Model ControlNeXt: Powerful and Efficient Control for Image and Video Generation

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:42:45.280474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:ccf53e9f3a931c2d9fa59654408d0ddb167f4beb8769eca9a78f2b21d7665b68

Observation 87101d13-8322-4bb1-9546-2c05cb45c359 · outbound

This paper cites Movie Gen: A Cast of Media Foundation Models.

Open-Sora Plan: Open-Source Large Video Generation Model Movie Gen: A Cast of Media Foundation Models

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-23T08:42:45.177050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:b869eb1c66cc338846890a32e23c5eccc6e8fac4e4cdf40f6bda90b6b5f2f261

Observation 75907270-ba0a-48fd-912c-f78279ef2539 · outbound

This paper cites Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism.

Open-Sora Plan: Open-Source Large Video Generation Model Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-23T08:42:45.182411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:d6ab23c3a19168604ed0d9401d6f7b73f1a8fc26f636f7ad4be29f5b1fb887a1

Observation 1097b401-e7e9-4cf0-b933-4909e9ad59b9 · outbound

This paper cites Denoising Diffusion Implicit Models.

Open-Sora Plan: Open-Source Large Video Generation Model Denoising Diffusion Implicit Models

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-23T08:42:45.253335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:75acdc5f7d63ad63ee5229f8a0965a0847a7676d163062cf6516766a9d979aa9

Observation 3b692f4f-ce5c-4b04-a6f3-7ff30684302e · outbound

This paper cites AnyText: Multilingual Visual Text Generation And Editing.

Open-Sora Plan: Open-Source Large Video Generation Model AnyText: Multilingual Visual Text Generation And Editing

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:42:45.286343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:ce3234f886373975a9a0b8a1f5a5f3c539ee260fb3edc3ec901f5e98ed261fe5

Observation e108ea2e-3b48-4b8c-91e9-463ea9edbf50 · outbound

This paper cites Tarsier: Recipes for Training and Evaluating Large Video Description Models.

Open-Sora Plan: Open-Source Large Video Generation Model Tarsier: Recipes for Training and Evaluating Large Video Description Models

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:42:45.194961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:78e4e06d6947d4356b3de56387173210a84eb646cafc6caed1970e797e316ee4

Observation 43616cc8-96ec-4f90-a7e8-090755816a2f · outbound

This paper cites FiTv2: Scalable and Improved Flexible Vision Transformer for Diffusion Model.

Open-Sora Plan: Open-Source Large Video Generation Model FiTv2: Scalable and Improved Flexible Vision Transformer for Diffusion Model

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:42:45.171364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:74d8dea1d066ad82eb8f535bf1cfc036d9f0807135af6887dfeebff44c70d28a

Observation 0a694148-1dd5-4d2d-9aa8-7f77d94d869b · outbound

This paper cites Easyanimate: A high-performance long video generation method based on transformer architecture.

Open-Sora Plan: Open-Source Large Video Generation Model Easyanimate: A high-performance long video generation method based on transformer architecture

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:42:45.247823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:28f60c053cba6675addf9ed6e479dfd92c151564262c4d92a96e6de1df692422

Observation 9ab36c5c-65c7-413a-809e-11efef6a3673 · outbound

This paper cites mT5: A massively multilingual pre-trained text-to-text transformer.

Open-Sora Plan: Open-Source Large Video Generation Model mT5: A massively multilingual pre-trained text-to-text transformer

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:42:45.201690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:f68919b5bd212fd86f20538f582739fa899f5c3dfe013bb9278033f795fcbc0d

Observation b28c722c-3d13-4928-b454-f282b5979b9f · outbound

This paper cites Qwen2 Technical Report.

Open-Sora Plan: Open-Source Large Video Generation Model Qwen2 Technical Report

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-23T08:42:45.141234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:8f6e963bf73c87ee3d797920ff699acc98a58e9ea15e8012e6cb31c2b54525ad

Observation fa1e2b3a-1aca-45ce-bfec-f549f0c2e0ff · outbound

This paper cites mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality.

Open-Sora Plan: Open-Source Large Video Generation Model mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-23T08:42:45.159395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:1e5deac8eea7c8ea4fd9e6d5e221c4cd4c36594767307417864ee2760308c216

Observation 6553fbeb-7e7e-4905-a0d0-8a312fc7ef82 · outbound

This paper cites Yi: Open Foundation Models by 01.AI.

Open-Sora Plan: Open-Source Large Video Generation Model Yi: Open Foundation Models by 01.AI

Reference 25

Resolution
metadata mismatch
local_arxiv, observed 2026-05-23T08:42:45.235533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:949c2af9bf2cddf633bb94c65d95e6b6bfc853a8b9dd83e04865a5aef9d9c476

Observation 13c5c85e-0dde-4253-a59c-690dc05874c7 · outbound

This paper cites ChronoMagic-Bench: A Benchmark for Metamorphic Evaluation of Text-to-Time-lapse Video Generation.

Open-Sora Plan: Open-Source Large Video Generation Model ChronoMagic-Bench: A Benchmark for Metamorphic Evaluation of Text-to-Time-lapse Video Generation

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:42:45.147833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:0a1fb13032ba6449efc8fbaaeeeacb4647096b713a67b75d5ae7998dc01efaa9

Observation f88a77b2-41f0-446a-bf18-a96b632ecaa3 · outbound

This paper cites Efros, Eli Shechtman, and Oliver Wang.

Open-Sora Plan: Open-Source Large Video Generation Model Efros, Eli Shechtman, and Oliver Wang

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T08:42:45.630623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:73de3904b06a003557d4fa26daf2b99fcb7c543b1234ee05e32b30412a3a228a

Observation 517148f7-01cb-45a4-8633-55d02be40235 · outbound

This paper cites PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel.

Open-Sora Plan: Open-Source Large Video Generation Model PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-23T08:42:45.224881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:e24ef038d1125374a02ccab0e5f8ce800c0832b76c13c9cbc0a294d3638efe36

Observation 89506578-e208-4f5b-80e6-2ac5e98685cc · outbound

This paper cites Allegro: Open the Black Box of Commercial-Level Video Generation Model.

Open-Sora Plan: Open-Source Large Video Generation Model Allegro: Open the Black Box of Commercial-Level Video Generation Model

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:42:45.213636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:b9155692db3647a572197157e7066bbc1a091ba8a8455d81fb3a248a1a2f0444

Observation c9741512-0adf-4e01-b09d-e322610482dd · outbound

This paper cites LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment.

Open-Sora Plan: Open-Source Large Video Generation Model LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-23T08:42:45.207252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:fa91f6423dcaf1f1b7899bb602290c467cfcab26ff28738736fe5ff148be8887

Pith citing papers

Observation 3b71d077-572e-4feb-af2a-6c581261a4ee · inbound

Latte: Latent Diffusion Transformer for Video Generation cites this paper.

Latte: Latent Diffusion Transformer for Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-13T21:45:35.801465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T21:45:35.754742Z digest=sha256:3af5dcb95563a1887e5a73d1a160fed77dc2c5802db61e3799492484a3624b54

Observation e1812281-c146-4237-9171-db0ae9f8da03 · inbound

Open-Sora: Democratizing Efficient Video Production for All cites this paper.

Open-Sora: Democratizing Efficient Video Production for All Open-Sora Plan: Open-Source Large Video Generation Model

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:01:51.686526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T12:01:51.366667Z digest=sha256:72a7c6c4ffcf1d2ceef6ec184faaa1baaa8b67db70a70e30b1260adf2448e023

Observation 93b9804f-b0e7-498c-809b-3508074ba2ad · inbound

Cosmos World Foundation Model Platform for Physical AI cites this paper.

Cosmos World Foundation Model Platform for Physical AI Open-Sora Plan: Open-Source Large Video Generation Model

Reference 112

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.652103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:65e3633c8425f662c5dcb5704458a207e31039e6d5c04882d1366e7e31d6a988

Observation a67e297b-5765-4bc4-8670-189dbb410da3 · inbound

History-Guided Video Diffusion cites this paper.

History-Guided Video Diffusion Open-Sora Plan: Open-Source Large Video Generation Model

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-16T12:00:14.747474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-16T12:00:14.672729Z digest=sha256:a8e77ac919f34dbff78050ca2d291203a018148c9c3e70ef453c01d3ee75779e

Observation 7475e662-ce52-48f3-97e8-f00428f6dfd3 · inbound

Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model cites this paper.

Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model Open-Sora Plan: Open-Source Large Video Generation Model

Reference 69

Resolution
metadata mismatch
local_arxiv, observed 2026-05-19T08:02:23.993679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-19T08:02:23.002090Z digest=sha256:4d1ab15b46de8585c187bdc0e29927b1c26f4d5f8b3276b0f0993c5d9604569c

Observation 8e5d2d71-1997-413f-9bcb-532db01d877e · inbound

Wan: Open and Advanced Large-Scale Video Generative Models cites this paper.

Wan: Open and Advanced Large-Scale Video Generative Models Open-Sora Plan: Open-Source Large Video Generation Model

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-22T23:07:14.329855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T23:05:32.595632Z digest=sha256:6a070c988ff6abeab1a9ab07c4c396a5bfb904378b5d21fd8bffceb36075bcc0

Observation f7993094-b263-4cc6-9285-452044a86869 · inbound

ImgEdit: A Unified Image Editing Dataset and Benchmark cites this paper.

ImgEdit: A Unified Image Editing Dataset and Benchmark Open-Sora Plan: Open-Source Large Video Generation Model

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-12T18:17:45.278588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T18:17:45.123690Z digest=sha256:5c074932cdb84c7f2906abbbf937d3c23db35d08505a929473bc003d8c9f2923

Observation f0f704d5-9fa9-4cae-a9b1-2f9a887b1cc9 · inbound

Latent Wavelet Diffusion For Ultra-High-Resolution Image Synthesis cites this paper.

Latent Wavelet Diffusion For Ultra-High-Resolution Image Synthesis Open-Sora Plan: Open-Source Large Video Generation Model

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-19T12:02:16.758598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T12:00:37.025335Z digest=sha256:9646117f3a46d513b132bd821ae99e65fba01503d5aa011a575ab3d1fa85ddb8

Observation db94e39f-31d7-4bf0-82b8-f27eccc9f687 · inbound

UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation cites this paper.

UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-12T17:34:27.100126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T17:34:26.951644Z digest=sha256:dc3baaddf8dd3a2b56031d66c98ab8fcc29c196619de5d2313c81d7be216f7d1

Observation 29c61da0-d32e-4460-a099-f21d17ad952a · inbound

Show-o2: Improved Native Unified Multimodal Models cites this paper.

Show-o2: Improved Native Unified Multimodal Models Open-Sora Plan: Open-Source Large Video Generation Model

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-12T18:51:16.016370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T18:51:15.428692Z digest=sha256:2467e4d8111dd2075ec5f38859ff9c6d0466f570a1b2ac149569d70bbc05ca62

Observation 34ca8adf-9162-44e3-8901-b29d53285c1a · inbound

GenHSI: Controllable Generation of Human-Scene Interaction Videos cites this paper.

GenHSI: Controllable Generation of Human-Scene Interaction Videos Open-Sora Plan: Open-Source Large Video Generation Model

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-05-19T07:27:09.022234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T07:25:41.219753Z digest=sha256:4c06c015818ec21708c56348d7052d027828f4191028d1e6607364b6f44c1a0b

Observation ac7bf415-ab17-4405-abdd-1cb849e57583 · inbound

HumanGenesis: Agent-Based Geometric and Generative Modeling for Synthetic Human Dynamics cites this paper.

HumanGenesis: Agent-Based Geometric and Generative Modeling for Synthetic Human Dynamics Open-Sora Plan: Open-Source Large Video Generation Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T20:49:50.891081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:49:50.891081Z digest=sha256:15c362ada63f4dd57fa1feef5f3dffeef6d62bb6edb00bce98925264430f6280

Observation 51874436-f855-475c-bd13-030ec8abd31e · inbound

Hierarchical Fine-grained Preference Optimization for Physically Plausible Video Generation cites this paper.

Hierarchical Fine-grained Preference Optimization for Physically Plausible Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T20:20:22.655498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:20:22.655498Z digest=sha256:7d44cb86e970c993edec421dd8bf4daf8830a037e8e398e0240e1e0b53bb0556

Observation d1b3b800-be56-4d7a-928b-a459b3f49f4a · inbound

From Black Box to Transparency: Enhancing Automated Interpreting Assessment with Explainable AI in College Classrooms cites this paper.

From Black Box to Transparency: Enhancing Automated Interpreting Assessment with Explainable AI in College Classrooms Open-Sora Plan: Open-Source Large Video Generation Model

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T20:17:47.794995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:17:47.794995Z digest=sha256:ca5a0379b8272c48efc54b19dd9c00b5bf32e41c9f4a98bfcd7f5d6f595d5fb4

Observation 735067cc-6a02-40f4-bfd8-7d3d24a5651b · inbound

Matrix-game 2.0: An open-source real-time and streaming interactive world model cites this paper.

Matrix-game 2.0: An open-source real-time and streaming interactive world model Open-Sora Plan: Open-Source Large Video Generation Model

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-18T22:36:53.250702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T22:36:31.044743Z digest=sha256:746f912a3b68521888db59f9ad2c757c3dc21cc541f602e2c518b55242606c74

Observation c549c3bf-6b12-4dea-a045-4f94bc930185 · inbound

Self-Forcing++: Towards Minute-Scale High-Quality Video Generation cites this paper.

Self-Forcing++: Towards Minute-Scale High-Quality Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:39:54.090184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T22:39:53.995700Z digest=sha256:27c55018d77954e65c9be5cd0cc11cafff900ac4cc264ba29f89d27d25cd25f0

Observation c586ab46-a947-48f8-b992-f7cbc7d348af · inbound

Taming Text-to-Sounding Video Generation via Advanced Modality Condition and Interaction cites this paper.

Taming Text-to-Sounding Video Generation via Advanced Modality Condition and Interaction Open-Sora Plan: Open-Source Large Video Generation Model

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T12:38:57.889197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T12:38:57.889197Z digest=sha256:b07521063e78bf45b35b03a7f41796b9c1374494c46fff39c2dea90e705d3574

Observation fbe4347f-4299-401f-a2fa-d910dfadb650 · inbound

Scaling Up Occupancy-centric Driving Scene Generation: Dataset and Method cites this paper.

Scaling Up Occupancy-centric Driving Scene Generation: Dataset and Method Open-Sora Plan: Open-Source Large Video Generation Model

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-04T08:07:44.374359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:07:44.374359Z digest=sha256:3046de9b99138df2c73cd5a7f3d0cc0ce5a500c9e2d53785c90f8c2ab91754f2

Observation eb804562-27f2-4c48-9e03-96e84d56e588 · inbound

Timeripple: Accelerating vDiTs by Understanding the Spatio-Temporal Correlations in Latent Space cites this paper.

Timeripple: Accelerating vDiTs by Understanding the Spatio-Temporal Correlations in Latent Space Open-Sora Plan: Open-Source Large Video Generation Model

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T22:12:44.379396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:12:44.379396Z digest=sha256:bbe3ecc8e1c499b1b206b3680306a1a66ccba9e3627f81c393df1cc1739d5794

Observation 91adbb46-3923-4cf1-b182-efd37378e8f5 · inbound

Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation cites this paper.

Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-05-16T18:17:55.190470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T18:17:54.943863Z digest=sha256:db3772bc19f533084e321a1a838d1b0be833b68a9c55f07f1c3875ce01d69d7b

Observation c079eb38-0b79-4445-9ada-c3c83337446e · inbound

TAGRPO: Boosting GRPO on Image-to-Video Generation with Direct Trajectory Alignment cites this paper.

TAGRPO: Boosting GRPO on Image-to-Video Generation with Direct Trajectory Alignment Open-Sora Plan: Open-Source Large Video Generation Model

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T11:36:14.203846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T11:36:14.203846Z digest=sha256:ddd0b49abb32a75b83dcaccc9c816104a240cebe3b8b3637aa0849389e542826

Observation a41442bd-9f94-40f2-9480-41d0d7a8d03f · inbound

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation cites this paper.

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-21T17:32:01.761159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:32:01.642256Z digest=sha256:dc61e9e0f19f9018bea491c4c028cc4add821af10114f14f6a239442e1570c46

Observation d57d7f1c-8e50-4a92-b9cf-4cdff5ed935e · inbound

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation cites this paper.

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-22T11:51:29.881813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T11:48:00.633421Z digest=sha256:23b4a8bfdb1f29b13ed10daa174bc792ffc9181e3b6f4937efe35834714b40c6

Observation fcbfb9e7-9349-41bd-a5a7-bff81398a621 · inbound

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation cites this paper.

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T05:30:26.409689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:30:26.409689Z digest=sha256:930b46a5fe9575451cc13876f004f7c3a3f8da18eb141e5d3d5d1e1bc79be8bf

Observation 45099116-c6a5-4335-87a8-373de93c005e · inbound

LUVE : Latent-Cascaded Ultra-High-Resolution Video Generation with Dual Frequency Experts cites this paper.

LUVE : Latent-Cascaded Ultra-High-Resolution Video Generation with Dual Frequency Experts Open-Sora Plan: Open-Source Large Video Generation Model

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T00:09:56.989756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:09:56.989756Z digest=sha256:5c6be6b5bdbea2563deb74834338b9d341f4bbe0bc620f0a5006da9bfa9e3789

Observation 1d540e82-3936-4497-ac33-bff8b77d236b · inbound

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation cites this paper.

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:20:22.412366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T22:20:16.320171Z digest=sha256:5c743b2047bc2fa4896c07de5bbb07e4f729fadea097489cc05ae3a231625291

Observation 8f59fcb4-a8b4-4292-b82f-53cefe45ce9e · inbound

MultiAnimate: Pose-Guided Image Animation Made Extensible cites this paper.

MultiAnimate: Pose-Guided Image Animation Made Extensible Open-Sora Plan: Open-Source Large Video Generation Model

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:20:16.114115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T19:20:00.030459Z digest=sha256:d0c691b3a37d5bf2a67ca54df3f0a8e8fd14ef3a948eaab8c2623ab37fe53b3a

Observation c334a3e2-1f8c-4c95-af5a-ba56ccca7387 · inbound

FrameDiT: Diffusion Transformer with Matrix Attention for Efficient Video Generation cites this paper.

FrameDiT: Diffusion Transformer with Matrix Attention for Efficient Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-15T13:30:01.683924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T13:27:23.641294Z digest=sha256:81ab13f454519af8440dfaba657fb7fe6be841ca28887a1aa72e29f8e1ad34a6

Observation d6abab1b-6424-44f2-addb-636396f40e24 · inbound

Attention Sparsity is Input-Stable: Training-Free Sparse Attention for Video Generation via Offline Sparsity Profiling and Online QK Co-Clustering cites this paper.

Attention Sparsity is Input-Stable: Training-Free Sparse Attention for Video Generation via Offline Sparsity Profiling and Online QK Co-Clustering Open-Sora Plan: Open-Source Large Video Generation Model

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:59:53.221692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T08:55:52.757489Z digest=sha256:f68c8f5932f4056aa3d4b2cbe06d5507fb8cc6f9c9c913ddaff1eb7b0c747236

Observation ad104016-13cd-4aaa-8a9a-7bf7cc47d74b · inbound

Evolution of Video Generative Foundations cites this paper.

Evolution of Video Generative Foundations Open-Sora Plan: Open-Source Large Video Generation Model

Reference 85

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:05:51.871148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T18:41:38.616611Z digest=sha256:9d77625f61a30d636d890f9993d5818dddec358899e3d25026eadc32b3657d52

Observation 8353aabe-6aa1-4242-8a43-366ddb86beb0 · inbound

Latent-Compressed Variational Autoencoder for Video Diffusion Models cites this paper.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Open-Sora Plan: Open-Source Large Video Generation Model

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.345396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:aef8ba07e3741c9f8e3c3eaf39dec6cc30496ac03c1da64d2390767148164833

Observation 117af888-d99e-4b7e-97e1-8f0468250913 · inbound

TS-Attn: Temporal-wise Separable Attention for Multi-Event Video Generation cites this paper.

TS-Attn: Temporal-wise Separable Attention for Multi-Event Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-10T02:53:29.936813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T02:45:10.577070Z digest=sha256:7d989481d283c19df3b5572efdfe9dfea125c87a2b2b3fe751f2c423c2b92859

Observation 496cc55d-fc71-46da-9f4e-2656bcc3591f · inbound

Reshoot-Anything: A Self-Supervised Model for In-the-Wild Video Reshooting cites this paper.

Reshoot-Anything: A Self-Supervised Model for In-the-Wild Video Reshooting Open-Sora Plan: Open-Source Large Video Generation Model

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-09T23:04:17.911804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T23:00:45.496971Z digest=sha256:85a472a8930f03d1b35b1e382788b1150617ee2bd9c79c379a4c296c2ce6c11e

Observation 5c47183c-a2b4-4dae-977b-477328becd18 · inbound

HuM-Eval: A Coarse-to-Fine Framework for Human-Centric Video Evaluation cites this paper.

HuM-Eval: A Coarse-to-Fine Framework for Human-Centric Video Evaluation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:15.296025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-07T16:53:42.069330Z digest=sha256:4215621340158fc85d530087b871d5430af75ad233894940596da1e9e1858c64

Observation 0e3cd7dc-8108-4b00-bfba-1e50b668614b · inbound

UniCustom: Unified Visual Conditioning for Multi-Reference Image Generation cites this paper.

UniCustom: Unified Visual Conditioning for Multi-Reference Image Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:02:23.947271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T05:53:21.851578Z digest=sha256:42cd4956e56b07ad92765b06cef2eb97f886c9f7c21da700c2a83ac4e6a5457d

Observation 09296c9f-dca5-4d05-9874-c99f9e738eeb · inbound

UniCustom: Unified Visual Conditioning for Multi-Reference Image Generation cites this paper.

UniCustom: Unified Visual Conditioning for Multi-Reference Image Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:03:03.393757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-14T22:00:01.349754Z digest=sha256:b8e917a9dd7de92d396e4a5106066ddecdb802fc5602154c86fad0a3ad3c80f0

Observation 42c09cb4-dd10-4228-ab68-bfc0d2b1c5bb · inbound

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation cites this paper.

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-15T01:53:28.814379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T01:52:14.874049Z digest=sha256:e5a978baaa9a1a70ea500a71a424dd634b31355a5ffce68bd1d371a5a3947181

Observation 561fe0e7-753e-4f7e-87e9-6e818a40bc91 · inbound

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation cites this paper.

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-20T21:59:06.492205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T21:54:33.902256Z digest=sha256:7bde0f0d07b80cdefc1d80c29fd010740a6c66f946331548fcb5303be8db9b81

Observation 5cf4393d-2fdf-442c-ad75-9ab48901537f · inbound

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation cites this paper.

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-21T09:14:05.676048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T09:12:35.777810Z digest=sha256:2af24d6fea207c953ddf701064a039644bb48e9991b9b2c460ddaf80f4e1c134

Observation 3d9c9957-f512-4bb2-ae6d-2f839a11a89e · inbound

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation cites this paper.

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-06-30T21:45:05.537792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T21:44:21.427570Z digest=sha256:09df508027e7527449438c4d33f83b13f3d30ffb225613ecb1f0eafae249e470

Observation 4123b95b-e10e-415c-8c3b-891195c6e1c8 · inbound

Causal Forcing++: Scalable Few-Step Autoregressive Diffusion Distillation for Real-Time Interactive Video Generation cites this paper.

Causal Forcing++: Scalable Few-Step Autoregressive Diffusion Distillation for Real-Time Interactive Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-06-30T21:05:04.370537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T20:59:34.847496Z digest=sha256:405a261b90e74e19b730e4ffd01f2546a0bd113a09bd36a4f6c9bb04b4bcf90c

Observation b38240eb-3593-4591-a439-069ac24261e2 · inbound

AtlasVid: Efficient Ultra-High-Resolution Long Video Generation via Decoupled Global-Local Modeling cites this paper.

AtlasVid: Efficient Ultra-High-Resolution Long Video Generation via Decoupled Global-Local Modeling Open-Sora Plan: Open-Source Large Video Generation Model

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-20T18:28:53.070649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T18:26:57.735246Z digest=sha256:4e5751c5928b497fd2105c31abedf81a7b857ef3c5d9a885c816af1dd27a5b9f

Observation b19502a1-9499-4b81-be7d-1ae9bc8f3f4c · inbound

GeoWorld-VLM: Geometry from World Models for Vision-Language Models cites this paper.

GeoWorld-VLM: Geometry from World Models for Vision-Language Models Open-Sora Plan: Open-Source Large Video Generation Model

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-20T17:58:49.633267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T17:57:00.909897Z digest=sha256:9a19464c1f41a0e8962ad10d1905a56c1546e83caaf26d940b3abef3559c0923

Observation 983652c9-a2df-4043-9b13-39b245cc8cfe · inbound

GeoWorld-VLM: Geometry from World Models for Vision-Language Models cites this paper.

GeoWorld-VLM: Geometry from World Models for Vision-Language Models Open-Sora Plan: Open-Source Large Video Generation Model

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-06-30T19:05:00.622177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T19:02:05.125937Z digest=sha256:e2c02a6c781faccda314cc5bc03f2c9744b6dd4eddf4c2dabf4640af6c5779a5

Observation 46a88673-7abf-4a28-97a7-f456d19489c6 · inbound

Image-to-Video Diffusion: From Foundations to Open Frontiers cites this paper.

Image-to-Video Diffusion: From Foundations to Open Frontiers Open-Sora Plan: Open-Source Large Video Generation Model

Reference 65

Resolution
verified exact
local_arxiv, observed 2026-05-20T15:08:24.883653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T15:06:02.084336Z digest=sha256:08e28263359973677f63a467074f9dcb18750dbe60fd86330e1bebe45ce2a889

Observation db3e7051-d569-400f-90df-0843942277d4 · inbound

Rebalancing Reference Frame Dominance to Improve Motion in Image-to-Video Models cites this paper.

Rebalancing Reference Frame Dominance to Improve Motion in Image-to-Video Models Open-Sora Plan: Open-Source Large Video Generation Model

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-20T05:58:05.003928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T05:57:44.096941Z digest=sha256:e7ab4a85079a656b72129cc2dfec2f294e689bd23a7c556f3a7ba8c0d0ca8c97

Observation 6c63cccb-5ed3-49e9-8df7-22536dfe62a5 · inbound

Rebalancing Reference Frame Dominance to Improve Motion in Image-to-Video Models cites this paper.

Rebalancing Reference Frame Dominance to Improve Motion in Image-to-Video Models Open-Sora Plan: Open-Source Large Video Generation Model

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-21T07:59:50.822126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T07:55:40.616049Z digest=sha256:24aa836b3044f484c78b12a4aad83ed130f28891ae57db543c78c86a99ac7946

Observation 9d50b069-b3f8-425e-9464-7a6d68970a09 · inbound

Rebalancing Reference Frame Dominance to Improve Motion in Image-to-Video Models cites this paper.

Rebalancing Reference Frame Dominance to Improve Motion in Image-to-Video Models Open-Sora Plan: Open-Source Large Video Generation Model

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-06-30T18:35:00.140064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T18:33:44.480873Z digest=sha256:e9bd07fec4e13d254fc29fbb82525fc8501b622478e940828a78d40708c65aaf

Observation 30566c93-ccf7-451f-b941-6c9f46557dd5 · inbound

ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration cites this paper.

ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration Open-Sora Plan: Open-Source Large Video Generation Model

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:24:43.421730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T07:21:16.861996Z digest=sha256:fa57a853911d5c69494b45b41890378cfffe15450ce0bedce0e09b7f7cb0ae67

Observation 26177d40-ab86-4205-ae34-1ff7727a98c6 · inbound

SCOPE: Simulating Cross-game Operations in Playable Environments for FPS World Models cites this paper.

SCOPE: Simulating Cross-game Operations in Playable Environments for FPS World Models Open-Sora Plan: Open-Source Large Video Generation Model

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-25T05:00:21.318377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-25T04:59:59.545967Z digest=sha256:e9eda2744acc56c0160d3930db2655bf6f8937f158bdb77e628c94868a748f2a

Observation 6e138333-21b0-44c4-ad68-3eb99b09bdea · inbound

PixelWizard: Towards Efficient High-Fidelity Video Generation at Ultra-Large Spatial Resolution cites this paper.

PixelWizard: Towards Efficient High-Fidelity Video Generation at Ultra-Large Spatial Resolution Open-Sora Plan: Open-Source Large Video Generation Model

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-06-29T22:34:02.309608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T22:24:50.679950Z digest=sha256:dc0b5908590809e75200e1806a94faa6270dd72480c2f0c376dc905171c3efcc

Observation 238a5de7-a209-46dc-8fb9-b4130bb0ca7c · inbound

OSP-Next: Efficient High-Quality Video Generation with Sparse Sequence Parallelism, HiF8 Quantization, and Reinforcement Learning cites this paper.

OSP-Next: Efficient High-Quality Video Generation with Sparse Sequence Parallelism, HiF8 Quantization, and Reinforcement Learning Open-Sora Plan: Open-Source Large Video Generation Model

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-06-29T13:13:27.484506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T13:06:02.201607Z digest=sha256:9b4e44e22befdb7d6477c890124e7d981847af7a3d86607bef8aafced22cfd03

Observation 4eac45e4-b455-4b48-b2d3-4414c949a202 · inbound

minWM: A Full-Stack Open-Source Framework for Real-Time Interactive Video World Models cites this paper.

minWM: A Full-Stack Open-Source Framework for Real-Time Interactive Video World Models Open-Sora Plan: Open-Source Large Video Generation Model

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:23:14.798086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T08:22:48.074724Z digest=sha256:1f0bd60a4aaf589022929f0d814b129d782fa09f744e1a69f8405c2cfa544520

Observation 71bcf46d-cfc7-4508-808c-ab6ea7be2319 · inbound

Veda: Scalable Video Diffusion via Distilled Sparse Attention cites this paper.

Veda: Scalable Video Diffusion via Distilled Sparse Attention Open-Sora Plan: Open-Source Large Video Generation Model

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.737493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T07:54:13.390911Z digest=sha256:aea7f19c42e15c1e75d4809e0419835702ca858124a10cc2ce1f9927e6a01a47

Observation 1dacd92b-7423-4076-867e-5f796b5062b9 · inbound

OmniMem: Scalable and Adaptive Memory Retrieval for Long Video Generation cites this paper.

OmniMem: Scalable and Adaptive Memory Retrieval for Long Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.456297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T07:56:40.001739Z digest=sha256:451371f187756799feda775225c3c6b274d421bd059a85bff08cb27bef9cc2a3

Observation 103510e3-8a9d-4276-a83a-94fcbaeb3d46 · inbound

TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation cites this paper.

TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-06-28T23:12:46.609111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T23:11:58.580004Z digest=sha256:79715d235eaf93db318c3ae5d95c8913521ba2ed7de26514d06989a77754e105

Observation 44eff456-1881-481d-b7f1-f1449f59fedf · inbound

MedSyn2: Flexible Control of 3D CT Generation via Text and Semantically-Defined Segmentation Prompts cites this paper.

MedSyn2: Flexible Control of 3D CT Generation via Text and Semantically-Defined Segmentation Prompts Open-Sora Plan: Open-Source Large Video Generation Model

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-01T20:46:14.037384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T17:43:27.624762Z digest=sha256:cc1fcf184655d289b433f266f0fc85e3e4744da73aa7e21036d6bb6e8d375cf5

Observation e0d77bf7-fee1-411d-945c-75d4c94353cd · inbound

AAD-1: Asymmetric Adversarial Distillation for One-Step Autoregressive Video Generation cites this paper.

AAD-1: Asymmetric Adversarial Distillation for One-Step Autoregressive Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:06:29.348745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T10:24:32.525652Z digest=sha256:720b0110f9b532d55e4a9503cda586ccaf59b96d4784d3a84518391a552b094d

Observation d7ee5ce3-63ce-44b1-85be-8ce4a0e85ce8 · inbound

OmniGen-AR: AutoRegressive Any-to-Image Generation cites this paper.

OmniGen-AR: AutoRegressive Any-to-Image Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-07-03T00:47:29.682145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T17:05:16.883488Z digest=sha256:8eb7d025a78ecc40ad4fe4bde7c9b6acf3567ed3539091a3c4236f1be87535ba

Observation 3fb949d8-fc4e-4748-830f-5487295ce08c · inbound

CineDance: Towards Next-Generation Multi-Shot Long-Form Cinematic Audio-Video Generation cites this paper.

CineDance: Towards Next-Generation Multi-Shot Long-Form Cinematic Audio-Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 37

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T00:07:28.147823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T17:30:25.371658Z digest=sha256:d505a0bc4e23bb2a4d8b152022d1bb1770e61368c35b205b826cd3b033fb103c

Observation ae5fb0e8-e200-4e62-afa5-2656ef49ffb2 · inbound

ARGUS: Stacked Multi-View Identity Mosaic Injection for Subject-Preserving Video Generation cites this paper.

ARGUS: Stacked Multi-View Identity Mosaic Injection for Subject-Preserving Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-07-03T08:17:45.938052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T10:46:56.871174Z digest=sha256:bc54c77e1c04dcbba41c85f932aaf5196d5b27e2cb468c067bc1613d1dbfe452

Observation 30d96edb-58f0-4993-b494-00e0ee84dfb9 · inbound

Towards Error-Free Long Video Generation cites this paper.

Towards Error-Free Long Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 24

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T08:59:42.140260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T10:49:36.838492Z digest=sha256:a35520c9e920559f97947f49286c8f60d479db14aec583112d0e9236e219b6fe

Observation fd2dd144-f5ed-48d7-bd30-c705d67c59cc · inbound

SteerVTE: Seamless Video Text Editing with Style and Glyph Control cites this paper.

SteerVTE: Seamless Video Text Editing with Style and Glyph Control Open-Sora Plan: Open-Source Large Video Generation Model

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-07-04T10:29:44.579365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T08:52:23.330736Z digest=sha256:4a90e1c6ac14eeb587343b83a9ad5709b5944542db17dd1fe8ce4a089b844c43

Observation 2c0422b0-87e0-41cb-a819-1fc281dd944f · inbound

Ocean4D: Generative Underwater 4D Reconstruction via Medium-Aware Video Diffusion cites this paper.

Ocean4D: Generative Underwater 4D Reconstruction via Medium-Aware Video Diffusion Open-Sora Plan: Open-Source Large Video Generation Model

Reference 33

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T09:49:44.719698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T09:24:28.715994Z digest=sha256:a47ecbd6853a04498108bfc730b7ac6e10b3e7d081009391ea13bc3f5b5c6f00

Observation 548cd4ee-59b1-4b66-a817-74f855e5af7b · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments Open-Sora Plan: Open-Source Large Video Generation Model

Reference 36

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T18:25:57.796851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T02:03:45.564122Z digest=sha256:e7e993ea3f8084f582de691310c26cca57f63fad1cabffb1e5ac21d11289c39b

Observation dcb2d73f-4780-4efe-b59f-fc19156a57af · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments Open-Sora Plan: Open-Source Large Video Generation Model

Reference 36

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T09:35:40.622607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T06:25:58.872140Z digest=sha256:517eab86530b7afdc8118871196421ae2fb33fbfa2dd6ef99f7222f5e4235a3b

Observation 45bbaf8f-1b8a-43ac-9088-e28e9d807d92 · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments Open-Sora Plan: Open-Source Large Video Generation Model

Reference 36

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T20:57:22.827511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-02T20:52:28.444524Z digest=sha256:5bdce2c133885adb12612c44da9d7466270320cbeb6ffc310c6f5fee0940ccce

Observation 37edbace-e95a-4a85-9977-caeb8d07700c · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments Open-Sora Plan: Open-Source Large Video Generation Model

Reference 36

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T22:49:00.808533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-03T22:44:16.272541Z digest=sha256:ec3d50a49189ef90d78d62df50d042713545b71e5feaf1b2481fc99d3be03c81

Observation 62134171-34c6-40f0-8a8e-6954c13ea35e · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments Open-Sora Plan: Open-Source Large Video Generation Model

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-14T17:14:19.770867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T17:14:19.770867Z digest=sha256:52af9497af841886b552af02fa8c8e5fb03d935e0fd4ac94e5bec522947d7bdd

Observation 270c9132-6a86-46cf-af6f-429fb0db3b20 · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments Open-Sora Plan: Open-Source Large Video Generation Model

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T10:00:11.778050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:00:11.778050Z digest=sha256:6663049d80578b07941eb09df91f4644b33efe541082c146bec4ba0514b831b7

Observation 16c8e2f8-2374-4f49-821f-563958934997 · inbound

AVTok: 1D Unified Tokenization for Holistic Audio-Video Generation cites this paper.

AVTok: 1D Unified Tokenization for Holistic Audio-Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 39

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T09:45:39.991657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T06:16:13.420481Z digest=sha256:9350ad5edd90c8daa0bc1c69a886b5b7d5fb3f03bf2bb787dcafd3d3ef7b1530

Observation 4541a181-69bd-4412-b9db-0d2e2263fd25 · inbound

Bridging Video Understanding and Generation in a Unified Framework cites this paper.

Bridging Video Understanding and Generation in a Unified Framework Open-Sora Plan: Open-Source Large Video Generation Model

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-07-01T10:05:40.385063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T05:57:54.653504Z digest=sha256:3863b1f1948812a9ce63bf792a12bc75da61f502e6a12ef7bb9e07e8821c5480

Observation 3aac3edf-2356-4e02-8fbc-131764494162 · inbound

Enhancing Video Physical Consistency via Role-aware Joint Training and Modality-decoupled Denoising cites this paper.

Enhancing Video Physical Consistency via Role-aware Joint Training and Modality-decoupled Denoising Open-Sora Plan: Open-Source Large Video Generation Model

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-11T15:44:22.210969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T15:44:22.210969Z digest=sha256:52ec4b857c39de668a62833e5c29f7e63be631839a8db8c247d95a9183c3a214

Observation 3902703d-b4b2-4ca6-aa26-2dbce20c91fb · inbound

MobileWan: Closing the Quality Gap for Mobile Video Diffusion cites this paper.

MobileWan: Closing the Quality Gap for Mobile Video Diffusion Open-Sora Plan: Open-Source Large Video Generation Model

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-07-08T14:44:59.777376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-08T14:37:46.957265Z digest=sha256:10c36edd5518bdbc98ca0a871f90f5e2068c786d6a62bdcc90a92a1fa2f5d90b

Observation 717b6340-37b6-40f5-8f41-4c2c4397acae · inbound

MobileWan: Closing the Quality Gap for Mobile Video Diffusion cites this paper.

MobileWan: Closing the Quality Gap for Mobile Video Diffusion Open-Sora Plan: Open-Source Large Video Generation Model

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T08:21:47.178949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:21:47.178949Z digest=sha256:f392e96b264bbf21ad71b6831fe025cf865e9636783068616109bb9306e41ed1

Observation 7b453b4f-0005-4ac4-9868-a818568201f4 · inbound

AlayaWorld: Long-Horizon and Playable Video World Generation cites this paper.

AlayaWorld: Long-Horizon and Playable Video World Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-07-08T10:54:49.237180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-08T10:53:28.412288Z digest=sha256:012f43bef82f7d8abaf870a0227c3f159a871874c8435dd881abfeef58f90a98

Observation 179c17f1-6dbd-495c-9634-cc37859011d1 · inbound

Kaleido: Algorithm-Hardware Co-Design for Video Diffusion Transformers by Exploiting Latent Space Correlations cites this paper.

Kaleido: Algorithm-Hardware Co-Design for Video Diffusion Transformers by Exploiting Latent Space Correlations Open-Sora Plan: Open-Source Large Video Generation Model

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T03:51:47.084326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:51:47.084326Z digest=sha256:8ebadea7b88f533e2ad99031e612d3210831de325af72f2b38a1585ff2af24ae

Observation 287bf623-14d0-4835-8281-4bec30b8362c · inbound

FlashDecoder: Real-Time Latent-to-Pixel Streaming Decoder with Transformers cites this paper.

FlashDecoder: Real-Time Latent-to-Pixel Streaming Decoder with Transformers Open-Sora Plan: Open-Source Large Video Generation Model

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T00:50:14.834226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:50:14.834226Z digest=sha256:0b3af314b72717c95009a5e5a9a1c8b6adbb0285ea873a1ceecb997a7407e52a

Observation 0e637dd5-8ac4-4be2-ad7b-9371d273c54b · inbound

CODA: Algorithm-Hardware Co-design for Edge Video Diffusion via NMP-Enabled Compute-Cache Operator Disaggregation cites this paper.

CODA: Algorithm-Hardware Co-design for Edge Video Diffusion via NMP-Enabled Compute-Cache Operator Disaggregation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T00:49:37.841847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:49:37.841847Z digest=sha256:d1949fc6b92c7caaae92cf601847758d2046609c573826a38c7c936c35a25544

Observation 23d5cbc2-0dcf-4b75-9d79-a8077e71a2ef · inbound

DSTAR: Accelerating Diffusion Transformers via Spatial and Temporal Redundancy Reduction cites this paper.

DSTAR: Accelerating Diffusion Transformers via Spatial and Temporal Redundancy Reduction Open-Sora Plan: Open-Source Large Video Generation Model

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T22:14:35.439256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:14:35.439256Z digest=sha256:4adb1ef955300ed2eeba05d789c798d954bd3510d3aa96495219064db769f356

Observation d7e5f709-fdab-4b74-9cad-ab967fb82720 · inbound

HOMIE: Human-object Centric Video Personalization via Multimodal Intelligent Enhancement cites this paper.

HOMIE: Human-object Centric Video Personalization via Multimodal Intelligent Enhancement Open-Sora Plan: Open-Source Large Video Generation Model

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:22.050100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:22.050100Z digest=sha256:f97a6c1d80541caa8e7af5db77c69d8032b1261cf976757b36d701c768d22414

Observation 21201fe6-024b-481b-ac1b-ce2d3c57562e · inbound

GroupVideo: Multi-Identity Customized Text-to-Video Generation cites this paper.

GroupVideo: Multi-Identity Customized Text-to-Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T08:45:49.939631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:45:49.939631Z digest=sha256:d982e4379ab2355fd9387243ccf50ff4a5eca6759d122a8837bc972b5101b134

Observation af7c9b65-50c6-400d-ac44-94205630c58f · inbound

GraphVid: Interactive Graph-Controllable Video Generation cites this paper.

GraphVid: Interactive Graph-Controllable Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T07:05:02.494777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T07:05:02.494777Z digest=sha256:ef92be434cb0d64b8cdc04e89c7b9c6539689006018f71e09d5bdaf944a90049

Observation dd99ed6c-18d3-4986-877c-6966f0a820d0 · inbound

Streaming Multi-Agent Autoregressive Diffusion Model with World State Registers cites this paper.

Streaming Multi-Agent Autoregressive Diffusion Model with World State Registers Open-Sora Plan: Open-Source Large Video Generation Model

Reference 116

Resolution
unresolved
no resolver link, observed 2026-08-01T07:02:50.696486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:02:50.696486Z digest=sha256:7417e9b8b8eaf0a599b761d3803c271ba6db6834d912c0a45719af97bfd803da

Observation 2bbebcd1-a9cc-4bc4-b1f6-08c7e4cabf20 · inbound

VIPER: Visual In-Context Physics Reasoning for Physically Plausible Video Generation cites this paper.

VIPER: Visual In-Context Physics Reasoning for Physically Plausible Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-30T21:29:21.999038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:29:21.999038Z digest=sha256:c717894adf94f169e099bfe9f72f0dcb86d59361e279feba53a93f66bcc1fd47

Observation 83e902c7-8848-429f-909e-90e227d75ea1 · inbound

Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer cites this paper.

Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer Open-Sora Plan: Open-Source Large Video Generation Model

Reference 202

Resolution
unresolved
no resolver link, observed 2026-07-31T08:51:26.765760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T08:51:26.765760Z digest=sha256:cfe09348c5c8e6cceb5e2c536ea29990cfa61ed423f7ca67091751f3019ec05e

Observation 606c0e24-a62f-4a16-b2b5-800b23144fae · inbound

PhiZero: A World Model Built Around Physical Language cites this paper.

PhiZero: A World Model Built Around Physical Language Open-Sora Plan: Open-Source Large Video Generation Model

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-31T01:50:30.651120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T01:50:30.651120Z digest=sha256:abe810ea7753899fc6b587aea5a2432d5d1f953937847b9ea823566a6879b3bf

Observation 20c6ace6-e135-4bb5-a263-204f179e002a · inbound

Illuminating Visual Identity in Universal Multimodal Embeddings cites this paper.

Illuminating Visual Identity in Universal Multimodal Embeddings Open-Sora Plan: Open-Source Large Video Generation Model

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T20:43:16.957129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:43:16.957129Z digest=sha256:58080cc4ec5df0b4abc42a86328591183f857e11aaa128f6d79b74d434d6de59

Observation 76259ac3-ccd3-4139-aeb9-b9efd11f0efc · inbound

UniMoCa: Unifying Motion and Camera Controls as Visual Proxies for Faithful Human Video Generation cites this paper.

UniMoCa: Unifying Motion and Camera Controls as Visual Proxies for Faithful Human Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T17:55:19.674234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T17:55:19.674234Z digest=sha256:40e5ea74865c572f41582757578ad495571b12e0c157ca911323a31c8f706d74