Pith. sign in

Paper Citation Record · LEDGER

Vidu S1: A Real-Time Interactive Video Generation Model

As of 11 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 2 inbound Pith citation observations for arXiv:2607.03118.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.03118 v2

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T08:56:31.626723Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T04:52:49.585526Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T04:52:50.130247Z

Reference resolution

61 of 61 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved60
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3e363ebb-bca1-4d0b-a364-4ccab1a3f003 · outbound

This paper cites Video generation models as world simulators.

Vidu S1: A Real-Time Interactive Video Generation Model Video generation models as world simulators

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.463504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.463504Z digest=sha256:0002a403e2cd5d4dedd6d0957b01637a7407a14671be1e65ffaf139d73025c3a

Observation 3aeb8d28-103a-4275-8e24-ceb48a86870c · outbound

This paper cites Veo: A text-to-video generation system.

Vidu S1: A Real-Time Interactive Video Generation Model Veo: A text-to-video generation system

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.467458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.467458Z digest=sha256:7d5e89de925c29b21bf9a341a9f631510f61af09a2373d47430dc81a663ea115

Observation cab3bf8a-6dcd-4df4-9eef-844027cd223d · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Vidu S1: A Real-Time Interactive Video Generation Model Wan: Open and Advanced Large-Scale Video Generative Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.470703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.470703Z digest=sha256:d73e135d318037a73fa9c8d4eea19e4652a7e06e0fc888cc4fe3c53994530f8d

Observation d3c08625-a7aa-4c0a-be1f-03b1b8be0acf · outbound

This paper cites Seedance 2.0: Advancing Video Generation for World Complexity.

Vidu S1: A Real-Time Interactive Video Generation Model Seedance 2.0: Advancing Video Generation for World Complexity

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.474238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.474238Z digest=sha256:a73d3efc6ec52508987085208f95e08f511dd2e64cde02b1d7c550fd90644be7

Observation ecccb5ee-80b0-48a0-a499-e5cf55ef0736 · outbound

This paper cites Diffusion forcing: Next-token prediction meets full-sequence diffusion.

Vidu S1: A Real-Time Interactive Video Generation Model Diffusion forcing: Next-token prediction meets full-sequence diffusion

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.477580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.477580Z digest=sha256:9effcf5d631cc47e8f683fba91275e0681f11968611747d6fb84311c912448cb

Observation b6abb58f-af3c-4794-80ac-d74398402727 · outbound

This paper cites Fifo-diffusion: Generating infinite videos from text without training.

Vidu S1: A Real-Time Interactive Video Generation Model Fifo-diffusion: Generating infinite videos from text without training

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.480537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.480537Z digest=sha256:00fd15b868c7f3da4c13f901cdd495be8aa8a9d3145164ec943c43495121c129

Observation ff47484f-91e5-47ee-ace7-21a04e51327f · outbound

This paper cites Streamingt2v: Consistent, dynamic, and extendable long video generation from text.

Vidu S1: A Real-Time Interactive Video Generation Model Streamingt2v: Consistent, dynamic, and extendable long video generation from text

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.483719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.483719Z digest=sha256:40934b794f92a587754dd0a726d7926241ce3711ad770e9892d23ef8a36f611f

Observation 581655c9-ad42-4b91-8ac9-1c790b4fdbe3 · outbound

This paper cites History-Guided Video Diffusion.

Vidu S1: A Real-Time Interactive Video Generation Model History-Guided Video Diffusion

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.486576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.486576Z digest=sha256:d666d5aab3c13d2f94701e403a22239ea8f47fe420561e45ac287236ac07b279

Observation c4b0d531-7a5c-4974-a016-ad469f72aa68 · outbound

This paper cites Rolling Forcing: Autoregressive Long Video Diffusion in Real Time.

Vidu S1: A Real-Time Interactive Video Generation Model Rolling Forcing: Autoregressive Long Video Diffusion in Real Time

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.489618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.489618Z digest=sha256:c1a4c4fa212452c81fb6130d490ff33b05e159040a361d67b189255aaabd0174

Observation 0c05cea9-6547-4cc4-ac7f-d63ab2c9fc8e · outbound

This paper cites Ar-diffusion: Asynchronous video generation with auto-regressive diffusion.

Vidu S1: A Real-Time Interactive Video Generation Model Ar-diffusion: Asynchronous video generation with auto-regressive diffusion

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.492622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.492622Z digest=sha256:dd5a6475b632f2ff74d78bfa4b8d7684dbee4a595f31a4598ab5d1b3af8df57f

Observation b2c57829-ae31-4cc4-ba59-fdf2bc436448 · outbound

This paper cites Progressive autoregressive video diffusion models.

Vidu S1: A Real-Time Interactive Video Generation Model Progressive autoregressive video diffusion models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.495691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.495691Z digest=sha256:0429fa37c37ac5748b745aa36999055b1ca3c59c4d52f76ab687ed1554adaa43

Observation 5d57c7ad-8073-405f-8070-3b48af4375bd · outbound

This paper cites Streamdiffusionv2: A streaming system for dynamic and interactive video generation.

Vidu S1: A Real-Time Interactive Video Generation Model Streamdiffusionv2: A streaming system for dynamic and interactive video generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.498709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.498709Z digest=sha256:155d48da7aa810ee277a0f8dbb8b30b865c105cfeedcc2fd6a2033056807cdc4

Observation c4a4b9e2-570e-40f6-b829-1a6faac2be07 · outbound

This paper cites From slow bidirectional to fast autoregressive video diffusion models.

Vidu S1: A Real-Time Interactive Video Generation Model From slow bidirectional to fast autoregressive video diffusion models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.501590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.501590Z digest=sha256:756143ef40bc9c8bbb87af95b307e823dab140ddb7e48572cf48fd89026fbeea

Observation 067a425a-2ee4-429b-8fc8-d83be035a332 · outbound

This paper cites Self forcing: Bridging the train-test gap in autoregressive video diffusion.

Vidu S1: A Real-Time Interactive Video Generation Model Self forcing: Bridging the train-test gap in autoregressive video diffusion

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.504457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.504457Z digest=sha256:c3bec67b8259c9b4505b8420a73649234911149a619549907653241fc455132f

Observation 3248f336-7196-4c2b-a7a5-984225c59fa5 · outbound

This paper cites Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation.

Vidu S1: A Real-Time Interactive Video Generation Model Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.507401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.507401Z digest=sha256:8806d2097cfb6ddff3f909261bdadcdc0b12fd725790755826f12f57e6c924bf

Observation 348b9833-f2b8-4e1b-a326-4bcb07919cf1 · outbound

This paper cites Causal Forcing++: Scalable Few-Step Autoregressive Diffusion Distillation for Real-Time Interactive Video Generation.

Vidu S1: A Real-Time Interactive Video Generation Model Causal Forcing++: Scalable Few-Step Autoregressive Diffusion Distillation for Real-Time Interactive Video Generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.510408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.510408Z digest=sha256:f736b0e6af36ca05b3a0e3576a215ca0956ce115a6d11940c9aade0616081f0e

Observation bc26c285-1f33-44d8-af4e-02ca1c9744fc · outbound

This paper cites LongLive: Real-time Interactive Long Video Generation.

Vidu S1: A Real-Time Interactive Video Generation Model LongLive: Real-time Interactive Long Video Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.513461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.513461Z digest=sha256:c8b6e7442781664e60e1b31b2dd5c979a8ab9c4af4b55a88584dffd21e9511b0

Observation a82db0fd-a0c7-46b3-b959-ff2f213dfb0c · outbound

This paper cites MAGI-1: Autoregressive Video Generation at Scale.

Vidu S1: A Real-Time Interactive Video Generation Model MAGI-1: Autoregressive Video Generation at Scale

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.516388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.516388Z digest=sha256:a861a25ab0f944a07545133ad9a40fc0b489695ffd74038391bf7e9503b3da27

Observation f7ef0785-6436-4487-a2d4-6e82a188ba99 · outbound

This paper cites SkyReels-V2: Infinite-length Film Generative Model.

Vidu S1: A Real-Time Interactive Video Generation Model SkyReels-V2: Infinite-length Film Generative Model

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.519369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.519369Z digest=sha256:10d546be20262ea6668686807a41144881f66ebc675d632d146a5e75f2826676

Observation 2218d08e-c25b-4132-993c-bcb47a358fec · outbound

This paper cites Packing input frame context in next-frame prediction models for video generation.

Vidu S1: A Real-Time Interactive Video Generation Model Packing input frame context in next-frame prediction models for video generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.522348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.522348Z digest=sha256:4ecf0102180b9814b82062cdf3967ba56cb582b74886c66a7b7901b42b4efd7a

Observation f9925a27-de13-409c-9838-6f113ba0d756 · outbound

This paper cites Wan-S2V: Audio-Driven Cinematic Video Generation.

Vidu S1: A Real-Time Interactive Video Generation Model Wan-S2V: Audio-Driven Cinematic Video Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.525012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.525012Z digest=sha256:c90258eadf236a441a736037cd48644581c7b539d38001151c90b305161eb319

Observation 2bd93fa1-c131-48d6-95ee-b48d84c7e2b8 · outbound

This paper cites Turbodiffusion: Accelerating video diffusion models by 100-200 times.

Vidu S1: A Real-Time Interactive Video Generation Model Turbodiffusion: Accelerating video diffusion models by 100-200 times

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.527895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.527895Z digest=sha256:6c83f832a8f3a0887709b97c376fdab503ce5ad9cb4773d270dd050ff51fe431

Observation 91a812e8-d2b7-461d-ba52-e2324aa14959 · outbound

This paper cites TurboServe: Serving Streaming Video Generation Efficiently and Economically.

Vidu S1: A Real-Time Interactive Video Generation Model TurboServe: Serving Streaming Video Generation Efficiently and Economically

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.530403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.530403Z digest=sha256:cc2f09a210644946f82372fa2cfc9cab6e2086a07b786feb70c846a942c4f2e4

Observation 80dde387-c348-401a-89c8-e55c6bc8a88a · outbound

This paper cites Qwen3-Omni Technical Report.

Vidu S1: A Real-Time Interactive Video Generation Model Qwen3-Omni Technical Report

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.533158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.533158Z digest=sha256:29786e9d707a00b64e23adfcbb40b84ef05142856e7292ed82c3a1795e1e753c

Observation c5e46572-31eb-402d-8272-29e4bf9de353 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Vidu S1: A Real-Time Interactive Video Generation Model Gemini: A Family of Highly Capable Multimodal Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.535865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.535865Z digest=sha256:3ba0241596229c5e096ff56b33af5b80faf4f8f2360cf08cec562ea0179938fc

Observation 90abec4b-4c0b-4ca5-830a-35c121473a6d · outbound

This paper cites Ca2-vdm: Efficient autoregres- sive video diffusion model with causal generation and cache sharing, 2025.

Vidu S1: A Real-Time Interactive Video Generation Model Ca2-vdm: Efficient autoregres- sive video diffusion model with causal generation and cache sharing, 2025

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.538797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.538797Z digest=sha256:2e216138377b882c921d0b9aa53dc2e201c892d18e7ac04bc77acb435744c8ef

Observation 844b100c-1434-4ba1-be53-fb5578ab5542 · outbound

This paper cites Pyramidal flow matching for efficient video generative modeling.

Vidu S1: A Real-Time Interactive Video Generation Model Pyramidal flow matching for efficient video generative modeling

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.541421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.541421Z digest=sha256:9e97b371b1b065243803d19ab71da61b4cf0592d72e47cb56438034d11fb31d3

Observation f53fc15a-4014-4084-926d-6038ff82afdd · outbound

This paper cites One-step diffusion with distribution matching distillation.

Vidu S1: A Real-Time Interactive Video Generation Model One-step diffusion with distribution matching distillation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.543850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.543850Z digest=sha256:2579c19e680c5e747d57f07a704c58bdd1c3be1516553f5e86b836d67a79e007

Observation e483012a-357b-4ffd-a522-b4d0a5dd4aa3 · outbound

This paper cites Phased consistency models.

Vidu S1: A Real-Time Interactive Video Generation Model Phased consistency models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.546327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.546327Z digest=sha256:f555ce730473de174aaa521d8550beae9a0f7458235db02dc4b6d6c457ade615

Observation 32310bdd-18a0-4d66-8659-30afe28ceb70 · outbound

This paper cites Efficient streaming language models with attention sinks.

Vidu S1: A Real-Time Interactive Video Generation Model Efficient streaming language models with attention sinks

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.548773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.548773Z digest=sha256:18195521bf1b9f79c8c61ab96b49b85dcc096498480150dd3f8db5c4fd5d6349

Observation a3526f36-03ed-4c48-8697-14a7388cfa1f · outbound

This paper cites Infinity-rope: Action-controllable infinite video generation emerges from autoregressive self-rollout.

Vidu S1: A Real-Time Interactive Video Generation Model Infinity-rope: Action-controllable infinite video generation emerges from autoregressive self-rollout

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.551331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.551331Z digest=sha256:34b3416bd9958fef52e1a095cca0a5209cf0cff0b4b61138c3a04d200ddc24f7

Observation 20b37558-ec6d-4ac8-856c-26e2e753c268 · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.

Vidu S1: A Real-Time Interactive Video Generation Model Roformer: Enhanced transformer with rotary position embedding

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.553831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.553831Z digest=sha256:a0c0ee84d4cad3e3ee3a5569ec7d4b256f6c873262c063b964867e5832463a6a

Observation d3d86871-2b0a-4c2d-b6c6-f19a4c8bd6ab · outbound

This paper cites Deep forcing: Training-free long video generation with deep sink and participative compression.

Vidu S1: A Real-Time Interactive Video Generation Model Deep forcing: Training-free long video generation with deep sink and participative compression

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.556217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.556217Z digest=sha256:5eb904180fa0121b00998d25f39c97b8516d1f3c89975cc06fff6a66af8df88b

Observation 3a63176c-696c-445e-ba39-027af13592af · outbound

This paper cites Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion.

Vidu S1: A Real-Time Interactive Video Generation Model Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.558808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.558808Z digest=sha256:44fd2f99fa9fdc2a195f1334bdc8ab6c1ee446da3bb2e47071aa87cdfb11a8a8

Observation fc8ac0a1-ce5a-4ddc-a3f6-eeeeb1e2d13f · outbound

This paper cites Grounded Forcing: Bridging Time-Independent Semantics and Proximal Dynamics in Autoregressive Video Synthesis.

Vidu S1: A Real-Time Interactive Video Generation Model Grounded Forcing: Bridging Time-Independent Semantics and Proximal Dynamics in Autoregressive Video Synthesis

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.561574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.561574Z digest=sha256:e51869b4948edfd2ae00498dfa632334dbc8f819d0a0c4d35005d5604865e9d5

Observation f8ad4ad4-25dd-4bc8-9e75-e6d9a8d17dd9 · outbound

This paper cites Memrope: Training-free infinite video generation via evolving memory tokens.

Vidu S1: A Real-Time Interactive Video Generation Model Memrope: Training-free infinite video generation via evolving memory tokens

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.564278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.564278Z digest=sha256:39c223ec6854fb3b4dc31b7fcb7641fb3202a495f97e7751ca74de29cfdc15db

Observation 40026c4a-3386-4c42-9dce-a3c30f010f97 · outbound

This paper cites Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length.

Vidu S1: A Real-Time Interactive Video Generation Model Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.566881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.566881Z digest=sha256:4a1beb5320ef8f8195d828eb80a3f125b1e1a1348c8bcb45dad81bd3f026c12d

Observation 9b39dc43-2355-46b4-a07e-7ba9e4f8bba1 · outbound

This paper cites Causal-rCM: A Unified Teacher-Forcing and Self-Forcing Open Recipe for Autoregressive Diffusion Distillation in Streaming Video Generation and Interactive World Models.

Vidu S1: A Real-Time Interactive Video Generation Model Causal-rCM: A Unified Teacher-Forcing and Self-Forcing Open Recipe for Autoregressive Diffusion Distillation in Streaming Video Generation and Interactive World Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.569666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.569666Z digest=sha256:f74fd67f22f35460a1896b7b7dfdf3143a394eb418cf50597d7b583c2f34741d

Observation 4d6cf2d9-c8ad-4f51-90cd-55ef574df26e · outbound

This paper cites LPM 1.0: Video-based Character Performance Model.

Vidu S1: A Real-Time Interactive Video Generation Model LPM 1.0: Video-based Character Performance Model

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.572316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.572316Z digest=sha256:63227aef92e2f5d33e32ac6de7ca456eeb4ebf4ceb11fdd2e96683acb6fbba83

Observation a5fbe3e1-45f0-46e9-9495-94e6005945bc · outbound

This paper cites Efficient attention methods: Hardware-efficient, sparse, compact, and linear attention.

Vidu S1: A Real-Time Interactive Video Generation Model Efficient attention methods: Hardware-efficient, sparse, compact, and linear attention

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.575254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.575254Z digest=sha256:ee864bc25766be8b86ad8dfc7f4f6c1731e3a15ccab8a1dfac595225822d5809

Observation aea88014-b7a7-4c7a-b921-e7557d3be766 · outbound

This paper cites Sageattention: Accurate 8-bit attention for plug-and-play inference acceleration.

Vidu S1: A Real-Time Interactive Video Generation Model Sageattention: Accurate 8-bit attention for plug-and-play inference acceleration

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.577869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.577869Z digest=sha256:e1f4ccf063ae439863ee97722a93641d4d03a3bf3e282aec807489056c3e4a8b

Observation 89b775ef-041d-4a9c-b2e2-24aa4c7dd394 · outbound

This paper cites Sageattention2: Efficient attention with thorough outlier smoothing and per-thread int4 quantization.

Vidu S1: A Real-Time Interactive Video Generation Model Sageattention2: Efficient attention with thorough outlier smoothing and per-thread int4 quantization

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.580166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.580166Z digest=sha256:33e8caa11394edc62750fb460afcddca1f42a300e942eeda8c9c83998275d156

Observation c63f6e16-8a37-4f22-ad1a-41449ec5fce3 · outbound

This paper cites SageAttention2++: A More Efficient Implementation of SageAttention2.

Vidu S1: A Real-Time Interactive Video Generation Model SageAttention2++: A More Efficient Implementation of SageAttention2

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.582638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.582638Z digest=sha256:dddcaceffe11c057d5289092309ccb5d77d4033fca0cfaf87d7d69863891c4ba

Observation 12bb74d1-4e59-4bca-9baa-a160ba739091 · outbound

This paper cites Sageattention3: Microscaling fp4 attention for inference and an exploration of 8-bit training.

Vidu S1: A Real-Time Interactive Video Generation Model Sageattention3: Microscaling fp4 attention for inference and an exploration of 8-bit training

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.585274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.585274Z digest=sha256:57a8b722d7674579809d738aa7f5d9af22b6a356c8c683c803aa1f81adf2317f

Observation 89b897d3-425f-466e-9771-2ae98351a798 · outbound

This paper cites Sagebwd: A trainable low-bit attention.

Vidu S1: A Real-Time Interactive Video Generation Model Sagebwd: A trainable low-bit attention

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.587582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.587582Z digest=sha256:1e34e02d700fd6e4c9136b1baca4ce1f0d2da4245a4c0184b3455e4a5b86a124

Observation e2343944-edae-4aab-9f22-0a80487b7aca · outbound

This paper cites Spargeattention: Accurate and training-free sparse attention accelerating any model inference.

Vidu S1: A Real-Time Interactive Video Generation Model Spargeattention: Accurate and training-free sparse attention accelerating any model inference

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.589966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.589966Z digest=sha256:b895a3dde1aa724fdd964b28ad32cfa4cec73b5ea12e2b81a9a069028bbd450b

Observation 585f5161-1ea3-4582-bcb9-015ff28b0062 · outbound

This paper cites Spargeattention2: Trainable sparse attention via hybrid top-k+ top-p masking and distillation fine-tuning.

Vidu S1: A Real-Time Interactive Video Generation Model Spargeattention2: Trainable sparse attention via hybrid top-k+ top-p masking and distillation fine-tuning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.592261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.592261Z digest=sha256:a04389fdcee03c74996d9186323f3788615c8f4b05643474cfa9b6a84f81408e

Observation 26d5b5f4-4764-4b0e-a758-e474bc62ee61 · outbound

This paper cites Gonzalez, Jun Zhu, and Jianfei Chen.

Vidu S1: A Real-Time Interactive Video Generation Model Gonzalez, Jun Zhu, and Jianfei Chen

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.594594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.594594Z digest=sha256:fd3762d015df1d213f1b42cd3f5b647c4d25cdd4b87a7849abd0cd2a42a25147

Observation e94ba322-1a22-4e46-817b-4fe618e2438c · outbound

This paper cites Sla2: Sparse-linear attention with learnable routing and qat.

Vidu S1: A Real-Time Interactive Video Generation Model Sla2: Sparse-linear attention with learnable routing and qat

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.596847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.596847Z digest=sha256:3ea04e2314335d73a77e79a58a9ec19e48011ce7f5b42855d8098a13253549c6

Observation 880d7b0c-8379-4b62-ba22-fee057930146 · outbound

This paper cites DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models.

Vidu S1: A Real-Time Interactive Video Generation Model DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.599291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.599291Z digest=sha256:0b035fb1df0fc4c3316e1b46a03531bd87c48dd8192e0867bfb5257953c0b830

Observation ec607be8-1c1d-4669-bcc2-c93d41decb29 · outbound

This paper cites Flow-guided one-shot talking face generation with a high-resolution audio-visual dataset.

Vidu S1: A Real-Time Interactive Video Generation Model Flow-guided one-shot talking face generation with a high-resolution audio-visual dataset

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.601964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.601964Z digest=sha256:bf72d1d206142e240681d195942234f38c1670f8addc68fbf4828f2e4f924234

Observation b0ef7be3-eaa4-4688-8e47-b50b0d1cae7e · outbound

This paper cites Heygen ai video avatar.

Vidu S1: A Real-Time Interactive Video Generation Model Heygen ai video avatar

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.604304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.604304Z digest=sha256:a83e03ed0895fa125a2b520053f452f860ba9194eddc212ebeedb1343bf12d5d

Observation a1cfbc37-775c-4c06-b731-a8d685c99c37 · outbound

This paper cites Lemonslice studio: Create talking and singing ai avatar videos.

Vidu S1: A Real-Time Interactive Video Generation Model Lemonslice studio: Create talking and singing ai avatar videos

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.606615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.606615Z digest=sha256:7c7efe82ed9d07539c4be1c84b762935baf5e28d221bb98c52216e7e5d279195

Observation 21712d40-31e3-45b6-a168-1421154120c9 · outbound

This paper cites Klingavatar 2.0 technical report, 2025.

Vidu S1: A Real-Time Interactive Video Generation Model Klingavatar 2.0 technical report, 2025

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.611583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.611583Z digest=sha256:86e4245546aa7126f03a69f9b58ecceb587ed5cc2bfd1ee869016a788e6f5e90

Observation 0d4e942e-947d-46e0-95e1-092e8fcd4b2d · outbound

This paper cites OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation.

Vidu S1: A Real-Time Interactive Video Generation Model OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.614075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.614075Z digest=sha256:5bf8433ac1244d0f8f6b8cbaf3242489b8ef4304ba37709129cc7f5bbad57685

Observation 8f59d1de-8a7f-4be0-ae84-aaea78b0a6da · outbound

This paper cites Hallo3: Highly dynamic and realistic portrait image animation with video diffusion transformer.

Vidu S1: A Real-Time Interactive Video Generation Model Hallo3: Highly dynamic and realistic portrait image animation with video diffusion transformer

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.616682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.616682Z digest=sha256:0f21454ad372baafecede3141d0adc9ec9073095a5ae7588caa25836b3085c5f

Observation 62419ab1-da7b-4530-bc99-fb8a4f8b7e7d · outbound

This paper cites StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation.

Vidu S1: A Real-Time Interactive Video Generation Model StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.619017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.619017Z digest=sha256:f1644cd270c2d496e9e55adf8c15b63a3f682708c6dc452e878c1496244c7714

Observation cf1e41b7-bd69-4204-bff0-6aa2dc81e491 · outbound

This paper cites Arcface: Additive angular margin loss for deep face recognition.

Vidu S1: A Real-Time Interactive Video Generation Model Arcface: Additive angular margin loss for deep face recognition

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.621637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.621637Z digest=sha256:4293035d859945d842daea2c6645b9abc734635b37d83e336f646708930b3ad3

Observation e8267e36-cfb7-4f07-ad67-2fe614302a9a · outbound

This paper cites A lip sync expert is all you need for speech to lip generation in the wild.

Vidu S1: A Real-Time Interactive Video Generation Model A lip sync expert is all you need for speech to lip generation in the wild

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.624150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.624150Z digest=sha256:5c38f966b8b1e760ba6a031713dba64f62fd600c80fa2edcbee6ea0b3729036e

Observation 0a2c29d2-6d73-401b-937d-778037eea38a · outbound

This paper cites Exploring video quality assessment on user generated contents from aesthetic and technical perspectives.

Vidu S1: A Real-Time Interactive Video Generation Model Exploring video quality assessment on user generated contents from aesthetic and technical perspectives

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.626723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.626723Z digest=sha256:e9781037d9775b06d090e08e6bf596b05c6df5bf3fde3ecfb17a550e34cb1905

Observation dea6cf42-a1ea-4d65-850d-f2713d97858c · outbound

This paper cites an unresolved cited work.

Vidu S1: A Real-Time Interactive Video Generation Model Unresolved cited work

Reference 2026

Resolution
parse uncertain
no resolver link, observed 2026-08-02T08:56:31.609050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.609050Z digest=sha256:7ae930f0ac2835172d0b6f2184e3db2383f42f1043a3ba8507ad9d919f1e3ba4

Pith citing papers

Observation 07857c41-bc51-4ef3-80d4-f446320f17d1 · inbound

Visko Orbis 1.0: A Live Model for Real-Time Interactive Long Video Generation cites this paper.

Visko Orbis 1.0: A Live Model for Real-Time Interactive Long Video Generation Vidu S1: A Real-Time Interactive Video Generation Model

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-30T23:47:57.729629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T23:47:57.729629Z digest=sha256:51d070f7c3e5331197afe07bd2018aeba4d10c1cc2a274199cfddd88c6d672fe

Observation 8c4e5496-cf18-4387-9ff8-e09321f7374c · inbound

JoyAI-Video-Edit: Real-Time Open-Ended Video Editing with Autoregressive Diffusion cites this paper.

JoyAI-Video-Edit: Real-Time Open-Ended Video Editing with Autoregressive Diffusion Vidu S1: A Real-Time Interactive Video Generation Model

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-05T04:52:50.219179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T04:52:49.585526Z digest=sha256:a6955ae173358820b633894aaf9f2c956126caef1048b21e8c87cf47204b45a8