Pith. sign in

Paper Citation Record · LEDGER

Loong: Generating Minute-level Long Videos with Autoregressive Language Models

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 44 inbound Pith citation observations for arXiv:2410.02757.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.02757 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 44 of 44 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:56:00.753842Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:59:42.133975Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 060d1fed-0a77-4bea-a450-04fdd56235eb · inbound

Long Video Diffusion Generation with Segmented Cross-Attention and Content-Rich Video Data Curation cites this paper.

Long Video Diffusion Generation with Segmented Cross-Attention and Content-Rich Video Data Curation Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T04:32:44.517208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:32:44.517208Z digest=sha256:6bca50961512ae0f0001bad9b11514661d99ad4da11e53ddac970385b3eaa616

Observation f3383fad-0396-41ba-885f-78bf3f574bc7 · inbound

ARCON: Advancing Auto-Regressive Continuation for Driving Videos cites this paper.

ARCON: Advancing Auto-Regressive Continuation for Driving Videos Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T22:11:52.923496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:11:52.923496Z digest=sha256:d72edc08d75f24d4aba6225638e3bd488f0e6c1b76c227595b4636503fd93de9

Observation 0a0c2acc-ae6e-42e1-9068-93d1cba5aadb · inbound

Liquid: Language Models are Scalable and Unified Multi-modal Generators cites this paper.

Liquid: Language Models are Scalable and Unified Multi-modal Generators Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-11T21:35:40.560765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:35:40.560765Z digest=sha256:081f6f40a419321496c3eb4534079e907451c2f8b92c5e6de5ed4c7c55d7a01b

Observation 4573d81c-d98a-4cca-8581-86e496381cb2 · inbound

Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis cites this paper.

Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-11T21:29:34.608744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:29:34.608744Z digest=sha256:1a9ad5a118d9d33dd09edfaee0ce215aa718ad51c629ae8bb6d60919f15bab3b

Observation a68f103a-d157-45b2-89d8-ff723c089931 · inbound

Divot: Diffusion Powers Video Tokenizer for Comprehension and Generation cites this paper.

Divot: Diffusion Powers Video Tokenizer for Comprehension and Generation Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-11T21:30:08.521692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:30:08.521692Z digest=sha256:b0da39e65fe98e0b69fce4d730813c7e0fb4e3575ab63fccd3b207c39a430a77

Observation 13f00a1f-ca3e-4c96-9bb6-99320135d3d4 · inbound

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity cites this paper.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.519752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.519752Z digest=sha256:e92eead1b6bddf92beed8cab5772cb5d562a99202063bfb3961ed340d964a92c

Observation 34dda6b8-76dd-4155-9252-3833f216b457 · inbound

Do Language Models Understand Time? cites this paper.

Do Language Models Understand Time? Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 179

Resolution
unresolved
no resolver link, observed 2026-08-11T12:47:17.968968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:47:17.968968Z digest=sha256:1f81c3f7ac5fec6750cca8ba145b2f108b77aaee85a7dfffa31377f5e58e75b4

Observation 43ecb43a-4735-4712-b75f-d7a2087f9dc3 · inbound

Autoregressive Video Generation without Vector Quantization cites this paper.

Autoregressive Video Generation without Vector Quantization Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T15:07:39.816112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-17T15:07:39.718555Z digest=sha256:bf0a306eaaca2a587f7ac3caef4ec5643fdee8c3c8bfea8867656e2b36f65d4a

Observation 2c05405a-903f-4e5e-9a2f-a9429b5b35c4 · inbound

Parallelized Autoregressive Visual Generation cites this paper.

Parallelized Autoregressive Visual Generation Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-11T11:42:31.780286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:42:31.780286Z digest=sha256:36c71af1a0dcedafed4daa968a25a80d334bccafeb82e3db4d1c8f9127ebf705

Observation 6dadcf4f-79bf-4c73-9cf3-98155c27e32f · inbound

Is Your World Simulator a Good Story Presenter? A Consecutive Events-Based Benchmark for Future Long Video Generation cites this paper.

Is Your World Simulator a Good Story Presenter? A Consecutive Events-Based Benchmark for Future Long Video Generation Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T13:15:28.569256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:15:28.569256Z digest=sha256:b2cffc1792bdc53dfd849d1c18e2e9c4e40ffc4ea04a4ef9f73e976328e6795b

Observation 80859059-388d-4add-84e8-eae0e44422db · inbound

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion cites this paper.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.123934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.123934Z digest=sha256:21dccfdefc4c92d772a020be0f3791aaae48a178489f0a13edac421fd077cf0c

Observation 0ce9a84c-a9e9-45a9-acb3-932d6896425c · inbound

Eliciting In-context Retrieval and Reasoning for Long-context Large Language Models cites this paper.

Eliciting In-context Retrieval and Reasoning for Long-context Large Language Models Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T20:33:22.669062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T20:33:22.669062Z digest=sha256:dbaaa6fad67d5daf4645b430a75679178dad9eabfb5481d3ef4781e13f03c1b6

Observation beed3940-cae6-45a7-bc5a-b0fac0372341 · inbound

A Survey of Interactive Generative Video cites this paper.

A Survey of Interactive Generative Video Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-16T04:56:00.753842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:56:00.753842Z digest=sha256:5156b85e2c3e5dc54542ad583bc8e94257074ff01eb7a41922592e29eefd461a

Observation 234ef67b-9fed-4f69-b73b-0b10d0e4e238 · inbound

Long-Context State-Space Video World Models cites this paper.

Long-Context State-Space Video World Models Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T14:03:20.688743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:03:20.688743Z digest=sha256:2711d51128a9a9ce9bf6e9f95978fa4cde6922d71b3da0438421a06d23172b0b

Observation 65cb59c6-2c2e-4798-9077-12b5e9ee947f · inbound

Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots cites this paper.

Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:17.046537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:00:17.046537Z digest=sha256:e438dbffeafd693da2aacbd19c12eb3d44188fa7d97bc9d3834bfa03dda0de6d

Observation ecf31cbd-af7b-4861-b57c-eee64139f826 · inbound

Video World Models with Long-term Spatial Memory cites this paper.

Video World Models with Long-term Spatial Memory Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:41.605178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:41.605178Z digest=sha256:b44cc753c3f67bbcf6c820feb0b638d8318d9d472e0c2c8ff494024adfef9650

Observation 4dad1b58-0de1-4a3d-a9d1-2327fcd5b564 · inbound

Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion cites this paper.

Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:36:53.131734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-11T01:36:53.029590Z digest=sha256:5abf851060aec28dd9171f39010ab8a0f88fbda1bc3882bd77588cd09d73a4cb

Observation 7e651edd-6035-493a-89c6-4aad685075c8 · inbound

SpectralAR: Spectral Autoregressive Visual Generation cites this paper.

SpectralAR: Spectral Autoregressive Visual Generation Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T04:17:51.201816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:17:51.201816Z digest=sha256:3cfdc02ea235fb18082964b83edd496924dfbda72d5eb09d4ca0d8d98639aceb

Observation f9215a7e-f11a-4acd-a52a-ec5a86ff0863 · inbound

UltraVideo: High-Quality UHD Video Dataset with Comprehensive Captions cites this paper.

UltraVideo: High-Quality UHD Video Dataset with Comprehensive Captions Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T00:30:58.539477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:30:58.539477Z digest=sha256:26a8b32885866a6e1d415e1ea2ff0fee6a3f69839f7f5fb4db8f955c446acd8f

Observation 7f265ae3-e3f9-4289-ab4c-441826bbe344 · inbound

VideoMAR: Autoregressive Video Generatio with Continuous Tokens cites this paper.

VideoMAR: Autoregressive Video Generatio with Continuous Tokens Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:30:28.919613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:30:28.919613Z digest=sha256:10d01d3c65dfe29d3f240a96aa8b14f535bb573315104c2db9b4ea268edc4cc2

Observation 08562ac5-5180-4707-9435-0fbf8578180c · inbound

A Survey on Long-Video Storytelling Generation: Architectures, Consistency, and Cinematic Quality cites this paper.

A Survey on Long-Video Storytelling Generation: Architectures, Consistency, and Cinematic Quality Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T18:51:29.715057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:51:29.715057Z digest=sha256:8cc19a7f839731f4784e59b3e658d7cf38c196e3bb3fe7a1365a218c903eb309

Observation 7f4c72cf-7aea-430b-bf5e-81a1e9075b20 · inbound

TokensGen: Harnessing Condensed Tokens for Long Video Generation cites this paper.

TokensGen: Harnessing Condensed Tokens for Long Video Generation Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T15:28:30.643014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:28:30.643014Z digest=sha256:8ae3c11d738110a56e85a973a1897cb0fc70bb849550dc90be2178007d9c214f

Observation 4df67f3b-21dd-4fca-895e-6b84f1582585 · inbound

Lumina-mGPT 2.0: Stand-Alone AutoRegressive Image Modeling cites this paper.

Lumina-mGPT 2.0: Stand-Alone AutoRegressive Image Modeling Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T18:22:42.207651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:22:42.207651Z digest=sha256:3fb9746ccfcf204baa132c57622304f12796b93074502e80e21a36740da2ffa6

Observation 18149d3f-4904-48fb-97af-b879614ffba6 · inbound

Enhancing Scene Transition Awareness in Video Generation via Post-Training cites this paper.

Enhancing Scene Transition Awareness in Video Generation via Post-Training Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:03.279456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:03.279456Z digest=sha256:c9a89c4cd3490161311dce98f2c28b062084584f749a7de80eccf5ac973d8a40

Observation e17afe89-c307-4c59-a8a1-eae9de4956e7 · inbound

HumanGenesis: Agent-Based Geometric and Generative Modeling for Synthetic Human Dynamics cites this paper.

HumanGenesis: Agent-Based Geometric and Generative Modeling for Synthetic Human Dynamics Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T20:49:52.758882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:49:52.758882Z digest=sha256:169abd6e8d3469d5eb9e105eb8d1273ac1fb407820c6cf69bdec2420b956dd1b

Observation 05edc1df-c0f8-4a7c-bebf-b1e5fa8eb76a · inbound

Rolling Forcing: Autoregressive Long Video Diffusion in Real Time cites this paper.

Rolling Forcing: Autoregressive Long Video Diffusion in Real Time Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 97

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T11:15:29.226122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-16T11:15:29.102090Z digest=sha256:03e207a8b997fd483ebabfff541a17d62acf42ad8d4cdf39ae5940401b6b9c58

Observation 4965dc40-66e5-46c6-a47e-84b38ec35294 · inbound

RAPO++: Cross-Stage Prompt Optimization for Text-to-Video Generation via Data Alignment and Test-Time Scaling cites this paper.

RAPO++: Cross-Stage Prompt Optimization for Text-to-Video Generation via Data Alignment and Test-Time Scaling Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 81

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T05:15:54.297371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T05:13:42.934115Z digest=sha256:9ec4e20c269434e50af7d67849c58aa9a42702f206a196a05414a7eb27687d56

Observation 6b987e18-a487-43ca-a8c6-39df490ec4bb · inbound

Inferix: A Block-Diffusion based Next-Generation Inference Engine for World Simulation cites this paper.

Inferix: A Block-Diffusion based Next-Generation Inference Engine for World Simulation Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:19:05.018987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-17T05:14:36.280155Z digest=sha256:d1ced9db44d3cc830147224cb989d43ca01d2c72118eb370f722057135226a01

Observation ff0dd921-4190-4d2d-bbac-46e2f790ad5c · inbound

BIFE: Better Interaction, Fewer Errors for Minute-Long Video Generation cites this paper.

BIFE: Better Interaction, Fewer Errors for Minute-Long Video Generation Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T19:41:58.807489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:41:58.807489Z digest=sha256:ac3f8f49e61bd15e608f9150118ba88d49f43b991909f5029a04d284ecf97e76

Observation ad757ea3-19b1-4a42-b533-57ba00fbb443 · inbound

Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion cites this paper.

Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 93

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T07:07:29.831479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T07:02:38.876518Z digest=sha256:828e4dc038a94d520829b0e8593092cd091bfd512acf36acb26e0980fabf49be

Observation 4d353ac6-2594-423e-b748-f3962774c70e · inbound

GeoWorld: Geometric World Models cites this paper.

GeoWorld: Geometric World Models Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-21T11:40:03.337727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-21T11:39:15.308355Z digest=sha256:800151b73bd55450cd688102017fb198a637a72a255cc47f1eb8ab8141a09b1e

Observation db661cb8-0b07-4279-8e00-1d6053650965 · inbound

EduVQA: Towards Concept-Aware Assessment of Educational AI-Generated Videos cites this paper.

EduVQA: Towards Concept-Aware Assessment of Educational AI-Generated Videos Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-21T11:54:09.103453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-21T11:52:09.262497Z digest=sha256:d71203b0846ab78678635b57abb707b050b3958ea465d368190893d6a551f2d5

Observation a5275a79-fece-4cb6-871f-5115c91336d4 · inbound

Video Generation Models as World Models: Efficient Paradigms, Architectures and Algorithms cites this paper.

Video Generation Models as World Models: Efficient Paradigms, Architectures and Algorithms Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 66

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T01:38:36.372424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-14T01:35:14.878069Z digest=sha256:27df6a102f37c17f6483337f414132dbcb00b3748ef951380a96eb4a3bb8a82a

Observation 60e5bb63-2705-4696-9a2a-b466ed02e487 · inbound

INSPATIO-WORLD: A Real-Time 4D World Simulator via Spatiotemporal Autoregressive Modeling cites this paper.

INSPATIO-WORLD: A Real-Time 4D World Simulator via Spatiotemporal Autoregressive Modeling Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:25:53.605274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T19:08:56.588282Z digest=sha256:0064dfa5540ece69f10bcd287caa1eadaea24ff42100b33e84320f8f77fe148f

Observation abc3491d-9480-4f38-b430-594dc4f6170b · inbound

Mutual Forcing: Dual-Mode Self-Evolution for Fast Autoregressive Audio-Video Character Generation cites this paper.

Mutual Forcing: Dual-Mode Self-Evolution for Fast Autoregressive Audio-Video Character Generation Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:14.905365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-07T16:54:23.108142Z digest=sha256:1a767e1550337d36aee6241b2565e84d90faca5a1bb8e51ccd3a6fb58c37a940

Observation d4980422-aaba-423f-8a2a-c00d56f982f0 · inbound

Head Forcing: Long Autoregressive Video Generation via Head Heterogeneity cites this paper.

Head Forcing: Long Autoregressive Video Generation via Head Heterogeneity Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:33:32.264156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T02:31:48.593354Z digest=sha256:725d96e242f8eba90c0f5d25a34290eddde9abe67adc31f58871fec98264e8da

Observation 007b082c-3c63-4872-99d7-ab14f035742e · inbound

OmniMem: Scalable and Adaptive Memory Retrieval for Long Video Generation cites this paper.

OmniMem: Scalable and Adaptive Memory Retrieval for Long Video Generation Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:14.485825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T07:56:40.001739Z digest=sha256:085733dcb8fa1f583e5f43d3ab2b5da734c872388cf4f1ce822c5db6d5b4be44

Observation a6dc219e-b016-4691-93c1-1586790ca6a8 · inbound

Lumos-Nexus: Efficient Frequency Bridging with Homogeneous Latent Space for Video Unified Models cites this paper.

Lumos-Nexus: Efficient Frequency Bridging with Homogeneous Latent Space for Video Unified Models Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:16:00.543053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T22:56:21.783415Z digest=sha256:5bff8e31f2b2aff8424c4a209844ad15dd77fa6eaa832ebf6a076f42e29f66aa

Observation e517b4e0-dba0-4f18-a000-163c857f7277 · inbound

Streaming Video Generation with Streaming Force Control cites this paper.

Streaming Video Generation with Streaming Force Control Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:07:12.425890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T22:14:32.465663Z digest=sha256:a251dddcbeab9c042905ff510888eb14c3e0f4af86dfdbc32f3cef3d2b7e1168

Observation 4339dd7b-7c9d-463d-b2de-e6ea4fb0fd9c · inbound

Towards Error-Free Long Video Generation cites this paper.

Towards Error-Free Long Video Generation Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 37

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T08:59:42.135414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T10:49:36.838492Z digest=sha256:faf27494ae6ecc17b2444c6d2ef50efd2c0bc4c6dc0d9b725f07ee70db9789be

Observation 7624b5da-d2a8-474a-b6c4-83e8f3175c03 · inbound

Bridging Video Understanding and Generation in a Unified Framework cites this paper.

Bridging Video Understanding and Generation in a Unified Framework Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:05:40.345090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-01T05:57:54.653504Z digest=sha256:59c5a4a595ae058e07f34d9bdf818b08b0c5685766dc79acd387faf33b79f216

Observation d0ac97ba-c813-4b6a-9385-ec4ecb60ab98 · inbound

Flex-Forcing: Towards a Unified Autoregressive and Bidirectional Video Diffusion Model cites this paper.

Flex-Forcing: Towards a Unified Autoregressive and Bidirectional Video Diffusion Model Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-12T01:59:48.820112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:59:48.820112Z digest=sha256:22ba6b7906dbe3c36137e6a16fb86f511f78445cc44283a188977cc9f0779f1a

Observation 43049303-1777-48e5-a9e9-f660684b8476 · inbound

In-Context Forcing: Uncovering Context Effects in Autoregressive Video Diffusion cites this paper.

In-Context Forcing: Uncovering Context Effects in Autoregressive Video Diffusion Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T17:36:05.895482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:36:05.895482Z digest=sha256:cdd4c32f8286fa9b0d11bde08e380a219c63da50d441101e53a52e8f13c0d374

Observation c4ca32f3-efd9-4ae0-96ae-c2f06e8ce615 · inbound

Stream Forcing: Constructing Unified Training Trajectory for Robust Streaming Video Generation cites this paper.

Stream Forcing: Constructing Unified Training Trajectory for Robust Streaming Video Generation Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T14:26:02.614896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:26:02.614896Z digest=sha256:1e539a25388c42f9761f86489d20b3ec2265f625d05911d1285149eef4cd1196