Pith. sign in

Paper Citation Record · LEDGER

Compositional Video Synthesis by Temporal Object-Centric Learning

As of 7 August 2026, this Paper Citation Record lists 73 of 73 outbound references and 0 inbound Pith citation observations for arXiv:2507.20855.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.20855 v1

Coverage vector

measured 73 of 73 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:16:44.217929Z

measured 73 of 73 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

73 of 73 outbound references displayed

  • verified exact0
  • verified fuzzy66
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4a53892a-e5ca-4463-a595-7d52443aef09 · outbound

This paper cites Slamp: Stochastic latent appearance and motion pre- diction.

Compositional Video Synthesis by Temporal Object-Centric Learning Slamp: Stochastic latent appearance and motion pre- diction

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.808932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:42.664655Z digest=sha256:1c0d6fa06bb5056be9e4fea2112bedfa8322df9b486e345c01c698abbf9f6a32

Observation eb8ca687-a7ee-4de6-87da-0115022032ca · outbound

This paper cites Stretchbev: Stretching future instance prediction spatially and temporally.

Compositional Video Synthesis by Temporal Object-Centric Learning Stretchbev: Stretching future instance prediction spatially and temporally

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.800561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:42.726015Z digest=sha256:c963b1a127e039d089a0cd6f37b1ac546da0e645689cd13367f6b77eebd8f7a5

Observation 1d62f899-9ec3-4a13-b283-b9a7c11f328d · outbound

This paper cites Slot-guided adaptation of pre-trained diffusion models for object-centric learning and compositional generation.

Compositional Video Synthesis by Temporal Object-Centric Learning Slot-guided adaptation of pre-trained diffusion models for object-centric learning and compositional generation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.792481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:42.800382Z digest=sha256:b6ae234208da305f3170a8638fb295ad59f8b165efd7e3069afc7fa883c18b78

Observation b98a97f7-cbe1-49a6-a34e-55b60c85acfd · outbound

This paper cites Self- supervised Object-centric Learning for Videos.

Compositional Video Synthesis by Temporal Object-Centric Learning Self- supervised Object-centric Learning for Videos

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.784142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:42.882950Z digest=sha256:28f7e4e74d3057ea12393951b7a34b7fdc2603165ab3331a3ebc10e15465212b

Observation fa7374ae-00d6-49ec-bd27-1016364b0d8b · outbound

This paper cites Systematic generalization: What is required and can it be learned? In Proc.

Compositional Video Synthesis by Temporal Object-Centric Learning Systematic generalization: What is required and can it be learned? In Proc

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.776126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:42.996989Z digest=sha256:ce7115a6d9ee2493b1217e2223b1cbf1906050674a3de50717db6923b15879e6

Observation 95eb09a6-88ba-4d1d-b393-f1c94fef3206 · outbound

This paper cites Object discovery from motion- guided tokens.

Compositional Video Synthesis by Temporal Object-Centric Learning Object discovery from motion- guided tokens

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.768077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:43.070604Z digest=sha256:5349409661feeb9f5a2b7ef2f55b69d2593c50d4569c94f7569677aa29c1578e

Observation 7533f941-d8d2-41d5-915f-96c2da1f9353 · outbound

This paper cites Lumiere: A space-time diffusion model for video generation.

Compositional Video Synthesis by Temporal Object-Centric Learning Lumiere: A space-time diffusion model for video generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T13:16:43.142294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:16:43.142294Z digest=sha256:378bea0827555aabe1feff1aec63008080ec476134020fd3ffc4b21ccf953ee9

Observation 904ad9e4-e0ac-4776-8d6f-6d8f52199e08 · outbound

This paper cites Invariant slot attention: Object discovery with slot-centric reference frames.

Compositional Video Synthesis by Temporal Object-Centric Learning Invariant slot attention: Object discovery with slot-centric reference frames

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.754950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:43.213561Z digest=sha256:ea114c652cbbaeabe0a3526d32bf299df86a399be1a7031fcddc77eb289dd999

Observation 564176c5-8881-49e7-bf47-90de864fcdac · outbound

This paper cites Align your latents: High-resolution video synthesis with latent diffusion models.

Compositional Video Synthesis by Temporal Object-Centric Learning Align your latents: High-resolution video synthesis with latent diffusion models

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.747508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:43.244544Z digest=sha256:12c8daa8f1a655899ecaf1880b7e951350b2089733e80e570f83c6f89a1c60a5

Observation 7a2ab687-3873-4bdf-a200-01ed3cfdfe65 · outbound

This paper cites Emerging properties in self-supervised vision transformers.

Compositional Video Synthesis by Temporal Object-Centric Learning Emerging properties in self-supervised vision transformers

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.740093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:43.317792Z digest=sha256:b5cc7cbbf60db5c5a3fa2339bdad0c578c2308d55f43b34f74269637e50a0acf

Observation 707f5e7d-a837-45ad-96ec-57af4f0a52be · outbound

This paper cites Pixart-α: Fast training of diffusion trans- former for photorealistic text-to-image synthesis.

Compositional Video Synthesis by Temporal Object-Centric Learning Pixart-α: Fast training of diffusion trans- former for photorealistic text-to-image synthesis

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.732572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:43.429828Z digest=sha256:175e70dc186084e898884b60245a115a79f2fa40bec3b2d5e60b7220610ff4c8

Observation 59be87bd-7ec8-4363-b3bc-ad85fe3548b2 · outbound

This paper cites Vision transformers need registers.

Compositional Video Synthesis by Temporal Object-Centric Learning Vision transformers need registers

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.725128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:43.547864Z digest=sha256:d378c2ee8881675401b6a5b0944909c659b53edd510d911407084c7ea5fd0c41

Observation 9d9630db-6022-4695-9b67-8a9adce56d21 · outbound

This paper cites Diffusion models beat GANs on image synthesis.

Compositional Video Synthesis by Temporal Object-Centric Learning Diffusion models beat GANs on image synthesis

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.717663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:43.664202Z digest=sha256:257ee858dc54d64f52a53cd583c7cfe63bf0db2748eb9b8149f51eee42386353

Observation 8c2f57a6-5277-4fe7-93d7-3b620ce15ed6 · outbound

This paper cites Betrayed by attention: A simple yet ef- fective approach for self-supervised video object segmenta- tion.

Compositional Video Synthesis by Temporal Object-Centric Learning Betrayed by attention: A simple yet ef- fective approach for self-supervised video object segmenta- tion

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.710132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:43.782017Z digest=sha256:d4d5895311e68bd18806c08df5f4ee02569c5f678b348e939b739a3df2a233d5

Observation 74e803c0-7c62-43c1-b6b8-4245ad24607d · outbound

This paper cites SAVi++: Towards end-to-end object-centric learning from real-world videos.

Compositional Video Synthesis by Temporal Object-Centric Learning SAVi++: Towards end-to-end object-centric learning from real-world videos

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.702365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:43.921894Z digest=sha256:f76b51eba41b28376a483e0c4d18846ba73b791782e6d721c831e62b32ca11b3

Observation 5f54637c-c83b-4f9a-901e-cc3a1f26adbf · outbound

This paper cites Attend, infer, re- peat: Fast scene understanding with generative models.

Compositional Video Synthesis by Temporal Object-Centric Learning Attend, infer, re- peat: Fast scene understanding with generative models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.694512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:43.928307Z digest=sha256:df0641de30cc1eb02fc6b6ae0df4228a54b51d90d363613fc39eda343d410c25

Observation 29f9374a-5e7b-447c-a0ca-b79e2474428c · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

Compositional Video Synthesis by Temporal Object-Centric Learning Scaling rectified flow transformers for high-resolution image synthesis

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.687199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.034897Z digest=sha256:58d10f8f505df3392a6920f6969d8243378e8a78e60183815021aca101847d9e

Observation e848e0dd-2fde-4e9c-9db3-7ee48f205314 · outbound

This paper cites The PASCAL visual object classes (VOC) challenge.

Compositional Video Synthesis by Temporal Object-Centric Learning The PASCAL visual object classes (VOC) challenge

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.679727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.050887Z digest=sha256:6e93e4abeda66a05ca1213cce65189719f28586bccf97b711f6eb1f50131910e

Observation 2c509ebe-7e2b-4dbc-ad6d-22c102cae5ff · outbound

This paper cites Connectionism and cognitive architecture: A critical analysis.

Compositional Video Synthesis by Temporal Object-Centric Learning Connectionism and cognitive architecture: A critical analysis

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.672196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.057618Z digest=sha256:ae90ae1316b2da4910e6c6ddf7e362b69df44a16ba7478582f97fd0571748e6b

Observation 534e61bc-8442-41ee-9984-7d04bdbb1073 · outbound

This paper cites Understanding the diffi- culty of training deep feedforward neural networks.

Compositional Video Synthesis by Temporal Object-Centric Learning Understanding the diffi- culty of training deep feedforward neural networks

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.664636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.085924Z digest=sha256:8dd2bb75479c87d5538ebf078acc4120511ca1998bb5a95e9448b9b0e062af21

Observation 77e30ea2-d3d8-4895-a2cf-b9afd15ebbfc · outbound

This paper cites Multi-object representation learning with iterative variational inference.

Compositional Video Synthesis by Temporal Object-Centric Learning Multi-object representation learning with iterative variational inference

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.656961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.088511Z digest=sha256:e41b64576344a8168818b604a54f85324a5e2a75057c8fc90480790bc2c9dcd6

Observation 58c9c3fa-53e4-411f-a45c-a0f2267e8589 · outbound

This paper cites On the Binding Problem in Artificial Neural Networks.

Compositional Video Synthesis by Temporal Object-Centric Learning On the Binding Problem in Artificial Neural Networks

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T13:16:44.091002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:16:44.091002Z digest=sha256:165a7534072bc50ed1843ed99dd73c57f578e169a054aa9c5da84fd34e7b8bf5

Observation 1477d4d0-274e-45d9-b34f-952f1459db35 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equi- librium.

Compositional Video Synthesis by Temporal Object-Centric Learning Gans trained by a two time-scale update rule converge to a local nash equi- librium

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.648983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.093688Z digest=sha256:01bed092e13f32484ca6d1607b06f540967360cf2ec75a01eb8443bf2cfe5f2a

Observation 7d369d06-fc9f-4b2d-a042-6b0553d869a1 · outbound

This paper cites Denoising diffu- sion probabilistic models.

Compositional Video Synthesis by Temporal Object-Centric Learning Denoising diffu- sion probabilistic models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.641276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.096080Z digest=sha256:fafa23b7f733170f94602bed608c22981ffb1bb686adb9764af9e1ba167b118a

Observation 8f00cd9d-d67b-4eee-9010-f96cbbebff78 · outbound

This paper cites Imagen Video: High Definition Video Generation with Diffusion Models.

Compositional Video Synthesis by Temporal Object-Centric Learning Imagen Video: High Definition Video Generation with Diffusion Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T13:16:44.098463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:16:44.098463Z digest=sha256:6ea87097302c9a4577e085093e9c006f7d39d5f3bb39c05622abf743ae23b419

Observation 0b8f085c-cbb8-4239-b402-210cb3196108 · outbound

This paper cites Video dif- fusion models.

Compositional Video Synthesis by Temporal Object-Centric Learning Video dif- fusion models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.633260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.101111Z digest=sha256:6fde8bc2143bcae6eb3b19a9160bc9acaa765e6abc696f6c366cabfa3ec743d1

Observation e8569ae3-5caf-4512-a0f6-1878af20ae95 · outbound

This paper cites Object-centric slot diffusion.

Compositional Video Synthesis by Temporal Object-Centric Learning Object-centric slot diffusion

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.625355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.103608Z digest=sha256:e364949ac69edf2f73b0fe2d4cd37aed9b33bf07e915ecfc12944e35014d7880

Observation 27f4d3f3-0c29-4eb7-b51f-4772803cad9c · outbound

This paper cites CLEVR: A diagnostic dataset for compositional language and elementary visual reasoning.

Compositional Video Synthesis by Temporal Object-Centric Learning CLEVR: A diagnostic dataset for compositional language and elementary visual reasoning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.617467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.106047Z digest=sha256:6ca6bbbd72da6f90b57e4c646a922d2dd1cf2c763a905fd931f07999f04c290b

Observation 77265028-906b-4e4e-aba8-e68f208e99f6 · outbound

This paper cites ClevrTex: A Texture-Rich Benchmark for Unsupervised Multi-Object Segmentation.

Compositional Video Synthesis by Temporal Object-Centric Learning ClevrTex: A Texture-Rich Benchmark for Unsupervised Multi-Object Segmentation

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.609002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.108500Z digest=sha256:7a865c247886b420ac6b202c65fdb303c78cd3d7b09c49f80c26b9541abbb2fa

Observation 1ed7a1e7-71f1-4bc6-815d-02ac1b9c81a6 · outbound

This paper cites Con- ditional object-centric learning from video.

Compositional Video Synthesis by Temporal Object-Centric Learning Con- ditional object-centric learning from video

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.600628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.111133Z digest=sha256:71c6d2724aba0bcb8b28c360adf3db99fb2cd90c30a892582da61371211d047c

Observation 30f913bd-9f30-4e1b-a0dc-deaa3d39e6d0 · outbound

This paper cites Sequential attend, infer, repeat: Generative mod- elling of moving objects.

Compositional Video Synthesis by Temporal Object-Centric Learning Sequential attend, infer, repeat: Generative mod- elling of moving objects

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.592327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.113684Z digest=sha256:e7f49e9054b1cd3817e6f3407d64f8f7b7c849548a126f750192155551b555b4

Observation 17a89202-5cc8-4743-b9af-57f04207f6ea · outbound

This paper cites Structured object-aware physics prediction for video modeling and planning.

Compositional Video Synthesis by Temporal Object-Centric Learning Structured object-aware physics prediction for video modeling and planning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.584228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.116118Z digest=sha256:903f4c0507ad22f77e6f0861f4331cd9a198ff2c04d29048fec4264a9f01efb2

Observation e6fb7c66-938a-4e30-912d-14fc1de3b904 · outbound

This paper cites Hierarchical compact clustering attention (coca) for unsupervised object-centric learning.

Compositional Video Synthesis by Temporal Object-Centric Learning Hierarchical compact clustering attention (coca) for unsupervised object-centric learning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.575661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.118397Z digest=sha256:be2c2ae553587d9d224b718ebc3d7ee7ab527e69562c0328fb42834d6d37eda3

Observation 5ddaa5c3-53a9-47ec-b5bc-1618ee2143fb · outbound

This paper cites an unresolved cited work.

Compositional Video Synthesis by Temporal Object-Centric Learning Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T13:16:44.120650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:16:44.120650Z digest=sha256:045942bc0f2ede16747f2425bc3dd622268b4b05a86f9c733f6baf9d4c6a49d2

Observation 834353d7-8e1f-4680-a4b7-7da091f144c5 · outbound

This paper cites Building machines that learn and think like people.

Compositional Video Synthesis by Temporal Object-Centric Learning Building machines that learn and think like people

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.562155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.123075Z digest=sha256:7c5b235defc4fe6c854b913ee534249e54cadeb3f01904f64e07dae85da10fcc

Observation 0d7f7806-0acb-42db-a5c5-f5a61b987d54 · outbound

This paper cites an unresolved cited work.

Compositional Video Synthesis by Temporal Object-Centric Learning Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:16:44.554659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.125394Z digest=sha256:949d7f11a5624f6595f1719d1ce4bde36ad1bc972d07ce96629ce2019c44b629

Observation eba8958f-3f05-401b-a4ab-d739601f684b · outbound

This paper cites Microsoft COCO: Common objects in context.

Compositional Video Synthesis by Temporal Object-Centric Learning Microsoft COCO: Common objects in context

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.547077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.127855Z digest=sha256:777aaa23beee8f95a98295b2ba09456261a00c57e3f5693fec48761d23b16b5b

Observation 3545ae88-31fd-481a-8f73-0a4d3a4493e6 · outbound

This paper cites Improving generative imagination in object-centric world models.

Compositional Video Synthesis by Temporal Object-Centric Learning Improving generative imagination in object-centric world models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.539180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.130351Z digest=sha256:c2643bf6e8e53ef56fedf20e839a3f6636186354a01b5107d9316e9b20c0afaa

Observation 0e3c82ae-6228-47e3-9f6e-e060c820ef0a · outbound

This paper cites Object- centric learning with slot attention.

Compositional Video Synthesis by Temporal Object-Centric Learning Object- centric learning with slot attention

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.531327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.132668Z digest=sha256:e4478fc0244ca67d884f732c53dc8f7c5c9926dff66c595f63dd5776d707146c

Observation 32da4fa8-bedc-423f-8fe7-39cb16b2ee4e · outbound

This paper cites Decoupled weight decay regularization.

Compositional Video Synthesis by Temporal Object-Centric Learning Decoupled weight decay regularization

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.523772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.135327Z digest=sha256:b8238e1c07b527ba12a8ac591d7ab9cf88c155ec0b4ffa0eecf5e05a0bd4cf2e

Observation 5adc5170-5997-46c5-9570-514199a6fa17 · outbound

This paper cites Temporally consistent object-centric learning by contrasting slots.

Compositional Video Synthesis by Temporal Object-Centric Learning Temporally consistent object-centric learning by contrasting slots

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.515469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.138069Z digest=sha256:719c1fb1915e1b6f54d7f3989ff72cfc91747089f084947c070cae100d070839

Observation 2fc8e8f7-388a-47d0-a9ff-cdd64dffacbf · outbound

This paper cites T2i-adapter: Learning adapters to dig out more controllable ability for text-to- image diffusion models.

Compositional Video Synthesis by Temporal Object-Centric Learning T2i-adapter: Learning adapters to dig out more controllable ability for text-to- image diffusion models

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.507184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.140512Z digest=sha256:f5a0a23c44fb99c24be5c4237d4e062ab4e4f24bfc764c0ef9f9bb5262b902e8

Observation 00dd651f-33cd-41ae-b79d-05ca2e1c01a3 · outbound

This paper cites Segmentation of moving objects by long term video analysis.

Compositional Video Synthesis by Temporal Object-Centric Learning Segmentation of moving objects by long term video analysis

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.499643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.142834Z digest=sha256:8e29cc20c7fa927fb6aa2ac124b7f20d2767a4fe17963b6b83d2d854ba8f5cf9

Observation 7f7c8fda-9ce4-40af-8397-092a650ea45d · outbound

This paper cites Dinov2: Learning robust visual features without supervision.

Compositional Video Synthesis by Temporal Object-Centric Learning Dinov2: Learning robust visual features without supervision

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.491763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.145544Z digest=sha256:0e7baff7ec70d22c361645e884ef869babd10f97427f8ee5be8ef8b0bf68081b

Observation 24f67be7-53c4-4945-9708-08e717ecb150 · outbound

This paper cites A benchmark dataset and evaluation methodology for video object segmentation.

Compositional Video Synthesis by Temporal Object-Centric Learning A benchmark dataset and evaluation methodology for video object segmentation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.484295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.148072Z digest=sha256:752ac7cd87dc42b7d9eca7b089e44d9063017b693fc3460eb09015bd5cc4e1b0

Observation a7e07070-f2f3-4527-867b-36a12b29eaf6 · outbound

This paper cites Sdxl: Improving latent diffusion models for high-resolution image synthesis.

Compositional Video Synthesis by Temporal Object-Centric Learning Sdxl: Improving latent diffusion models for high-resolution image synthesis

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.476532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.150674Z digest=sha256:d6ef0fc2dda4316d94b944c39fc7977170edff80171c76f790502fb34ed506ee

Observation 19ade172-0b41-4ab7-91ad-6c6fd7a0acf7 · outbound

This paper cites Rethinking image-to-video adaptation: An object-centric perspective.

Compositional Video Synthesis by Temporal Object-Centric Learning Rethinking image-to-video adaptation: An object-centric perspective

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.468916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.153056Z digest=sha256:e6bd69770fad2248374c3bf2df66cb1438346a9f4a63801367633f595ccc6b64

Observation 701887ea-ba26-49eb-8e74-852cdb8f9523 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Compositional Video Synthesis by Temporal Object-Centric Learning Learning transferable visual models from natural language supervision

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.461242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.155324Z digest=sha256:c6cdb51911bcad17a3756c99983eba17c2664b0605daedd936a42ff7623a475d

Observation 9b7f2095-64c9-4937-bfd2-2b70c1e364d3 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Compositional Video Synthesis by Temporal Object-Centric Learning Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T13:16:44.157831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:16:44.157831Z digest=sha256:7700e97a30a28ed312c139984a53c76f591527b344ca41361cb7938fe63d3064

Observation 21d5838a-89ce-4f63-b000-97a33195b8a6 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Compositional Video Synthesis by Temporal Object-Centric Learning High-resolution image synthesis with latent diffusion models

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.453298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.160495Z digest=sha256:968ed602a7dafd40b90e8b51f3af2800a1059deca6159954cd8c6d43a1f41842

Observation 73df8951-2164-4f06-ac6b-f976ac6425ea · outbound

This paper cites Photorealistic text-to-image JOURNAL OF LATEX CLASS FILES, VOL.

Compositional Video Synthesis by Temporal Object-Centric Learning Photorealistic text-to-image JOURNAL OF LATEX CLASS FILES, VOL

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.444819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.162837Z digest=sha256:9da86460b9c4d47c3386181885ca724af8e90850a4a69ba3c12b265e7abf100e

Observation c56b3bf3-90b6-4d05-b79c-9b3d626a4569 · outbound

This paper cites Toward causal representation learning.

Compositional Video Synthesis by Temporal Object-Centric Learning Toward causal representation learning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.437460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.165116Z digest=sha256:8b2ee90da97024de1f681896569ba46739899abcec2937bd425d9f41adfa8c01

Observation 79c76354-ae9d-4e40-8cb6-4c0501799db6 · outbound

This paper cites Bridging the gap to real-world object-centric learning.

Compositional Video Synthesis by Temporal Object-Centric Learning Bridging the gap to real-world object-centric learning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.430034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.167520Z digest=sha256:384b7e3777dd78ac004768d6ea693793177d29e83f4bbdd5a238bb8197432a33

Observation 3ab10f8a-824f-47d9-91f1-a7c1d2ab0205 · outbound

This paper cites Make-a-video: Text-to-video generation without text-video data.

Compositional Video Synthesis by Temporal Object-Centric Learning Make-a-video: Text-to-video generation without text-video data

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.422188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.169814Z digest=sha256:cf5a48e00a546b66a6dfc3b95340ba96ffe3e85af21ce37be4510434aa516ba6

Observation 8a7ca348-9b49-46d5-8aa7-70c0098f49b6 · outbound

This paper cites Illiterate dall- e learns to compose.

Compositional Video Synthesis by Temporal Object-Centric Learning Illiterate dall- e learns to compose

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.414080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.172171Z digest=sha256:1bd725295f2c0bf4c54b52cd5b079520ab8b62c3bde2f40ce898eccdf416c47c

Observation 9d97a6e9-7422-43a5-8a2c-5a424caf13f7 · outbound

This paper cites Simple unsu- pervised object-centric learning for complex and natural- istic videos.

Compositional Video Synthesis by Temporal Object-Centric Learning Simple unsu- pervised object-centric learning for complex and natural- istic videos

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.406073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.175236Z digest=sha256:dcea45eefc8be7c4dd03b96a11dc857c0f96c9eaf18dcadcd4a1e804711b2877

Observation 425ccc73-ca74-4252-bcb6-d1e1c1ad69e4 · outbound

This paper cites Guided latent slot diffusion for object-centric learning.

Compositional Video Synthesis by Temporal Object-Centric Learning Guided latent slot diffusion for object-centric learning

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.398611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.177916Z digest=sha256:2d93d8cd8ecd50f921c4b58eaeb6f7a6cf569d429d18487940a0177acf3f420c

Observation 7ac35c2b-f2d0-487e-bede-3d8421ad2798 · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

Compositional Video Synthesis by Temporal Object-Centric Learning Deep unsupervised learning using nonequilibrium thermodynamics

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.390287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.180311Z digest=sha256:cec4edd103c4150f9f91a77d07a7de414cd204b4e6e8968216fe738d2d597919

Observation bd9ce06d-2791-4be4-8d9d-68151011e5e9 · outbound

This paper cites Core knowl- edge.

Compositional Video Synthesis by Temporal Object-Centric Learning Core knowl- edge

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.382246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.182654Z digest=sha256:471538b89eecf51eae956e3ac82b7e7176157e0430285c75f98c00f38d9d36e9

Observation b0fbe04d-6cdf-42a2-a350-706fcfb4e2c0 · outbound

This paper cites Mind games: Game engines as an architecture for intuitive physics.

Compositional Video Synthesis by Temporal Object-Centric Learning Mind games: Game engines as an architecture for intuitive physics

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.375066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.185445Z digest=sha256:8f13579ea6cd9a2d9e73003b9c1410e5fc86d504aa291b34ca90d1129659f812

Observation af1b4f84-5a23-4632-aedb-2b80edaab6a8 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

Compositional Video Synthesis by Temporal Object-Centric Learning Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T13:16:44.188237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:16:44.188237Z digest=sha256:1f4758af2f1faa28a730d27ef7309b745863adc80d19143ec578e6d836f6f23d

Observation 823ef0a0-2908-4923-874a-54a70b1d1b13 · outbound

This paper cites Attention is all you need.

Compositional Video Synthesis by Temporal Object-Centric Learning Attention is all you need

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.367744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.191030Z digest=sha256:afe0b2f709278217541c623c109a83727e29df482a5f35aa92e20bde123370e2

Observation 47df23b6-9d0d-4d93-b779-f07af27600c2 · outbound

This paper cites Phenaki: Variable length video generation from open domain textual descrip- tions.

Compositional Video Synthesis by Temporal Object-Centric Learning Phenaki: Variable length video generation from open domain textual descrip- tions

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.360071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.193315Z digest=sha256:f427b82432b1142cf1a462c3125dd5cd6e46547b7c2fa83f6bcf94437a58a8d5

Observation df72c4e7-c5a8-41bf-b7c3-552be28f8323 · outbound

This paper cites Videocomposer: Compositional video syn- thesis with motion controllability.

Compositional Video Synthesis by Temporal Object-Centric Learning Videocomposer: Compositional video syn- thesis with motion controllability

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.352046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.195673Z digest=sha256:b98c839895a78e37c96235afe95bf77f3ce8adfe474c121f8e52130595c4d287

Observation edcc823e-9d61-4eba-8783-04ba52c8c529 · outbound

This paper cites Bovik, Hamid R.

Compositional Video Synthesis by Temporal Object-Centric Learning Bovik, Hamid R

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.344281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.198465Z digest=sha256:da4b6159e213f186ee7f5691d6cb61688e34cf681328f70a59077c7ad455071f

Observation ee08bb91-16ae-485c-939d-fdc14c5731c6 · outbound

This paper cites Tune-a-video: One-shot tun- ing of image diffusion models for text-to-video generation.

Compositional Video Synthesis by Temporal Object-Centric Learning Tune-a-video: One-shot tun- ing of image diffusion models for text-to-video generation

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.336566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.200811Z digest=sha256:e291cf12f39504831dfbff0ae657c7f7270c98f41357c54cabc76d0963a163ec

Observation 430bb969-8803-40e6-b8c4-8ed33041241e · outbound

This paper cites SlotFormer: Unsupervised visual dynamics simulation with object-centric models.

Compositional Video Synthesis by Temporal Object-Centric Learning SlotFormer: Unsupervised visual dynamics simulation with object-centric models

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.328676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.203102Z digest=sha256:d2b1a5f94399be69090b06ac8d8b45e36ede27388a839bf6803c14a14a7d451e

Observation e0d99ffa-f4c9-4e52-b5e7-f53cfec49e3b · outbound

This paper cites Slotdiffusion: Object-centric generative model- ing with diffusion models.

Compositional Video Synthesis by Temporal Object-Centric Learning Slotdiffusion: Object-centric generative model- ing with diffusion models

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.319819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.205454Z digest=sha256:b9d91e67e2c6c14fb082b4d132a0d3f6fd9fdc81ff6995c4925ef9c32a39ecb5

Observation 049bc11e-77fc-4b40-946f-fe0b3b0f08ef · outbound

This paper cites Segment- ing moving objects via an object-centric layered represen- tation.

Compositional Video Synthesis by Temporal Object-Centric Learning Segment- ing moving objects via an object-centric layered represen- tation

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.311397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.208326Z digest=sha256:c48e63b75d7388579731b1b4177bc0cafb394e0d76dfba73301c4e6182e90ad3

Observation ab3234e6-fe68-4713-a9ce-81bc16a5decc · outbound

This paper cites Youtube-vos: A large-scale video object segmentation benchmark.

Compositional Video Synthesis by Temporal Object-Centric Learning Youtube-vos: A large-scale video object segmentation benchmark

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.303372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.210743Z digest=sha256:7c557b97a7e02492a5145d9454a5f62a1fc75d93bc603ed02a3b2b8bcf898968

Observation 94a6aae8-6199-45ce-943e-2c72a445de2b · outbound

This paper cites Video instance segmentation.

Compositional Video Synthesis by Temporal Object-Centric Learning Video instance segmentation

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.295284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.213086Z digest=sha256:9311b029fc28ebf6a39f167d737639a59ef6b8c6daf255f3e7c207e8211e693e

Observation 53ca6954-638a-4aff-90f2-0f387802af24 · outbound

This paper cites Efros, Eli Shechtman, and Oliver Wang.

Compositional Video Synthesis by Temporal Object-Centric Learning Efros, Eli Shechtman, and Oliver Wang

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.287579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.215584Z digest=sha256:8e90e3da7304297438cc9a71832d100be2bba2b02bdc54c9569dcb5bf7909c88

Observation cdae38e8-f316-438b-ab0c-21b77609e9d0 · outbound

This paper cites Controlvideo: Training-free controllable text-to-video generation.

Compositional Video Synthesis by Temporal Object-Centric Learning Controlvideo: Training-free controllable text-to-video generation

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:16:44.278578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:16:44.217929Z digest=sha256:f35b82b42c54201ba45dd26b68d95cbe60f78f2cc8ce5b129a26820f94675823

Pith citing papers

No inbound Pith citation observations are available.