Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-14T02:16:45.554252Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 98 of 98 outbound references and 100 inbound Pith citation observations for arXiv:2509.20328.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-14T02:16:45.554252Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T08:20:10.044650Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T12:15:01.137692Z
98 of 98 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation bef73ba4-bf2d-41db-8f33-085b65e67f4f · outbound
Video models are zero-shot learners and reasoners A Survey on Large Language Models for Code Generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4f41cd84-640e-4566-b877-8b0cad30b2a1 · outbound
Video models are zero-shot learners and reasoners Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 70f7216d-28ca-4688-ad18-9e2d03da39b4 · outbound
Video models are zero-shot learners and reasoners Weaver: Foundation Models for Creative Writing
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 33061bc7-b30a-4fc4-bdee-440561bd778b · outbound
Video models are zero-shot learners and reasoners Multilingual Machine Translation with Large Language Models: Empirical Results and Analysis
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1f8e1a78-c68f-4e95-adc5-ac257f4defc3 · outbound
Video models are zero-shot learners and reasoners The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1e4b0727-f9e6-4e01-ace3-8b1d8e75b720 · outbound
Video models are zero-shot learners and reasoners Agent Laboratory: Using LLM Agents as Research Assistants
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f01a4d9f-b346-4837-a3ad-f958976b4d7e · outbound
Video models are zero-shot learners and reasoners Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 348889bd-e45a-4be7-be25-ee6416d71be5 · outbound
Video models are zero-shot learners and reasoners Emergent Abilities of Large Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ef3203b8-916f-402c-b6dd-615aa45ae9b8 · outbound
Video models are zero-shot learners and reasoners A Survey on In-context Learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a12b7684-345e-4894-930a-8cb96a6a7e17 · outbound
Video models are zero-shot learners and reasoners Large language models are zero-shot reasoners.Advances in neural information processing systems, 35: 22199–22213
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 07e10f6a-0a33-436b-8da5-831d03b9d893 · outbound
Video models are zero-shot learners and reasoners Segment anything
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 31b2fcc2-bc71-4d27-a17d-8578eacd7d05 · outbound
Video models are zero-shot learners and reasoners SAM 2: Segment Anything in Images and Videos
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6e7e27cb-928d-4b5a-a0ac-b0f675908781 · outbound
Video models are zero-shot learners and reasoners You only look once: Unified, real-time object detection
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 599fb6f1-94b3-435c-aaf5-cf4dcb48c2cc · outbound
Video models are zero-shot learners and reasoners YOLOv11: An Overview of the Key Architectural Enhancements
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 88ce5459-4164-4a44-985c-3275a995d336 · outbound
Video models are zero-shot learners and reasoners From Generation to Generalization: Emergent Few-Shot Learning in Video Diffusion Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6db6f449-ee4c-4925-96a1-26781f0063bc · outbound
Video models are zero-shot learners and reasoners Taskonomy: Disentangling task transfer learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cb05e227-021d-4edc-b7b1-623258346c5a · outbound
Video models are zero-shot learners and reasoners RealGeneral: Unifying Visual Generation via Temporal In-Context Learning with Video Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 296b9fa4-276c-4e55-9db6-add78bb5826a · outbound
Video models are zero-shot learners and reasoners Visualcloze: A universal image generation framework via visual in-context learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ae17744b-1166-4a7c-b573-310a46abdc7b · outbound
Video models are zero-shot learners and reasoners Images speak in images: A generalist painter for in-context visual learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 02ae06f0-e66e-4bea-9b6e-cb5f3c7669ab · outbound
Video models are zero-shot learners and reasoners Test- time visual in-context tuning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e9fd3736-1373-466c-bc23-fbf9e0aba84c · outbound
Video models are zero-shot learners and reasoners PixWizard: Versatile Image-to-Image Visual Assistant with Open-Language Instructions
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b3e254ba-791e-47a7-9ea8-d5cb42603a52 · outbound
Video models are zero-shot learners and reasoners Omnigen: Unified image generation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c66a504e-7a54-40b2-9460-8134ed831263 · outbound
Video models are zero-shot learners and reasoners One diffusion to generate them all
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cc555498-692d-46d6-8300-b1e47755234d · outbound
Video models are zero-shot learners and reasoners Dreamix: Video Diffusion Models are General Video Editors
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c4cebdcd-5d64-4fac-b9bd-0e9abc855778 · outbound
Video models are zero-shot learners and reasoners Scalingproperties of diffusion models for perceptual tasks
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4a4b10f8-090a-45da-8050-70aa98552be8 · outbound
Video models are zero-shot learners and reasoners Video as the New Language for Real-World Decision Making
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b033b0c7-d2a0-4392-960a-97a413181f01 · outbound
Video models are zero-shot learners and reasoners Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 35:24824–24837
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4edf34c7-13a5-44d1-8391-472090970482 · outbound
Video models are zero-shot learners and reasoners Large language models are human-level prompt engineers
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 56fcc7f1-f44a-42d9-93c1-8b4f9da2b148 · outbound
Video models are zero-shot learners and reasoners Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing.ACM computing surveys, 55(9):1–35
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation abbdbd02-e4b8-4427-b500-ab4bafa21be2 · outbound
Video models are zero-shot learners and reasoners Vertex AI Veo Prompt Rewriter.https://cloud.google.com/vertex-ai/ generative-ai/docs/video/turn-the-prompt-rewriter-off#prompt-rewriter
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d24e5fc7-f1b8-4028-9193-a70fe8262488 · outbound
Video models are zero-shot learners and reasoners Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 28ab6dec-88e1-475e-8054-73ef29444676 · outbound
Video models are zero-shot learners and reasoners Lmsys org text-to-video leaderboard.https://lmarena.ai/leaderboard/t ext-to-video, September 2025
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8376f9d4-3c43-4aba-a2c3-94e1c960dc6d · outbound
Video models are zero-shot learners and reasoners Veo 2 announcement.https://blog.google/technology/google-labs/vide o-image-generation-update-december-2024/
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 909cac90-fa50-496e-b8b9-e400b1cc3e07 · outbound
Video models are zero-shot learners and reasoners Veo 2 launch.https://developers.googleblog.com/en/veo-2-video-gen eration-now-generally-available/
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6954cf1d-45d1-4eeb-ae65-5546e6d0dbe1 · outbound
Video models are zero-shot learners and reasoners Veo 3 announcement.https://blog.google/technology/ai/generative-m edia-models-io-2025/
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ae0d721a-5a0d-48e1-b0cd-12dad6264034 · outbound
Video models are zero-shot learners and reasoners Veo 3 launch
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 83e868e0-cbe2-4cd9-9898-a5030c80e7c3 · outbound
Video models are zero-shot learners and reasoners Holistically-nested edge detection
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 662da419-d5dc-44b8-90e1-60a2fc0f1461 · outbound
Video models are zero-shot learners and reasoners IntPhys: A Framework and Benchmark for Visual Intuitive Physics Reasoning
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 71f277fe-c5d9-4842-a170-1f6c587c8070 · outbound
Video models are zero-shot learners and reasoners Bear, Elias Wang, Damian Mrowca, Felix J
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6f0f6882-7402-49be-a0c9-578581e2ce98 · outbound
Video models are zero-shot learners and reasoners Benchmarking progress to infant-level physical reasoning in ai
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e67d384f-5157-445e-8efd-cab7b1cd49a2 · outbound
Video models are zero-shot learners and reasoners GRASP: A novel benchmark for evaluating language GRounding And Situated Physics understanding in multimodal language models
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 88cbd2c3-72a9-4dc1-b803-bc9136d59245 · outbound
Video models are zero-shot learners and reasoners Physion++: Evaluating physical scene understanding that requires online inference of different physical properties.Advances in Neural Information Processing Systems, 36
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d4da3c3e-8496-43bd-8189-9107dfec051c · outbound
Video models are zero-shot learners and reasoners Videophy: Evaluating physical commonsense for video generation
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0bb14a58-a96f-4d6b-b8f7-35fef17abfa4 · outbound
Video models are zero-shot learners and reasoners LLMPhy: Parameter-Identifiable Physical Reasoning Combining Large Language Models and Physics Engines
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 67e73ec4-8a3f-4025-a1c0-ba330521b86e · outbound
Video models are zero-shot learners and reasoners Towards world simulator: Crafting physical commonsense- based benchmark for video generation
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b46f4546-d677-490e-88a4-123891b1c737 · outbound
Video models are zero-shot learners and reasoners How Far is Video Generation from World Model: A Physical Law Perspective
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 11174655-76a8-4901-8b29-a845e33e090a · outbound
Video models are zero-shot learners and reasoners Do generative video models understand physical principles?
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 658d0d9f-56e1-443c-aea4-d5a136be12ee · outbound
Video models are zero-shot learners and reasoners Generative Physical AI in Vision: A Survey
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cd3b184b-2f2c-4dc2-b44e-bb5f7f648a13 · outbound
Video models are zero-shot learners and reasoners Visual cognition in multimodal large language models.Nature Machine Intelligence, pages 1–11
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f958c9a8-4a82-449e-bf1f-b2406c8a9c42 · outbound
Video models are zero-shot learners and reasoners Evaluating Newtonian Mechanics in Video Generative Models with Real Physical Systems
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 69a5a721-c0c7-4f28-9e95-6fb80f7ebe5c · outbound
Video models are zero-shot learners and reasoners Intuitive physics understanding emerges from self-supervised pretraining on natural videos
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 92c33d57-4426-4dd7-aaa5-6de33846f577 · outbound
Video models are zero-shot learners and reasoners Visual Jenga: Discovering Object Dependencies via Counterfactual Inpainting
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c478925f-bc14-4f6d-8c21-88a2922abdb7 · outbound
Video models are zero-shot learners and reasoners Human-level concept learning through probabilistic program induction.Science, 350(6266):1332–1338
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 96208e04-2720-4f38-9ca4-5006ef5ab42c · outbound
Video models are zero-shot learners and reasoners V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 04c05d67-43d5-4323-94cd-b5b60de22f9b · outbound
Video models are zero-shot learners and reasoners Nano Banana: Gemini Image Generation Overview.https://gemini.google/ov erview/image-generation/
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 819c2c9b-35fe-4b93-ba1e-7ad3611949f1 · outbound
Video models are zero-shot learners and reasoners Text-to-image diffusion models are zero shot classifiers.Advances in Neural Information Processing Systems, 36:58921–58937
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cc1b3590-b6be-4fa7-9507-ff61347bde4f · outbound
Video models are zero-shot learners and reasoners Intriguing properties of generative classifiers
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d0b8f05f-29c6-4fde-828f-30cdf38962e9 · outbound
Video models are zero-shot learners and reasoners Peekaboo: Text to Image Diffusion Models are Zero-Shot Segmentors
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c52699cf-9069-4241-a370-53baac5fbf99 · outbound
Video models are zero-shot learners and reasoners Text2video-zero: Text-to-image diffusion models are zero-shot video generators
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5d6a7d5a-7fb4-475c-a741-d2b810d252b4 · outbound
Video models are zero-shot learners and reasoners Dense extreme inception network: Towards a robust CNN model for edge detection
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3048c93f-6aae-4f61-b8df-6a717a2cf6b6 · outbound
Video models are zero-shot learners and reasoners Dense extreme inception network for edge detection.Pattern Recognition, 139:109461
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 66e36985-accc-4d25-b3eb-a604f46f84a8 · outbound
Video models are zero-shot learners and reasoners LVIS: A dataset for large vocabulary instance segmentation
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 90b5de49-3b14-4f96-be59-40327be201fc · outbound
Video models are zero-shot learners and reasoners Emu edit: Precise image editing via recognition and generation tasks
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 85e9d970-c35b-4f80-8a49-45453bb18c8a · outbound
Video models are zero-shot learners and reasoners Diffusion Model-Based Video Editing: A Survey
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e1cbb1d1-877c-44da-8271-f7955efe7489 · outbound
Video models are zero-shot learners and reasoners VEGGIE: instructional editing and reasoning video concepts with grounded generation
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7d99fe5e-b18c-43bb-a476-59b7987e8e6a · outbound
Video models are zero-shot learners and reasoners Pathways on the image manifold: Image editing via video generation
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 562bad6e-290e-434a-bb7f-6e19ffefdb60 · outbound
Video models are zero-shot learners and reasoners Kiva: Kid-inspired visual analogies for testing large multimodal models
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 82df084a-7573-43b6-8f1a-caa524c8f283 · outbound
Video models are zero-shot learners and reasoners ImageNet classification with deep convolutional neural networks.Advances in neural information processing systems, 25
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e5ac8b7b-b2a7-462c-b506-d6c6c3b849da · outbound
Video models are zero-shot learners and reasoners The broader spectrum of in-context learning
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 19f0b701-2afd-4426-896c-e7054f8baa9e · outbound
Video models are zero-shot learners and reasoners Performance vs
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 32f7ea35-5384-41ba-9d67-137c080d06f2 · outbound
Video models are zero-shot learners and reasoners Shortcut learning in deep neural networks.Nature Machine Intelligence, 2(11):665–673
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7b0824cf-c67a-4100-95d4-2fb9af23cecd · outbound
Video models are zero-shot learners and reasoners LLM inference prices have fallen rapidly but unequally across tasks, march 2025
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation df72a32f-a177-4c83-8b72-91c9b6d470e8 · outbound
Video models are zero-shot learners and reasoners Self-Consistency Improves Chain of Thought Reasoning in Language Models
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e8623ba1-15bd-4c91-b129-e55e413cfbd1 · outbound
Video models are zero-shot learners and reasoners Self-refine: Iterative refinement with self-feedback.Advances in Neural Information Processing Systems, 36:46534–46594
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5bd3cd9d-904f-446a-a968-c1c1e5f1c99e · outbound
Video models are zero-shot learners and reasoners OpenAI o1 System Card
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c3da3f50-69ff-463b-9fb1-17618de2aa3e · outbound
Video models are zero-shot learners and reasoners Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8c1e15e0-3515-4b12-9dc9-f0a7d598cb99 · outbound
Video models are zero-shot learners and reasoners Training language models to follow instructions with human feedback.Advances in neural information processing systems, 35: 27730–27744
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f1b08678-78a0-49f1-b3c5-0af7bc882d9c · outbound
Video models are zero-shot learners and reasoners Instruction Tuning with GPT-4
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2490a2bf-1199-4680-b469-ec23e4abd2c4 · outbound
Video models are zero-shot learners and reasoners Sparse gradient regularized deep retinex network for robust low-light image enhancement.IEEE Transactions on Image Processing, 30:2072–2086
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 129d23ba-5da9-4787-a816-bca8f930feac · outbound
Video models are zero-shot learners and reasoners Under- standing the limits of vision language models through the lens of the binding problem.Advances in Neural Information Processing Systems, 37:113436–113460
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 57826310-2a63-4706-a644-d42444dbfa92 · outbound
Video models are zero-shot learners and reasoners Unresolved cited work
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b6c999c7-920d-4a66-8efc-081d7d7e94b2 · outbound
Video models are zero-shot learners and reasoners The intelligent eye
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9c7b4fd0-bcc6-4d1c-b8be-2081d169ced8 · outbound
Video models are zero-shot learners and reasoners ImageNet-trained CNNs are biased towards texture; increasing shape bias improves accuracy and robustness
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e2ea640d-bf28-41bb-a0b7-198814fce6af · outbound
Video models are zero-shot learners and reasoners Objaverse-xl: A universe of 10m+ 3d objects.Advances in Neural Information Processing Systems, 36:35799–35813
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3540fdfa-a7d1-4c1c-8c83-db344f8fa203 · outbound
Video models are zero-shot learners and reasoners On the Measure of Intelligence
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0117641c-90ad-48ed-9529-5ca00c388108 · outbound
Video models are zero-shot learners and reasoners Lawrence Zitnick
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9dee6f10-07ab-47ed-90d3-c01d8328272c · outbound
Video models are zero-shot learners and reasoners Lawrence Zitnick
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b2f09b07-9408-4916-9701-b83fc85fe857 · outbound
Video models are zero-shot learners and reasoners Lawrence Zitnick and Piotr Dollár
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 22148205-b5f1-43a3-98d7-96f2b2b93fc6 · outbound
Video models are zero-shot learners and reasoners Superedge: Towards a generalization model for self-supervised edge detection.CoRR
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a2962c2f-0e77-42b3-aef7-fcbac00a5771 · outbound
Video models are zero-shot learners and reasoners Maze dataset.https://pypi.org/project/maze-dataset/0.3.4/
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ab715fa7-ffc0-4b98-b2c7-296648a4e7e3 · outbound
Video models are zero-shot learners and reasoners 15 Video models are zero-shot learners and reasoners
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a2c54ea0-4e4c-4a4d-b816-275a92d5c5d2 · outbound
Video models are zero-shot learners and reasoners Diffusion classifiers understand compositionality, but conditions apply.arXiv preprint arXiv:2505.17955, 2
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cddfd6bc-0e33-40b3-91b4-c2c9ba743516 · outbound
Video models are zero-shot learners and reasoners Force prompting: Video generation models can learn and gen- eralize physics-based control signals
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f114cad5-3a23-437f-b7fc-9982b8b31f6e · outbound
Video models are zero-shot learners and reasoners Motion Prompting: Controlling Video Generation with Motion Trajectories
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7856f664-b4a5-4927-b09e-3cfdc3488dd3 · outbound
Video models are zero-shot learners and reasoners target image
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7ba182f8-0dc4-48e4-af9f-9607b4e24084 · outbound
Video models are zero-shot learners and reasoners Focus on the object {color}
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b8e22924-356d-4135-82a2-70431328e59a · outbound
Video models are zero-shot learners and reasoners For example, if the target image shows a dog, and the choices show a cat, the object types are considered different
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 70fc672a-27aa-4c97-86d8-861cd9af2f00 · outbound
Video models are zero-shot learners and reasoners Final Answer: [answer]
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 00d0a81b-6dc7-4915-9968-81d4395ce2ad · inbound
Are Video Models Emerging as Zero-Shot Learners and Reasoners in Medical Imaging? Video models are zero-shot learners and reasoners
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5deacf6f-7f5c-4b23-8a73-47a005454c83 · inbound
Epipolar Geometry Improves Video Generation Models Video models are zero-shot learners and reasoners
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9980435e-eb6a-4904-aba3-28082a0df574 · inbound
Thinking with Video: Video Generation as a Promising Multimodal Reasoning Paradigm Video models are zero-shot learners and reasoners
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6e03a2a8-bdb5-47e1-b374-98b9b0fffd5d · inbound
Target-Bench: Can Video World Models Achieve Mapless Path Planning with Semantic Targets? Video models are zero-shot learners and reasoners
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bb196b9e-e861-4f40-abcd-3af5ff1f6414 · inbound
PhysChoreo: Physics-Controllable Video Generation with Part-Aware Semantic Grounding Video models are zero-shot learners and reasoners
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a76490f6-51a5-4fbc-a34f-6206a431792d · inbound
Generative Action Tell-Tales: Assessing Human Motion in Synthesized Videos Video models are zero-shot learners and reasoners
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0aeee703-2ac3-4b4b-beac-aedc66bd6505 · inbound
Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation Video models are zero-shot learners and reasoners
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 297246b8-0278-49a7-862c-6a00eee4c807 · inbound
VideoCoF: Unified Video Editing with Temporal Reasoner Video models are zero-shot learners and reasoners
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ebe77d37-8928-48fa-8bc9-0858cb0472b3 · inbound
WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling Video models are zero-shot learners and reasoners
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c9dc0b19-4e8f-4c0e-a3cb-92c671ce8b7e · inbound
WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling Video models are zero-shot learners and reasoners
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e49a4b71-0c03-4438-a892-7a4551b8a785 · inbound
mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs Video models are zero-shot learners and reasoners
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f80e95d7-ac90-4e57-ba48-0adcc831f0e5 · inbound
End-to-End Training for Autoregressive Video Diffusion via Self-Resampling Video models are zero-shot learners and reasoners
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c56be35-e66c-4d41-bbbf-538729048392 · inbound
Kling-Omni Technical Report Video models are zero-shot learners and reasoners
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation eb08817f-69c5-4bf1-a53c-249ee7977af4 · inbound
DriveLaW:Unifying Planning and Video Generation in a Latent Driving World Video models are zero-shot learners and reasoners
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c412de83-6508-4305-90d1-588dfac72819 · inbound
Rewriting Video: Text-Driven Reauthoring of Video Footage Video models are zero-shot learners and reasoners
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8cff95ed-22f9-420a-a778-817638137559 · inbound
Walk through Paintings: Egocentric World Models from Internet Priors Video models are zero-shot learners and reasoners
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e1b603e-3425-4ea3-a639-ce4b70dcfe05 · inbound
MentisOculi: Revealing the Limits of Reasoning with Mental Imagery Video models are zero-shot learners and reasoners
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42e9d9c3-35f6-4d26-9826-aeb083b53ec1 · inbound
PerpetualWonder: Long-Horizon Action-Conditioned 4D Scene Generation Video models are zero-shot learners and reasoners
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f02c4f4e-950a-42bb-bc05-791687929a38 · inbound
DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos Video models are zero-shot learners and reasoners
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d4271ad1-e621-4140-b318-795f780d4f68 · inbound
Olaf-World: Orienting Latent Actions for Video World Modeling Video models are zero-shot learners and reasoners
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b18cd75-4f0e-42f7-9e8b-5b7d318d266e · inbound
Demystifying Video Reasoning Video models are zero-shot learners and reasoners
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 235c9488-a005-442c-aae0-b67a4bbaab26 · inbound
Demystifying Video Reasoning Video models are zero-shot learners and reasoners
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b78b9f2-52b7-4f9b-81f4-541eb2d1fe70 · inbound
Generative Control as Optimization: Time Unconditional Flow Matching for Adaptive and Robust Robotic Control Video models are zero-shot learners and reasoners
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ba64a69c-d493-454d-91a5-a60a73667133 · inbound
Pretrained Video Models as Differentiable Physics Simulators for Urban Wind Flows Video models are zero-shot learners and reasoners
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a5cc8712-4230-45de-b03c-9803796d2771 · inbound
Stepper: Stepwise Immersive Scene Generation with Multiview Panoramas Video models are zero-shot learners and reasoners
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 46d4b097-e5b1-45cb-9639-013698813b3f · inbound
LivingWorld: Interactive 4D World Generation with Environmental Dynamics Video models are zero-shot learners and reasoners
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 925b46ba-2822-46e1-8030-7a14036a260e · inbound
OpenWorldLib: A Unified Codebase and Definition of Advanced World Models Video models are zero-shot learners and reasoners
Reference 132
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 243eeb01-d7d0-4e85-8212-34c02326e6c8 · inbound
OpenWorldLib: A Unified Codebase and Definition of Advanced World Models Video models are zero-shot learners and reasoners
Reference 132
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e17e2b26-c9dd-43f8-9081-405113acdb3f · inbound
Neural Computers Video models are zero-shot learners and reasoners
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1cf11679-e11a-4b4b-a1b7-60848a6e8e6b · inbound
VAG: Dual-Stream Video-Action Generation for Embodied Data Synthesis Video models are zero-shot learners and reasoners
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 984ba074-04f9-4fe4-b1f8-ff33c0703bb3 · inbound
Rays as Pixels: Learning A Joint Distribution of Videos and Camera Trajectories Video models are zero-shot learners and reasoners
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 87becae0-5a4f-45a6-85cd-7f1ed152188d · inbound
Rays as Pixels: Learning A Joint Distribution of Videos and Camera Trajectories Video models are zero-shot learners and reasoners
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4aeb11e-a523-4b42-9660-0de421449b27 · inbound
GTASA: Ground Truth Annotations for Spatiotemporal Analysis, Evaluation and Training of Video Models Video models are zero-shot learners and reasoners
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 05dc2edd-95c2-4580-b303-04baf8f8cb4e · inbound
LMMs Meet Object-Centric Vision: Understanding, Segmentation, Editing and Generation Video models are zero-shot learners and reasoners
Reference 185
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ac169a5b-0bba-4a6a-b831-076d4fa5b746 · inbound
VibeFlow: Versatile Video Chroma-Lux Editing through Self-Supervised Learning Video models are zero-shot learners and reasoners
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1b480e70-761c-44a2-9e30-df8eb095ca8e · inbound
Motif-Video 2B: Technical Report Video models are zero-shot learners and reasoners
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 735057b6-9c5a-4d3b-b333-b96c82f07fab · inbound
Motif-Video 2B: Technical Report Video models are zero-shot learners and reasoners
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 23898561-f7fc-4679-b269-768414dbec70 · inbound
ViPS: Video-informed Pose Spaces for Auto-Rigged Meshes Video models are zero-shot learners and reasoners
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0c5c945e-32b6-4df0-b193-7bbd5026eb61 · inbound
ViPS: Video-informed Pose Spaces for Auto-Rigged Meshes Video models are zero-shot learners and reasoners
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 32b338c2-0eae-413b-b304-250bb2a6c282 · inbound
Grokking of Diffusion Models: Case Study on Modular Addition Video models are zero-shot learners and reasoners
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 47e19029-4357-4292-be91-2dfa8b99743b · inbound
Memorize When Needed: Decoupled Memory Control for Spatially Consistent Long-Horizon Video Generation Video models are zero-shot learners and reasoners
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4daf5438-bf57-48ba-a6d8-7872d39fc58c · inbound
How Far Are Video Models from True Multimodal Reasoning? Video models are zero-shot learners and reasoners
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3b411e37-1b73-450a-b891-3cdfb1d70d9e · inbound
Image Generators are Generalist Vision Learners Video models are zero-shot learners and reasoners
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 45cefcfd-2ef4-4145-93d7-257f374f74b8 · inbound
Image Generators are Generalist Vision Learners Video models are zero-shot learners and reasoners
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 18e25512-3d4f-42c9-ad82-d92c2eefc363 · inbound
Open-Source Image Editing Models Are Zero-Shot Vision Learners Video models are zero-shot learners and reasoners
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0b894adf-48d8-4de1-b4e2-34205f2d0048 · inbound
Eulerian Motion Guidance: Robust Image Animation via Bidirectional Geometric Consistency Video models are zero-shot learners and reasoners
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cf7d03c2-7f76-4988-a158-239e9d3dcd04 · inbound
Eulerian Motion Guidance: Robust Image Animation via Bidirectional Geometric Consistency Video models are zero-shot learners and reasoners
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 389e8009-68a0-4f98-92c5-77ccaefd2b8c · inbound
Eulerian Motion Guidance: Robust Image Animation via Bidirectional Geometric Consistency Video models are zero-shot learners and reasoners
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 04aa48ef-e24c-40e1-88ea-b60f033b90f2 · inbound
Eulerian Motion Guidance: Robust Image Animation via Bidirectional Geometric Consistency Video models are zero-shot learners and reasoners
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 697e63d6-4fd3-42be-9818-d8fab8f63292 · inbound
CollabVR: Collaborative Video Reasoning with Vision-Language and Video Generation Models Video models are zero-shot learners and reasoners
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5be5ac0f-af33-4fd4-8469-02e0762a63c6 · inbound
Perception Without Engagement: Dissecting the Causal Discovery Deficit in LMMs Video models are zero-shot learners and reasoners
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 32b52f5c-3574-4a18-aa09-c33c549bf975 · inbound
Do multimodal models imagine electric sheep? Video models are zero-shot learners and reasoners
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 145f546c-a99d-4267-aeb7-7ef1b2de59ae · inbound
Progressive Photorealistic Simplification Video models are zero-shot learners and reasoners
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 425ab599-d4ba-4dc6-8df2-a395cc433f46 · inbound
WorldReasonBench: Human-Aligned Stress Testing of Video Generators as Future World-State Predictors Video models are zero-shot learners and reasoners
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 99de3e62-1ccb-4dcf-81dd-0fa80f3c5d32 · inbound
Video Models Can Reason with Verifiable Rewards Video models are zero-shot learners and reasoners
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5d5396a0-1d46-41f5-8166-401dbc58a4e7 · inbound
Soap2Soap: Long Cinematic Video Remaking via Multi-Agent Collaboration Video models are zero-shot learners and reasoners
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 030c2cfc-0d4b-4ecf-acc3-44152f3c6e53 · inbound
RoboFlow4D: A Lightweight Flow World Model Toward Real-Time Flow-Guided Robotic Manipulation Video models are zero-shot learners and reasoners
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bacafae6-d485-41d0-bd49-fa5299c96311 · inbound
GeoFlow: Enforcing Implicit Geometric Consistency in Video Generation Video models are zero-shot learners and reasoners
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 69863998-34f2-4cb3-8c98-2f504caefe43 · inbound
PhyWorld: Physics-Faithful World Model for Video Generation Video models are zero-shot learners and reasoners
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1bffb555-5f83-4d1b-98dd-642440800660 · inbound
Scalable, Energy-Efficient Optical-Neural Architecture for Multiplexed Deepfake Video Detection Video models are zero-shot learners and reasoners
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 29bc56db-8e81-4077-b4cb-7a4792f2284f · inbound
VGenST-Bench: A Benchmark for Spatio-Temporal Reasoning via Active Video Synthesis Video models are zero-shot learners and reasoners
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6c87aaae-e097-4264-b3d9-d151a9ac0836 · inbound
MotiMotion: Motion-Controlled Video Generation with Visual Reasoning Video models are zero-shot learners and reasoners
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e8e8d0ad-deb8-487e-8684-fefbe74f49a3 · inbound
Are Video Models Zero-Shot Learners and Reasoners in Education? EduVideoBench, A Knowledge-Skills-Attitude Benchmark for Educational Video Generation Video models are zero-shot learners and reasoners
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 74fa3fa4-a446-41b2-b1b6-4b3f4c61aac9 · inbound
Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players Video models are zero-shot learners and reasoners
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 76e35556-fa49-42d9-8c92-057d85e24d5d · inbound
StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement Video models are zero-shot learners and reasoners
Reference 110
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1dc5ba28-b9fc-4ef2-bf5e-601f76404a6c · inbound
OptiWorld: Optimal Control for Video World Generation under Physical Constraints Video models are zero-shot learners and reasoners
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1b146f1f-0cb9-4312-8b2c-a0c5c6d7bdc2 · inbound
AlbedoEdit: Unified Instance-Level Video Editing with Albedo Guidance Video models are zero-shot learners and reasoners
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 871081c3-33ae-4a0e-ba68-9dc2fedbbb32 · inbound
VLMs are Good Teachers for Video Reasoning via Adaptive Test-Time Optimization Video models are zero-shot learners and reasoners
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 997a36fb-569a-42a8-b392-62f02be18087 · inbound
VLMs are Good Teachers for Video Reasoning via Adaptive Test-Time Optimization Video models are zero-shot learners and reasoners
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0ed2161f-86dc-48bf-80ec-fcf968de0120 · inbound
Cosmos 3: Omnimodal World Models for Physical AI Video models are zero-shot learners and reasoners
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 52cef885-4b75-4b3c-8992-f7eeac9d89a1 · inbound
PointAction: 3D Points as Universal Action Representations for Robot Control Video models are zero-shot learners and reasoners
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 59b46e92-a43b-454f-8eab-708b51a08e82 · inbound
OmniTryOn: Video Try-On Anything at Once! Video models are zero-shot learners and reasoners
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 851a4ec4-c9d4-4a56-922d-7e0f3139c5a7 · inbound
Data-Driven Automation Video models are zero-shot learners and reasoners
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6ce2e47a-a03b-4ff9-835b-5cad02df67d7 · inbound
Quo Vadis, Visual In-Context Learning? A Unified Benchmark Across Domains and Tasks Video models are zero-shot learners and reasoners
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2b1b0719-1c48-459f-b10e-4a5201d8595e · inbound
WorldOlympiad: Can Your World Model Survive a Triathlon? Video models are zero-shot learners and reasoners
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6aad9cc6-d226-435e-90f7-3bc87b10614d · inbound
VICX: Generalizable Robot Manipulation via Video Generation and In-Context Operator Network Video models are zero-shot learners and reasoners
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3cf84998-9a19-450d-a02b-c440560e3faf · inbound
World Model Self-Distillation: Training World Models to Solve General Tasks Video models are zero-shot learners and reasoners
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2765f2ef-669a-46b6-8033-9fd7646e9c92 · inbound
WEAVER, Better, Faster, Longer: An Effective World Model for Robotic Manipulation Video models are zero-shot learners and reasoners
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 900e083e-6dc4-4831-9237-075a15147f80 · inbound
NEXUS: Neural Energy Fields for Physically Consistent Contact-Rich 3D Object Dynamics Video models are zero-shot learners and reasoners
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 24a22b57-39c7-4bfd-bb47-1c06a3136b9f · inbound
Physics-IQ Verified Video models are zero-shot learners and reasoners
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dfa7dfc3-070c-45e7-bb53-d6ad772fdd4d · inbound
FLAT: Feedforward Latent Triangle Splatting for Geometrically Accurate Scene Generation Video models are zero-shot learners and reasoners
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0b66c0aa-f0f9-44b7-a1ef-ad9f4fd0a122 · inbound
PhysRAG: Enhancing Physics-Awareness in Video Generation via Retrieval-Augmented Generation Video models are zero-shot learners and reasoners
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5759d396-7558-45f8-b65a-d45dedb64175 · inbound
A Good Talk Does not Look Like a Summary, It Teaches You! Measuring Takeaways from Paper-to-Video Talks Video models are zero-shot learners and reasoners
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 981ed76b-0c7d-4df0-a74e-595e213c2758 · inbound
Bridging Video Understanding and Generation in a Unified Framework Video models are zero-shot learners and reasoners
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e8ddcdf1-dfc9-41a5-b97d-2a8bbc1f8a1d · inbound
Global Pose Control for Generative View Synthesis in Normalized Object Coordinate Space Video models are zero-shot learners and reasoners
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0db48952-a698-4bfa-939b-d25a5512eb4c · inbound
Self-Improving Diffusion Classifiers with Minority Preference Optimization Video models are zero-shot learners and reasoners
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8277d48-383d-42a7-9771-029834d3907c · inbound
Video Generation Models Are Inherent Lighting Estimators Video models are zero-shot learners and reasoners
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c660030-abd4-41db-ab64-750e25cad946 · inbound
Deform360: A Massive Multi-view Visuotactile Dataset for Deformable World Models Video models are zero-shot learners and reasoners
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa4ea3b1-c690-484e-a86f-6a14c1f70746 · inbound
Gen4U: Unifying Video Generation and Understanding via Diffusion Video models are zero-shot learners and reasoners
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8657b9ef-67e8-4500-a5f8-38582f8953d2 · inbound
OpenCoF: Learning to Reason Through Video Generation Video models are zero-shot learners and reasoners
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8e37c9d8-47b1-4b2e-aff0-3e4f726cd7f5 · inbound
Video Generation Models are General-Purpose Vision Learners Video models are zero-shot learners and reasoners
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb1be0d7-326c-4df3-92ca-1544e8a006fb · inbound
From Pixels to States: Rethinking Interactive World Models as Game Engines Video models are zero-shot learners and reasoners
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c482106-cb9a-469d-9c97-c392dce55833 · inbound
Privacy-Aware Synthetic Video Benchmarking and Relational Evaluation for Worker-Under-Suspended-Load Detection Video models are zero-shot learners and reasoners
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6792c75c-d84d-4134-b62d-4259ad71952f · inbound
Apple-$\pi$: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence Video models are zero-shot learners and reasoners
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3250ba63-0662-4c3d-a75d-c3d83c42d811 · inbound
Between Safe Boundaries: Exploiting Temporal Consistency for Jailbreaking Text-To-Video Generation Models Video models are zero-shot learners and reasoners
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 127f1619-a1a4-436b-bd5c-7b1487b9bdc1 · inbound
Diffusion ReRoll: Revisable Denoising for Robotic Sequential Prediction Video models are zero-shot learners and reasoners
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9846cd65-116c-4b57-9aa4-0b11093d535f · inbound
FilmBench: A Film-Grade Benchmark for Cinematic Video Generation Video models are zero-shot learners and reasoners
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0654ae9d-0209-4ca9-b583-f60db8edc336 · inbound
Visual prompt engineering for video models Video models are zero-shot learners and reasoners
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bf596d3-03c3-4ea5-b621-d9e868882fec · inbound
Ripple: Real-Time Streaming Audio-Video Generation With Cross-Modal Recurrent Memory Video models are zero-shot learners and reasoners
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06ddc16a-919d-4051-9d31-01dab965ebab · inbound
World Action Planner: Generalizable Decision-Making with Action-Conditioned World Models Video models are zero-shot learners and reasoners
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.