Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:12:37.935546Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 71 of 71 outbound references and 2 inbound Pith citation observations for arXiv:2505.13851.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:12:37.935546Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T16:07:49.790348Z
A source-named dated measurement, never combined with another source.
Source: cited_works
71 of 71 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9735a573-7a8c-4871-9572-e3aec62cc635 · outbound
A Challenge to Build Neuro-Symbolic Video Agents In 2019 IEEE 58th conference on decision and control (CDC), pages 5338--5343
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6831c3f7-23a4-4c98-b754-46db89b2556a · outbound
A Challenge to Build Neuro-Symbolic Video Agents Principles of Model Checking
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation fa7957a3-740f-450a-a5d2-5c3cfd51d01f · outbound
A Challenge to Build Neuro-Symbolic Video Agents Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 53a35f3f-d6a7-4351-ac3d-15ad2981c3a0 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Align your latents: High-resolution video synthesis with latent diffusion models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation bc7292ec-8497-4bd6-8fb1-21fc3a270c1c · outbound
A Challenge to Build Neuro-Symbolic Video Agents Hourvideo: 1-hour video-language understanding
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e138589a-bdaa-4835-8546-1350f997621a · outbound
A Challenge to Build Neuro-Symbolic Video Agents Langchain, 2022
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8ff190ce-6749-488c-b5f4-47a00bcbc4f0 · outbound
A Challenge to Build Neuro-Symbolic Video Agents ComPhy: Compositional Physical Reasoning of Objects and Events from Videos
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0f2c29c-4a05-499b-90b5-46ffa52a3be2 · outbound
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7567b5ad-4d7e-4c92-b450-9a7cf46da630 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Sora as an agi world model? a complete survey on text-to-video generation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41cd61bd-5fbe-40b1-86d6-ef7acece1221 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Towards neuro-symbolic video understanding
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b92ae639-df77-4e2b-b643-75e8bf5ceb04 · outbound
A Challenge to Build Neuro-Symbolic Video Agents We'll Fix it in Post: Improving Text-to-Video Generation with Neuro-Symbolic Feedback
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e81aebdd-a2e9-4d82-97ae-ecd6cf9e27ee · outbound
A Challenge to Build Neuro-Symbolic Video Agents Real-Time Privacy Preservation for Robot Visual Perception
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc4d308a-aecf-4269-8762-95f95dcfa416 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Parks research shows consumers are after integrated smart locks and security cameras
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a497cb8d-e782-4998-ba07-9f13565de877 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Towards interpretable video anomaly detection
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation bc736d57-a10b-4d3f-8286-96d77fc033e0 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Structure and content-guided video synthesis with diffusion models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b4e23337-d344-4b8a-ad19-e4a2875b2b69 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Convolutional two-stream network fusion for video action recognition
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f1ccf334-3cc6-4ed8-8712-5a9590d99c4e · outbound
A Challenge to Build Neuro-Symbolic Video Agents Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 60f6e5dd-3ebf-49d7-9ad1-57a97330a47e · outbound
A Challenge to Build Neuro-Symbolic Video Agents Slowfast networks for video recognition
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6948090a-3e78-45bc-b4ce-68050aead013 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Ego4d: Around the world in 3,000 hours of egocentric video
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e33e310-d468-4f3d-9a4f-e9a68882dcd4 · outbound
A Challenge to Build Neuro-Symbolic Video Agents The efficacy of neural planning metrics: A meta-analysis of pkl on nuscenes
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation eaafed71-be0e-4534-99f8-7f53631625ac · outbound
A Challenge to Build Neuro-Symbolic Video Agents Stabletoolbench: Towards stable large-scale benchmarking on tool learning of large language models, 2024
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1b4c00d-c025-4014-927a-7ba6fb8797a5 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Denoising diffusion probabilistic models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aabb3fd2-ea12-48ac-a8e0-b8e93c839fd8 · outbound
A Challenge to Build Neuro-Symbolic Video Agents CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24809207-4feb-476d-ac2f-75e755ac1b33 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Vbench: Comprehensive benchmark suite for video generative models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbc2d06c-2170-4dbd-991e-607de1ab895a · outbound
A Challenge to Build Neuro-Symbolic Video Agents Adaptive Video Understanding Agent: Enhancing efficiency with dynamic frame sampling and feedback-driven reasoning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dce12e36-8334-4b06-a80c-f6469c6690be · outbound
A Challenge to Build Neuro-Symbolic Video Agents Safe autonomy under perception uncertainty using chance-constrained temporal logic
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 29778a5f-0ac2-4fef-b8a9-12399fdaef42 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Tsaftaris, and Aggelos K
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b337289-ff70-4f8c-b2f5-2096e55c4c3b · outbound
A Challenge to Build Neuro-Symbolic Video Agents Temporal-logic-based reactive mission and motion planning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8097b98c-0bf9-4b2a-b7be-5b3080cf7bcf · outbound
A Challenge to Build Neuro-Symbolic Video Agents A neural-symbolic approach to computer vision
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation bed02405-314a-4689-b528-ff84e779a6d9 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Pika ai: Free video generator with scene ingredients, 2024
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 738b471e-f590-44a0-90d8-0a6cb0d28167 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Human-related anomalous event detection via spatial-temporal graph convolutional autoencoder with embedded long short-term memory network
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2335cd9b-cab6-4e81-86b0-eae069542e70 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a59c1b95-db6d-49fd-b702-dba9feb0bc97 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Representation learning on visual-symbolic graphs for video understanding
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3423089d-d6fc-48d0-a420-25ebbf955fa8 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Medioni, Isaac Cohen, Fran c ois Br \' e mond, Somboon Hongeng, and Ramakant Nevatia
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0bf424f4-1809-4da7-b4f2-bfce4cccd102 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Formal methods to comply with rules of the road in autonomous driving: State of the art and grand challenges
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4f7d6e52-cc40-4282-be7b-4d799373607b · outbound
A Challenge to Build Neuro-Symbolic Video Agents 8 must-follow social listening trends for 2025, January 2025
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8d32f684-f89d-43a9-ac55-5f4ff25a60f7 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Learning audio-video modalities from image captions
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 00895aae-9a46-4c63-bb57-0e619cff48cf · outbound
A Challenge to Build Neuro-Symbolic Video Agents Federal motor vehicle safety standards; automatic emergency braking systems for light vehicles
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 12bb363c-e1fd-4cde-997f-bebe1db94f8e · outbound
A Challenge to Build Neuro-Symbolic Video Agents GPT-4 Technical Report
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cda27a41-ce49-4ecb-bfe6-5a9dd190e11f · outbound
A Challenge to Build Neuro-Symbolic Video Agents Video generation models as world simulators, 2024
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6e63a228-7da0-4400-b57e-6d8210ccbae3 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Gorilla: Large Language Model Connected with Massive APIs
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 733bf8e6-fa5a-423a-a37e-8a52d352a409 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Toolllm: Facilitating large language models to master 16000+ real-world apis, 2023
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59ab042d-7e39-4940-ae2b-4dbc2a09597a · outbound
A Challenge to Build Neuro-Symbolic Video Agents Rapidapi hub
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 15021405-cc0f-485b-805f-2e4339d76eff · outbound
A Challenge to Build Neuro-Symbolic Video Agents Introducing gen-3 alpha: A new frontier for video generation, 2024
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9581fbea-01b9-458a-ad3e-b2384b5c4265 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Early detection of combustion instability by neural-symbolic analysis on hi-speed video
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 03be9951-6b02-4ba4-b82c-30f37c80e96b · outbound
A Challenge to Build Neuro-Symbolic Video Agents Neuro-Symbolic Evaluation of Text-to-Video Models using Formal Verification
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d91204f2-d990-4edd-816d-cc5a803fe063 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Linear temporal logic motion planning for teams of underactuated robots using satisfiability modulo convex programming
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4f4b3b50-f84e-40d4-a588-6f95915a78aa · outbound
A Challenge to Build Neuro-Symbolic Video Agents Shultz and Richard D
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e423d689-7017-4bce-a6ee-1baa6651bd2a · outbound
A Challenge to Build Neuro-Symbolic Video Agents VideoAgent: Self-Improving Video Generation
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 112858a6-bb3b-4076-8e75-be239a51a244 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Scalability in perception for autonomous driving: Waymo open dataset
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 358a5ee7-c886-4bf2-a34c-7676c203f0ba · outbound
A Challenge to Build Neuro-Symbolic Video Agents Gemini: A Family of Highly Capable Multimodal Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 907d582c-7805-4444-8023-bc218adb79cd · outbound
A Challenge to Build Neuro-Symbolic Video Agents Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d9b4d22-5926-4e9f-9397-a2d0c1a00ac1 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Video classification with channel-separated convolutional networks
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation cb95453b-dea5-4942-b1d3-40dfeabc63af · outbound
A Challenge to Build Neuro-Symbolic Video Agents Twilio: Cloud communications platform, 2025
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3e8ddd96-625e-4479-9c75-e9d041c65650 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Phenaki: Variable Length Video Generation From Open Domain Textual Description
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 514fada2-65fe-4e86-934b-60dffde979c4 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Lave: Llm-powered agent assistance and language augmentation for video editing
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e5fffafe-c83f-4919-ad93-5698cd2fc9b0 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Videoagent: Long-form video understanding with large language model as agent
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a314f256-c948-471a-81b8-34c797c3002d · outbound
A Challenge to Build Neuro-Symbolic Video Agents InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0dd12723-69b4-407d-8032-fcf61d90973d · outbound
A Challenge to Build Neuro-Symbolic Video Agents Discovqa: Temporal distortion-content transformers for video quality assessment
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f1711f3a-f577-461d-853a-92bae6e38ed1 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Exploring video quality assessment on user generated contents from aesthetic and technical perspectives
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 94821730-d236-48d2-b876-4f1f0e4b04b3 · outbound
A Challenge to Build Neuro-Symbolic Video Agents A graph-based framework to bridge movies and synopses
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f6951e0d-5ee5-407f-b184-b3c3ac23eb14 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Msr-vtt: A large video description dataset for bridging video and language
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aac87d70-054d-4842-839b-9054d1c40561 · outbound
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 75b9bfce-61d0-4dde-86ec-cac5cf08f154 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Specification-Driven Video Search via Foundation Models and Formal Verification
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 606c09b5-9861-4122-9d55-163c5dfabbf2 · outbound
A Challenge to Build Neuro-Symbolic Video Agents CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee854281-90a0-4b39-964b-60e5e8eef2b3 · outbound
A Challenge to Build Neuro-Symbolic Video Agents React: Synergizing reasoning and acting in language models
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e208524f-70b6-4513-b00a-b63f287e9a97 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Neural-symbolic VQA: disentangling reasoning from vision and language understanding
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 30c2dae2-e768-49f9-87b2-c7c9314d5eab · outbound
A Challenge to Build Neuro-Symbolic Video Agents A probabilistic graphical model based on neural-symbolic reasoning for visual relationship detection
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ee7345f8-4105-46e6-a41e-11b703efe8b8 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 787a5d85-77ac-482f-b443-ed0e42b9169a · outbound
A Challenge to Build Neuro-Symbolic Video Agents I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 276b431e-0376-4591-973d-015d548758e9 · outbound
A Challenge to Build Neuro-Symbolic Video Agents Abnormal event detection by a weakly supervised temporal attention network
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 867abc38-38c6-4c4c-9dcd-95d244eecd0f · inbound
Incentivizing Vision Language Models to Search for Long Video Question Answering A Challenge to Build Neuro-Symbolic Video Agents
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6c0e8fb-65bf-418e-bb05-2d0f6145cbbc · inbound
SGA: Plug&Play Geometric Verification for Educational Video Synthesis A Challenge to Build Neuro-Symbolic Video Agents
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.