Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:46:46.420145Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 100 of 111 outbound references and 4 inbound Pith citation observations for arXiv:2505.12098.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:46:46.420145Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:16:47.490821Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T13:56:59.272327Z
100 of 111 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 61a35551-7f2a-4690-942f-dad9a23d3306 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation A survey of ai text-to-image and ai text-to-video generators,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b8f57f8-2501-4f08-9e6e-c24c02ac98f9 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation A survey on video diffusion models,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e44c3a8-09d5-46dd-9c68-fe981fa4d090 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Evaluation of text-to-video generation models: A dynamics perspective,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15dd866e-086f-4f47-83ab-5347401615be · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation InternLM-XComposer: A Vision-Language Large Model for Advanced Text-image Comprehension and Composition
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15736eb3-5dd8-4477-91e5-21bddb86a63b · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation mplug-owl3: Towards long image-sequence understanding in multi-modal large language models,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60c06238-d86d-4ead-a420-149f21fbe5ea · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca2b3726-1037-499f-9fc2-c8bb07261c1a · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Aigv-assessor: Benchmarking and evaluating the perceptual quality of text-to-video generation with lmm,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99c91dc4-bbf9-4826-bda9-d1e73bff9f72 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Evaluating and improving compositional text-to-visual generation,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58aecaa8-fdb7-4a39-826b-71ceb23667e1 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation VBench: Comprehensive benchmark suite for video generative models,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5ad89dd-09f6-486f-ae60-2f0bf0f1592f · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9793d22f-e635-431c-aa52-446efa8aa278 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Eval- crafter: Benchmarking and evaluating large video generation models,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a74f312d-709c-426c-86ff-64f71f99afc7 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Fetv: A benchmark for fine- grained evaluation of open-domain text-to-video generation,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efe373d1-982b-4f13-be01-6203fd93c494 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Q-Eval-100K: Evaluating Visual Quality and Alignment Level for Text-to-Vision Content
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ca14115-61a2-4657-a9ba-fb77ed92e633 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Measuring the Quality of Text-to-Video Model Outputs: Metrics and Dataset
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47c68798-8b00-467d-a50b-6009b943bfeb · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Benchmarking Multi-dimensional AIGC Video Quality Assessment: A Dataset and Unified Model
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c867d89c-cfd2-446f-b1b3-820443e436ff · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Subjective-aligned dataset and metric for text-to-video quality assessment,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5dddfb8-e313-44b3-80c5-4a77023e37df · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Gaia: Rethinking action quality assessment for ai-generated videos,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c02f201-e318-4cea-95e2-fbbe8a392221 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation LMME3DHF: Benchmarking and Evaluating Multimodal 3D Human Face Generation with LMMs
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9be9ceb2-d771-4ee6-8171-43822aac8f3a · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Quality assessment of in-the-wild videos,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c41296fd-8525-42d1-ae45-26c8bb5b14cf · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation A deep learning based no-reference quality assessment model for ugc videos,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d7c3b6b-9ed3-4db4-af6b-4d82902141dc · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Fast-vqa: Efficient end-to- end video quality assessment with fragment sampling,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 491b2055-99c1-4323-b4df-01a08c10689f · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Exploring video quality assessment on user generated contents from aesthetic and technical perspectives,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 705c6c3a-6ce4-4d84-9a0d-d502e1533332 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Geneval: An object-focused framework for evaluating text-to- image alignment,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fefb36c9-4029-4bff-8b82-3b649d8dc024 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Methodology for the subjective assessment of the quality of television pictures,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 642c64cc-0541-436c-b14f-fde4909a6721 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Improved training of wasserstein gans,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb565e60-b0e0-4449-8dbb-dc2d3829b8d4 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Towards Accurate Generative Models of Video: A New Metric & Challenges
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f85bd524-f68f-455c-84d2-dc9821871fbc · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Visual instruction tuning,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e42132e7-9319-4fac-a383-c50f264635ce · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Making a “completely blind
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6482f19-a5df-4152-8113-e86a493021df · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Aigcoiqa2024: Perceptual quality assessment of ai generated omnidirectional images,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ac8e209-4e15-4fe0-8852-1362bf47cf05 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Finevq: Fine- grained user generated content video quality assessment,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fe07f86-cec6-48dc-920f-13b7e973806b · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Quality Assessment for AI Generated Images with Instruction Tuning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efec826d-5ebc-450d-b188-a0d626f9f1cc · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation HarmonyIQA: Pioneering Benchmark and Model for Image Harmonization Quality Assessment
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74bee7d7-868f-4ce4-80fb-9050a8314d73 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Blindly assess quality of in-the-wild videos via quality-aware pre-training and motion perception,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5c48b8f-e08e-414f-994a-5413c6cf487c · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Pick-a-pic: An open dataset of user preferences for text-to-image generation,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 558d7112-d39b-42ab-b3ca-99b2ec0977f7 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Learning without human scores for blind image quality assessment,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94afa634-d82f-41bf-b106-4bb9f72c6dcc · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation No-reference image quality assessment in the spatial domain,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a64aec0-a3db-44a4-b562-8a18a208db06 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6c507e3-dc62-4164-a420-9c01f9638d0d · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Jimeng ai
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d753434-345b-4f4f-9da8-3d7d708f9c0c · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Vidu ai
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d64f5add-52ef-4c8e-b85d-d840da771d9f · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Lora: Low-rank adaptation of large language models,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50b0051f-8ca3-4046-a3a8-b833e1ecee52 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2945d9b-871a-4074-8e6d-e0bf0d337c35 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Swin transformer: Hierarchical vision transformer using shifted windows,
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 456be38c-fd37-4916-8909-1122959912a0 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94b9038b-a310-496d-8777-1b5c45d280a9 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Blind image quality estimation via distortion aggravation,
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27a26964-fa7c-4fb5-869c-05b65e8847a7 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Blind quality assessment based on pseudo- reference image,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d1f400b7-5a34-49de-aaec-d07c14a38e67 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Blind image quality assessment based on high order statistics aggregation,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 13363b51-a1d3-46b6-acf4-f736835c2e97 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation CLIPScore: A Reference-free Evaluation Metric for Image Captioning
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d046d088-d194-4c38-aa56-1731522c3a33 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Blip: Bootstrapping language-image pre-training for unified vision- language understanding and generation,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f31a70b2-0af7-48ff-995b-babaaf0dd569 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Laion-5b: An open large-scale dataset for training next generation image-text models,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4c3d2a88-1a51-4370-afd6-d445633d07e2 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Imagereward: Learning and evaluating human preferences for text-to-image generation,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0a09fa6c-6547-4cdd-94fb-06c22389c9b4 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Human preference score: Better aligning text-to-image models with human preference,
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6c5980de-0401-4a87-bc22-9002791225ac · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation EvalMuse-40K: A Reliable and Fine-Grained Benchmark with Comprehensive Human Annotations for Text-to-Image Generation Model Evaluation
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05ecb803-3790-4fc2-8c0d-8daf8f1fbfef · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d4bc4fc-0be7-4b65-a30a-ecc2c620d7b3 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c114032-4fad-4103-8b82-879326445598 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c1e9db0-baf2-4b5d-876e-623b9dd10707 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Qwen2.5-VL Technical Report
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb615db2-9e93-4f70-85ec-ce592d6f28c0 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Llama 3.2: Revolutionizing edge ai and vision with open, customizable models,
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 970e3531-6178-436b-9a17-85be54aa8254 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation CogAgent: A Visual Language Model for GUI Agents
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efe47176-b4ff-43eb-87c1-804b72632c54 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e651e428-6d02-48fa-a6a5-4c574a25ea3e · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation LLaVA-OneVision: Easy Visual Task Transfer
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff836913-1b37-4ce2-aa6f-0e068e386471 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Gemini1.5-pro
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d482f168-e05a-4ca4-aeca-9b013a45b568 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Claude3.5
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9144f4d9-4333-4048-91b2-af05fb14161e · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Grok2 vision
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9243cd1b-c85a-4e4f-a461-f9e47bb5fae0 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Chatgpt-4o
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation df98fcd3-02db-4a78-86a4-45cbea9d647f · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Pixverse: Ai video creation platform
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 553f17d8-6a13-4630-b229-d8620dc11c56 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Wanxiang
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ba761b63-d2da-4fe2-bde6-7130fe9a28f7 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Hailuo ai
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 512faa2f-f5b9-477a-8f56-0af0f4310f3b · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Team, “Sora.”https://openai.com/research/video-generation-models-as-world-simulators ,
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e11c36e8-7e6e-4681-8621-271b32ffcb65 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e37b01b3-705c-4b39-8c5c-4b5570bfb9dd · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Introducing gen-3 alpha: A new frontier for video generation
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d3e6c0fd-1c2c-4398-a063-509f1150adbe · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Kling ai
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ece6ebe9-88e5-4fa7-b2cf-ab5b365c8dbd · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Team, “Gemo.” https://www.genmo.ai, 2024
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1827f4e8-bbc8-4fa2-9245-2eccf06fba5d · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df667d26-fcdb-4311-96eb-eb9b0164927f · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Team, “Xunfei.” https://typemovie.art/, 2024
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 11fcc47c-0da7-4c0d-b2fc-199df9ef65b3 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Pyramidal flow matching for efficient video generative modeling,
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b19dd4e6-a8d4-42f3-a353-c4f399307d56 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Wan: Open and Advanced Large-Scale Video Generative Models
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d787cc19-6e4b-48af-8967-89d90ccde02e · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Allegro: Open the Black Box of Commercial-Level Video Generation Model
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cb38cda-cf41-44bc-b751-6bf9f3ae61a2 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08df4b1d-f2ad-402c-8e47-dd0efceea38b · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 172303cf-eef3-4237-acd8-a9e7b98addaf · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Easyanimate: A high-performance long video generation method based on transformer architecture,
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c15cdad3-2c9c-4dbf-9b58-b6c45adcc72b · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ba89d31-8912-4380-a8e3-fe8a242c480d · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Hotshot-XL
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 64d0ee34-a58c-4238-9783-39a4d8531532 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Latte: Latent diffusion transformer for video generation,
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 69fd7429-2994-43d3-842e-b96fd21aec47 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a4fc38e-3dd6-4096-acf5-f0942e7a6183 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Text2Video-Zero: Text-to-Image Diffusion Models are Zero-Shot Video Generators
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87c5e10e-66cb-4f8a-8f30-d3e460b0ffb3 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Autoregressive Video Generation without Vector Quantization
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b662b758-6f11-440e-947c-bb0e0e3ccbff · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation ModelScope Text-to-Video Technical Report
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf06afa6-8297-4d57-9058-49757e28e44d · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Tune-a- video: One-shot tuning of image diffusion models for text-to-video generation,
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ba8d69a9-4038-4117-a7cb-d31f50b5300a · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation LTX-Video: Realtime Video Latent Diffusion
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 176ff004-7ab7-4b43-bcd6-2f60616b7757 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Latent Video Diffusion Models for High-Fidelity Long Video Generation
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 296fd56a-f4ba-4111-a16c-f4911da5b5f2 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Zeroscope
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 21f0a260-cfb8-41ff-9399-a8c1eee5c20c · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation World Model on Million-Length Video And Language With Blockwise RingAttention
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a5f7f94-2b68-479b-971f-84ac4b686245 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Responsible research with crowds: pay crowdworkers at least minimum wage,
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f7c16bbd-9349-474b-8510-06ffa80dae80 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Toward verifiable and reproducible human evaluation for text-to-image generation,
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation af42f055-f406-45d5-b940-ce19f78f00ea · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Internvid: A large-scale video-text dataset for multimodal understanding and generation,
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d3ff9fb9-4325-4710-9bb2-44708319fb72 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Msr-vtt: A large video description dataset for bridging video and language,
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ed33fdc9-709d-48e4-aa0f-03205f982a49 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Frozen in time: A joint video and image encoder for end-to-end retrieval,
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3cc1b1da-32d8-4335-bbf1-b9f1b8fc9346 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Tgif: A new dataset and benchmark on animated gif description,
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 96c01551-6739-4012-8700-4e3227a3def4 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation T2i-compbench: A comprehensive benchmark for open- world compositional text-to-image generation,
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6b1a57a9-e7b9-4f94-9a8e-eff47ad23972 · outbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation LMM4LMM: Benchmarking and Evaluating Large-multimodal Image Generation with LMMs
Reference 101
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea652983-a7d5-42ec-beb5-76a4e9f08b94 · inbound
DFBench: Benchmarking Deepfake Image Detection Capability of Large Multimodal Models LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f03c7216-4506-48bb-a9b4-1fe2e7189b63 · inbound
PhyGround: Benchmarking Physical Reasoning in Generative World Models LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 71868fec-ce45-41ca-801b-15d7a07607cc · inbound
LongVQUBench: Benchmarking Long-Term Video Quality Understanding of Vision-Language Models LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7de1b9a1-3abf-4470-b2b2-10d662cad8e8 · inbound
Natural Language Camera Movement Understanding LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.