Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:40:44.329318Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:2506.04715.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:40:44.329318Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
58 of 58 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a07b24ca-6e2a-4635-858e-02e5d81925f1 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4d8a437-10dc-4032-a342-f2c0d627b82a · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model PaliGemma: A versatile 3B VLM for transfer
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9ff83c0-0d01-4117-a313-954579a3319b · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Video generation models as world simulators.OpenAI Blog, 1:8, 2024
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 83054b39-1872-43d2-8751-5e7388374fbd · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Gaia: Rethinking action quality assessment for ai-generated videos.Advances in Neural Information Processing Systems, 37:40111–40144, 2024
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation cde9eb70-cf40-4b76-a9b0-568517b381e5 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model FineVQ: Fine-Grained User Generated Content Video Quality Assessment
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70292eb2-ec04-43d3-98fd-bddd339d4244 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Slowfast networks for video recognition
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f107fd90-7332-4310-944a-47692c976595 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model LMM-VQA: Advancing Video Quality Assessment with Large Multimodal Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f27c5f0d-be89-4559-9385-25e53a47fc21 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model The Llama 3 Herd of Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe7a4433-8d72-4cca-8287-d3516d4b07e8 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f15939d0-1231-49d1-9fdd-eb51e5b1d537 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Deep residual learning for image recognition
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29cdfd3c-1b3f-48eb-a91c-d9b440d1da04 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51e2a8d0-1b30-411c-8bc5-695271c9e2cc · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0acd423a-8bc7-45bd-82bb-d9ac9d4f0e62 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Lora: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b8ec7eb7-bf5b-4406-ba70-9c726a0943e9 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model VBench++: Comprehensive and Versatile Benchmark Suite for Video Generative Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60577905-30eb-4f96-9fbe-9621bceebc64 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model T2vbench: Benchmarking temporal dynamics for text-to- video generation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7a2f419a-51e4-492a-b327-16ce279e1381 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model VQA$^2$: Visual Question Answering for Video Quality Assessment
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07b76e28-bad1-4919-8c28-a32924d4c17f · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model MANTIS: Interleaved Multi-Image Instruction Tuning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28be2090-27a2-493a-bb34-c42e933bb7af · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model The Kinetics Human Action Video Dataset
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2479dc1-eedd-4002-b112-57effcd4eaec · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model HunyuanVideo: A Systematic Framework For Large Video Generative Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3546953f-b69f-4dfe-ae73-3a59cb2b9460 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Subjective-aligned dataset and metric for text-to-video qual- ity assessment
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation bf598dba-e315-46f0-9637-a073a898f36c · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 86903391-b6d1-4232-96c8-5b9b5d6993f0 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d2a30562-a971-4681-bb31-80e0f58f61c2 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a2c6da0-da65-4cc0-a606-15fdd29e4e80 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Unmasked teacher: Towards training-efficient video foundation models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 52c7be92-9738-4ca7-9001-f73a2614b322 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Evaluating text-to-visual generation with image-to-text gen- eration
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 50455e1e-506f-456d-af73-905e59297120 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Evalcrafter: Benchmarking and eval- uating large video generation models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a43245b0-3b1d-4c7b-bba0-18ceef46d07d · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Fetv: A bench- mark for fine-grained evaluation of open-domain text-to- video generation.Advances in Neural Information Process- ing Systems, 36, 2024
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6d2ec5a0-aad9-4d02-927a-461ea1eabbd5 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model A convnet for the 2020s
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68e3b078-b8b3-463e-8df2-e46017142904 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Video swin transformer
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fcb6ded-57e1-4926-93ef-bd9d10289684 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Aigc- vqa: A holistic perception metric for aigc video quality assessment
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33b35cc4-aae1-44b0-a66c-a5948ce78f64 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bca5590-1785-4951-950f-004c454deaa0 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Open-Sora 2.0: Training a Commercial-Level Video Generation Model in $200k
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 888614ba-6a5d-42af-a596-16d8e859fe87 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac16df1a-4f5a-4c46-9528-58febc3545f8 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model T2veval: Benchmark dataset and objective evaluation method for t2v-generated videos, 2025
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation de7c3ed7-5a84-453a-ac76-80d65a94535e · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Comprehensive subjective and objective evaluation method for text-generated video.arXiv preprint arXiv:2501.08545, 2025
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63dfae0d-6ccf-4033-98da-c0ba011c92b3 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Learning transferable visual models from natural language supervi- sion
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5f906f3-4809-4d6c-bff4-c63f59be90d3 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Make-A-Video: Text-to-Video Generation without Text-Video Data
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48e0d44d-f7df-49e8-840a-43702c398847 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model A deep learning based no-reference quality assessment model for ugc videos
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 66e8f99e-4d92-4084-908c-0a625085a732 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3401f142-855d-4df7-b9ee-8fa22fea7c2f · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model AIGV-Assessor: Benchmarking and Evaluating the Perceptual Quality of Text-to-Video Generation with LMM
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1add9b73-f19b-4630-bd41-a83353015b22 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Videomae v2: Scaling video masked autoencoders with dual masking
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e871c747-e2c3-4241-9229-d221f6bd513f · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa25a1f6-2f0c-4953-900c-79ed1d994ee9 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model An ensemble approach to short-form video quality assess- ment using multimodal llm
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e23de760-b9c2-480a-878d-6b13cfd6ef92 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model GODIVA: Generating Open-DomaIn Videos from nAtural Descriptions
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb7d9e39-8d5e-4bc2-b803-cd00513ada45 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Fast- vqa: Efficient end-to-end video quality assessment with frag- ment sampling
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation dfa46c75-9ae7-41ab-9d46-a60028823639 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Exploring video quality assessment on user gener- ated contents from aesthetic and technical perspectives
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b5b8cca9-1ef2-4526-bf60-e1b449fd3dea · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef889838-8e80-4eb4-b263-6611978d116c · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Q-instruct: Improving low-level visual abilities for multi-modality foundation models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97ca7cfb-dd89-4015-892b-afcd42b9f595 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Grit: A gener- ative region-to-text transformer for object understanding
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7a8f4aa1-f205-479f-a0f0-c7e1f99203cd · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Ntire 2025 xgc quality assessment challenge: Methods and results.CVPR Workshop, 2025
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f0a91cb3-f3a7-4f3b-b077-ae9e4c66857c · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Imagere- ward: Learning and evaluating human preferences for text- to-image generation.Advances in Neural Information Pro- cessing Systems, 36:15903–15935, 2023
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4ac5347-e102-48d8-8855-c34593cfa4f2 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Imagere- ward: Learning and evaluating human preferences for text- to-image generation.Advances in Neural Information Pro- cessing Systems, 36, 2024
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1cb4c45e-3c98-40cd-809e-609daf22a195 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Qwen2.5 Technical Report
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b071af26-8d4c-4a8a-aa99-61b0b8eaa9f5 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Chronomagic-bench: A bench- mark for metamorphic evaluation of text-to-time-lapse video generation.Advances in Neural Information Processing Sys- tems, 37:21236–21270, 2024
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9ffb76a0-0386-44f1-bab9-528dcac89c26 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Sigmoid loss for language image pre-training
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86fbed20-9b37-4249-8fee-373b6ee76418 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Q-Eval-100K: Evaluating Visual Quality and Alignment Level for Text-to-Vision Content
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 03188bb0-db5f-41f1-a1fd-eb982ab514e2 · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Judging llm-as-a-judge with mt-bench and chatbot arena.Advances in Neural Information Processing Systems, 36:46595–46623, 2023
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f0b9a4e7-9957-4256-9693-8c4568e6499d · outbound
Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model Open-sora: Democratizing efficient video production for all, march 2024.URL https://github
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
No inbound Pith citation observations are available.