Frontier multimodal models judge Chinese short-video misinformation inconsistently, and their veracity ratings shift when videos carry verified or authoritative channel identities.
arXiv preprint arXiv:2503.09387 , year=
2 Pith papers cite this work. Polarity classification is still indexing.
2
Pith papers citing it
years
2026 2representative citing papers
ViCoStream is a new coordinated pipeline framework for streaming VideoLLMs that achieves 134 FPS video throughput and less than 50 ms TTFT on A100 while keeping accuracy near full-history baselines.
citing papers explorer
-
Probing Multimodal Large Language Models on Cognitive Biases in Chinese Short-Video Misinformation
Frontier multimodal models judge Chinese short-video misinformation inconsistently, and their veracity ratings shift when videos carry verified or authoritative channel identities.
-
ViCoStream: Streaming VideoLLMs Can Run Beyond 100 FPS with Stage-Wise Coordinated Inference
ViCoStream is a new coordinated pipeline framework for streaming VideoLLMs that achieves 134 FPS video throughput and less than 50 ms TTFT on A100 while keeping accuracy near full-history baselines.