Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T12:18:32.900270Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 2 inbound Pith citation observations for arXiv:2508.03722.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T12:18:32.900270Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T11:42:33.784811Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-12T06:41:37.028986Z
33 of 33 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8d5c627a-a6db-43ee-a226-5bb0c96c6d11 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4eb8c3b6-26c8-493a-b570-d4b6c0db845e · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Tulu 3: Pushing Frontiers in Open Language Model Post-Training
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efd91f36-55f8-451b-912c-67610fb9ae5e · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Chain-of-thought prompting elicits reasoning in large language models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23b59b22-e7a1-4a78-8c88-f34c8c64a3e5 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Train- ing language models to follow instructions with human feedback
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8206c12d-31a4-40cd-8a65-113b73225169 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 407afac2-5f17-40be-adcf-141bf72293d0 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors R., Shravan Venkatraman, Vigya Sharma, Santhosh Malarvannan, and Modigari Narendra
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8bee82be-d806-41f9-a2a2-d50e6b8b2026 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Tacfn: Transformer-based adaptive cross-modal fusion network for multimodal emotion recognition.Artificial Intelligence Research, 2, 2023
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 74b3c003-b1f4-4aff-bef3-b35b768feb6d · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Multimodal language analysis in the wild: CMU-MOSEI dataset and interpretable dynamic fusion graph
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 08dc0430-9e1c-4952-8d6f-886f58d8298d · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Livingstone and Frank A
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 23f35a65-b2f2-4331-855f-2cd9752c836f · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Multimodal emotion recognition with vision–language prompting and modality dropout
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fd14d689-c8ec-4d1b-81ad-326227f8d045 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Visual and textual prompts for enhancing emotion recognition in video
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0b3c981a-614c-4c59-b32c-700f34dd850d · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Ov-mer: Towards open-vocabulary multi- modal emotion recognition
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4645c05e-13d6-4053-a3ff-b54cca8f6139 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Af- fectgpt: A new dataset, model, and benchmark for emotion understanding with multimodal large language models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5e7e6434-b61d-4182-b307-d66b066735de · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors From System 1 to System 2: A Survey of Reasoning Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec32d509-c19c-43eb-a525-ecec7b3af040 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Visual-RFT: Visual Reinforcement Fine-Tuning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acbf79e9-a453-445c-81e2-2cf1ccce4be1 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Visual Programming: Compositional visual reasoning without training
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d07a0b31-6cee-4585-ba9b-6525e8515787 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Visual chain-of-thought prompting for knowledge- based visual reasoning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d99ad617-03b0-4534-a7ea-5f6a91e237b6 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors LaMI: Augmenting Large Language Models via Late Multi-Image Fusion
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5f6af8b3-4203-4346-9353-5555f5361ecf · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors An explainable vision question answer model via diffusion chain-of-thought
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 12e8b7a5-2f19-4031-b883-ac4889c056a2 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors The relation between valence and arousal in subjective experience.Psychological bulletin, 139(4):917, 2013
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7ac25b27-2593-4994-b885-9b9f2105ae1d · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Arousal, valence, and mem- ory for detail.Memory, 12(2):237–247, 2004
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d515ee12-2acc-4932-a8b1-c7eca222a843 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Facial action coding system.Environmental Psychology & Nonverbal Behavior, 1978
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 883dd76c-e4e3-45f0-9bf0-fd9b040bdf6b · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Gemini: A Family of Highly Capable Multimodal Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 540bedeb-187a-448c-9dfc-dcadab01b2d4 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Representation Learning with Contrastive Predictive Coding
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a883165-e486-469f-bb15-7a22a935a461 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Balanced contrastive learning for long-tailed visual recognition
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 361c99b7-16e4-41ab-86b2-ed2c3d85bde2 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Mer 2024: Semi-supervised learning, noise robustness, and open-vocabulary multimodal emotion recognition
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c71cc0df-bb19-4809-950a-07bf92d0baa4 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Mer 2023: Multi-label learning, modality robustness, and semi-supervised learning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9d69c8ff-0d2c-4ea8-93f5-18dd6cb9654c · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Emotion-llama: Multimodal emotion recognition and reasoning with instruction tuning.Advances in Neural Information Processing Systems, 37:110805–110853, 2024
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation aa515d7c-71c5-416b-a031-6d593afef563 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ef91280-3fa7-4007-aa6a-39f50dfe993e · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c920637-e1db-4474-889f-be9987090f8c · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b8c1a8a-da01-478c-b1a1-083df0888033 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f53d9fd-8034-4b53-97fe-4a5538355a79 · outbound
Multimodal Video Emotion Recognition with Reliable Reasoning Priors Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a5087e3-3222-4fec-a5e0-b7b40d215aa8 · inbound
WINELL: Wikipedia Never-Ending Updating with LLM Agents Multimodal Video Emotion Recognition with Reliable Reasoning Priors
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4134853e-9692-4730-a6db-6da003818775 · inbound
Position: Life-Logging Video Streams Make the Privacy-Utility Trade-off Inevitable Multimodal Video Emotion Recognition with Reliable Reasoning Priors
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.