Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:02:24.778729Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 74 of 74 outbound references and 2 inbound Pith citation observations for arXiv:2508.02429.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:02:24.778729Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T11:19:34.140857Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-15T13:55:53.243562Z
74 of 74 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 97ea4dbf-5925-42e3-87e0-8cbc9aab1064 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting A review of affective computing: From unimodal analysis to multimodal fusion,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11db27f3-e04a-4364-a9a2-a32fe4fb8e4a · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting A systematic review on affective computing: Emotion models, databases, and recent advances,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation de78fbb7-3a1d-4f76-84a8-a9aca737d6bd · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting An effective data fusion methodology for multi-modal emotion recognition: A survey,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b7525118-6794-41e5-a078-930cd45e178b · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Surveying the mllm landscape: A meta-review of current surveys,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1b077f7-53f5-49e0-b917-d81568e64559 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Learning by comparing: Boosting multi- modal affective computing through ordinal learning,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e75fac92-5937-4d9f-99cf-2ab6992642da · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Llm-based nlg evaluation: Current status and challenges,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f3f1bbd2-9bca-4223-b3f0-c63677aba923 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting The Dawn of LMMs: Preliminary Explorations with GPT-4V(ision)
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81d76dc5-c191-4d76-9687-2d6e99d50007 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Visual instruction tuning,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0fe48994-3171-4398-b5a2-43388ce28bca · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Gemini: A Family of Highly Capable Multimodal Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd611aaf-b1b6-4e25-96de-8140eb598eae · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Qwen-vl: A versatile vision-language model for understanding, localization,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation df5a41c4-2648-4435-a407-e74517b4513f · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting A survey on multimodal large language models,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b0d994dc-4d41-410e-91a1-54be9d9d8e9a · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting EmotionQueen: A Benchmark for Evaluating Empathy of Large Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54470714-89de-4455-97f4-f8f45267dceb · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Eemo-bench: A benchmark for multi-modal large lan- guage models on image evoked emotion assessment,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1c5d6d0-bdc4-4b2a-8647-0b82c892956b · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Can Large Language Models Help Multimodal Language Analysis? MMLA: A Comprehensive Benchmark
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42d0d87b-261b-41e3-8ed8-a2c60f5f6119 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting The Future of MLLM Prompting is Adaptive: A Comprehensive Experimental Evaluation of Prompt Engineering Methods for Robust Multimodal Performance
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1b62d41-cbfa-452e-a7a0-e9afabb2e9b6 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting MOSI: Multimodal Corpus of Sentiment Intensity and Subjectivity Analysis in Online Opinion Videos
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21db1851-7af0-41ee-ba55-5898ff420f84 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Ch- sims: A chinese multimodal sentiment analysis dataset with fine-grained annotation of modality,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3309977f-1b35-4894-a074-94440234caa7 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Make acoustic and visual cues matter: Ch-sims v2. 0 dataset and av-mixup consistent module,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d915f76c-e091-4c4e-8496-012f69e9870b · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0855267a-b265-4c1c-838b-74924086cff2 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting UR-FUNNY: A Multimodal Language Dataset for Understanding Humor
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd73413d-c3a3-48cf-b9f0-4314da5146bc · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Generated Knowledge Prompting for Commonsense Reasoning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fc8751e-6fb9-4930-8744-859f0a979492 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Self-attentive feature-level fusion for multimodal emotion detection,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation adfaf75b-99c2-44bb-a848-01b9d6828b6c · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Emotion recognition using feature-level fusion of facial expressions and body gestures,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 090ae2fc-4055-4976-84ac-48158ff0fd74 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Decision-level fusion method for emotion recognition using multimodal emotion recog- nition information,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1e10bca3-bf67-4ee0-9c05-2e71ca3b791f · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Deep learning-based late fusion of multi- modal information for emotion classification of music video,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 368a5029-19d1-4bcd-9ea0-746739f59219 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting A joint cross-attention model for audio-visual fusion in dimensional emotion recognition,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 404f1533-fea9-4bfe-a18d-5545045a8f36 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Speech emotion recognition with co-attention based multi-level acoustic information,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b284676b-38d0-44fd-978f-2acca5ce5b4a · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Omni-Emotion: Extending Video MLLM with Detailed Face and Audio Modeling for Multimodal Emotion Analysis
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74964bcf-796f-4350-a856-84deb9c68674 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting OmniVox: Zero-Shot Emotion Recognition with Omni-LLMs
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0fec81d-4e6d-40dc-9b2f-8d553f64eefc · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Mellm: Exploring llm-powered micro-expression understanding enhanced by subtle motion perception,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3f3bde7-2787-49c7-9fca-fcaec3ec39ca · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Learning transferable visual models from natural language supervision,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c35e7560-2238-4b01-b812-b28bcbea21f0 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Videomae: Masked autoen- coders are data-efficient learners for self-supervised video pre-training,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d32b479c-3c7f-4f03-9026-72d988877b7b · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting BEATs: Audio Pre-Training with Acoustic Tokenizers
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7731c3b-89bc-49ca-a089-6d1ee0720fdf · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Blip-2: Bootstrapping language- image pre-training with frozen image encoders and large language models,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20b4fd2d-8357-4fa8-905d-821aa6e42340 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting SALMONN: Towards Generic Hearing Abilities for Large Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5af11d9b-7bd0-4ea7-ac40-7c30cab0911d · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11f2de67-3088-49d6-8256-7b9f687d8eef · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa56c0e6-2f13-4c61-8dc4-f259ad81b243 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting HumanOmni: A Large Vision-Speech Language Model for Human-Centric Video Understanding
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a779eae7-fa98-4831-9eb1-a1dd5b7829b9 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Ola: Pushing the frontiers of omni-modal language model with progressive modality alignment,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ce2dff1b-3623-49ca-8b58-96642e13d7f8 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Qwen2.5-Omni Technical Report
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23714ce4-ca41-49be-a199-fdb7fe05d0c0 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Learning emotional prompt features with multiple views for visual emotion analysis,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fd4ab5fb-87cc-402c-988d-3457955a13c0 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Visual and textual prompts in vllms for enhancing emotion recognition,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 86ed16c1-d44c-4cd9-80cf-0b421f060ff1 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Multimodal language analysis in the wild: Cmu-mosei dataset and interpretable dynamic fusion graph,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b4ec7c38-d45f-40db-9361-da55708317ab · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0dd638df-6b46-4131-9a35-4098b317b9dd · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Emotion-llama: Multimodal emotion recognition and reasoning with instruction tuning,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a138cafe-fb39-4808-a7e6-ba7c95b95a7d · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting PandaGPT: One Model To Instruction-Follow Them All
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42933506-5da1-4470-bf52-7a1a9093cb09 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Lora: Low-rank adaptation of large language models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 23935ab7-5e3e-4cd0-bf5e-dfdfc236d63f · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Multimodal information bottleneck: Learning minimal sufficient unimodal and multimodal representations,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f4ade031-da1b-491c-a62b-695a1d283a13 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Injecting multimodal informa- tion into pre-trained language model for multimodal sentiment analysis,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9364ce03-6343-4e38-8aba-e1f5a2243737 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Towards Explainable Fusion and Balanced Learning in Multimodal Sentiment Analysis
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 07d8b9bb-8567-4c82-b9d9-4716b492c995 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Hgtfm: Hierarchical gating-driven transformer fusion model for robust multimodal sentiment analysis,
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 19298fd8-686c-4506-ab67-daa8ca63f8cf · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting End-to-end Semantic-centric Video-based Multimodal Affective Computing
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 721e03b3-f413-49f6-b0d5-41838fea01f9 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50dd4fc4-8c65-4259-999c-7920a4cd450d · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef9dcab4-4f63-40a7-a256-d0d59273805d · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Divide, conquer and combine: Hierarchical feature fusion network with local and global perspectives for multimodal affective computing,
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e1fd8b1b-21fa-48b1-9fa3-0f4edd6ef6ae · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Qwen2.5-Coder Technical Report
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85140fb3-e5f7-4992-910b-cca5bdd125be · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Robust speech recognition via large-scale weak supervi- sion,
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c2e31f1-ad87-48ef-89f8-083ee16eb6b6 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Qwen2.5-VL Technical Report
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f549136-c179-4339-8972-b9a90ce551b8 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting A survey on vision transformer,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4e5f5b46-67cf-4e38-9266-19f98ad5916a · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Sigmoid loss for language image pre-training,
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8b1132f2-adf1-4797-9d46-f0ba22dcad40 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting LLaVA-OneVision: Easy Visual Task Transfer
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b95fbbae-b9c3-43f3-8cf9-7f2c29168dbd · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Oryx MLLM: On-Demand Spatial-Temporal Understanding at Arbitrary Resolution
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39714fe6-3995-4275-b61f-95aee8211602 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Qwen2 Technical Report
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5548e265-db02-49f1-82f6-2416ffd59237 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Imagebind: One embedding space to bind them all,
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3fa0f8c0-2f9e-4206-a2a3-56bda65f93b1 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Hubert: Self-supervised speech representation learning by masked prediction of hidden units,
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 263337ce-aea8-41fb-9da2-b727ba36fc0d · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Mae-dfer: Efficient masked au- toencoder for self-supervised dynamic facial expression recognition,
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0b18373b-846d-452e-9c72-040a0914caa9 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Eva: Exploring the limits of masked visual representation learning at scale,
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 74da8b45-edc1-4c74-897f-a14243688431 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting LLaMA: Open and Efficient Foundation Language Models
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cde85087-202c-4242-87b8-d29ff4cbc18a · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Unresolved cited work
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 917be28d-0b19-4dab-b624-cd745e496966 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Unresolved cited work
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 311dbcf4-fc63-4d0f-8ad0-e80cfbc0cb9b · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Unresolved cited work
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation abe15f6e-c07c-4049-80bb-ed134b901b3a · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Unresolved cited work
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 99597bf0-f020-44c7-84c0-680b7b5a62ee · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Its core architecture follows the Thinker-Talker design
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 27a231c1-e83b-44f8-b2fd-748b9bd889d0 · outbound
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting Its key innovation lies in the ability to simultaneously process visual and speech information in human-centric scenes
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8b65755c-ea84-4a75-8437-4557e50e4303 · inbound
QASA: Quality-Aware Semantic Augmentation for Robust Multimodal Sentiment Analysis Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73165972-423e-47ac-9fdc-f2fb77c25237 · inbound
C2F-Thinker: Coarse-to-Fine Reasoning with Hint-Guided Reinforcement Learning for Multimodal Sentiment Analysis Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.