Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-11T01:53:00.127636Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 0 inbound Pith citation observations for arXiv:2607.06611.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-11T01:53:00.127636Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
53 of 53 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7f693040-bd69-46b8-ae57-62bf2b178b88 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts The evolution of sentiment analysis and conversational AI: Techniques applications and future research directions
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c797137c-7de2-4edd-ae15-66f2ec2735aa · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Sentiment analysis and emotion recognition from speech using universal speech representations
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f07b8f8e-8391-4d93-8498-6176ad613e87 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts wav2vec 2.0: A framework for self-supervised learning of speech representations, in: Proceedings of NeurIPS, pp
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d21193c9-b6e7-431c-bbec-bdcb507027ae · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts TweetEval: Unified benchmark and comparative evaluation for tweet classification, in: Findings of EMNLP, pp
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1e7a7188-cd28-45ff-9ddc-32c0a3f60164 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Sentiment Analysis of Customer Feedback and Reviews in E-Commerce Systems, in: Proceedings of ICTCS, pp
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9f5f687a-49a2-4fcc-beff-cb3be8a79c5b · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts IEMOCAP: Interactive emotional dyadic motion capture database
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6bf36bf0-23ef-4bd0-a977-ea0adbef5d99 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts The MSP-Podcast Corpus
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5581cd06-56bc-48e7-a1d8-cc58f3a6080f · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts German’s next language model, in: Proceedings of COLING, pp
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 639c29ce-3311-40e5-ac9e-dc5278c8b5b3 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts WavLM: Large-Scale Self-Supervised Pre-Training for Full Stack Speech Processing
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b7b2eaf5-d8ad-4360-9fbb-d7fe89dfd425 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts BERT: Pre-training of deep bidirectional transformers for language understanding, in: Proceedings of NAACT-HLT, pp
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0f7eca62-d1b6-4a2c-ab23-17fd0297fd19 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Optimized Sentiment Analysis in Tagalog Speech Using PCA and BRNN on Prosodic Suprasegmental and MFCC Features, in: Proceedings of ICTC, pp
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0f0b3717-bb94-4ffc-aecc-b9e4d329bfec · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts A review on speech emotion recognition: A survey, recent advances, challenges, and the influence of noise
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cdeb7a1d-0ee1-46fa-9d83-315bb24b73f7 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Teacher-Student Training and Triplet Loss for Facial Expression Recognition under Occlusion, in: Proceedings of ICPR, pp
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7e0b00e8-8c0e-40c6-acd1-5426bd551698 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts AST: Audio Spectrogram Transformer, in: Proceedings of INTERSPEECH, pp
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 173827fc-beb8-4ce7-b67f-01a05ee034c6 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Distilling the Knowledge in a Neural Network
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fc45812c-5ef5-406c-bc89-7e90096498fc · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c9ae9372-df68-4ce1-ac83-bdbdeda3aba4 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts LoRA: Low-Rank Adaptation of Large Language Models, in: Proceedings of ICLR
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 44ae0fd6-1df5-4da5-99f2-d10fdf2153b6 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech, in: Pro- ceedings of ICML, pp
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2daecc1d-f41f-4b82-973e-41c02b5f2660 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts faster-whisper: Faster Whisper transcription with CTranslate2.https://github.com/SYSTRAN/faster-whisper
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b98508a2-01cb-49c5-bd1f-3374212a27c6 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Incorporating end-to-end speech recognition models for sentiment analysis, in: Proceedings of ICRA, pp
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7a028ca4-f932-4d96-92e0-96691cde939e · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Unimodal-driven distillation in multimodal emotion recognition with dynamic fusion, in: Proceedings of ICME, pp
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5bd5a4c2-13a7-4ecd-ac3f-a003d5b00c58 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Speech Emotion Recognition With ASR Transcripts: a Comprehensive Study on Word Error Rate and Fusion Techniques, in: Proceedings of SLT, pp
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8ac18c4f-f752-4748-87f7-d42d163a878e · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Decoupled multimodal distilling for emotion recognition, in: Proceedings of CVPR, pp
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3aea66e4-3850-4ebb-9bcd-e5852175771b · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts A survey of deep learning-based multimodal emotion recognition: Speech, text, and face
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2ebb9d24-e0ba-4c28-9175-b8ad767de9a0 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Development of interactive English e-learning video entertainment teaching environment based on virtual reality and game teaching emotion analysis
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 470f6a7e-e1b9-4267-a186-e558dc4a4df9 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Unifying distillation and privileged information, in: Proceedings of ICLR
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ffdce282-5df0-4fcb-bfb4-1c1f1d883b29 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts ScaleVLAD: Improving Multimodal Sentiment Analysis via Multi-Scale Fusion of Locally Descriptors
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation becdda28-330f-4f03-83d9-c406ffb7824f · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Audio sentiment analysis by heterogeneous signal features learned from utterance-based parallel neural network, in: Proceedings of AffCon@AAAI, pp
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ffef5c22-058d-4bc6-b934-a8361559a09d · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts CamemBERT: a tasty French language model, in: Proceedings of ACL, pp
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b74604ae-6a73-4ab8-b296-85bad7aa1bed · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Learning using generated privileged information by text-to-image diffusion models, in: Proceedings of ICPR, pp
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9d77120e-2752-4f43-a185-3bdf66a764c7 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Verbal sentiment analysis and detection using recurrent neural network, in: Advanced Data Mining Tools and Methods for Social Computing
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5c34b6f5-2a3b-4f4d-af1c-0af8fd4e45ba · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Bridging Modalities: Knowledge Distillation and Masked Training for Translating Multi-Modal Emotion Recognition to Uni-Modal, Speech-Only Emotion Recognition
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bfaf5578-4fc6-4d6f-8a12-8647fb491780 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7fcf8860-9b95-48d0-8dcb-d715a2158b6e · outbound
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8adb56cc-3895-4df3-8298-bf67eef50be9 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts No Language Left Behind: Scaling Human-Centered Machine Translation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e3810431-9f63-4532-bbbb-7601d50ccc77 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts RoBERTuito: a pre-trained language model for social media text in Spanish, in: Proceedings of LREC, pp
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f4e197aa-3d68-4436-b2cc-caa9b2cea59c · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations, in: Proceedings of ACL, pp
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 31401229-8535-41aa-abe5-79ab315dbe2a · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Scaling speech technology to 1,000+languages
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2b27edf8-5f60-414a-839d-285db828943c · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Robust Speech Recognition via Large-Scale Weak Supervision, in: Proceedings of ICML, pp
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7687e83f-2f1a-478f-8084-fbcba3c8d3b2 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Cascaded cross-modal transformer for audio-textual classification
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 17fe5ee7-83eb-411a-814d-bb49c472a88e · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Cascaded cross-modal transformer for request and complaint detection, in: Proceedings of ACMMM, pp
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f3cdb596-e0bf-4499-b398-13ab84f207f4 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts SepTr: Separable Transformer for Audio Spectrogram Processing, in: Proceedings of INTER- SPEECH, pp
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 050246b2-a0f1-436a-8777-1513cb80f3a5 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts An integrated approach for mental health assessment using emotion analysis and scales
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 13152bb5-235f-490d-a9b3-8ff72df03477 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Leveraging pre-trained language model for speech sentiment analysis, in: Proceedings of INTERSPEECH, pp
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9a154783-b248-45ab-9fcc-c64df45744a8 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts A comparative study on Bengali speech sentiment analysis based on audio data, in: Proceedings of BigComp, pp
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3ad7cd62-36fd-4bd7-8197-18dbb4cbf17d · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Multimodal transformer for unaligned multimodal language sequences, in: Proceedings of ACL, pp
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d4331dbd-6328-40ab-8fa6-2c0ad0ab1bc1 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Hierarchical cross-modal attention and dual audio pathways for enhanced multimodal sentiment analysis
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d41049fb-4d5a-4183-8be6-3675fee23ca4 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Learning using privileged information: Similarity control and knowledge transfer
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2b030a15-2b87-45b2-8590-d35b6fbd9979 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts A Fine-tuned Wav2vec 2.0/HuBERT Benchmark For Speech Emotion Recognition, Speaker Verification and Spoken Language Understanding
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 55ee1f27-4db3-4afe-ac07-226c743311bf · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Enriching multimodal sentiment analysis through textual emotional descriptions of visual-audio content, in: Proceedings of AAAI, pp
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 53ca2bb2-41f4-4adb-9b3f-f91edeec52d4 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts A Self-Adjusting Fusion Representation Learning Model for Unaligned Text-Audio Sequences
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d1bbf595-ef6c-43fa-a12b-4ec23e26d35b · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Multimodal speech emotion recognition using audio and text, in: Proceedings of SLT, pp
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7c0dee6f-1b7b-4c3b-971c-47b38a7c9102 · outbound
Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts Personality-aware multimodal driver emotion recognition towards intelligent connected vehicles
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
No inbound Pith citation observations are available.