Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:07:25.935688Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 1 inbound Pith citation observation for arXiv:2505.20007.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:07:25.935688Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:07:23.506269Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T14:07:26.629736Z
34 of 34 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 85b01d42-30f7-49ea-b4fc-04c115e0578d · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 81b16ec7-6582-4c28-86f0-6b083741b15a · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c62b6e52-2cd3-4500-8546-f3c16ae7c59c · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a1410ac3-3a42-43f1-b9c4-34b3631ebd7d · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Dataset The provided challenge data consist of recordings from the MSP-Podcast dataset [23]
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9fff3cdf-f16f-422f-ae8d-60ccb7663244 · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Half of these models were trained us- ing only WCE loss, while the remaining half were trained with the additional batch balancing and SML loss
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f77ffcb1-d802-4a98-a13d-f96582f07e1d · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Notably, the pro- posed architecture benefits from the combination of multiple modalities, with its worst performance occurring when only a single modality is used
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 86cb8479-fc75-4ff5-8695-511c8774c39f · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model It is also supported by FAPESP (BI0S #2020/09838-0 and Ho- rus #2023/12865-8)
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f7f809d7-b7e8-49cb-a12b-e003f6c165ee · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Speech emotion recognition from voice messages recorded in the wild,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c2d137f5-ecde-482a-94f4-8723403ca5f7 · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Acoustic Emotion Recognition for Affective Computer Gaming,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f97c11c4-f9da-4bbf-b0ff-d872d2b3b93e · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Speech emotion recognition using machine learning — A systematic review,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7455f5c0-df4e-455c-a364-aacf3dd60a6d · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Speech Emotion Recognition in Neurological Disorders Using Convolutional Neu- ral Network,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation aa7754ff-ee80-4b51-97b2-e64fd05fb0d5 · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Automatic Assessment of Depression From Speech via a Hierarchical Attention Transfer Network and Attention Autoencoders,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f6b8a47f-37a0-4a9f-a109-23c45955205f · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Negative Emotion Recognition using Deep Learning for Thai Language,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 36976519-c9c0-42c3-a844-50b01338b691 · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Negative emotions detection as an indicator of dialogs quality in call centers,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2c8ce4c1-fad4-4ff1-9edc-ed23994d1870 · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Using Paralinguistic Cues in Speech to Recognise Emotions in Older Car Drivers,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 257c00ae-18b7-4a35-acd2-3978261093c2 · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Affective Human-Robotic Interac- tion,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 67bd16b8-eb3b-4c0b-b3cf-29a25965f6c9 · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Multimodal emotion recognition using cross-modal attention and 1d convolutional neural networks
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 438fe95a-3f1c-40d2-ba53-e49f31e543e6 · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Speech emotion recognition combining acoustic features and linguistic information in a hy- brid support vector machine-belief network architecture,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 30f96a3f-3b1a-4e22-bc5a-46487e6c690b · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Wavlm: Large-scale self-supervised pre-training for full stack speech processing,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11753f69-f739-4061-9930-c3d3f32e05c7 · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Hubert: Self-supervised speech representation learning by masked prediction of hidden units,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 548b8aa6-054a-4591-ab42-50cf786f7968 · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Multimodal Emotion Recognition,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 92362960-d51c-4aeb-a1e2-e5be6b502f46 · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model 1st Place Solution to Odyssey Emotion Recognition Challenge Task1: Tackling Class Imbalance Problem,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c417e31a-acc9-48b6-9e38-603fdb8651eb · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Naturalspeech 3: zero-shot speech synthesis with factorized codec and diffusion models,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9efa3cfc-be32-4808-9d16-fd2946423efa · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Emotion recogni- tion through multiple modalities: face, body gesture, speech,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 719ee80e-8b33-41e2-b265-214c4917f5ad · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Amphion: An Open-Source Audio, Music and Speech Generation Toolkit
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 456ed2e1-8f2a-4837-82b3-a2bb1ed2a09b · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Using transformers for mul- timodal emotion recognition: Taxonomies and state of the art review,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d96251dc-d414-423a-af73-20c24b5721a1 · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Deep neural networks for emotion recognition com- bining audio and transcripts,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b90d3cd5-7c2c-433e-a73e-9ace91887ee0 · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Multimodal emotion recognition with transformer-based self supervised feature fusion,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9443443c-afb9-4378-8e89-84eef59d008b · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model The geneva minimalistic acoustic parameter set (gemaps) for voice research and affective computing,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8eb908dc-9e08-4f78-9e9d-1596ea8d4e10 · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Stacked generalization,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3280dc54-e0eb-480d-8e42-a32c1dc5b4d1 · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model The interspeech 2025 challenge on speech emotion recognition in naturalistic conditions,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9ec5e433-fa1c-449d-ad93-b4f637906853 · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01748e3d-fb67-41d2-b111-67b40c60117f · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Odyssey 2024 - speech emotion recognition challenge: Dataset, baseline framework, and results,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a07c5487-f316-4b0a-9500-72d1415fb878 · outbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model EMOVOME: A Dataset for Emotion Recognition in Spontaneous Real-Life Speech
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81b16ec7-6582-4c28-86f0-6b083741b15a · inbound
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.