Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T17:58:27.925157Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 0 inbound Pith citation observations for arXiv:2507.19356.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T17:58:27.925157Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
20 of 20 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 65a1ec94-ccd5-4f1a-9375-d501dab9dd1e · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization Integrating emotion recognition with speech recognition and speaker diarisation for conversations,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4ae37139-1c2c-42de-b31d-67004035000d · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization RoBERTa: A Robustly Optimized BERT Pretraining Approach
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c0807a3-afe2-4d28-ba97-8f7d905e066b · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a00965dc-6ad0-49e7-9e78-a44e24b5a903 · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization Cross-Modality Gated Attention Fusion for Multimodal Sentiment Analysis
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8abb5c52-f9d6-4b3e-9112-00a3efa89468 · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization IEMOCAP: In- teractive emotional dyadic motion capture database,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6c6fd2b4-234e-4d8e-b18e-e88208b4279a · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization WhisperX: Time accurate speech transcription of long form audio,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4bae0ed-fb30-40e4-a72a-f4fa25a3730a · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization Speech emotion recognition with ASR transcripts: A comprehensive study on word error rate and fusion techniques,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66813e96-3786-467d-9511-4e03379b63fa · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization emoDARTS: Joint Optimisation of CNN & Sequential Neural Network Architectures for Superior Speech Emotion Recognition
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ea78dee3-a181-4a97-94fb-15f08dc95cfe · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization Speech sentiment analysis via pre- trained features from end-to-end ASR models,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4094e105-2344-4330-86bd-c97793158f5b · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization learning discriminative features from spectrograms using center loss for speech emotion recognition
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45de3c0f-aed2-4757-806d-e376c58a24bf · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization Multimodal multi-loss fusion network for sentiment analysis,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 31112366-bbc2-44f5-be01-9b8ad3113a6c · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization Speech emotion recognition in dyadic dialogues with attentive interaction modeling,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9dded3b1-4def-439f-8428-588568e78806 · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization Temporal context in speech emotion recognition,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8ab3af26-85b1-4931-92d5-6e9de366e262 · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization Transcribe-to-diarize: Neural speaker diarization for unlimited number of speakers using end-to-end speaker-attributed ASR,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0d1ed5e0-e93d-41d3-86b1-1d85510c30f6 · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization Streaming Speaker-Attributed ASR with Token-Level Speaker Embeddings
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 083e3fea-e673-4231-a515-c77139045883 · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization Speech emotion diarization: Which emotion appears when?,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f3862308-facf-4b79-8a18-660760e1d4c9 · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization Meeting recognition with continuous speech separation and transcription-supported diarization,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 58eb620d-d29f-4bf9-8404-c0571fff8df3 · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization pyannote.audio: neural building blocks for speaker diarization,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c950f8c1-2d87-46f9-8716-5ef2086e7a65 · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df8c584f-eb5a-4005-b8a8-38828189d440 · outbound
Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization 2019, Accepted at EMNLP-IJCNLP 2019]
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
No inbound Pith citation observations are available.