Pith. sign in

Paper Citation Record · LEDGER

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning

As of 13 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2411.19803.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.19803 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T13:35:28.550194Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

19 of 19 outbound references displayed

  • verified exact1
  • verified fuzzy17
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 87d9ad96-4d1c-42b6-a0eb-5c8f1854490b · outbound

This paper cites Emotion recognition combining acoustic and linguistic features based on speech recognition results[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Emotion recognition combining acoustic and linguistic features based on speech recognition results[C]

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.886347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.461456Z digest=sha256:a060a30be30e8fa5bee2245981ed341d283886c111ed75740b41dcb6a5f8249c

Observation 908c9d4e-0cfc-490f-8897-0b7a95af226d · outbound

This paper cites Deep implicit distribution alignment networks for cross-corpus speech emotion recognition[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Deep implicit distribution alignment networks for cross-corpus speech emotion recognition[C]

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.870600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.467592Z digest=sha256:cb97e7945dca52fc227451feab4b2a27ef3d2b2d641121aa66b206033850dd38

Observation ed9329f7-274f-4a1b-acc8-6abcdaa1aeec · outbound

This paper cites Exploring wav2vec 2.0 fine tuning for improved speech emotion recognition[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Exploring wav2vec 2.0 fine tuning for improved speech emotion recognition[C]

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.855026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.472486Z digest=sha256:f18d3e9958a34fd56c8a157febc367fe88c048f1c3a44ae76a60b0603747f631

Observation a2f1f8a3-12d1-4b71-92af-fdd301833b95 · outbound

This paper cites A Comprehensive Exploration of Fine-Tuning WavLM for Enhancing Speech Emotion Recognition[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning A Comprehensive Exploration of Fine-Tuning WavLM for Enhancing Speech Emotion Recognition[C]

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.839547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.477509Z digest=sha256:9fc194bd8545b7fb66d6a7ab004aa450385acf002996a2bad07c70154ca926b8

Observation c1fe7058-3456-488b-b213-abdc0b1d33de · outbound

This paper cites Revealing Emotional Clusters in Speaker Embeddings: A Contrastive Learning Strategy for Speech Emotion Recognition[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Revealing Emotional Clusters in Speaker Embeddings: A Contrastive Learning Strategy for Speech Emotion Recognition[C]

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.823518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.482252Z digest=sha256:40ec8787ee9823018191df2cf32a49a76f41faae963e0b6a5335f2b6998a8299

Observation 78951810-22e1-4f3c-9da0-4e3089592eae · outbound

This paper cites IEMOCAP: Interactive emotional dyadic motion capture database[J].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning IEMOCAP: Interactive emotional dyadic motion capture database[J]

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.808189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.487355Z digest=sha256:8b31da041b2412464f43a67bb7a040090fa9b49a43456e4bab96f8a1e3133048

Observation 68df4ff5-db8f-44bf-82aa-01abd2b52ee6 · outbound

This paper cites The CASIA audio emotion recognition method for audio/visual emotion challenge 2011[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning The CASIA audio emotion recognition method for audio/visual emotion challenge 2011[C]

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.792596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.492850Z digest=sha256:f347bf1af157bef222ecddf3caf9c337e5e9913cfa2419ea9b7d7ea266941437

Observation a24e56bf-b372-4e91-af90-d57b559d7a02 · outbound

This paper cites Hubert: Self-supervised speech representation learning by masked prediction of hidden units[J].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Hubert: Self-supervised speech representation learning by masked prediction of hidden units[J]

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.776679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.497455Z digest=sha256:8d41d4c3a168140ed5792234c1a43084d2b84d347ee51026c55996f21e65ffb6

Observation f166aee7-3d95-447d-b6b5-4c640f29c041 · outbound

This paper cites Wavlm: Large-scale self-supervised pre-training for full stack speech processing[J].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Wavlm: Large-scale self-supervised pre-training for full stack speech processing[J]

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.761468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.502117Z digest=sha256:f95096f683c6afb9a66720b9db7fbaf206d126a76de7a9717641759d70ef582a

Observation f1ae18d8-651a-4b66-ad8f-b08d92f33ee8 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations[J].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning wav2vec 2.0: A framework for self-supervised learning of speech representations[J]

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.744898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.506783Z digest=sha256:fa8c275785344c4cc1bb21b94760233e0b41539f7c2072cb4cd1e82b48b953c6

Observation cfe810fc-43d6-48a3-a627-6466166fd8f7 · outbound

This paper cites Speech emotion recognition using self-supervised features[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Speech emotion recognition using self-supervised features[C]

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.729302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.511457Z digest=sha256:260b474ccf964eddb86a9b0ef65eb7957caf4eeca7cf247fc79a54dbdb036ef6

Observation 5444e15b-5d3c-47b1-99d8-a235998fe7eb · outbound

This paper cites ECAPA-TDNN: Emphasized Channel Attention, Propagation and Aggregation in TDNN Based Speaker Verification.

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning ECAPA-TDNN: Emphasized Channel Attention, Propagation and Aggregation in TDNN Based Speaker Verification

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T13:35:28.516037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:35:28.516037Z digest=sha256:11d58a896d393deff5d7cccb38d37931fd9f800068eb13b3f5f2755ff68665c8

Observation 24e17b1d-7691-4ee7-9e4a-3ea6f758267b · outbound

This paper cites Speech-based emotion recognition with self-supervised models using attentive channel-wise correlations and label smoothing[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Speech-based emotion recognition with self-supervised models using attentive channel-wise correlations and label smoothing[C]

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.713049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.521467Z digest=sha256:7c1178879b1720d973412b02af40a1628ceef2df81e8be8f0e33eb2368194f35

Observation 243d1a48-66b0-4c87-b8ee-ddcbf3c12125 · outbound

This paper cites Contrastive unsupervised learning for speech emotion recognition[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Contrastive unsupervised learning for speech emotion recognition[C]

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.694407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.526364Z digest=sha256:678766f7556128024daa26d9bbe0f519fd504610d1c1bedab72fb910e2ab8230

Observation 7ab4b816-627c-4f34-90c5-e42d2f9a46a0 · outbound

This paper cites MCM-CSD: Multi-Granularity Context Modeling with Contrastive Speaker Detection for Emotion Recognition in Real-Time Conversation[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning MCM-CSD: Multi-Granularity Context Modeling with Contrastive Speaker Detection for Emotion Recognition in Real-Time Conversation[C]

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.678448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.530843Z digest=sha256:ff9eaf3e620920b96dbfade5cf8c3f276991a1babe7605142964e0af848bf006

Observation 6de5e7d8-b3f2-4ff0-9efc-ae630032846a · outbound

This paper cites Self-attention encoding and pooling for speaker recognition.

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Self-attention encoding and pooling for speaker recognition

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-12T13:35:28.594134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.535310Z digest=sha256:eb302ca87e2089a173c9a87b3374d67fbb3091d6c821afd439c0c5096e640ed3

Observation 1a12576f-c3c8-470d-a1ff-12533273f591 · outbound

This paper cites Dual-tbnet: Improving the robustness of speech features via dual-transformer-bilstm for speech emotion recognition[J].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Dual-tbnet: Improving the robustness of speech features via dual-transformer-bilstm for speech emotion recognition[J]

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.662155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.540367Z digest=sha256:ce6b02fefab39a1fc51ec9c023333dbb0f63a25ec538fda817a45d75ed5d539c

Observation 137d70b1-5395-48d2-95ba-63954eca09a5 · outbound

This paper cites Temporal modeling matters: A novel temporal emotional modeling approach for speech emotion recognition[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Temporal modeling matters: A novel temporal emotional modeling approach for speech emotion recognition[C]

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.645491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.545503Z digest=sha256:bea67da2b26549b5502e317bb2ecd15482d297f3d26487bafcb122df828f8cca

Observation 00da6090-390c-4739-871c-fe190ed922e3 · outbound

This paper cites emodarts: Joint optimisation of cnn & sequential neural network architectures for superior speech emotion recognition[J].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning emodarts: Joint optimisation of cnn & sequential neural network architectures for superior speech emotion recognition[J]

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.629417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:35:28.550194Z digest=sha256:fb4c991a74e71def11f7ba41aefdcf22cb2dedb7969fc9bc36c8e77690fb25ca

Pith citing papers

No inbound Pith citation observations are available.