Pith. sign in

Paper Citation Record · LEDGER

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning

As of 13 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2411.19803.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.19803 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T13:35:28.550194Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

19 of 19 outbound references displayed

  • verified exact1
  • verified fuzzy17
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 87d9ad96-4d1c-42b6-a0eb-5c8f1854490b · outbound

This paper cites Emotion recognition combining acoustic and linguistic features based on speech recognition results[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Emotion recognition combining acoustic and linguistic features based on speech recognition results[C]

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.886347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.461456Z digest=sha256:9a3d55168c00bdf6fc8fb986b64c57fa907337b2c566e11b03df2cd1e05a9650

Observation 908c9d4e-0cfc-490f-8897-0b7a95af226d · outbound

This paper cites Deep implicit distribution alignment networks for cross-corpus speech emotion recognition[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Deep implicit distribution alignment networks for cross-corpus speech emotion recognition[C]

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.870600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.467592Z digest=sha256:d109f8911ec6781f4ca01fa4cb1a889406de57204b74e382734cb2cfc84d2f0d

Observation ed9329f7-274f-4a1b-acc8-6abcdaa1aeec · outbound

This paper cites Exploring wav2vec 2.0 fine tuning for improved speech emotion recognition[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Exploring wav2vec 2.0 fine tuning for improved speech emotion recognition[C]

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.855026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.472486Z digest=sha256:ba780d30452a01e339c5762e1b36177c6095ff861d912484b41b583a231d3836

Observation a2f1f8a3-12d1-4b71-92af-fdd301833b95 · outbound

This paper cites A Comprehensive Exploration of Fine-Tuning WavLM for Enhancing Speech Emotion Recognition[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning A Comprehensive Exploration of Fine-Tuning WavLM for Enhancing Speech Emotion Recognition[C]

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.839547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.477509Z digest=sha256:b720d78dbfcf26e9ea4ead81f6571b019372df45c86fcbabc87be05f6fdd38e0

Observation c1fe7058-3456-488b-b213-abdc0b1d33de · outbound

This paper cites Revealing Emotional Clusters in Speaker Embeddings: A Contrastive Learning Strategy for Speech Emotion Recognition[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Revealing Emotional Clusters in Speaker Embeddings: A Contrastive Learning Strategy for Speech Emotion Recognition[C]

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.823518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.482252Z digest=sha256:f707a0e78a9fbeb46fac8fb3efdd816d71db2c3c3c70076d5201ec1238bea173

Observation 78951810-22e1-4f3c-9da0-4e3089592eae · outbound

This paper cites IEMOCAP: Interactive emotional dyadic motion capture database[J].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning IEMOCAP: Interactive emotional dyadic motion capture database[J]

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.808189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.487355Z digest=sha256:dcd6ce26e5597f103fe636ebb59630beb32aa1e53ccc5c8bd8f6992c0202ef30

Observation 68df4ff5-db8f-44bf-82aa-01abd2b52ee6 · outbound

This paper cites The CASIA audio emotion recognition method for audio/visual emotion challenge 2011[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning The CASIA audio emotion recognition method for audio/visual emotion challenge 2011[C]

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.792596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.492850Z digest=sha256:0eb25f9eebdda1a31a6125ade585d0f92c589a31f192541ce66915c0f063f6b6

Observation a24e56bf-b372-4e91-af90-d57b559d7a02 · outbound

This paper cites Hubert: Self-supervised speech representation learning by masked prediction of hidden units[J].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Hubert: Self-supervised speech representation learning by masked prediction of hidden units[J]

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.776679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.497455Z digest=sha256:7a1cae71935231f11d5253042e683e638d8662fe8f716858279c81a95d7181e3

Observation f166aee7-3d95-447d-b6b5-4c640f29c041 · outbound

This paper cites Wavlm: Large-scale self-supervised pre-training for full stack speech processing[J].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Wavlm: Large-scale self-supervised pre-training for full stack speech processing[J]

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.761468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.502117Z digest=sha256:c39e4e0e9705f9a2d5e29fe1970162c95b88e3890b3f983e819e83ae5d28bbbb

Observation f1ae18d8-651a-4b66-ad8f-b08d92f33ee8 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations[J].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning wav2vec 2.0: A framework for self-supervised learning of speech representations[J]

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.744898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.506783Z digest=sha256:ad008c11534cda99e01b1570a7025dadd7620e76562515aaebcb49ed9cf1765c

Observation cfe810fc-43d6-48a3-a627-6466166fd8f7 · outbound

This paper cites Speech emotion recognition using self-supervised features[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Speech emotion recognition using self-supervised features[C]

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.729302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.511457Z digest=sha256:ab9fc6e61385972725e5cd0efc58a3b17f3d4643c321f703b8d7fa4662707c0e

Observation 5444e15b-5d3c-47b1-99d8-a235998fe7eb · outbound

This paper cites ECAPA-TDNN: Emphasized Channel Attention, Propagation and Aggregation in TDNN Based Speaker Verification.

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning ECAPA-TDNN: Emphasized Channel Attention, Propagation and Aggregation in TDNN Based Speaker Verification

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T13:35:28.516037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:35:28.516037Z digest=sha256:11d58a896d393deff5d7cccb38d37931fd9f800068eb13b3f5f2755ff68665c8

Observation 24e17b1d-7691-4ee7-9e4a-3ea6f758267b · outbound

This paper cites Speech-based emotion recognition with self-supervised models using attentive channel-wise correlations and label smoothing[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Speech-based emotion recognition with self-supervised models using attentive channel-wise correlations and label smoothing[C]

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.713049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.521467Z digest=sha256:8f4f53c50b203ff01a9c80fa2611b7b1916e9445597e0e902929baeebc1b2b50

Observation 243d1a48-66b0-4c87-b8ee-ddcbf3c12125 · outbound

This paper cites Contrastive unsupervised learning for speech emotion recognition[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Contrastive unsupervised learning for speech emotion recognition[C]

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.694407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.526364Z digest=sha256:90cf3686cbb87984a16ea24263d0574d15cf6b3ea6a47969ce59f61a26ce981e

Observation 7ab4b816-627c-4f34-90c5-e42d2f9a46a0 · outbound

This paper cites MCM-CSD: Multi-Granularity Context Modeling with Contrastive Speaker Detection for Emotion Recognition in Real-Time Conversation[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning MCM-CSD: Multi-Granularity Context Modeling with Contrastive Speaker Detection for Emotion Recognition in Real-Time Conversation[C]

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.678448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.530843Z digest=sha256:9a89620f81638c84a5d3269f4ed3b69f522aa5088cca394b5eef03634d9e34d2

Observation 6de5e7d8-b3f2-4ff0-9efc-ae630032846a · outbound

This paper cites Self-attention encoding and pooling for speaker recognition.

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Self-attention encoding and pooling for speaker recognition

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-12T13:35:28.594134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.535310Z digest=sha256:d1d99d26dc19b1138c026118877466413024f97b0a2329aee68869e973111969

Observation 1a12576f-c3c8-470d-a1ff-12533273f591 · outbound

This paper cites Dual-tbnet: Improving the robustness of speech features via dual-transformer-bilstm for speech emotion recognition[J].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Dual-tbnet: Improving the robustness of speech features via dual-transformer-bilstm for speech emotion recognition[J]

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.662155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.540367Z digest=sha256:14eb4756e40158a4e9b8cf3b5bc490eca988ee86bd2454a14940c19c2b490a58

Observation 137d70b1-5395-48d2-95ba-63954eca09a5 · outbound

This paper cites Temporal modeling matters: A novel temporal emotional modeling approach for speech emotion recognition[C].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning Temporal modeling matters: A novel temporal emotional modeling approach for speech emotion recognition[C]

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.645491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.545503Z digest=sha256:16edadd4b0fbf2e59f06fa2039e5e1fdd14eae548fe0ffe58f7583fc8e77fe2e

Observation 00da6090-390c-4739-871c-fe190ed922e3 · outbound

This paper cites emodarts: Joint optimisation of cnn & sequential neural network architectures for superior speech emotion recognition[J].

A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning emodarts: Joint optimisation of cnn & sequential neural network architectures for superior speech emotion recognition[J]

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:35:28.629417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:35:28.550194Z digest=sha256:ea39bae6b49c3c4316f706c0aae782ac63e557f4a5c0bd9498d71786cade056e

Pith citing papers

No inbound Pith citation observations are available.