Pith. sign in

Paper Citation Record · LEDGER

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning

As of 7 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 2 inbound Pith citation observations for arXiv:2507.04048.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.04048 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:01:29.438826Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:01:29.331354Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T01:07:00.272497Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact2
  • verified fuzzy29
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 92b8aeb1-c476-4ab4-a49c-c1763b421afb · outbound

This paper cites this is a sound of.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning this is a sound of

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.777066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.327598Z digest=sha256:f20922014e3e0d435bb997d5927f0533f418e959cdf2cba73cf86d4c7b1d2934

Observation ad2f2715-09be-4f1c-89e4-ceed8bf8eea9 · outbound

This paper cites CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.331354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.331354Z digest=sha256:f4ac4a1117b9b06155c2f5e7e46a2fcb131dc2570e255f3da048ec2c966d2234

Observation 79d52062-635a-4fc2-8219-aea16cf4813c · outbound

This paper cites Happy” or “Sad.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Happy” or “Sad

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.768125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.334984Z digest=sha256:6f9d20da4008d22acb91c5c4a2114eb6e8827992f90ed409187280061cebba4e

Observation 670fde26-2dd7-438f-99dc-9ffda49bd000 · outbound

This paper cites excited” and “happy.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning excited” and “happy

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.759674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.338492Z digest=sha256:4ee59886c65202f2d70af5b93beac9e0ce66cb4a2ad58f506f1c82ce25deaac5

Observation 9b8f5ce6-b8e1-4eca-8a2c-603583d0ac75 · outbound

This paper cites Integrating Acoustic Context Prompt Tuning (ACPT), CLEP-DG improves general- ization across diverse acoustic environments without additional labeled speech data.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Integrating Acoustic Context Prompt Tuning (ACPT), CLEP-DG improves general- ization across diverse acoustic environments without additional labeled speech data

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.751978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.341488Z digest=sha256:39893407db7c025124a6f32f5591f0ad6ec4090267012b490713d96d9bfb593f

Observation 8c4ede18-eee3-4a24-a63f-310a23eb363b · outbound

This paper cites Affective computing: A review,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Affective computing: A review,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.744069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.344546Z digest=sha256:c5a3b66f6e7304a8425612ed936b69c06bdcb528efa4c968bd2ee5ade1d64c77

Observation 99d1de42-a775-4ee7-ab23-3589916be9f4 · outbound

This paper cites Preece, Y.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Preece, Y

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.736494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.347248Z digest=sha256:c67ae5876878b2d63ce1577ef50ebe1b146ab5acea28039895712a4fb3133275

Observation f8212573-553b-4109-88da-8f8b7a38ef16 · outbound

This paper cites Clap learning audio concepts from natural language supervision,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Clap learning audio concepts from natural language supervision,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.349695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.349695Z digest=sha256:73a573a6a05f9db52fcf23e48f9d40666e2fe2624db744019e851ab774715914

Observation 370f554c-f3d9-4e8c-95ce-f3c59a0f963c · outbound

This paper cites Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.352747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.352747Z digest=sha256:5b1360fe8ae541abefc478157583ef3a8aa145864e1afe791958a53758866de9

Observation 94dcbb3d-0e1a-4fbb-ab3a-f09af57abd36 · outbound

This paper cites ParaCLAP -- Towards a general language-audio model for computational paralinguistic tasks.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning ParaCLAP -- Towards a general language-audio model for computational paralinguistic tasks

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:01:29.518461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.355551Z digest=sha256:e1072c96952b2abf6fe7e7cf29a71f103e3fbf7460449b9ec1e112301e23318b

Observation 4cba598e-f394-4aaa-b568-172460786805 · outbound

This paper cites GEmo-CLAP: Gender-Attribute-Enhanced Contrastive Language-Audio Pretraining for Accurate Speech Emotion Recognition.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning GEmo-CLAP: Gender-Attribute-Enhanced Contrastive Language-Audio Pretraining for Accurate Speech Emotion Recognition

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:01:29.507173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.358691Z digest=sha256:5627b6b3db711a55d6870b08989ff6f905fd825366a9e05f061f9752aa362032

Observation dfa20ce6-0dce-4dfa-9608-140020328e92 · outbound

This paper cites Cross-modal features interaction-and-aggregation network with self-consistency train- ing for speech emotion recognition,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Cross-modal features interaction-and-aggregation network with self-consistency train- ing for speech emotion recognition,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.719068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.361954Z digest=sha256:6814a681236fa4e42fe931e64aca501f36a118a8561884cc83ea3689cb92f6e2

Observation 7797ec94-1d13-421d-a9c2-632b61bfbf16 · outbound

This paper cites Conditional prompt learning for vision-language models,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Conditional prompt learning for vision-language models,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.711360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.364310Z digest=sha256:eecaa18ba341494cbbfb3b52e3fd553c41e027f5666b282a743523678844312f

Observation 2ce0f5f2-f8ce-4c16-95f4-35219796bc82 · outbound

This paper cites Learning to prompt for vision-language models,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Learning to prompt for vision-language models,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.703426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.367220Z digest=sha256:821ee810f9f1d8e53bfc59d93fc8963b9261e190d09bacfa187541ba58ca2188

Observation b888d2f4-3420-418e-990a-d53cfaca40a6 · outbound

This paper cites Diagnosing and Rectifying Vision Models using Language.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Diagnosing and Rectifying Vision Models using Language

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.369990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.369990Z digest=sha256:698f826b673fa110dea056dc1a0c0f48dd3b98242488f0760d492451486ba89a

Observation c28a94be-cb91-447c-9971-1e4bccd0005e · outbound

This paper cites Using language to ex- tend to unseen domains.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Using language to ex- tend to unseen domains

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.695918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.372739Z digest=sha256:d63694a89c4955ee4603f0c279c4120fe12eaf5ec4244b05482f754a8145fffa

Observation 64e1a4bd-0bca-4c0e-b7af-3d4419df860d · outbound

This paper cites Promptstyler: Prompt-driven style generation for source-free domain generaliza- tion,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Promptstyler: Prompt-driven style generation for source-free domain generaliza- tion,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.688417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.375131Z digest=sha256:c816e08223a17ef69666c4307e9489824f4982148260c1fdfc93d9dbda5de36a

Observation dff51774-32af-4eae-bbd5-54d657626344 · outbound

This paper cites Dpstyler: dynamic prompt- styler for source-free domain generalization,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Dpstyler: dynamic prompt- styler for source-free domain generalization,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.680801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.377373Z digest=sha256:8f53d13a8f897b444793796489ded29d22f249cd1376855d371320076194e1b6

Observation 8816ffcc-4c9a-4caa-8900-cebb04616ce6 · outbound

This paper cites Whisper: Robust speech recognition via large-scale weak supervision,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Whisper: Robust speech recognition via large-scale weak supervision,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.672462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.380093Z digest=sha256:ba5bff1895cc77e0d3a1de43049ff50c58e1f95e8dfd4ca449b59b500e1a8d61

Observation 8fb4169a-d501-497a-b5f9-9597b10410fd · outbound

This paper cites Wavlm: Large-scale self- supervised pre-training for full stack speech processing,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Wavlm: Large-scale self- supervised pre-training for full stack speech processing,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.382456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.382456Z digest=sha256:c4f94d0dab2c58cfcafeb05011cc9421dbfc2f5237907abc5bed74fd892f313c

Observation 70ad7e6a-9751-4580-a8ab-441ee5be647c · outbound

This paper cites Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.384916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.384916Z digest=sha256:03ff88f4fa90fdbe62b275e474eb54af8a33b4657aa22b0ba4b1973e9157453b

Observation 01573aa2-c727-433e-883e-63f65e8768f0 · outbound

This paper cites Wav2clip: Learning robust audio representations via contrastive learning,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Wav2clip: Learning robust audio representations via contrastive learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.655603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.387390Z digest=sha256:4eeda93979b68ffa63bbe7245ffe0f024b6ec6746eba9044a9779e51c2436aa9

Observation 2625ad7b-fcfd-4752-9299-1440e23b9889 · outbound

This paper cites Audioclip: Ex- tending clip to audio for zero-shot learning,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Audioclip: Ex- tending clip to audio for zero-shot learning,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.648806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.390000Z digest=sha256:c930fb51dbb880ca2ffa6fdd4839ce64afa72ab06fe1b16098c0a8f5828cb472

Observation 91557793-16a5-440d-985b-1a8294ce4506 · outbound

This paper cites Compa-clap: Composi- tional prompts for improving audio-text alignment,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Compa-clap: Composi- tional prompts for improving audio-text alignment,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.641902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.392561Z digest=sha256:c717830e890b7b2e8ea530825ff969712d6183fa4ed4a0b65f9afb4ad2fbebe8

Observation cc81ca27-3f19-46a3-b1bd-790c4c42a003 · outbound

This paper cites Deep Convolutional Ranking for Multilabel Image Annotation.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Deep Convolutional Ranking for Multilabel Image Annotation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.395015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.395015Z digest=sha256:0d6a9d856b1f1b7072f4bf2a81ec1929b033238c68038c8996db535a0f8e5483

Observation aa93ad76-4c87-46f2-9331-c24d3baab56a · outbound

This paper cites Arcface: Additive angular margin loss for deep face recognition,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Arcface: Additive angular margin loss for deep face recognition,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.634308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.397824Z digest=sha256:16cf518d27a7515e3135e3ed0affebbd4a088939e7b34825d77da482e4f2513e

Observation d08ccacf-c164-47bd-98c7-9b4f6ed5406b · outbound

This paper cites IEMOCAP: In- teractive emotional dyadic motion capture database,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning IEMOCAP: In- teractive emotional dyadic motion capture database,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.626550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.400110Z digest=sha256:b16f9e82920b0fc2000c7bdae191c5a919fc1336d38a7ecb19591a29996037e7

Observation 055f513e-b28d-43f8-a98a-adbf5f065974 · outbound

This paper cites MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.618864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.402579Z digest=sha256:494ce902434e2237e951753b98922636906265827b4ecc41b1be66e9663909e5

Observation b6872e72-4806-4a77-8d43-e4957be45e4e · outbound

This paper cites MEAD: A Large-Scale Audio-Visual Dataset for Affective Un- derstanding and Emotional Expression Analysis,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning MEAD: A Large-Scale Audio-Visual Dataset for Affective Un- derstanding and Emotional Expression Analysis,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.611138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.405056Z digest=sha256:509b3491229f345aba76ea7109faa2384b7bfad4808d70d7fecf4203278153d4

Observation 3d296c53-f88e-4c98-ac46-035dbca314f6 · outbound

This paper cites CMU-MOSEI: A Dataset for Multimodal Sentiment Analysis and Emotion Recognition,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning CMU-MOSEI: A Dataset for Multimodal Sentiment Analysis and Emotion Recognition,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.603415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.407680Z digest=sha256:8372e844bc536c666ce062ce475d8a4bebce74402a38acc901591172fc961194

Observation d0869239-af0e-4643-b230-96a001f5cd9f · outbound

This paper cites Wav2clip: Learning robust audio representations from clip,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Wav2clip: Learning robust audio representations from clip,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.595828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.410016Z digest=sha256:55555ad7c1a255c30d3d1fb2da02c216fc1b6a5bb2855e556eb4149b05623d21

Observation 3a01696e-7f0d-4ddf-99fb-bd60b513d358 · outbound

This paper cites AudioCLIP: Extending CLIP to Image, Text and Audio.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning AudioCLIP: Extending CLIP to Image, Text and Audio

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.412385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.412385Z digest=sha256:5ed5bf739571c26913da3fe48a1627afa8e5b077fcc1fdbf5a6eef6df8e43d45

Observation ca24a637-25ba-4b48-aaa4-e835a7fa2ce1 · outbound

This paper cites CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.415153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.415153Z digest=sha256:db98b376a28826cef5eb6084f2a47047545e073bbf293605374d013a0a709ee5

Observation 770b239e-4e4b-4de8-91cf-9a3cc21ec562 · outbound

This paper cites The Ryerson Audio-Visual Database of Emotional Speech and Song (RA VDESS),.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning The Ryerson Audio-Visual Database of Emotional Speech and Song (RA VDESS),

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.587583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.418479Z digest=sha256:9b6a99f1eb7a65f8fb86579a90f05d789542078667b5e319271c5b10a9a9e7f7

Observation 55cbc3c2-9c45-482e-8f43-e23149d54465 · outbound

This paper cites Toronto emotional speech set (tess)-younger talker happy,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Toronto emotional speech set (tess)-younger talker happy,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.580223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.420933Z digest=sha256:a9f673e9f986770a5c35c4309e873f2baf47eedecb4567f0fe194971294ed5b7

Observation b01a99ee-2eac-4821-b6cc-aa66c965a4e4 · outbound

This paper cites SA VEE: Surrey Audio-Visual Expressed Emotion,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning SA VEE: Surrey Audio-Visual Expressed Emotion,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.572715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.423438Z digest=sha256:bb6d0e1241f55d446aaf57cbc1b9316bd734d5d55d411ed29cfd4a62b192e83e

Observation 75b1e647-c221-40d3-9a69-e280d3211535 · outbound

This paper cites DST: Deformable Speech Transformer for Emotion Recognition,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning DST: Deformable Speech Transformer for Emotion Recognition,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.564895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.426195Z digest=sha256:3ef3292e6ac505660b53afdfb4053bdb1ae80232ef736dc84737c87745b55167

Observation 5f541385-8e8c-430e-b204-058de5de6653 · outbound

This paper cites Tem- poral modeling matters: A novel temporal emotional modeling approach,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Tem- poral modeling matters: A novel temporal emotional modeling approach,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.556456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.429206Z digest=sha256:344296ff8f8bf3dba62dd7bf95c076cd833615f1b44918ddfbbaf954ba4067d5

Observation c969f048-ef1b-4930-b924-17b2e368aec5 · outbound

This paper cites The ryerson audio-visual database of emotional speech and song (ravdess): A dynamic, multimodal set of facial and vocal expressions in north american english,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning The ryerson audio-visual database of emotional speech and song (ravdess): A dynamic, multimodal set of facial and vocal expressions in north american english,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.431567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.431567Z digest=sha256:7e3c684a7b2239be9c4693d4c236113036fae618121bf19393e8d75cd3fcf9f4

Observation 129db151-23ea-4960-96d6-925a35c3d28f · outbound

This paper cites Speaker-dependent audio- visual emotion recognition.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Speaker-dependent audio- visual emotion recognition

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.543120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.434143Z digest=sha256:e46c53e44b2c2d9961baa4b2cab20c783414d271a52db96dc1c06f59cd8ff987

Observation 4dd6b5c7-e19c-4e85-bd4a-1ff104e030e5 · outbound

This paper cites Panns: Large-scale pretrained audio neural networks for audio pattern recognition,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Panns: Large-scale pretrained audio neural networks for audio pattern recognition,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.535020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:01:29.436495Z digest=sha256:4622e352e5f13aff7347cf63033f88fc9144e529b86b5966bc9b50dd0e2b288f

Observation 2d292fc6-b95e-4061-932f-ee7dd5174719 · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.438826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.438826Z digest=sha256:37067f396b0acd35dbfa43f7dee5e0df52793b8eb438b36bfc7c9741096e1421

Pith citing papers

Observation ad2f2715-09be-4f1c-89e4-ceed8bf8eea9 · inbound

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning cites this paper.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.331354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.331354Z digest=sha256:f4ac4a1117b9b06155c2f5e7e46a2fcb131dc2570e255f3da048ec2c966d2234

Observation 9226301f-dbc6-4bb4-a138-ba43f41028d8 · inbound

AffectCodec: Emotion-Preserving Neural Speech Codec for Expressive Speech Modeling cites this paper.

AffectCodec: Emotion-Preserving Neural Speech Codec for Expressive Speech Modeling CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:07:00.273850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T01:04:54.506749Z digest=sha256:9febf167e4a06c9849d4ec7711666eaa5d710d52ea02732e06f833da3c886228