Pith. sign in

Paper Citation Record · LEDGER

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning

As of 20 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2412.20707.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.20707 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T23:17:22.404354Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

38 of 38 outbound references displayed

  • verified exact2
  • verified fuzzy33
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7e4890a5-752f-4351-86bb-1063ac2411dc · outbound

This paper cites A multilingual framework based on pre-training model for speech emotion recognition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning A multilingual framework based on pre-training model for speech emotion recognition,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.236730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.214287Z digest=sha256:48ffb1905fef477fe9d1dcf4affed5ef48b734fa6d61f5bdc5423745f6ba9de2

Observation 7a8dd181-99c8-48f1-b14e-9c642ed57327 · outbound

This paper cites Speech emotion recognition using self-supervised features,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Speech emotion recognition using self-supervised features,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.216513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.220243Z digest=sha256:238a688b74f1c7b000c5dce1422c7cc2c6f74a77da0bb87727f083619a618f58

Observation 0f0fbd59-fd4f-4a89-9376-7e87bb499867 · outbound

This paper cites Multimodal Emotion Recognition using Transfer Learning from Speaker Recognition and BERT-based models.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Multimodal Emotion Recognition using Transfer Learning from Speaker Recognition and BERT-based models

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-10T23:17:22.500873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.225885Z digest=sha256:0678f811e51a576ec7addacad371bf1c8648efac7aa6d041dff95ce5cfe4353a

Observation 79fdec54-7ebf-4b23-bcb9-c73f79528433 · outbound

This paper cites Two-stage finetuning of wav2vec 2.0 for speech emotion recognition with ASR and gender pretraining,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Two-stage finetuning of wav2vec 2.0 for speech emotion recognition with ASR and gender pretraining,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.195810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.231782Z digest=sha256:74042f83830c8fb7a94d168b5fd40636750d6bf2f9ea6ccd536e76c28100eb38

Observation 6dbf03d7-0821-497e-abc5-a7c3cc6292a3 · outbound

This paper cites A critical review of state-of-the- art chatbot designs and applications,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning A critical review of state-of-the- art chatbot designs and applications,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.168237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.237201Z digest=sha256:b136b90c31a764844d9d1b37755f37be62c98836cd1c8b2d886486fa87bebaf5

Observation 710fe9a4-5f3d-4d00-b5c2-de1b00e836b6 · outbound

This paper cites 5IDER: Unified query rewriting for steering, intent carryover, disfluencies, entity carryover and repair,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning 5IDER: Unified query rewriting for steering, intent carryover, disfluencies, entity carryover and repair,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.145921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.242770Z digest=sha256:58dd3ba5bc776db4b65fadec573587a7df27f01807d862f1da02de0a037589b5

Observation 07df7389-5682-4096-8ac9-ba6e29496c3b · outbound

This paper cites Cross- lingual/cross-channel intent detection in contact-center conversations,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Cross- lingual/cross-channel intent detection in contact-center conversations,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.125229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.248415Z digest=sha256:02b0041251916cf0968ee1e11f6ea63662dce26b6128af1cc04423b94c2e5c14

Observation 256dc594-4132-4e5b-8010-d8b773311bc8 · outbound

This paper cites Automated neural nursing assistant (ANNA): An over-the-phone system for cogni- tive monitoring,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Automated neural nursing assistant (ANNA): An over-the-phone system for cogni- tive monitoring,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.104554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.253302Z digest=sha256:18541c0c35a3fef3ef7bf6980630bb2b7e17f1d8dab6f6c52668f8317a945872

Observation b93b4a2d-fa71-4f20-8695-dbeb479022a6 · outbound

This paper cites Whisper-based transfer learning for alzheimer disease classification: Leveraging speech segments with full transcripts as prompts,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Whisper-based transfer learning for alzheimer disease classification: Leveraging speech segments with full transcripts as prompts,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.085668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.258651Z digest=sha256:5dd8e69655610ef9b5f37d4d86b0d50fe87ae050cc4b3e28dd3836859f843e50

Observation 8985799d-a91e-4e82-b842-9c51ebce62b9 · outbound

This paper cites Cross-lingual alzheimer’s disease detection based on paralinguistic and pre-trained features,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Cross-lingual alzheimer’s disease detection based on paralinguistic and pre-trained features,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.063241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.263723Z digest=sha256:133aab4faa7977c5a5181c67641cd0a5eabfaea6df7b5bafd19e2debd0a1c0e5

Observation 3a9f7c62-16b3-470f-a5d0-0001e400133e · outbound

This paper cites Exploiting emotion information in speaker embeddings for expressive text-to-speech,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Exploiting emotion information in speaker embeddings for expressive text-to-speech,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.040858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.268653Z digest=sha256:f77e712e18d154747b7576797a39c2cad05d1aab0afcdeb183f97414c5ed5643

Observation b53dbe3c-341f-4a2d-8fd5-7bdff34ad2b0 · outbound

This paper cites Fusing ASR outputs in joint training for speech emotion recognition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Fusing ASR outputs in joint training for speech emotion recognition,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.016350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.273899Z digest=sha256:3468082050a471d19db6281408d3be765db9c491ad8bd468a74bf76bded6276d

Observation dbc7a5f6-9e06-4bad-aa56-096815562573 · outbound

This paper cites Speaker-aware training of speech emotion classifier with speaker recognition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Speaker-aware training of speech emotion classifier with speaker recognition,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.994690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.278978Z digest=sha256:a65d9a72bb5dda0470f897bf3272412b84e62170d64756c59dcadd3146d164cb

Observation 033dcf7b-d786-4ddf-95ec-adb14de1245d · outbound

This paper cites Speech emotion recognition combining acoustic features and linguistic information in a hybrid support vector machine-belief network architecture,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Speech emotion recognition combining acoustic features and linguistic information in a hybrid support vector machine-belief network architecture,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.976220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.283316Z digest=sha256:4a18fd17ccbd52c7736c3171cfd00463dd9e080c72f668fab97ddb2c16ee94fe

Observation 9624075a-b886-457c-8280-683a005e5be0 · outbound

This paper cites Mmer: Multimodal multi-task learning for speech emotion recognition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Mmer: Multimodal multi-task learning for speech emotion recognition,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.955535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.288098Z digest=sha256:5d13da3b157c26da9af044a15f17baaae7fad30c8a0105b1de8dbe57e7022d63

Observation 6e475d38-d491-4e50-84f0-7a1bd40f4e49 · outbound

This paper cites Speech emotion recognition using decomposed speech via multi-task learning,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Speech emotion recognition using decomposed speech via multi-task learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.931812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.292405Z digest=sha256:9a5fa7e56908160028f99eb9b7ee93c062f8c56bbebee3f920748cc0bfe26912

Observation 370fad4d-56e1-4f10-bb79-27d6e2e5ccbf · outbound

This paper cites Multi-task learning based end-to-end speaker recognition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Multi-task learning based end-to-end speaker recognition,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.912334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.297144Z digest=sha256:b3c79ab1a707fe4176fb7e8b15a7b3b4217549423527d57c286d90647a9d84e5

Observation 029539ea-7a8b-4d82-9485-fbc6fc9dfa7d · outbound

This paper cites Mutitask learning based muti-examples keywords spotting in low resource condition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Mutitask learning based muti-examples keywords spotting in low resource condition,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.892600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.301649Z digest=sha256:6b98db7505a12c01920184d8a368859ab26f1bc2d102779ebdb2cdc359405f8d

Observation 692e9aec-eb37-4f04-87f7-149c263a8c96 · outbound

This paper cites MMER: Multimodal multi-task learning for speech emotion recogni- tion,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning MMER: Multimodal multi-task learning for speech emotion recogni- tion,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.869114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.305904Z digest=sha256:bc45be21e44831ca0e3bef692c63b869f4b3a9cfb30a7e687c65f3d2bb37ed86

Observation c8a87b47-d21e-4ce1-9a57-baac1f7080d0 · outbound

This paper cites A review of speech emotion recognition: Datasets, features, and machine learning algorithms,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning A review of speech emotion recognition: Datasets, features, and machine learning algorithms,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.849947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.310575Z digest=sha256:270d2010b3943726f1702ebc72933648a778d5f24bb8e19b853e0f0fdf759c50

Observation a4398042-aed0-456f-9c08-9d63cb487125 · outbound

This paper cites Gra- dient surgery for multi-task learning,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Gra- dient surgery for multi-task learning,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.828679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.315120Z digest=sha256:09c1b82c5069e084b969bd533742c67002d6cd7e0b3e684f6ccff6539c618cef

Observation 50eb2c70-50a2-437c-b745-c271dadcd72c · outbound

This paper cites A Brief Review of Deep Multi-task Learning and Auxiliary Task Learning.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning A Brief Review of Deep Multi-task Learning and Auxiliary Task Learning

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-10T23:17:22.473148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.320384Z digest=sha256:1820f247685dce78a105af0671c80814e69cba5bd1215a85abaec10660bad2b1

Observation b3e33fbf-3f5f-43fb-8183-7085f5148971 · outbound

This paper cites Exploring large scale pre-trained models for robust machine anomalous sound detection,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Exploring large scale pre-trained models for robust machine anomalous sound detection,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.811053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.325674Z digest=sha256:14a61fe0eff6397f08c439e356af931be6885bddd9d78a5bd2b84bbaa06e904a

Observation 6b178ba7-f6d3-4aee-8c32-56a30ee016dd · outbound

This paper cites Ef- ficient multi-task auxiliary learning: Selecting auxiliary data by feature similarity,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Ef- ficient multi-task auxiliary learning: Selecting auxiliary data by feature similarity,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.792409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.330852Z digest=sha256:4024fc2972f3e21cc83504b01a457796f32153cef5241d5be2829d1fd080babf

Observation cd313358-5553-4dad-b853-c26c2a1b9c43 · outbound

This paper cites Towards discriminative representations and unbiased predictions: Class-specific angular softmax for speech emotion recognition.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Towards discriminative representations and unbiased predictions: Class-specific angular softmax for speech emotion recognition

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.772343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.336168Z digest=sha256:1c34a964883127624db6b41c451ff5ec4b35e8927aabcf71ff56dfb67684b25b

Observation abb88b1c-2745-4d6e-a0cf-9712a643f370 · outbound

This paper cites HuBERT: Self-supervised speech representation learning by masked prediction of hidden units,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning HuBERT: Self-supervised speech representation learning by masked prediction of hidden units,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.752145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.341020Z digest=sha256:d84375b9913dbec3cd7ac2647bab2ccc5bc364c7221fdc8f14b5bce772c55698

Observation 1b776ba8-d3fd-4fbe-8d87-9717ece76e1e · outbound

This paper cites WavLM: Large-scale self-supervised pre- training for full stack speech processing,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning WavLM: Large-scale self-supervised pre- training for full stack speech processing,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T23:17:22.346074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:17:22.346074Z digest=sha256:94d6124c8ecaf8b60dedfccd267048804c4f8a76e8bd0e2e8cb2d3a92b602cb7

Observation b18588ce-ea09-4d63-be53-4c8701ac5d8c · outbound

This paper cites Connectionist temporal classification: labelling unsegmented sequence data with re- current neural networks,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Connectionist temporal classification: labelling unsegmented sequence data with re- current neural networks,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.713652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.350923Z digest=sha256:14db52f6243eaf714d5a915fe270c46cc7f7f4e31cc6ea516e51ace1d3570d49

Observation cccd2cc8-763b-49c1-9b5d-7c8e048f58c7 · outbound

This paper cites IEMOCAP: Interactive emotional dyadic motion capture database,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning IEMOCAP: Interactive emotional dyadic motion capture database,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.694026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.355879Z digest=sha256:60613b761faea3309b1d990e2e0f2786d79eb7d276ebfa054456559730b86c4f

Observation e37e29a3-5858-466d-a04a-79493be0a9d4 · outbound

This paper cites Mingling or misalignment? temporal shift for speech emotion recognition with pre-trained representations,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Mingling or misalignment? temporal shift for speech emotion recognition with pre-trained representations,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.672207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.360783Z digest=sha256:7fa33831fa485ee7d751b2ad958a114869e0461679c0f5f485c2678bd14ed1d1

Observation 30a51ffa-d905-4a99-84e0-b44b1a954969 · outbound

This paper cites Temporal modeling matters: A novel temporal emotional modeling approach for speech emotion recognition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Temporal modeling matters: A novel temporal emotional modeling approach for speech emotion recognition,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.651732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.365523Z digest=sha256:1c073bbacec87aab67f3d4d84b9a1e616b0dafaa0d0e0b6b6003ed5a4200051a

Observation d32ad3da-e89f-49a6-9041-82bc06da6272 · outbound

This paper cites DST: Deformable speech transformer for emotion recognition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning DST: Deformable speech transformer for emotion recognition,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.632156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.370421Z digest=sha256:0a32e0cc8fccb94a67a9960ba64e364c2ef4001edeb819f42e9b92fb90cd2811

Observation e909e910-155c-4bfc-86ff-9e951f380b52 · outbound

This paper cites DWFormer: Dynamic window transformer for speech emotion recognition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning DWFormer: Dynamic window transformer for speech emotion recognition,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.606479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.375867Z digest=sha256:5dc7bf597ebe8700f1c590661581e466ec067f0e4b821ce6b23947767dd4f268

Observation e96bf5fa-1564-44b2-bfa4-405c7d08e299 · outbound

This paper cites Multiple acoustic features speech emotion recognition using cross-attention transformer,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Multiple acoustic features speech emotion recognition using cross-attention transformer,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.585148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.380966Z digest=sha256:0cb6b1f518309d69d52311385838728e857a66217ad999c521f160485dbd6f20

Observation 70ad94d0-280f-49f8-bb53-8246e3f93987 · outbound

This paper cites PyTorch: An imperative style, high-performance deep learning library,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning PyTorch: An imperative style, high-performance deep learning library,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.562507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.385801Z digest=sha256:c95c7c2c2846caf24e81e1db19be1936d408fb2fb7864b1c61332581d019245c

Observation f49a7b98-3f6d-4ddc-a411-d9134287950a · outbound

This paper cites SpeechBrain: A General-Purpose Speech Toolkit.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning SpeechBrain: A General-Purpose Speech Toolkit

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T23:17:22.392521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:17:22.392521Z digest=sha256:e0217fb01db4084b1dae0369d80fba6c2890fffe881f1b0a057835bd8a7e8889

Observation 04aa0dd0-e364-402f-8a16-fa1fe7d14b07 · outbound

This paper cites Improving automatic speech recognition performance for low-resource languages with self-supervised models,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Improving automatic speech recognition performance for low-resource languages with self-supervised models,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T23:17:22.399386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:17:22.399386Z digest=sha256:17b2d8438f9fab2b2d4dba253d3a6efb48c91efab7522146f71394d721fd5e5c

Observation 7640c6a6-a954-4852-8296-f323df04bd98 · outbound

This paper cites Librispeech: an ASR corpus based on public domain audio books,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Librispeech: an ASR corpus based on public domain audio books,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.522365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T23:17:22.404354Z digest=sha256:97f6cdc0c6dc03d5c2964cef448b0e39f19805cb93253b7856065e60dace0bf4

Pith citing papers

No inbound Pith citation observations are available.