Pith. sign in

Paper Citation Record · LEDGER

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning

As of 20 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2412.20707.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.20707 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T23:17:22.404354Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

38 of 38 outbound references displayed

  • verified exact2
  • verified fuzzy33
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7e4890a5-752f-4351-86bb-1063ac2411dc · outbound

This paper cites A multilingual framework based on pre-training model for speech emotion recognition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning A multilingual framework based on pre-training model for speech emotion recognition,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.236730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.214287Z digest=sha256:0f5ec152f70a46599d23bb0c4da38ed5ee75501301433c38395dd65e289bf7a9

Observation 7a8dd181-99c8-48f1-b14e-9c642ed57327 · outbound

This paper cites Speech emotion recognition using self-supervised features,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Speech emotion recognition using self-supervised features,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.216513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.220243Z digest=sha256:47a1900d2eb752be9d66a9a740d9805e2d9b7e2e78be3b07e6c34470e4ea53a4

Observation 0f0fbd59-fd4f-4a89-9376-7e87bb499867 · outbound

This paper cites Multimodal Emotion Recognition using Transfer Learning from Speaker Recognition and BERT-based models.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Multimodal Emotion Recognition using Transfer Learning from Speaker Recognition and BERT-based models

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-10T23:17:22.500873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.225885Z digest=sha256:0f83de85bb40529b88fe740cabb311b7e0b7d611c55df6661a5c897b0f37e771

Observation 79fdec54-7ebf-4b23-bcb9-c73f79528433 · outbound

This paper cites Two-stage finetuning of wav2vec 2.0 for speech emotion recognition with ASR and gender pretraining,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Two-stage finetuning of wav2vec 2.0 for speech emotion recognition with ASR and gender pretraining,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.195810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.231782Z digest=sha256:7b8e5885fca87a893a0f19e13aa4cac57f7f1cead50024abcb46ef4ffc2927a2

Observation 6dbf03d7-0821-497e-abc5-a7c3cc6292a3 · outbound

This paper cites A critical review of state-of-the- art chatbot designs and applications,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning A critical review of state-of-the- art chatbot designs and applications,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.168237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.237201Z digest=sha256:7a8efb5884f99d2d39d3752d875ca425d264cf0e532fb88805b598be4f9797d6

Observation 710fe9a4-5f3d-4d00-b5c2-de1b00e836b6 · outbound

This paper cites 5IDER: Unified query rewriting for steering, intent carryover, disfluencies, entity carryover and repair,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning 5IDER: Unified query rewriting for steering, intent carryover, disfluencies, entity carryover and repair,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.145921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.242770Z digest=sha256:b4a085c717ee83eb4543f07320aba0a1096cd3d0135c5e11c159bc920bf9b800

Observation 07df7389-5682-4096-8ac9-ba6e29496c3b · outbound

This paper cites Cross- lingual/cross-channel intent detection in contact-center conversations,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Cross- lingual/cross-channel intent detection in contact-center conversations,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.125229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.248415Z digest=sha256:96bfb6682907efce8171d58fb061c162bdde6dbfa6cb2a2af43e4e5d08162367

Observation 256dc594-4132-4e5b-8010-d8b773311bc8 · outbound

This paper cites Automated neural nursing assistant (ANNA): An over-the-phone system for cogni- tive monitoring,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Automated neural nursing assistant (ANNA): An over-the-phone system for cogni- tive monitoring,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.104554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.253302Z digest=sha256:1f52eb623280bf9f53277c065cff7f4cb24a911194448fb4e8590abdf8bcd2b9

Observation b93b4a2d-fa71-4f20-8695-dbeb479022a6 · outbound

This paper cites Whisper-based transfer learning for alzheimer disease classification: Leveraging speech segments with full transcripts as prompts,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Whisper-based transfer learning for alzheimer disease classification: Leveraging speech segments with full transcripts as prompts,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.085668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.258651Z digest=sha256:143cfdb0bc897e4df0d8a60b68f546f8b9358a317e0586df4820941675a891a9

Observation 8985799d-a91e-4e82-b842-9c51ebce62b9 · outbound

This paper cites Cross-lingual alzheimer’s disease detection based on paralinguistic and pre-trained features,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Cross-lingual alzheimer’s disease detection based on paralinguistic and pre-trained features,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.063241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.263723Z digest=sha256:27ad2b898a1670e301aa8254fc8eba4a5b1c901f5fdb087143e5bda54b3fef47

Observation 3a9f7c62-16b3-470f-a5d0-0001e400133e · outbound

This paper cites Exploiting emotion information in speaker embeddings for expressive text-to-speech,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Exploiting emotion information in speaker embeddings for expressive text-to-speech,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.040858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.268653Z digest=sha256:6f1a244f8c8a132b49f20f330d64eb965a05725fc41839df78d2d986f22a3b12

Observation b53dbe3c-341f-4a2d-8fd5-7bdff34ad2b0 · outbound

This paper cites Fusing ASR outputs in joint training for speech emotion recognition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Fusing ASR outputs in joint training for speech emotion recognition,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:23.016350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.273899Z digest=sha256:e583f4ce811fc567e85ebd4f199aade81872717dcdcbd29463f26c85ee0e3bb1

Observation dbc7a5f6-9e06-4bad-aa56-096815562573 · outbound

This paper cites Speaker-aware training of speech emotion classifier with speaker recognition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Speaker-aware training of speech emotion classifier with speaker recognition,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.994690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.278978Z digest=sha256:eeafe91faea2edcb8e066770d3d164403eb5b18b6a0aa43b2a86b8222100bde9

Observation 033dcf7b-d786-4ddf-95ec-adb14de1245d · outbound

This paper cites Speech emotion recognition combining acoustic features and linguistic information in a hybrid support vector machine-belief network architecture,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Speech emotion recognition combining acoustic features and linguistic information in a hybrid support vector machine-belief network architecture,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.976220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.283316Z digest=sha256:39911874acc3097fae5685f733cfd3f1294efc753fa50b5ee12194c1f80a0f07

Observation 9624075a-b886-457c-8280-683a005e5be0 · outbound

This paper cites Mmer: Multimodal multi-task learning for speech emotion recognition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Mmer: Multimodal multi-task learning for speech emotion recognition,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.955535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.288098Z digest=sha256:af6d2a806fcf1e5b08b82fc854453108bfb0c390ff06546935aff2f0d43d4d86

Observation 6e475d38-d491-4e50-84f0-7a1bd40f4e49 · outbound

This paper cites Speech emotion recognition using decomposed speech via multi-task learning,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Speech emotion recognition using decomposed speech via multi-task learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.931812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.292405Z digest=sha256:f984beff5cd59c125ff86a8cf41bb788e71c98b5e0750fc6ed52cbef0b71317c

Observation 370fad4d-56e1-4f10-bb79-27d6e2e5ccbf · outbound

This paper cites Multi-task learning based end-to-end speaker recognition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Multi-task learning based end-to-end speaker recognition,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.912334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.297144Z digest=sha256:7f7d873a1348106995a7114eca0028b77f83fc22092169e0837e17ee3efe8f56

Observation 029539ea-7a8b-4d82-9485-fbc6fc9dfa7d · outbound

This paper cites Mutitask learning based muti-examples keywords spotting in low resource condition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Mutitask learning based muti-examples keywords spotting in low resource condition,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.892600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.301649Z digest=sha256:9a78e109f8b2a011ecfe0d3386c158e2ab3ee4ef43c4fcc18a39fa0305a09180

Observation 692e9aec-eb37-4f04-87f7-149c263a8c96 · outbound

This paper cites MMER: Multimodal multi-task learning for speech emotion recogni- tion,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning MMER: Multimodal multi-task learning for speech emotion recogni- tion,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.869114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.305904Z digest=sha256:9513ec9d68f74c87954782289b22c6b47073435c2374af891108aca082729a9b

Observation c8a87b47-d21e-4ce1-9a57-baac1f7080d0 · outbound

This paper cites A review of speech emotion recognition: Datasets, features, and machine learning algorithms,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning A review of speech emotion recognition: Datasets, features, and machine learning algorithms,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.849947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.310575Z digest=sha256:da5ce5add2a661c43ee0fdf90420cb6d2a3efbfbe3eb1203a1187c7db7e9e51e

Observation a4398042-aed0-456f-9c08-9d63cb487125 · outbound

This paper cites Gra- dient surgery for multi-task learning,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Gra- dient surgery for multi-task learning,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.828679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.315120Z digest=sha256:ed28abd40923e6d498fc127c79fd36cd3981cf54a80eef6852fb8ac6e8364b91

Observation 50eb2c70-50a2-437c-b745-c271dadcd72c · outbound

This paper cites A Brief Review of Deep Multi-task Learning and Auxiliary Task Learning.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning A Brief Review of Deep Multi-task Learning and Auxiliary Task Learning

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-10T23:17:22.473148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.320384Z digest=sha256:4fe05fb175d8d9f43029e9faad2d2f713c337e76bf1aabae87e59524ed7e987b

Observation b3e33fbf-3f5f-43fb-8183-7085f5148971 · outbound

This paper cites Exploring large scale pre-trained models for robust machine anomalous sound detection,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Exploring large scale pre-trained models for robust machine anomalous sound detection,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.811053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.325674Z digest=sha256:5cf3d2876c3d3111ed25cd76b585628f5c1f01d8021410860e3ca4d534dd827c

Observation 6b178ba7-f6d3-4aee-8c32-56a30ee016dd · outbound

This paper cites Ef- ficient multi-task auxiliary learning: Selecting auxiliary data by feature similarity,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Ef- ficient multi-task auxiliary learning: Selecting auxiliary data by feature similarity,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.792409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.330852Z digest=sha256:de4453ba81885dbc11bf2f9cd028306b1e5bd3eddac37f70bb464dbab8dc2bcd

Observation cd313358-5553-4dad-b853-c26c2a1b9c43 · outbound

This paper cites Towards discriminative representations and unbiased predictions: Class-specific angular softmax for speech emotion recognition.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Towards discriminative representations and unbiased predictions: Class-specific angular softmax for speech emotion recognition

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.772343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.336168Z digest=sha256:dfa1388af03886bf8d531e79b0789d3151ce968e92112f2f7a9bb2e028d5fa5d

Observation abb88b1c-2745-4d6e-a0cf-9712a643f370 · outbound

This paper cites HuBERT: Self-supervised speech representation learning by masked prediction of hidden units,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning HuBERT: Self-supervised speech representation learning by masked prediction of hidden units,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.752145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.341020Z digest=sha256:66002f8b5b1776c5ff9d4748d5179dbcd886c2112a827a618b0efb9c5ad29bf3

Observation 1b776ba8-d3fd-4fbe-8d87-9717ece76e1e · outbound

This paper cites WavLM: Large-scale self-supervised pre- training for full stack speech processing,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning WavLM: Large-scale self-supervised pre- training for full stack speech processing,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T23:17:22.346074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:17:22.346074Z digest=sha256:94d6124c8ecaf8b60dedfccd267048804c4f8a76e8bd0e2e8cb2d3a92b602cb7

Observation b18588ce-ea09-4d63-be53-4c8701ac5d8c · outbound

This paper cites Connectionist temporal classification: labelling unsegmented sequence data with re- current neural networks,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Connectionist temporal classification: labelling unsegmented sequence data with re- current neural networks,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.713652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.350923Z digest=sha256:64b6a335297f5e43f5d60b31235be863648aff0a21e6b85e8d0a8016abe1bc70

Observation cccd2cc8-763b-49c1-9b5d-7c8e048f58c7 · outbound

This paper cites IEMOCAP: Interactive emotional dyadic motion capture database,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning IEMOCAP: Interactive emotional dyadic motion capture database,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.694026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.355879Z digest=sha256:46cbe3c68e77e86adbd6b39e677571180f73cb8b4fe659dc86ec3100c3a3c0af

Observation e37e29a3-5858-466d-a04a-79493be0a9d4 · outbound

This paper cites Mingling or misalignment? temporal shift for speech emotion recognition with pre-trained representations,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Mingling or misalignment? temporal shift for speech emotion recognition with pre-trained representations,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.672207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.360783Z digest=sha256:e78e47225ffad6f9ca46dcfdf3f719bd1fa0131544f6af8dc61bc42ce8dc7451

Observation 30a51ffa-d905-4a99-84e0-b44b1a954969 · outbound

This paper cites Temporal modeling matters: A novel temporal emotional modeling approach for speech emotion recognition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Temporal modeling matters: A novel temporal emotional modeling approach for speech emotion recognition,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.651732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.365523Z digest=sha256:61b7108c6b318100116401bead05b4c4023c01839e4a7ed1886fbca0c99767aa

Observation d32ad3da-e89f-49a6-9041-82bc06da6272 · outbound

This paper cites DST: Deformable speech transformer for emotion recognition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning DST: Deformable speech transformer for emotion recognition,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.632156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.370421Z digest=sha256:5153f072c75d1f146135592c98652096da46e8017464c32efae08507b0b0dfda

Observation e909e910-155c-4bfc-86ff-9e951f380b52 · outbound

This paper cites DWFormer: Dynamic window transformer for speech emotion recognition,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning DWFormer: Dynamic window transformer for speech emotion recognition,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.606479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.375867Z digest=sha256:3d8e861f13a018eadc98181c050e37873078a41993e9f56eaf7d73dbe1858e83

Observation e96bf5fa-1564-44b2-bfa4-405c7d08e299 · outbound

This paper cites Multiple acoustic features speech emotion recognition using cross-attention transformer,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Multiple acoustic features speech emotion recognition using cross-attention transformer,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.585148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.380966Z digest=sha256:c14002f4a139b8f1e40347737b3f7de079384e412563df9d88afc5e224c0fc34

Observation 70ad94d0-280f-49f8-bb53-8246e3f93987 · outbound

This paper cites PyTorch: An imperative style, high-performance deep learning library,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning PyTorch: An imperative style, high-performance deep learning library,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.562507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.385801Z digest=sha256:7482f93e25e2b512c1acddacbc8b43ee4554db2952cf2c8e550f9085844794af

Observation f49a7b98-3f6d-4ddc-a411-d9134287950a · outbound

This paper cites SpeechBrain: A General-Purpose Speech Toolkit.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning SpeechBrain: A General-Purpose Speech Toolkit

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T23:17:22.392521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:17:22.392521Z digest=sha256:e0217fb01db4084b1dae0369d80fba6c2890fffe881f1b0a057835bd8a7e8889

Observation 04aa0dd0-e364-402f-8a16-fa1fe7d14b07 · outbound

This paper cites Improving automatic speech recognition performance for low-resource languages with self-supervised models,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Improving automatic speech recognition performance for low-resource languages with self-supervised models,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T23:17:22.399386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:17:22.399386Z digest=sha256:17b2d8438f9fab2b2d4dba253d3a6efb48c91efab7522146f71394d721fd5e5c

Observation 7640c6a6-a954-4852-8296-f323df04bd98 · outbound

This paper cites Librispeech: an ASR corpus based on public domain audio books,.

Metadata-Enhanced Speech Emotion Recognition: Augmented Residual Integration and Co-Attention in Two-Stage Fine-Tuning Librispeech: an ASR corpus based on public domain audio books,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:17:22.522365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T23:17:22.404354Z digest=sha256:c86243e9a99428f6b83df52e6f7562a3f2ad7ff1aa0c92dfffe9f5248f5e0ee7

Pith citing papers

No inbound Pith citation observations are available.