Pith. sign in

Paper Citation Record · LEDGER

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation

As of 7 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2508.07302.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.07302 v2

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T22:17:38.566540Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-18T16:27:37.596817Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T16:31:37.259420Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact4
  • verified fuzzy18
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 747a176c-8695-4ef4-a090-5eede55c3179 · outbound

This paper cites FastSpeech 2: Fast and High-Quality End-to-End Text to Speech.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation FastSpeech 2: Fast and High-Quality End-to-End Text to Speech

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.247156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.247156Z digest=sha256:876de7c6b69739ef4090768e7fc9b43e8bf1dadcd1e8a2a37f8070fb11e57d76

Observation cb5c16d8-bd7b-449b-8519-f84e610c92a5 · outbound

This paper cites AdaSpeech: Adaptive Text to Speech for Custom Voice.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation AdaSpeech: Adaptive Text to Speech for Custom Voice

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.258146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.258146Z digest=sha256:5aac338c911265da7b6a770fe9047fc86c2e6534575154e72e9997906c4009f1

Observation d3080bf9-3ac0-430d-91a3-617464b9f6d1 · outbound

This paper cites Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.858055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.274410Z digest=sha256:dd185a288a8595976018b4d3087dfa53d6ecd378b33ddbd8d9c5bb0ee4d0829d

Observation 5ec9cfc4-6a90-424a-9aa4-a7ac5a171124 · outbound

This paper cites Generspeech: Towards style transfer for generalizable out-of-domain text-to-speech,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Generspeech: Towards style transfer for generalizable out-of-domain text-to-speech,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.826565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.291568Z digest=sha256:0f8f112643c96f4d8252c5931482b4e0f8bf3513a6836e5071238c6df8cdf0f7

Observation 64a33c59-d966-4a47-b022-18342097d1b9 · outbound

This paper cites Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.301601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.301601Z digest=sha256:3c59ab8e9e584434b81fb0f1e9abfa32f8cfba574673bff575585ad040c4e9e7

Observation e7f00805-43b9-4391-b1f9-c4d742d2f610 · outbound

This paper cites XTTS: a Massively Multilingual Zero-Shot Text-to-Speech Model.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation XTTS: a Massively Multilingual Zero-Shot Text-to-Speech Model

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.311772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.311772Z digest=sha256:6fdddb1b4d458ac567f2f72896285c571829f78b9de7c1f747bea1f41c71c43e

Observation 59f5b7dc-2217-4e9f-bb9a-e6bfae52afd9 · outbound

This paper cites Zero-shot Cross-lingual Voice Transfer for TTS.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Zero-shot Cross-lingual Voice Transfer for TTS

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-05T22:17:39.030809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.318698Z digest=sha256:1bca0e2f32be99fc0a1c471e2efe30b0b7e936ae1f50a7d2f97cb8ce3c3db6e3

Observation 605ed599-7d83-4687-81d3-bacee00df06c · outbound

This paper cites Multilingual video dubbing—a tech- nology review and current challenges,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Multilingual video dubbing—a tech- nology review and current challenges,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.793134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.325689Z digest=sha256:2ed1f79690133f185a7804c839f81dae275210aa65b6010622eb953185f70d89

Observation 98aae8f0-105e-4709-aba3-f3d9113366d6 · outbound

This paper cites DSE-TTS: Dual Speaker Embedding for Cross-Lingual Text-to-Speech.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation DSE-TTS: Dual Speaker Embedding for Cross-Lingual Text-to-Speech

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-05T22:17:38.977395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.331461Z digest=sha256:7cea49b10202820d4800de6d014a631172a243cefaa14c2b836e9e56a96e1ff2

Observation a34f24ce-f9b9-4e55-9784-b403575276b2 · outbound

This paper cites Diclet-tts: Diffusion model based cross- lingual emotion transfer for text-to-speech – a study between english and mandarin,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Diclet-tts: Diffusion model based cross- lingual emotion transfer for text-to-speech – a study between english and mandarin,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.755773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.341071Z digest=sha256:a90be1661504738d1b45fdf89e60517d4f2b117a33671c337b70bcc8496d4a50

Observation dfd6b0ba-a7d5-4eaf-8b30-2dfe786e5db2 · outbound

This paper cites Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.347773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.347773Z digest=sha256:262d0d772212320edaaf8d7c3813aab7dcc46f9dea6a7ea5324214fb557a62d9

Observation 75e50fdd-80fb-4768-9237-e5027783d363 · outbound

This paper cites DelightfulTTS: The Microsoft Speech Synthesis System for Blizzard Challenge 2021.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation DelightfulTTS: The Microsoft Speech Synthesis System for Blizzard Challenge 2021

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-05T22:17:39.128510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.360877Z digest=sha256:15f55c9debd5bdf01c953e2f128e19d2dc1934ba36ca1c9ca839cab1a014f665

Observation c3cd9acd-ddc0-4b8c-aa05-0f5377f9ae59 · outbound

This paper cites iemotts: Toward robust cross- speaker emotion transfer and control for speech synthesis based on disentanglement between prosody and timbre,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation iemotts: Toward robust cross- speaker emotion transfer and control for speech synthesis based on disentanglement between prosody and timbre,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.723619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.374104Z digest=sha256:a1c635810c25bfd1d4fa4123d35d9d2975ebb69555fc78c84aab383e7f7c1f49

Observation ebd62e70-9ff8-444c-b5ad-d094344832c6 · outbound

This paper cites Zero-shot emotion transfer for cross-lingual speech synthesis,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Zero-shot emotion transfer for cross-lingual speech synthesis,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.694834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.382804Z digest=sha256:e5219268c4f2f6d4ed49407bdd50b50b0597994832e1d0cc57c9f755ddd02f8d

Observation ffc56a9f-a356-4bce-977d-4e0b7e8ef712 · outbound

This paper cites Zet- speech: Zero-shot adaptive emotion-controllable text-to-speech synthesis with diffusion and style-based models,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Zet- speech: Zero-shot adaptive emotion-controllable text-to-speech synthesis with diffusion and style-based models,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.668409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.396075Z digest=sha256:e889cf6869579ef4ea7b7fdfa447604ced020f911dda721823b6e53860a3e92a

Observation b5f15223-498d-4312-b7d4-6d7df3408c0e · outbound

This paper cites Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.408435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.408435Z digest=sha256:849866967507669b7c750d8641a763cbdfcaa1336a672163d418de357fbfd666

Observation b563c52b-a9a8-4dd0-b931-3a11f9ffc017 · outbound

This paper cites Metts: Multilingual emotional text-to-speech by cross- speaker and cross-lingual emotion transfer,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Metts: Multilingual emotional text-to-speech by cross- speaker and cross-lingual emotion transfer,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.590546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.432824Z digest=sha256:e2858ded421b3d5dd6bc341fb08bece76d8c4e0c61b61fa539b603b3c13bd2f8

Observation c289e77e-be12-49de-854d-42d9081d5e61 · outbound

This paper cites Cosyvoice 3: Towards in-the-wild speech generation via scaling-up and post-training,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Cosyvoice 3: Towards in-the-wild speech generation via scaling-up and post-training,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.550878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.448447Z digest=sha256:5d1cdbdf22f5d371de4d26b29c67fb32edc6d0ed4f83d78bd082409564ce191e

Observation f8971949-c88b-47e5-9cba-070cd6ad9ad7 · outbound

This paper cites Maskgct: Zero-shot text-to-speech with masked generative codec transformer,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Maskgct: Zero-shot text-to-speech with masked generative codec transformer,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.518076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.455700Z digest=sha256:0bbadbdc329bd4f96891c3d32663d90a91517b1f6fea90e4ffd5d0da8071a557

Observation 6d350465-2bb9-4d84-9e98-ca93d880dd75 · outbound

This paper cites Glm-4-voice: Towards intelligent and human-like end-to-end spoken chatbot,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Glm-4-voice: Towards intelligent and human-like end-to-end spoken chatbot,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.465254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.466878Z digest=sha256:a8a484cd9c4e78d1b748268aa8eaacbb691e29e61e265a29a8c89723378c2580

Observation fccb8667-09ba-4ccc-80fd-6853f0519b74 · outbound

This paper cites Speak foreign languages with your own voice: Cross-lingual neural codec language modeling,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Speak foreign languages with your own voice: Cross-lingual neural codec language modeling,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.626151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.476870Z digest=sha256:a5e70fcb8642c4f7341ceec2a7b0f2422337e75623800732b8bd5365d7c5e06e

Observation 5007cd8a-5f5a-4b75-a446-eaf0afcbbef1 · outbound

This paper cites Retrieval- augmented generation for knowledge-intensive nlp tasks,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Retrieval- augmented generation for knowledge-intensive nlp tasks,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.405977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.486028Z digest=sha256:db5bc99336a43658954f36a01b56ad224b33ecb6007c694e04e86837681d99b4

Observation 4c753727-01f5-46c9-b865-9e959a85ceaa · outbound

This paper cites AutoStyle-TTS: Retrieval-Augmented Generation based Automatic Style Matching Text-to-Speech Synthesis.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation AutoStyle-TTS: Retrieval-Augmented Generation based Automatic Style Matching Text-to-Speech Synthesis

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-05T22:17:38.830913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.493215Z digest=sha256:5a4431867395c60a8d28ddf6e90fe60911928788db9798be951bf055d3475e1f

Observation 1e42d927-5c44-4fff-a769-f3f74672a38c · outbound

This paper cites Wavrag: Audio-integrated retrieval augmented generation for spoken dialogue models,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Wavrag: Audio-integrated retrieval augmented generation for spoken dialogue models,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.343676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.503618Z digest=sha256:73010fa3c67454f7fceb79d7c27f1e38cf554a024c10d2a0b21acd977255306b

Observation cccfc76b-36da-420c-b2e4-c4131e596751 · outbound

This paper cites Codec does matter: Exploring the semantic shortcoming of codec for audio language model,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Codec does matter: Exploring the semantic shortcoming of codec for audio language model,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.304422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.511641Z digest=sha256:b2defa42a59c3ecc026547a74adddffda738654931fbdcc7e6dde2d97769736b

Observation ccf1bd42-12ea-41ad-adf4-7b8b96c96d3a · outbound

This paper cites Emo2vec: Learning emotional embeddings via multi-emotion cate- gory,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Emo2vec: Learning emotional embeddings via multi-emotion cate- gory,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.279034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.527270Z digest=sha256:25b18743ab9225f5bf02d47884549f277f85f03f76ed2623ced5f468b10d745e

Observation 41f923b0-f431-41f5-9199-848c59f8635e · outbound

This paper cites Dspgan: A gan-based universal vocoder for high-fidelity tts by time-frequency domain supervision from dsp,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Dspgan: A gan-based universal vocoder for high-fidelity tts by time-frequency domain supervision from dsp,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.259195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.538673Z digest=sha256:eb98998595dd06b2d3db010a082c1170908288b208de20c0ad5900cb9f6aeb72

Observation 673bf918-60b9-4a34-b0c1-4645917f6e50 · outbound

This paper cites Typhoon 2: A Family of Open Text and Multimodal Thai Large Language Models.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Typhoon 2: A Family of Open Text and Multimodal Thai Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.548736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.548736Z digest=sha256:3532f8f69f5f6934674b014899911b6314aba00c1531e60c590d460db5f577c5

Observation d50a04fe-1cb7-4954-97cf-7339f98a8a42 · outbound

This paper cites Wespeaker: A research and production oriented speaker embedding learning toolkit,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Wespeaker: A research and production oriented speaker embedding learning toolkit,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.234880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:17:38.555521Z digest=sha256:cbd470017829b12a75c489c9f9dcf2b4d5e81e43156e7696df7c8771cbc9cc34

Observation 36270e26-3b47-4623-b5ba-0d932e9939a4 · outbound

This paper cites Zipformer: A faster and better encoder for automatic speech recognition.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Zipformer: A faster and better encoder for automatic speech recognition

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.566540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.566540Z digest=sha256:92165731a83054a215f6e5ba438ea664843e2b575536e6ef957837f749909236

Pith citing papers

Observation a8438032-1c43-4fd7-a270-741ef50b734a · inbound

Towards Building Speech Large Language Models for Multitask Understanding in Low-Resource Languages cites this paper.

Towards Building Speech Large Language Models for Multitask Understanding in Low-Resource Languages XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:31:37.262268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T16:27:37.596817Z digest=sha256:3e7620a380ca38670d2fdc557138231f45f6755f2be2ca7cbe1cd6ff992f57c4