Pith. sign in

Paper Citation Record · LEDGER

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis

As of 7 August 2026, this Paper Citation Record lists 69 of 69 outbound references and 1 inbound Pith citation observation for arXiv:2507.02176.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02176 v1

Coverage vector

measured 69 of 69 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:40:16.437239Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:40:10.960719Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T20:40:17.636566Z

Reference resolution

69 of 69 outbound references displayed

  • verified exact4
  • verified fuzzy56
  • unresolved8
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fbd3ef56-0b7a-4d69-8241-d50fb04b296f · outbound

This paper cites Virtual chara cters are expected to possess unique identities that remain consi stent across time.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Virtual chara cters are expected to possess unique identities that remain consi stent across time

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:29.587245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:10.892442Z digest=sha256:cf8b9064812beb13ea3282ee740e7a22d1f3c532b1dbe0d05385774ef340201b

Observation 517abf04-5934-49ef-95b2-e23ab2515c83 · outbound

This paper cites Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:40:17.833571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:10.960719Z digest=sha256:08295bf1e565ff0b1a41450d8549d40279388e93ffa3a5f8fe688175be0cea4c

Observation f1f5725e-124c-47c6-963b-359f243ac434 · outbound

This paper cites We explore which mark- ers are represented in some widely used ASV embeddings, and measure the effect of confounding factors.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis We explore which mark- ers are represented in some widely used ASV embeddings, and measure the effect of confounding factors

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:28.996348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:11.033440Z digest=sha256:da00e360b05b0fa387f0483c86ffa70965302159a4fa36e12ae0ba5731ef11b1

Observation ce4bbb54-3138-4eec-b835-4404f612eb6d · outbound

This paper cites This reflects the lower bound of our metric, where we expect the smallest distances.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis This reflects the lower bound of our metric, where we expect the smallest distances

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:28.311331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:11.119053Z digest=sha256:d83cdeaa769a632bd1cc3c987647a17486c058950e4525f422c36b0f0c0e11bc

Observation 8bbfef3e-0a69-429e-9162-8340ecfa3475 · outbound

This paper cites This setting tests wh ether our metric can distinguish between speakers undistinguish - able with speech rate.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis This setting tests wh ether our metric can distinguish between speakers undistinguish - able with speech rate

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:28.102856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:11.225540Z digest=sha256:df5949010bc3a9d6587c4d84b9af64c1a94d88ed898aec23ceeda75123fe64ed

Observation 89f98804-a7b9-49a1-99e0-d0bacc9f6e2c · outbound

This paper cites The results in the top section of Table 3 show that rhythm distances are significantly larger between different speak ers, even those with similar speech rates.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The results in the top section of Table 3 show that rhythm distances are significantly larger between different speak ers, even those with similar speech rates

Reference 6

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T20:40:27.992700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:11.368279Z digest=sha256:d2b406003c40bf2298e82f9d1c2c9990eb57138ce7c2dcda73a41962f80303df

Observation 0ec6aae4-9475-4471-a80b-ff699a3f4477 · outbound

This paper cites We showed that ASV embeddings mainly encode speech identity markers relating to anatomy (e.g.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis We showed that ASV embeddings mainly encode speech identity markers relating to anatomy (e.g

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.880864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:11.464887Z digest=sha256:9f15d607864b55e0325e9542470332a3741bc25a8f196e347d19988a7ca4a4ad

Observation 5d5d9667-5627-4180-8746-56411436c248 · outbound

This paper cites An image is worth one word: Personalizing t ext-to- image generation using textual inversion,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis An image is worth one word: Personalizing t ext-to- image generation using textual inversion,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.756865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:11.537350Z digest=sha256:8c9b71789e377680dc683d07f05a75e7eda0eac54b711b4803fed9ead8ddcfe6

Observation 932e5017-ad30-4164-a120-a6d46a36daa7 · outbound

This paper cites Dreambooth: Fine tuning text-to-image d iffusion models for subject-driven generation,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Dreambooth: Fine tuning text-to-image d iffusion models for subject-driven generation,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.588974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:11.595403Z digest=sha256:52af9da003083ca1b6432a7462695559412de22b1e1d4aea1dfe0c0b9c5be1b5

Observation 0c62ff45-0144-4f18-aa42-7e8b96b9b4f1 · outbound

This paper cites ConsistentID: Portrait Generation with Multimodal Fine-Grained Identity Preserving.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis ConsistentID: Portrait Generation with Multimodal Fine-Grained Identity Preserving

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:11.679021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:11.679021Z digest=sha256:d659904bf7787ab1991111e5d5b6946cb92218f919527c838727b66e91915f11

Observation c89f5a9f-9e70-4ff9-a821-38f5722804a4 · outbound

This paper cites CharacterFactory: Sampling Consistent Characters with GANs for Diffusion Models.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis CharacterFactory: Sampling Consistent Characters with GANs for Diffusion Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:11.744848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:11.744848Z digest=sha256:2aa1f51df2c81ec92bee10550a8cff490f42abd7adcb509167f9330a288e70fd

Observation 5d17b08b-2cef-42f0-b20c-065a1b807fe3 · outbound

This paper cites Generat ing video game scripts with style,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Generat ing video game scripts with style,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.458551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:11.852395Z digest=sha256:7bc9d59440bc34d64823d36cf54d971a55f35308c665710818b11b8b450825c6

Observation 988925a3-6791-4cc9-b3ea-6d3176cae942 · outbound

This paper cites Meet your favorite character: Open-domai n chat- bot mimicking fictional characters with only a few utterance s,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Meet your favorite character: Open-domai n chat- bot mimicking fictional characters with only a few utterance s,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.351960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:11.916503Z digest=sha256:5692912005a8e819bc85b1ed5d45770a1180c9f9f9ff870a9c365e17a588f2d5

Observation ba599d21-aa4a-44e9-9de5-a9c72b8282d5 · outbound

This paper cites Information conveyed by vow- els.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Information conveyed by vow- els

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.170164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:11.991064Z digest=sha256:2da48b9bc937c8968284b7cb5828687d1cbb2f5b09218c45855564a38489c1de

Observation d6d6a64a-a140-4335-95bd-c7a4a8d02172 · outbound

This paper cites The perception of personal iden tity in speech: Evidence from the perception of twins’ speech,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The perception of personal iden tity in speech: Evidence from the perception of twins’ speech,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.987242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:12.078389Z digest=sha256:abdc2369267bc597d9f548f831c290758c0874db8720f22b088ac0775ec1b992

Observation e7c45ab1-d99e-4c5a-812e-3435e2d3b198 · outbound

This paper cites X-V ectors: Robust DNN embeddings for speaker recognition,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis X-V ectors: Robust DNN embeddings for speaker recognition,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.806842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:12.158681Z digest=sha256:a49b51201e7a762ed04041bf11fd0c24691e693609947ecbd6cead4753f068d2

Observation f7af280c-6e79-496f-a69c-c5d8babb5c99 · outbound

This paper cites ECAPA - TDNN: emphasized channel attention, propagation and aggre ga- tion in TDNN based speaker verification,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis ECAPA - TDNN: emphasized channel attention, propagation and aggre ga- tion in TDNN based speaker verification,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.629537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:12.245018Z digest=sha256:9a866525ea16ee11fa3689a74acbb08a26d26968cd367900bead7723e4d35ca0

Observation cc153765-41c9-49ce-b995-37e9c0f1de4d · outbound

This paper cites State-of-the-art speaker recogni tion with neural network embeddings in NIST SRE18 and speakers in the wild evaluations,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis State-of-the-art speaker recogni tion with neural network embeddings in NIST SRE18 and speakers in the wild evaluations,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.482722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:12.319383Z digest=sha256:399c43b179fa7029c2cf98f5b6b8c0d93a8a0bb0f6b38981f901f26fab91c5d7

Observation cb4977ac-8be1-44ea-ac18-7e54da19fefd · outbound

This paper cites Generalized end-to-end loss for speaker verifica- tion,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Generalized end-to-end loss for speaker verifica- tion,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.302718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:12.410015Z digest=sha256:09087a89799c2df0b075aad872c4977f537fbe1aba1ba85cfe9235e8f90f6a36

Observation b5ee12e4-44fe-4f82-9642-408fa023c43c · outbound

This paper cites We need variations in speech synthe- sis: Sub-center modelling for speaker embeddings,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis We need variations in speech synthe- sis: Sub-center modelling for speaker embeddings,

Reference 20

Resolution
verified exact
raw_fallback, observed 2026-08-06T20:40:17.374269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:12.504831Z digest=sha256:c8b2f5d793fbf955fb2428e6ca0a463d24ace4ace858cca2b332cd86eb96ca8a

Observation 7f5eebdb-5db5-4fd2-87f8-7537d1edc768 · outbound

This paper cites Predictions of subjective ratings and spoofing as- sessments of V oice Conversion Challenge 2020 submissions,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Predictions of subjective ratings and spoofing as- sessments of V oice Conversion Challenge 2020 submissions,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.121033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:12.603592Z digest=sha256:c47c6b1c49ef11ca4bbea981cf0e9fd1449e1d74f7d06a922b8445ada8391547

Observation b0d89eba-6eef-4933-97ed-e2d568c16f86 · outbound

This paper cites Transfer learning from speaker verificat ion to multi- speaker text-to-speech synthesis,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Transfer learning from speaker verificat ion to multi- speaker text-to-speech synthesis,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.902006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:12.679830Z digest=sha256:93da43da1ce31cb029c3888697f838e072c447b5add4ed88bafbf345d03cdbdf

Observation d59ce44c-d9f4-412b-bedd-6b5094c18c70 · outbound

This paper cites Zero-shot multi-speaker text-to-sp eech with state-of-the-art neural speaker embeddings,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Zero-shot multi-speaker text-to-sp eech with state-of-the-art neural speaker embeddings,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.728737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:12.740868Z digest=sha256:c731091a98ad2a587631791f337b19346b6a7e49e842c919cdb73a921dbd2a20

Observation 38425e51-2504-4a84-874e-c042e926c1d6 · outbound

This paper cites Y ourTTS: towards zero-shot multi- speaker tts and zero-shot voice conversion for everyone,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Y ourTTS: towards zero-shot multi- speaker tts and zero-shot voice conversion for everyone,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.600183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:12.837534Z digest=sha256:33b91c623f869522bcaa3049257ae7ba9e3b15a2df0091337bf1d3ae252c2e1b

Observation 1cb6588a-a323-4f5c-b318-afc26a02eae4 · outbound

This paper cites Investigating on incorporating pret rained and learnable speaker representations for multi-speaker mult i-style text-to-speech,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Investigating on incorporating pret rained and learnable speaker representations for multi-speaker mult i-style text-to-speech,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.420748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:12.936260Z digest=sha256:4be87c17948230d9eb9622246b9fa5d27637140f6588f1ca149356d2f2974270

Observation e3bf7555-52cb-4bf9-bd88-e292621223c0 · outbound

This paper cites Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:13.032052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:13.032052Z digest=sha256:06133a991321fd83037a9eddc04e694b9ef22d30a1ce4c7270ecacd043d1299c

Observation 01a602c7-7a3a-4992-bcca-2d20ea94f10d · outbound

This paper cites The Multi-Speaker Multi-Style V oice Clo ning Chal- lenge 2021,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The Multi-Speaker Multi-Style V oice Clo ning Chal- lenge 2021,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.237363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:13.128895Z digest=sha256:5e1298f0c441441ec9d7d91d1984a098ba8e3204ad4362d344f9fc4b26bc9aa1

Observation 4857260a-648a-4d2c-80a3-34fd08fb25de · outbound

This paper cites VoxGenesis: Unsupervised Discovery of Latent Speaker Manifold for Speech Synthesis.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis VoxGenesis: Unsupervised Discovery of Latent Speaker Manifold for Speech Synthesis

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:13.209463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:13.209463Z digest=sha256:758635bae23d707a4e91a786efaa10edb5d84770316eebf9503da06a8c5c4b30

Observation fc46cf74-cdd9-420c-b7f8-d04b70d048b7 · outbound

This paper cites Evaluating text-to-speech synthesis from a large discrete token-based speech language model,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Evaluating text-to-speech synthesis from a large discrete token-based speech language model,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.020012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:13.330030Z digest=sha256:57ed4add7f3fd958e16085f2c67b7869f3f429f0b6478a59e02205d0c36f1075

Observation f6e68c49-0f0f-4721-b3a1-898330ee7271 · outbound

This paper cites A comparison of discrete and soft speech units for improved voice conversion,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis A comparison of discrete and soft speech units for improved voice conversion,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:13.420536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:13.420536Z digest=sha256:2abd6ed85c688a96d0db200f9c70c66d2c36e4ca1520baf937dae58bc0c7439d

Observation 9fdce733-cd03-4761-b18d-ef7f05f91428 · outbound

This paper cites V oicebox: Text-guided multilingual univ ersal speech generation at scale,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis V oicebox: Text-guided multilingual univ ersal speech generation at scale,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:24.864555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:13.484801Z digest=sha256:196bc6289c5a4752e9cb428fa5b937d17b1cf9f8074ba3a01582eabf6933980e

Observation 67b8fbcc-5f6d-447d-bf18-4f153dc8ce3d · outbound

This paper cites Neural codec language models are zero-s hot text to speech synthesizers,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Neural codec language models are zero-s hot text to speech synthesizers,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:24.567138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:13.544795Z digest=sha256:750f6a316d74948a3830c6c1357d76d180edb2f8ff4c7b329b1c10aa4dd9424b

Observation fc5d7913-3f76-4168-90e3-16b024875a87 · outbound

This paper cites The Singing V oice Conversion Chall enge 2023,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The Singing V oice Conversion Chall enge 2023,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:24.362172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:13.610042Z digest=sha256:437c380b2ff9de4cdbb3912540af9deda15db6cd66e3c161f3ff58f5b3e4aec7

Observation f3ef1777-0b8f-4ab0-9ddc-4580ee3c651c · outbound

This paper cites Generative Data Augmentation Challe nge: Zero- shot speech synthesis for personalized speech enhancement ,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Generative Data Augmentation Challe nge: Zero- shot speech synthesis for personalized speech enhancement ,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:24.160692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:13.678518Z digest=sha256:d58e51493dcedfcebe530631d1adad6846290a69a54f0976b8ef6ab6b7d4e7a7

Observation 654da8cc-1fdc-4b24-b46c-06147b004493 · outbound

This paper cites Acoustic properties of voice timbre t ypes and their influence on voice classification,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Acoustic properties of voice timbre t ypes and their influence on voice classification,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:23.920263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:13.758589Z digest=sha256:5a36db82cfd5fd808716b5ec39deeb482b6eb0cfcfcf84ccaf45da531ef54d46

Observation 5c79d98a-4962-4f10-802e-f487f2340f69 · outbound

This paper cites Speech production patterns in producing l inguis- tic contrasts are partly determined by individual differen ces in anatomy,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Speech production patterns in producing l inguis- tic contrasts are partly determined by individual differen ces in anatomy,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:23.680062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:13.823753Z digest=sha256:f531afecaeeed00e0ba5940a57ff0d0d9035789177090142f2ce5cd2e602fd5e

Observation bc66290e-98e4-4bd5-981a-3190fe79369b · outbound

This paper cites an unresolved cited work.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:40:23.460801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:13.871818Z digest=sha256:1cec96547935e4c4b00666e49f451b94fef8aab53300f8f602c21ddfeca731f1

Observation cae77f0c-21b8-4b0b-8ad9-32e71334c683 · outbound

This paper cites A study of rhythm in London: Is syllable-timing a feature of Multicultural London English ?.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis A study of rhythm in London: Is syllable-timing a feature of Multicultural London English ?

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:23.289528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:13.928173Z digest=sha256:c35854d4cd49ad9981c42d499877ef6ba5e5c5af7a101539e49403a73205a1fa

Observation af5d4057-27d1-48f6-961b-011f8958c4ba · outbound

This paper cites The measurement of rhythm: a comparison of Sin- gapore and British English,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The measurement of rhythm: a comparison of Sin- gapore and British English,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:23.067362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:13.978443Z digest=sha256:73a81a0c9ad1048d2387cc87c824591c74caf377434c231b1920de081f302730

Observation 10783b19-ca12-4bc5-8432-b0df7d8a26ba · outbound

This paper cites Sociophonetics of phonotactic pheno mena in french,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Sociophonetics of phonotactic pheno mena in french,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:22.807797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:14.042470Z digest=sha256:4c9db5c8af58197b18f8749b46068b0e2a7fd4a57d2f976d102a1eb7e47b977b

Observation 4d90a904-75f6-4cf3-8437-f4ba2f576871 · outbound

This paper cites Individual di fferences in vowel production,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Individual di fferences in vowel production,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:22.620821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:14.108230Z digest=sha256:2378e119a669da110a8dfae568d2dd6c46726a9bcfab52b83e0a97983f2c11c4

Observation 13627fac-ceff-4ca7-9de0-37f89d1c47df · outbound

This paper cites Individual differences in speech pro- duction: V oice-onset-time,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Individual differences in speech pro- duction: V oice-onset-time,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:22.428908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:14.167152Z digest=sha256:572a2202e5d4a94a8a2ace83789f33d6eb68523ef74d51d4feaee7d2ba3b6e24

Observation 6c2ed39d-60f7-4358-bb49-930065f56271 · outbound

This paper cites The social life of phonetic s and phonology,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The social life of phonetic s and phonology,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:22.226893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:14.225970Z digest=sha256:f6820b956aba6b227ed9e0c8d26a4a6b40cb53843628f338cedccf0f82662a95

Observation a0a78c3e-a471-4b32-9051-7867fd9eefd9 · outbound

This paper cites Prosodic rhythm and Afric an American english,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Prosodic rhythm and Afric an American english,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:22.055134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:14.273503Z digest=sha256:8d166f6a2c211e4db703ae97141d325782fc5b9f8cfd7feb81da034e5c417acf

Observation dd9e3107-7a7a-413a-82ef-74ec4916a3e9 · outbound

This paper cites La liaison sans enchaˆ ınement,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis La liaison sans enchaˆ ınement,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.865653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:14.325252Z digest=sha256:187c8d64f811e71043fac259189145118fce0afd64bcdc510f980ceef682b60b

Observation f7e19809-8f4a-47d6-b320-721f6b2d4a4b · outbound

This paper cites Stuart-Smith, E.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Stuart-Smith, E

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.707914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:14.381549Z digest=sha256:8afc15345bea96e82c58b67291c113571fe34427d7a78217b22c11af1482ce62

Observation f2fee590-4f65-4a33-85bb-bb5f1f5ae99f · outbound

This paper cites The Blizzard Challenge 2023,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The Blizzard Challenge 2023,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.492072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:14.437994Z digest=sha256:24cdfdeecc63c6be74d2f594c9252c51cbfd3f5a61478e7cab166d1a6fb3ec6c

Observation 7c04876d-3c13-4c25-8171-85a9405c39fb · outbound

This paper cites Refining the evaluation of speech synthesis: A summ ary of the blizzard challenge 2023,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Refining the evaluation of speech synthesis: A summ ary of the blizzard challenge 2023,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.310864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:14.486042Z digest=sha256:58e46fa75c18b9d13e44f5c0f6d8aade3f4c4a02294209ce7e5622c377889fc9

Observation 457a1f57-16e9-4294-83bd-5d1340707f4c · outbound

This paper cites Good practices for evaluation of synthesized speech,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Good practices for evaluation of synthesized speech,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:14.536826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:14.536826Z digest=sha256:0937f0396057e58bcaa56c5ffe7ff1afb050f4dca9b034b47d88e517b3bf1ec7

Observation a8594ffb-9d77-4924-a681-932f2c5a1c65 · outbound

This paper cites Stuck in the MOS pit: A critical anal ysis of MOS test methodology in TTS evaluation,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Stuck in the MOS pit: A critical anal ysis of MOS test methodology in TTS evaluation,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.155580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:14.599962Z digest=sha256:0615f670386dc3e5420f20db51dff5d09f3255aaef282a53851da051bee9a792

Observation bf0a2f81-a772-4739-9fd1-56a402a46427 · outbound

This paper cites Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:40:16.930605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:14.655226Z digest=sha256:e1801b43ba4ad5a9bfc96bb862c3439eafba195ad1adeabf71209a74e7794506

Observation 26572b08-138c-4f2f-8822-ccf26c78472f · outbound

This paper cites MOS vs . AB: evaluating text-to-speech systems reliably using cluster ed stan- dard errors,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis MOS vs . AB: evaluating text-to-speech systems reliably using cluster ed stan- dard errors,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.015992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:14.709252Z digest=sha256:f76c1ce00357e8c2d8c934e7d63522f0076ad7fc41ef33f7500fcebb6f81f004

Observation 422c058a-9164-4776-a0b1-ed510877abc7 · outbound

This paper cites V oxSim: A perceptual voice similarity da taset,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis V oxSim: A perceptual voice similarity da taset,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:20.835953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:14.712118Z digest=sha256:86f664e98596f46fd6ef744530ebecbd4820cea0cedfc1bab30751b3adcb5391

Observation e7fd25d3-fd91-4c7c-832e-14a34f7cd461 · outbound

This paper cites Salmon: A suite for acous tic language model evaluation,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Salmon: A suite for acous tic language model evaluation,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:20.663293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:14.751685Z digest=sha256:648d8f37d96cbb6ae6ea6b9b04dec38f822a369e91ed3a177bb346e8874051a3

Observation 82dfec55-c787-4bd5-b723-5017a25bfc68 · outbound

This paper cites Automatic evaluation of speaker simila rity,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Automatic evaluation of speaker simila rity,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:20.472116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:14.887907Z digest=sha256:7bb59a6235603dc0a40cf59477ec924b34c80764b61b15b83e4d285e3d167088

Observation 36a86cc9-028d-4d05-b97d-d2dfadd18548 · outbound

This paper cites SVSNet: An end-to-end speaker voice simi larity assessment model,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis SVSNet: An end-to-end speaker voice simi larity assessment model,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:20.297397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:15.048606Z digest=sha256:5dfb2be9315f9517e74e11dee00673237cb784d72931b3f4edc3767e5bac3fcd

Observation 502ce3f8-0886-4a72-9e6a-87181406f5a4 · outbound

This paper cites The V oxCeleb Speaker Recognition Challe nge: a retrospective,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The V oxCeleb Speaker Recognition Challe nge: a retrospective,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:20.152485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:15.178145Z digest=sha256:956a16b4a053b48af7fe620f9e89ef381c3e4c9fb7f9ac08bb66b184ce89fc05

Observation d98022c9-0b21-4b97-a133-f6480c3df776 · outbound

This paper cites SpeechBrain: A General-Purpose Speech Toolkit.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis SpeechBrain: A General-Purpose Speech Toolkit

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:15.310440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:15.310440Z digest=sha256:2969c4254045cc18aabf66be4746d1022e24901691cd443e44ade5fc6f244eea

Observation b2034870-5e78-47c0-8f37-9038336a86b6 · outbound

This paper cites WavLM: large-scale self-supervised pr e-training for full stack speech processing,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis WavLM: large-scale self-supervised pr e-training for full stack speech processing,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.995774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:15.439938Z digest=sha256:6b71d50a845d551b272cbe4a937d6cf917824e8c2c0445ba5eca62215457327d

Observation 65b8ea99-700f-4336-a54f-b31272b74d15 · outbound

This paper cites The CMU Arctic speech databa ses,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The CMU Arctic speech databa ses,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.842760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:15.547259Z digest=sha256:a5acc21d4c11801c39dd7e60d04ac6277b980e141ec4aab00240a9b1839ae19f

Observation 56a9ff04-0491-43ae-bee3-ac8a523caf54 · outbound

This paper cites L2-ARCTIC: a non-native english speech c orpus,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis L2-ARCTIC: a non-native english speech c orpus,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.681892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:15.631101Z digest=sha256:94d928273d66bf2c4b009da6945b399018d20087450c8251eddddcbd23e50ac3

Observation 33f54206-915c-4f9a-b19d-ddd85c8d66db · outbound

This paper cites Lib- riSpeech: an ASR corpus based on public domain audio books,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Lib- riSpeech: an ASR corpus based on public domain audio books,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.533289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:15.701551Z digest=sha256:141b2cd73b076920597d787f320ef2b3bb0bfe15f023f88a28d24f795bcefd73

Observation e75cb6d1-b3e3-4a99-93e8-8046229d5370 · outbound

This paper cites Analysis of fun- damental frequency, jitter, shimmer and vocal intensity in chil- dren with phonological disorders,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Analysis of fun- damental frequency, jitter, shimmer and vocal intensity in chil- dren with phonological disorders,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.376682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:15.768394Z digest=sha256:192c17a8fa0353d1e9261f65ec7e20c435c6096964963218d474742b8d4057e3

Observation af70a177-120a-4554-bcd0-0fe20b521ab6 · outbound

This paper cites V ocal acoust ic analy- sis – jitter, shimmer and HNR parameters,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis V ocal acoust ic analy- sis – jitter, shimmer and HNR parameters,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.199832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:15.849729Z digest=sha256:8ba40e20d9eeafef043b9d8e55fc21acdce671172026255f873d76aa9d5c213f

Observation d80aa272-c9b8-486a-88f8-28b67beccaef · outbound

This paper cites V ariation of the acoustic parameters: f0, jitter, shimmer and alpha ratio in relation with different background noise levels,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis V ariation of the acoustic parameters: f0, jitter, shimmer and alpha ratio in relation with different background noise levels,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.053593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:15.945168Z digest=sha256:22e0f401ccfc7f3962d5fe5995cc1f130e529c56df0768b6438b6d658555c75d

Observation 6bc0d3b7-7283-4f48-948d-5ea32a83db7b · outbound

This paper cites Rhyth m modeling for voice conversion,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Rhyth m modeling for voice conversion,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:18.800943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:16.041293Z digest=sha256:ceca919a2242782dba5cb9a025fbd18bc503e98285fc12ebcde7ca65aea5200a

Observation f993c02a-30c5-4bb9-832e-82ceaf2ffa39 · outbound

This paper cites Opensmile: the munich versatile and fast open-source audio feature extractor,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Opensmile: the munich versatile and fast open-source audio feature extractor,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:18.468730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:16.162090Z digest=sha256:f0dbd44f3249e7f98bc66b1f10ad74fad43d253d912153811461fc5c272ac6f9

Observation 35d8cdf7-a228-477d-8e01-9ae7453031b1 · outbound

This paper cites ASRRL-TTS: Agile Speaker Representation Reinforcement Learning for Text-to-Speech Speaker Adaptation.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis ASRRL-TTS: Agile Speaker Representation Reinforcement Learning for Text-to-Speech Speaker Adaptation

Reference 68

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:40:16.704034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:16.321959Z digest=sha256:caa95120f2b47531d382d09dce00aa32cc2901a7c5571a485f4b24e303b670bc

Observation fd0d3d8b-c6c7-42cb-afba-aeecf2eaf8f8 · outbound

This paper cites All about audio equalization: Solu- tions and frontiers,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis All about audio equalization: Solu- tions and frontiers,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:18.192615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:16.437239Z digest=sha256:de7f5fb8c9cf20f2467b310b195ec9a9c856f6366ed733caa4a839bedd1f63ec

Pith citing papers

Observation 517abf04-5934-49ef-95b2-e23ab2515c83 · inbound

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis cites this paper.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:40:17.833571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:40:10.960719Z digest=sha256:08295bf1e565ff0b1a41450d8549d40279388e93ffa3a5f8fe688175be0cea4c