Pith. sign in

Paper Citation Record · LEDGER

StyleTTS: A Style-Based Generative Model for Natural and Diverse Text-to-Speech Synthesis

As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2205.15439.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2205.15439 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T19:25:16.247483Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T08:30:57.095678Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1eab5a8d-3828-48e2-bcf0-42473337183e · inbound

ProsodyFM: Unsupervised Phrasing and Intonation Control for Intelligible Speech Synthesis cites this paper.

ProsodyFM: Unsupervised Phrasing and Intonation Control for Intelligible Speech Synthesis StyleTTS: A Style-Based Generative Model for Natural and Diverse Text-to-Speech Synthesis

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:11.756716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:39:11.756716Z digest=sha256:d6417b41aad7ec0f8c6d56b439f76d302bb145a2ea7542fd38b68dedbcd66b18

Observation 983c872f-43f9-4b2c-a6a1-4569aa63bad7 · inbound

DrawSpeech: Expressive Speech Synthesis Using Prosodic Sketches as Control Conditions cites this paper.

DrawSpeech: Expressive Speech Synthesis Using Prosodic Sketches as Control Conditions StyleTTS: A Style-Based Generative Model for Natural and Diverse Text-to-Speech Synthesis

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:05.292855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:05.292855Z digest=sha256:34036e9aeb5c538e7884b9240aef6508d30a246299ac422accb7a36a63c7750c

Observation 46734ece-16fc-4464-b22c-f7d03c37f3cb · inbound

Continuous Autoregressive Modeling with Stochastic Monotonic Alignment for Speech Synthesis cites this paper.

Continuous Autoregressive Modeling with Stochastic Monotonic Alignment for Speech Synthesis StyleTTS: A Style-Based Generative Model for Natural and Diverse Text-to-Speech Synthesis

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T16:43:08.775875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T16:43:08.775875Z digest=sha256:33703582b7e1241ead820d931a326c2fe76197f03b3ba0cbc36de79b6d6538e5

Observation dbac7877-7128-4abe-8a4a-fe50e05afa1a · inbound

TCSinger 2: Customizable Multilingual Zero-shot Singing Voice Synthesis cites this paper.

TCSinger 2: Customizable Multilingual Zero-shot Singing Voice Synthesis StyleTTS: A Style-Based Generative Model for Natural and Diverse Text-to-Speech Synthesis

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:30:35.054489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:30:35.054489Z digest=sha256:90aa29f13602fcf0866a4bb55f2dab81342b5dead145a0551d4b8783e1f7fe62

Observation cedd6373-444c-41af-8de8-ad01e718ea8f · inbound

RapFlow-TTS: Rapid and High-Fidelity Text-to-Speech with Improved Consistency Flow Matching cites this paper.

RapFlow-TTS: Rapid and High-Fidelity Text-to-Speech with Improved Consistency Flow Matching StyleTTS: A Style-Based Generative Model for Natural and Diverse Text-to-Speech Synthesis

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T19:25:16.247483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:25:16.247483Z digest=sha256:7ace9d15de2b1bd24f2e123bc5a219f84715ab6ed6325348d8310426d8f332c6

Observation c4d5860e-7b97-4dcc-9c27-e2996a404932 · inbound

DMOSpeech 2: Reinforcement Learning for Duration Prediction in Metric-Optimized Speech Synthesis cites this paper.

DMOSpeech 2: Reinforcement Learning for Duration Prediction in Metric-Optimized Speech Synthesis StyleTTS: A Style-Based Generative Model for Natural and Diverse Text-to-Speech Synthesis

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T15:48:22.067217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:48:22.067217Z digest=sha256:adf34d325edc29d8f1b05829f39de0d1f74e74f0fdd75ee017ceac665a7f2c31

Observation c0ec5bb6-911c-431d-a0cb-8ee88c1898ba · inbound

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey cites this paper.

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey StyleTTS: A Style-Based Generative Model for Natural and Diverse Text-to-Speech Synthesis

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:30:57.097369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T16:36:33.264166Z digest=sha256:e1941ce8fe508ec94438100199389ab1b1b60019689d46589f0afde34483a93c

Observation 817657eb-e1b1-467f-bd95-45bd30334ac2 · inbound

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey cites this paper.

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey StyleTTS: A Style-Based Generative Model for Natural and Diverse Text-to-Speech Synthesis

Reference 132

Resolution
unresolved
no resolver link, observed 2026-07-12T22:04:31.302192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T22:04:31.302192Z digest=sha256:3f1283aca5f467a94c0d9cf1ef75e48ba7514bcf21e793403a311e95c4f58a1d