Pith. sign in

Paper Citation Record · LEDGER

HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2311.12454.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.12454 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:57:55.976737Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T12:26:37.491138Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1b46cae5-38b9-4e85-b7a8-66bdc50c31c2 · inbound

Seed-TTS: A Family of High-Quality Versatile Speech Generation Models cites this paper.

Seed-TTS: A Family of High-Quality Versatile Speech Generation Models HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:26:37.493296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T12:26:37.300599Z digest=sha256:303752f7b969289eaa3035817d1bb20f40314be05e8c108b3596227d53d4dce3

Observation 3cbbd4e7-709a-402c-b2e3-a61fcb7b3d06 · inbound

Rhythm Controllable and Efficient Zero-Shot Voice Conversion via Shortcut Flow Matching cites this paper.

Rhythm Controllable and Efficient Zero-Shot Voice Conversion via Shortcut Flow Matching HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:57:55.976737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:57:55.976737Z digest=sha256:d5fc45cae6a3df172e02212cbc01f4a10fa2cd153b379bcfc90dc52522bf6ec0

Observation 465a0c8c-bb1e-4e75-8021-eb5edcc90a2c · inbound

Towards Better Disentanglement in Non-Autoregressive Zero-Shot Expressive Voice Conversion cites this paper.

Towards Better Disentanglement in Non-Autoregressive Zero-Shot Expressive Voice Conversion HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T10:55:20.325815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:55:20.325815Z digest=sha256:62367263fce508a6837032dcda3e3ce0d7e2de7cdd2c4e32c2d4f924ba83354d

Observation 361f327e-a2bb-42b2-b314-ec4cf2d7f6ff · inbound

SemAlignVC: Enhancing zero-shot timbre conversion using semantic alignment cites this paper.

SemAlignVC: Enhancing zero-shot timbre conversion using semantic alignment HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T18:14:07.436262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:14:07.436262Z digest=sha256:508932fa4093203e47bb2e5b65c10aba1a25db7a121b058f666bd5c4bbe6c303

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T18:06:33.784245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:06:33.784245Z digest=sha256:ba6bcf5a5877ffed3e74af090489d0369035e162ee7c16bb26fa7db0c8d36bd0