Pith. sign in

Paper Citation Record · LEDGER

Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2306.03509.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.03509 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:07:03.766330Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:19:44.612291Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 53c5908d-2c1c-4b39-8fe2-588f9940ae35 · inbound

Seed-TTS: A Family of High-Quality Versatile Speech Generation Models cites this paper.

Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:26:37.342857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T12:26:37.300599Z digest=sha256:9f9a8e535fb67032f9e2020f5c2daecd2ed4886e0bbf10677df865ec568aabbe

Observation 2146cef1-0d9c-445e-8e79-3a3c1cf73e2a · inbound

ProMode: A Speech Prosody Model Conditioned on Acoustic and Textual Inputs cites this paper.

ProMode: A Speech Prosody Model Conditioned on Acoustic and Textual Inputs Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:03.766330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:03.766330Z digest=sha256:72120f8f87d4b5bf0fabff6c19296d893e5819259ec011f468ef679e3b2beea0

Observation 78e4abf9-eb85-44cf-9e23-1d155d9b1699 · inbound

UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models cites this paper.

UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T11:29:33.186280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:29:33.186280Z digest=sha256:c736738f7a718efb4c35f5891c7d815edbd06383306eb095a8b7ff0b6de604ef

Observation 8c88c9e2-06e9-4bc7-9437-475bd62904d4 · inbound

Two-Dimensional Quantization for Geometry-Aware Audio Coding cites this paper.

Two-Dimensional Quantization for Geometry-Aware Audio Coding Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:20:29.207437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-21T18:16:51.486807Z digest=sha256:f86b030ccda2157636ca26f933484e67503bf829183e57d5f662f8862b9689ac

Observation 18b57f1e-c3ff-45f2-985d-6beb4f49fef2 · inbound

ProsoCodec: Prosody-Oriented Speech Codec for Voice Conversion cites this paper.

ProsoCodec: Prosody-Oriented Speech Codec for Voice Conversion Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:19:44.613832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T11:49:04.308326Z digest=sha256:cb2a7faf0e4deadc78e28ad495091651606bef4ed56289d5110b28a19289f92d