Pith. sign in

Paper Citation Record · LEDGER

Zero-Shot Text-to-Speech from Continuous Text Streams

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2410.00767.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.00767 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T10:50:48.862070Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T06:19:09.583926Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1bcc38d5-d533-49fd-b2ad-725dfaec2251 · inbound

CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models cites this paper.

CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models Zero-Shot Text-to-Speech from Continuous Text Streams

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:19:09.585576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-13T06:19:09.507440Z digest=sha256:0f1eca8e5fe06b1fb61131b419b03b1c1cae2bb184908c72be6c82aa9a5e0af4

Observation fa140664-ae2b-4f37-9c28-9570a3db10ad · inbound

Interleaved Speech-Text Language Models for Simple Streaming Text-to-Speech Synthesis cites this paper.

Interleaved Speech-Text Language Models for Simple Streaming Text-to-Speech Synthesis Zero-Shot Text-to-Speech from Continuous Text Streams

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T10:50:48.862070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T10:50:48.862070Z digest=sha256:cbbdd457c13efdd1b3cd902227f594d52ff6f27c3dda33165879424a745c7a4c

Observation 0b239529-5d94-462f-9b54-9ad986574e10 · inbound

SpeakStream: Streaming Text-to-Speech with Interleaved Data cites this paper.

SpeakStream: Streaming Text-to-Speech with Interleaved Data Zero-Shot Text-to-Speech from Continuous Text Streams

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:22:30.726922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:22:30.726922Z digest=sha256:1ac577ab60a29984de6d860b06994b7d0cbf9c05210ff840956d7192f3a56434

Observation 041daa36-4dc0-4883-aa0a-bc187e0c0483 · inbound

Zero-Shot Streaming Text to Speech Synthesis with Transducer and Auto-Regressive Modeling cites this paper.

Zero-Shot Streaming Text to Speech Synthesis with Transducer and Auto-Regressive Modeling Zero-Shot Text-to-Speech from Continuous Text Streams

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:15:56.850541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:15:56.850541Z digest=sha256:6781d969a294a6fdbbc56794c9ca470dba64730eb6ecf1c1a9a556b982381b48

Observation b3ee0cd0-bd64-40b5-ae37-e300f621f096 · inbound

SimulS2ST-Omni: Data-Efficient Streaming Speech-to-Speech Translation via Explicit Trajectory Supervision cites this paper.

SimulS2ST-Omni: Data-Efficient Streaming Speech-to-Speech Translation via Explicit Trajectory Supervision Zero-Shot Text-to-Speech from Continuous Text Streams

Reference 218

Resolution
unresolved
no resolver link, observed 2026-08-01T11:43:06.771154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T11:43:06.771154Z digest=sha256:b145561eef554b2fef51f7e7cf92cd6993dd81beeb75a0276e4cda48f116fbbb