Pith. sign in

Paper Citation Record · LEDGER

Hi-Fi Multi-Speaker English TTS Dataset

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2104.01497.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2104.01497 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T17:10:04.049487Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T10:44:07.682416Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 04d6df58-72a5-42d1-b5e3-e1d6e6e605aa · inbound

Length-Aware Rotary Position Embedding for Text-Speech Alignment cites this paper.

Length-Aware Rotary Position Embedding for Text-Speech Alignment Hi-Fi Multi-Speaker English TTS Dataset

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T17:10:04.049487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:10:04.049487Z digest=sha256:89d5342e4635e7e1b755652c51b0c43558f6f610f08921ac9363745c54381c6b

Observation 7d15d4d7-08fe-4d0f-b947-d0bf53fb9d28 · inbound

StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs cites this paper.

StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs Hi-Fi Multi-Speaker English TTS Dataset

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:01:24.417213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T12:57:04.450462Z digest=sha256:bf0244d74582febe7fb51144286181e35d5d9ac38b5c682250706b4312149245

Observation 2e0b8a9f-af81-4ecb-b30a-e3aca0b1a9d7 · inbound

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning cites this paper.

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning Hi-Fi Multi-Speaker English TTS Dataset

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:35:18.340737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T08:34:56.898815Z digest=sha256:926056bd0a668c70f52ea06e6ad5852fadb6299da4da430496daf1ecd5611456

Observation bbcace2b-3698-4b78-9ce9-e17121cfad71 · inbound

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning cites this paper.

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning Hi-Fi Multi-Speaker English TTS Dataset

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-21T10:44:07.685113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T10:43:27.176535Z digest=sha256:aa0b18da905f53734b6184516c8829c95171080f3fe3b5a04e464abea4d998fa

Observation 9d510468-0d3f-47c4-8646-d3a90a8337c2 · inbound

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning cites this paper.

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning Hi-Fi Multi-Speaker English TTS Dataset

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-13T22:57:11.059962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:57:11.059962Z digest=sha256:47dd41bff0b6c1c0633ad86bf2c782a7f3f55182fabb0c160c73c9044e7f7706

Observation 0c217e5c-1c13-4177-b03e-0e6e0c8a3616 · inbound

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey cites this paper.

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey Hi-Fi Multi-Speaker English TTS Dataset

Reference 198

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:30:56.907217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T16:36:33.264166Z digest=sha256:2d23cbef06c4a87215a1b4141648554749cbbc2953fc3b7fb3d21c481e1d540f

Observation b2aba4bf-c4f9-47c9-bea6-b8ff609a08d8 · inbound

A Survey of Advancing Audio Super-Resolution and Bandwidth Extension from Discriminative to Generative Models cites this paper.

A Survey of Advancing Audio Super-Resolution and Bandwidth Extension from Discriminative to Generative Models Hi-Fi Multi-Speaker English TTS Dataset

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:27:53.861667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T20:26:59.049472Z digest=sha256:bcc57763b2084ae59fcb41c07d8a305cb0fdaa296881acab6bdb149e835cbb05

Observation 0c4ec660-1768-463d-8887-34e4cb624f23 · inbound

Designed Vocalizations Dataset: Sound-Designed Human and Animal Voices for Non-human Voice Conversion cites this paper.

Designed Vocalizations Dataset: Sound-Designed Human and Animal Voices for Non-human Voice Conversion Hi-Fi Multi-Speaker English TTS Dataset

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T08:57:40.034274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:57:40.034274Z digest=sha256:d470a00b952dd0c7131740348e73c9a51b5462d5ee30ed561c464786153d1304