Pith. sign in

Paper Citation Record · LEDGER

XY-Tokenizer: Mitigating the Semantic-Acoustic Conflict in Low-Bitrate Speech Codecs

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2506.23325.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.23325 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T11:40:59.738880Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T19:57:20.080166Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 827d65ca-cae9-4b43-9f3c-912e7370c341 · inbound

AudioCodecBench: A Comprehensive Benchmark for Audio Codec Evaluation cites this paper.

AudioCodecBench: A Comprehensive Benchmark for Audio Codec Evaluation XY-Tokenizer: Mitigating the Semantic-Acoustic Conflict in Low-Bitrate Speech Codecs

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T11:40:59.738880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:40:59.738880Z digest=sha256:96df648a14c1e13adfe2651d576e5e44f7c36b882cf769079e1705d7de0d9096

Observation 68ea7121-7178-46a9-853f-f6eea513ebac · inbound

Qwen3-TTS Technical Report cites this paper.

Qwen3-TTS Technical Report XY-Tokenizer: Mitigating the Semantic-Acoustic Conflict in Low-Bitrate Speech Codecs

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:24:56.099066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T19:24:56.057631Z digest=sha256:314d875b50d19238953fb14aecd9b736dedf3ca57d1ea16c44a13aac74f7d7b6

Observation 8868fc17-160d-4d5d-af6a-863c6956a272 · inbound

VERITAS: A Multi-Agent Co-Scientist for Verifiable Image-Derived Hypothesis Testing cites this paper.

VERITAS: A Multi-Agent Co-Scientist for Verifiable Image-Derived Hypothesis Testing XY-Tokenizer: Mitigating the Semantic-Acoustic Conflict in Low-Bitrate Speech Codecs

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-12T21:34:30.078813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T21:34:30.078813Z digest=sha256:b9c04655ed92a23fe621304a76b4d179795423316e93c75554ae169d1ecfa603

Observation 022ec4b3-f695-4420-9a06-905f81ca4397 · inbound

Why Your Tokenizer Fails in Information Fusion: A Timing-Aware Pre-Quantization Fusion for Video-Enhanced Audio Tokenization cites this paper.

Why Your Tokenizer Fails in Information Fusion: A Timing-Aware Pre-Quantization Fusion for Video-Enhanced Audio Tokenization XY-Tokenizer: Mitigating the Semantic-Acoustic Conflict in Low-Bitrate Speech Codecs

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:31:01.898524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:49:51.481873Z digest=sha256:04bda704e8765a9ff30dead060e56dc4334beb8f7abd632e408cba119f124e7e

Observation 919d90bb-c069-4a92-9139-adee8d952b25 · inbound

VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing cites this paper.

VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing XY-Tokenizer: Mitigating the Semantic-Acoustic Conflict in Low-Bitrate Speech Codecs

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:50:56.020011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T01:03:09.942984Z digest=sha256:6d15e278c58841aba98d1d491493f0660be69de53e50f229d5cb9e310fa0d537

Observation 594065c4-4506-4376-87c6-54f77b9c3ff2 · inbound

Reducing Linguistic Hallucination in LM-Based Speech Enhancement via Noise-Invariant Acoustic-Semantic Distillation cites this paper.

Reducing Linguistic Hallucination in LM-Based Speech Enhancement via Noise-Invariant Acoustic-Semantic Distillation XY-Tokenizer: Mitigating the Semantic-Acoustic Conflict in Low-Bitrate Speech Codecs

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:26:23.233947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T01:11:50.733585Z digest=sha256:46dae1b0afd70fb6f879bac8ea2336106767322e612d6dbe1a98b3163db38ced

Observation da31b60e-9d12-4c5a-ab33-74f6270c59ac · inbound

dots.tts Technical Report cites this paper.

dots.tts Technical Report XY-Tokenizer: Mitigating the Semantic-Acoustic Conflict in Low-Bitrate Speech Codecs

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:57:20.081577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T21:10:25.911203Z digest=sha256:e1a051559dc49c56d9268d79ca13d148188ecc829ff5a2231396e54a41fb5606

Observation 4c74f3f1-d424-4b63-a725-7f1931abe1aa · inbound

FlexiSLM: A Dynamic and Controllable Frame Rate Spoken Language Model cites this paper.

FlexiSLM: A Dynamic and Controllable Frame Rate Spoken Language Model XY-Tokenizer: Mitigating the Semantic-Acoustic Conflict in Low-Bitrate Speech Codecs

Reference 109

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T11:55:42.290110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-07-01T03:50:26.873406Z digest=sha256:e571c92e0825e739d761f0d2c24584de8713b611fbff0ad35102d7c07f286735