Pith. sign in

Paper Citation Record · LEDGER

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents

As of 16 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 0 inbound Pith citation observations for arXiv:2608.08638.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.08638 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:34:47.857279Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy24
  • unresolved22
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9f348351-f5c6-48eb-aebd-70385bd661f2 · outbound

This paper cites Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.690816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.690816Z digest=sha256:0fd0b19daa0ce0b6b7b3e9fbe091f5007a97d44e8ee5d5937abc7f4d11654a1c

Observation 2265625b-dbaf-4695-bae1-ced74b0f97e7 · outbound

This paper cites Transactions of the Association for Computational Linguistics , volume =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Transactions of the Association for Computational Linguistics , volume =

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.600345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.696570Z digest=sha256:97f18baca3233c802776cdf1c8214b6015bdc016f9da0955eb855d66bfe097f0

Observation 263b3929-e08d-4039-aa48-bbfce87a285c · outbound

This paper cites Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.589000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.700590Z digest=sha256:1527a568cb157e5a8ea191cdffb83bfe305ccf707d467cd7703247e830c76e8b

Observation 528bde7a-10be-4cd5-ae0d-7a3fa4d60659 · outbound

This paper cites an unresolved cited work.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:34:48.576877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.704386Z digest=sha256:630720b6b7fec9547ee8dd9b41a2873812b1d651dd487ca4b3e98a6c42fefc83

Observation ca9de21b-070a-4a41-8e40-9000f901dffa · outbound

This paper cites Advances in Neural Information Processing Systems , volume =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Advances in Neural Information Processing Systems , volume =

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.566541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.708558Z digest=sha256:2bc7cabdb0a05273bb5c011efea8b7b2060e3de5ab8bcd2ec0452413944130be

Observation c4a55a63-e3d1-44f3-9972-2cbbcf19b7fd · outbound

This paper cites Matcha-TTS: A fast TTS architecture with conditional flow matching.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Matcha-TTS: A fast TTS architecture with conditional flow matching

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.712299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.712299Z digest=sha256:6c4bebde3348cb2ec3d098f26e46da49ff3431e4e033022dff50ab3f999d46f6

Observation dc69aa60-8551-4c30-baf7-d3c52fff7bb2 · outbound

This paper cites an unresolved cited work.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:34:48.552579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.716159Z digest=sha256:c1b6646a2b27f85d9270a27666f1ce3ea0bb47ad75154f1f6ce8292e75cdcb3e

Observation cd0a88a1-8c3f-498e-823c-f267199e3654 · outbound

This paper cites CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.719682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.719682Z digest=sha256:f38e658eb21e967f6e8168fee3a25a2b45969a1619329b9de114f9d007ac8490

Observation b4764bab-f4c4-4c0d-8fea-fd213af4b19b · outbound

This paper cites CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.723267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.723267Z digest=sha256:872ed704fffd8d0a9c1f75f90afde84d4be7efe1b16f5ad52bf2d950564564a2

Observation c08ef51c-77f7-419c-9563-d504c0366d5c · outbound

This paper cites Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.541526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.727270Z digest=sha256:c46d880fbe523f8100a67ee1f6582a443790b7a2d94d3f0c14198e28b1ce1dc7

Observation a8005653-95e4-4bd2-b93d-aae2f764b6f8 · outbound

This paper cites arXiv preprint arXiv:2410.16048 , year =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents arXiv preprint arXiv:2410.16048 , year =

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.730831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.730831Z digest=sha256:7ed80242ff13f6650044d049c668341ed2c588ce5d427ce8c6c3dc052e91e37a

Observation c5be50df-da2b-48ed-b227-0d540de1b06c · outbound

This paper cites 2025 , url =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents 2025 , url =

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.530357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.734210Z digest=sha256:2e4550622f874da414a1660fa0d5a5ac06776040dad5e297ffb67e9975476ba6

Observation bfc28a2b-d9d1-43af-ae53-37b9ae64f760 · outbound

This paper cites 2025 , doi =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents 2025 , doi =

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.519301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.737637Z digest=sha256:a472848be42b92968f35c398815cbd2a6174e7d59bc2f1cc5f3c8b779e83e5bc

Observation 39ad39d4-2362-4965-9625-e2ab8828681e · outbound

This paper cites arXiv preprint arXiv:2509.06926 , year =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents arXiv preprint arXiv:2509.06926 , year =

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.741058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.741058Z digest=sha256:8d72f94023e1733f52ebda7666acb76764f43023beac09317139bed90ac06962

Observation 65323f6f-84f2-4476-8031-fa97c37987c5 · outbound

This paper cites Multimodal Latent Language Modeling with Next-Token Diffusion.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Multimodal Latent Language Modeling with Next-Token Diffusion

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.744852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.744852Z digest=sha256:bfdd9926538359253e0314a09f5c78892e461ccbe46dc0536805cb4aa11df158

Observation f18bc46a-a7fd-4235-a096-7c3f0a283207 · outbound

This paper cites High-Fidelity Audio Compression with Improved.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents High-Fidelity Audio Compression with Improved

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.507858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.748625Z digest=sha256:717f174e7c994a0f560ade52ed13877114820f55376a1be765c53c6d343a9db3

Observation 81886fc8-a880-426b-8207-de3d5d56cd95 · outbound

This paper cites VibeVoice Technical Report.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents VibeVoice Technical Report

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.751890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.751890Z digest=sha256:49359487493f7908e1f912a55dca029b04692a2be8f9895f5896a7a9828aaee8

Observation 36ae63b0-65a6-4bdf-832c-69dd7ac4aced · outbound

This paper cites 2025 , eprint =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents 2025 , eprint =

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.494830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.755814Z digest=sha256:f6010bfb1ac9bdd1edd6fae37bd1ea43d6de1b3e55568081a3256f74f434021e

Observation c2ec74ab-3073-4aa1-9268-44398cfcf990 · outbound

This paper cites 2025 , eprint =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents 2025 , eprint =

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.483310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.759287Z digest=sha256:776a8e3a0746521eae196b006946022eb754cd688761c9ccafb26c83a42e59e6

Observation 083af3b6-ba30-4d94-8c49-60a8b74fdbda · outbound

This paper cites 2025 , eprint =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents 2025 , eprint =

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.471583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.762805Z digest=sha256:8aedd9bc9334e072e27b18ab901e12b330cd9d92c601d4e8ad9e40b90edb60a2

Observation 428a9006-3935-4abc-b8e5-abb2d9ebcb9f · outbound

This paper cites 2025 , eprint =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents 2025 , eprint =

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.459216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.766120Z digest=sha256:a99f356fcabe7bfebb8801747824f76b6ddf7ccc9721904fa884a73d751110d1

Observation 73cbeffa-3907-46cf-955d-7867ebef71a8 · outbound

This paper cites The Eleventh International Conference on Learning Representations , year =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents The Eleventh International Conference on Learning Representations , year =

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.447427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.769256Z digest=sha256:39b65f2ff548dc21b319febd80d93bffa97baeb3a6c035c1d5ddd84963d010b3

Observation e02c77d7-d48f-4e73-9ca5-c67b6f3c0572 · outbound

This paper cites Classifier-Free Diffusion Guidance.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Classifier-Free Diffusion Guidance

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.772701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.772701Z digest=sha256:e41980dee65c37d329b0aeefd770da96506a8dcdcce731bd7882c24362877a00

Observation 46c16d2c-5daa-43bb-b59f-14e089c95063 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages =

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.436450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.776552Z digest=sha256:23fbbe180aff0da7ef561f9fdb9b1e006f65fa6dfedd433b734cfad933c68ba4

Observation 6c2bd392-48c0-45e8-a72f-6180f6937928 · outbound

This paper cites The Tenth International Conference on Learning Representations , year =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents The Tenth International Conference on Learning Representations , year =

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.424221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.779605Z digest=sha256:2d6a4dc530e3d6563976f629eebd796e55ec6582547c72f4f77f0ef61b600572

Observation 09873062-4a0f-47f8-b6b3-cfeb06f4c75f · outbound

This paper cites Proceedings of the 40th International Conference on Machine Learning , series =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Proceedings of the 40th International Conference on Machine Learning , series =

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.412154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.783172Z digest=sha256:f243916ba6bddb27d40a66f291eb360cd104a70f89b58df6ad8c8026588a6a87

Observation c295d162-8dfe-4232-96ed-089ee32ea4c7 · outbound

This paper cites Mean Flows for One-step Generative Modeling.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Mean Flows for One-step Generative Modeling

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.786352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.786352Z digest=sha256:eeacca4d264c7e973992ed0cfe8ff7517dfac921b62315e9f599c32add0af746

Observation b27b8a3b-f804-435f-86b3-827857d3dfb4 · outbound

This paper cites arXiv preprint arXiv:2510.07979 , year =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents arXiv preprint arXiv:2510.07979 , year =

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.789831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.789831Z digest=sha256:cc55f9618ed61e23021be74bad63bd867723d03d38e92bae8f6132d797f4e1d3

Observation 13f99b37-cf3c-4cff-b71b-5083e02514e2 · outbound

This paper cites The Thirteenth International Conference on Learning Representations , year =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents The Thirteenth International Conference on Learning Representations , year =

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.400368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.793206Z digest=sha256:f531bb5c092ad9aabdc9b15efd877641589bb2ff51cdb86ffdecd4a7ce01ff07

Observation ad0d5956-ddb7-45b3-ad18-48c0736bc789 · outbound

This paper cites Transactions on Machine Learning Research , year =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Transactions on Machine Learning Research , year =

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.796494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.796494Z digest=sha256:9fe7941a5eab784d01a84f649dac0079e68244dd78128918fa115469220cbeff

Observation 6f56e798-40f0-4c8f-a3d5-51565053e899 · outbound

This paper cites 2022 , doi =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents 2022 , doi =

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.799959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.799959Z digest=sha256:764e3fd0c79a8de5b4b497055fc6245afba000ba481ef88a64da90c58b7b0378

Observation 3aa0008a-61ab-4e8b-b154-266b1337adba · outbound

This paper cites 2020 , doi =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents 2020 , doi =

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.803174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.803174Z digest=sha256:068fca2ce20235a6fb1759ef09464383494a874afd3ced41821f17ccbc1d49c4

Observation 93a6d603-a4fd-45b2-8461-56b8219426b6 · outbound

This paper cites 2015 , doi =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents 2015 , doi =

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.368273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.806731Z digest=sha256:0386e8322b2e8a6fbe105c805910a02237ec84e698b08df7185640f8039fface

Observation 4f0bb643-69ee-44e1-bfe1-458b4d73fd3d · outbound

This paper cites Proceedings of the 40th International Conference on Machine Learning , series =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Proceedings of the 40th International Conference on Machine Learning , series =

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.355838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.810384Z digest=sha256:c1a9fc35a58092c7f3791814cf16c5287d88d96b7da31538efd811807925dea8

Observation 056454a2-9fe3-4340-a6e3-bd0ebda5ad3c · outbound

This paper cites an unresolved cited work.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:34:48.345096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.814011Z digest=sha256:f818c0fb8f7e6538b8db527a66b836caa351c0a5d5dd2ac5969013a6e096eee0

Observation c965efa6-0138-4ce3-99ca-a9a8277bcc45 · outbound

This paper cites 2023 , doi =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents 2023 , doi =

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.334685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.817524Z digest=sha256:aaf0f903fe714a6a8128b8704d70226e8aef30ab47fa8d75e00b966f6103938f

Observation c53016fa-7f3d-477f-9966-877847cbb790 · outbound

This paper cites 2022 , doi =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents 2022 , doi =

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.322397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.820763Z digest=sha256:74279b56b7b2b528049aa3d60cd10b0f2f703df1e6b4f0145401be66804937d0

Observation da649455-cd2b-4fc7-b64c-724111984689 · outbound

This paper cites 2026 , eprint =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents 2026 , eprint =

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.310203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.824208Z digest=sha256:3d35c39b09bf7521eacc576a2630d6d2ee11ce965c2a3298c37c4a16616cf0b3

Observation 00857867-259e-438d-97e5-5ea59ba9c2f5 · outbound

This paper cites an unresolved cited work.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:34:48.298673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.827519Z digest=sha256:236eaccd5020ddf7c52a3934a952e2291adf6c4b93732f1f68d01a1ca950dcce

Observation f6a00cd0-c484-42da-ae3f-df317a6ab822 · outbound

This paper cites FireRedTTS-2: Towards Long Conversational Speech Generation for Podcast and Chatbot.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents FireRedTTS-2: Towards Long Conversational Speech Generation for Podcast and Chatbot

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.830741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.830741Z digest=sha256:7ef086303a07f0f7a87b3d77bbbb17cbf53617197808de11411d125cbea64149

Observation d48ff219-ef5f-4097-b213-fb27449d39b1 · outbound

This paper cites ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.834426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.834426Z digest=sha256:bc3c9826379d62a03e08b7866b1a5018513cb3181eea97285c8b80dcb177bce5

Observation 118e73e2-b728-42ec-ab50-4485a16f7c18 · outbound

This paper cites IndexTTS2: A Breakthrough in Emotionally Expressive and Duration-Controlled Auto-Regressive Zero-Shot Text-to-Speech.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents IndexTTS2: A Breakthrough in Emotionally Expressive and Duration-Controlled Auto-Regressive Zero-Shot Text-to-Speech

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.839144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.839144Z digest=sha256:a242fba67eec7a44c3b2422d455ac5f05f6c5ad7fe105713797e890b7b07024f

Observation a709c555-245c-4f56-bd60-43e4691a48f2 · outbound

This paper cites CosyVoice 3: Towards In-the-wild Speech Generation via Scaling-up and Post-training.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents CosyVoice 3: Towards In-the-wild Speech Generation via Scaling-up and Post-training

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:47.842901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.842901Z digest=sha256:bc30997c9010d0d9ee16d728ad323aadf5f61ed5db0f7870ed9a5a3646a94d4c

Observation 80e12354-253c-4f35-aa32-75f8ade94d32 · outbound

This paper cites Interspeech 2022 , pages =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Interspeech 2022 , pages =

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.288224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.846901Z digest=sha256:93870786c7830d5aafafb0366bbed9558809bf0d283db6d39874d6f659d7d255

Observation 9bd9d3af-0c4b-4776-8f64-f9e3603f81f4 · outbound

This paper cites 2026 , eprint =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents 2026 , eprint =

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.278229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.850315Z digest=sha256:c9f09a0dbf52cda8bc6f2be5ac735725457de91bfa6e3633c970f880dbcd0492

Observation 9004cfd9-b545-4e17-9985-aa8d0974c1ac · outbound

This paper cites 2026 , howpublished =.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents 2026 , howpublished =

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:34:48.266572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T04:34:47.853760Z digest=sha256:f087da640d8ebaa69b0e0d7b07652c9b48ae0b87eba9c0269a016d963bf9bfe8

Observation c724480d-11d0-4999-bb45-1ecc651c3803 · outbound

This paper cites an unresolved cited work.

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents Unresolved cited work

Reference 47

Resolution
parse uncertain
no resolver link, observed 2026-08-14T04:34:47.857279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:34:47.857279Z digest=sha256:88ac8e2a973c3b16630753e7fcad0c429061d65512e810c38952ddf76b3b658e

Pith citing papers

No inbound Pith citation observations are available.