Pith. sign in

Paper Citation Record · LEDGER

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information

As of 15 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2505.15667.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15667 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:16:45.290095Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:16:42.871348Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T15:16:45.530174Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact1
  • verified fuzzy26
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5c1a7c5a-f82f-48c0-99d1-dadd05f77c3e · outbound

This paper cites an unresolved cited work.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:50.090889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:42.724908Z digest=sha256:c87a4a421319463e94ea8f4dcf5910869cc28072af064c0f49fbdd220cb5fdd1

Observation c65e3999-5a0d-403a-9d95-7cd30b799efe · outbound

This paper cites Are you kidding me?!.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Are you kidding me?!

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:49.909689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:42.799145Z digest=sha256:42c2b5c1eb11fa2e163662a0b7e0d7ca29ad14b22dc8b95cccf248837c79f24a

Observation 9bbcac71-74b0-4c85-aebf-96dc035cd577 · outbound

This paper cites Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:16:45.591375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:42.871348Z digest=sha256:9de3b31a265dfd71e9e67f19c89d732130d80bb6a13a04b362691bb34ab14dea

Observation 88a1359c-8961-4727-9298-ab4f85eb3114 · outbound

This paper cites Segmentation-Variant Codebooks We encode speech into continuous representations using the frozen HuBERT-large model [17].

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Segmentation-Variant Codebooks We encode speech into continuous representations using the frozen HuBERT-large model [17]

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:49.763315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:42.966193Z digest=sha256:2e2be2df114fe1928b294a1e475eb94a5271c3911599a1ca59d6700dbec49f91

Observation 5ae054a2-2fe4-47d5-81a6-1155e9aade0f · outbound

This paper cites Datasets and Alignment Process Naver Prosody Control [21] is a dataset designed for study- ing prosody control in TTS systems.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Datasets and Alignment Process Naver Prosody Control [21] is a dataset designed for study- ing prosody control in TTS systems

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:49.572574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:43.000083Z digest=sha256:c08c6e8ddb02e61f81da5c5cca1dbca43b7c73f9e535478bfde062e6714ba977

Observation 795a7989-4c17-4e90-b2f6-b2e979a45480 · outbound

This paper cites an unresolved cited work.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:49.350000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:43.040761Z digest=sha256:1deb1258826fa41aa278a33d6d3ea71bad2be0330bb4284cd10b0ce22113c29e

Observation ece5708e-0da5-4225-88ec-d0e7b934f5fb · outbound

This paper cites an unresolved cited work.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:49.060304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:43.092557Z digest=sha256:cabd3c03a4738dcdebc4756cfb3f745fa625e3debc10746f5bf9b8900c958f1a

Observation 0db052a4-8a80-459f-a483-56bdb38dbcff · outbound

This paper cites The higher the score, the better.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information The higher the score, the better

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.907843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:43.197519Z digest=sha256:9f4f01cccb3f0394d55e33bf9aec4e747aa8b94c91da0f4e2aae1ac53e9107f1

Observation 6e272bc9-7c89-410f-a156-539b597a0daa · outbound

This paper cites This study contributes to ongoing research on the use of DSUs, demonstrating their po- tential in speech representation learning and downstream tasks.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information This study contributes to ongoing research on the use of DSUs, demonstrating their po- tential in speech representation learning and downstream tasks

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.689125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:43.270137Z digest=sha256:f732525a4efcb936d2aba47e3120a2e3ec4f8d839d5585eed5345e911426d46d

Observation 88d2db23-b85e-426a-9398-377d80e93580 · outbound

This paper cites an unresolved cited work.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:48.496560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:43.351918Z digest=sha256:c50b4f573640fede2d5a84530f6fafa57b5b670cd5e7783ccbd48bd4b492d15c

Observation 3f653aa5-3026-4b40-a1f9-f0df0b20afb4 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:43.399696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:43.399696Z digest=sha256:d58d167ba20ecb2e0c08f2b4d54fb9b2bd005b3ea46264c1e9ddd0ec1408be46

Observation 57c35fe4-7054-47c1-9511-7031d4396df1 · outbound

This paper cites Exploring speech recognition, translation, and understanding with discrete speech units: A com- parative study,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Exploring speech recognition, translation, and understanding with discrete speech units: A com- parative study,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.391693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:43.437788Z digest=sha256:5bb3ee661cafbceb5114da9f90b80633b2d53a8c83b25b6ef11c65b8bedbbc3b

Observation 8b8bae58-1fa5-4470-8ac8-5a7e55fd40a7 · outbound

This paper cites On the use of self-supervised speech representations in spontaneous speech synthesis,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information On the use of self-supervised speech representations in spontaneous speech synthesis,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.247460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:43.481755Z digest=sha256:93721d75559a1b4409747f98618b19caf5268791024ed33e3f2fd4b4e1043e84

Observation 35b2e03c-73d3-467e-af37-12c1cad7df1d · outbound

This paper cites Selecttts: Syn- thesizing anyone’s voice via discrete unit-based frame selection,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Selecttts: Syn- thesizing anyone’s voice via discrete unit-based frame selection,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.126458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:43.540400Z digest=sha256:766239550123a6fc212a195cae11e372d29ee8bc8d9c428dd0e9507019a4d773

Observation 72b2e005-e738-4c3b-a928-d33964c18e1f · outbound

This paper cites Exploration of a self- supervised speech model: A study on emotional corpora,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Exploration of a self- supervised speech model: A study on emotional corpora,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.995252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:43.626702Z digest=sha256:171e9153468ae7497d3285dd6877641223e3eb5d43f8b81b3d8d3802a0c2cb63

Observation b188d8a3-253b-459b-801c-2ca318becb79 · outbound

This paper cites Exploring Wav2vec 2.0 fine tun- ing for improved speech emotion recognition,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Exploring Wav2vec 2.0 fine tun- ing for improved speech emotion recognition,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.865776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:43.724297Z digest=sha256:82e847f07e14e41d390970748618b405fd87bd1c9c618abd97011299acfb0734

Observation d36e856a-ce21-4066-b3b0-096d6b906caa · outbound

This paper cites Layer-wise analysis of self-supervised acoustic word embeddings: A study on speech emotion recognition,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Layer-wise analysis of self-supervised acoustic word embeddings: A study on speech emotion recognition,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.707895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:43.797913Z digest=sha256:9d684642a69fb1cf64bcb99966e50b7cda7e97410518448cd34280e11c30cff3

Observation d14d4f7f-73dd-4beb-aead-02aa54898e94 · outbound

This paper cites Crossmodal ASR error correc- tion with discrete speech units,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Crossmodal ASR error correc- tion with discrete speech units,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.557302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:43.897140Z digest=sha256:3e71c60016980192cc820cb02b4bbe5dd5e3c9650778bf17cdb0b2ff680aee70

Observation 9c85d235-3631-4b2a-8c01-d735bd2b04d9 · outbound

This paper cites DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:43.941022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:43.941022Z digest=sha256:1b1085678763b184569f2c0c7dd2df7337dcfad947dd07fac4d9738abe7b235d

Observation 8c79a01d-fdc2-4364-a74b-257daa0a1297 · outbound

This paper cites Analyzing acoustic word embeddings from pre-trained self-supervised speech models,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Analyzing acoustic word embeddings from pre-trained self-supervised speech models,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.423247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:43.974819Z digest=sha256:06ebf354f39fc45a60ce151b38d102ccca0a1fb5a00cc5882a395c1ee75af4c9

Observation a11d4300-c46c-4f5a-90ea-c8b1a0c6d2a0 · outbound

This paper cites Ex- presso: A benchmark and analysis of discrete expressive speech resynthesis,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Ex- presso: A benchmark and analysis of discrete expressive speech resynthesis,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.278145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:44.049595Z digest=sha256:0d4f2a83264a69f9a29843f51f1ac9f744ec69ddcb0d8321efdd0a569026d5c4

Observation 6b5ad4b3-752d-4953-a95f-8457fd82eab2 · outbound

This paper cites Emo-codec: An in-depth look at emotion preservation capacity of legacy and neural codec models with subjective and objective evaluations,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Emo-codec: An in-depth look at emotion preservation capacity of legacy and neural codec models with subjective and objective evaluations,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.171102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:44.115212Z digest=sha256:a70f5a0260eae5683e8ffe8d83122d1b03f33b1489bb452a4e33c0d91bd0f805

Observation bbf3e34f-2dda-491a-a99d-b68783173e64 · outbound

This paper cites Neural discrete represen- tation learning,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Neural discrete represen- tation learning,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.156454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.156454Z digest=sha256:d847df35def01202b1d0ce5ff9df6d09bdbea5c22f0c12081e82ffc22675ef54

Observation 4c0be852-d92c-423b-88e8-88c50b532cce · outbound

This paper cites vq-wav2vec: Self- supervised learning of discrete speech representations,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information vq-wav2vec: Self- supervised learning of discrete speech representations,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.047351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:44.199984Z digest=sha256:b4be9256474121522d7e770fe5a42211aa8ca4945dff6c81d91bab7d5d8bd7fa

Observation 8eafc58b-0597-44a9-ac22-ff1ceba174d8 · outbound

This paper cites High Fidelity Neural Audio Compression.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information High Fidelity Neural Audio Compression

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.240679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.240679Z digest=sha256:6a28a53ae3856aa86188e8d64e617834430a578d9fe5d96a8919fc93ccfbdcfb

Observation 4d15db86-6185-4789-8f19-c1e0798d2143 · outbound

This paper cites Soundstream: An end-to-end neural audio codec,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Soundstream: An end-to-end neural audio codec,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.302788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.302788Z digest=sha256:8bdd76d656f5c501e0428414278d45c11d547ff0b485763469fb8f7bcfd5490f

Observation a408aa65-4d83-4929-b89a-a68074fb7725 · outbound

This paper cites Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.430963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.430963Z digest=sha256:10da75d1df6917edbee42853f23ed3c3580b4fc6b2ea329289624b23d2dfc3e1

Observation a7b29f47-7768-4684-8867-28bd5bc6838c · outbound

This paper cites Phonetic analysis of self-supervised representations of english speech,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Phonetic analysis of self-supervised representations of english speech,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.901418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:44.504728Z digest=sha256:280a84d6862b5b1f294f753c648c16b223b81c408e96584cec28ad4f7bdcbc93

Observation 7d56b6eb-9f5e-466b-a8d7-2fe03012f0e7 · outbound

This paper cites Analysing discrete self supervised speech representation for spoken language modeling,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Analysing discrete self supervised speech representation for spoken language modeling,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.791664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:44.560689Z digest=sha256:05601df3fd80cb30655550d041be54d25213e1746dac849b79f9f8a89c84d9d9

Observation 89422f5a-da23-452f-81c1-21090ab9d992 · outbound

This paper cites The Interspeech 2024 Challenge on Speech Processing Using Discrete Units.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information The Interspeech 2024 Challenge on Speech Processing Using Discrete Units

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.632634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.632634Z digest=sha256:7599bc2d89322bf426de5b4c4250fec33bf28d39f51d3111c5185af13a0e79cd

Observation 65ef997c-b806-4173-ae93-9404b37028de · outbound

This paper cites Controlling prosody in end-to-end TTS: A case study on contrastive focus generation,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Controlling prosody in end-to-end TTS: A case study on contrastive focus generation,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.672049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:44.708653Z digest=sha256:67a51df9dc71c45b7b58a9ad13aee05a76b7ed23c21f2ec327e0bc5c5bbb8f03

Observation 11f3a1a4-3e78-4702-b64c-88b285f5efd8 · outbound

This paper cites IEMOCAP: Interactive emotional dyadic motion capture database,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information IEMOCAP: Interactive emotional dyadic motion capture database,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.767109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.767109Z digest=sha256:f180228d3366a6062265fd1f07686fd3b456c29e5fcd89d906bdcf6fd84fe96e

Observation d0956122-9443-4441-8c7b-08d9fa0bf926 · outbound

This paper cites Montreal forced aligner: Trainable text-speech align- ment using Kaldi,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Montreal forced aligner: Trainable text-speech align- ment using Kaldi,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.487028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:44.798029Z digest=sha256:85e3b6058c05d026079dc0597cc2d7cbbb2fd2f137b362c707523190e08de89c

Observation 1afc9f5f-e468-4347-ba72-1dba800550b7 · outbound

This paper cites The HTK book,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information The HTK book,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.363396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:44.842896Z digest=sha256:2f5953d9c892194517df9e1f6b4d72c339c4c72a35707f856f3476d96591d216

Observation c9acfe29-2a41-421d-9171-ed54aad14be2 · outbound

This paper cites k-means++: The advantages of careful seeding,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information k-means++: The advantages of careful seeding,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.900505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.900505Z digest=sha256:18fd67d959f678e5d6c43d0e1f6022b6985b09f533a75309e792645a1275feef

Observation e277a6ef-0e09-4456-a720-351c89d10147 · outbound

This paper cites Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.972228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.972228Z digest=sha256:3043441320f72506bc6bd8398cc82c2295248296bd91d1c4228db90fac1be3e4

Observation fa69722a-ed3e-4821-98c7-1861a4afbdca · outbound

This paper cites Su- perb: Speech processing universal performance benchmark,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Su- perb: Speech processing universal performance benchmark,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.199145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:45.013346Z digest=sha256:841dfcb66a03c11dcbcffd1908f1abac8fb76473ad22d1af527bf5694e1f6636

Observation e54207ef-db25-4e0b-9637-4f8ba3a6ea1d · outbound

This paper cites A Fine-tuned Wav2vec 2.0/HuBERT Benchmark For Speech Emotion Recognition, Speaker Verification and Spoken Language Understanding.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information A Fine-tuned Wav2vec 2.0/HuBERT Benchmark For Speech Emotion Recognition, Speaker Verification and Spoken Language Understanding

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:45.044605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:45.044605Z digest=sha256:dc09eb03f6d856420115acfa26f2c8aa8b43965e9749670c9178eb93a7210549

Observation 910719cb-7426-431a-a585-091dc3b4fa4f · outbound

This paper cites Speech emotion di- arization: Which emotion appears when?.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Speech emotion di- arization: Which emotion appears when?

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.101884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:45.083406Z digest=sha256:3e1947842f0e0218467580979c70b2e2bf2cbd41744ce6d6e9bc58cc138806d6

Observation 88baefd7-917e-467f-99c3-cda2827ab1bc · outbound

This paper cites EmphAssess : a Prosodic Benchmark on Assessing Emphasis Transfer in Speech-to-Speech Models.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information EmphAssess : a Prosodic Benchmark on Assessing Emphasis Transfer in Speech-to-Speech Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:45.122530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:45.122530Z digest=sha256:545391d9b37d3a6ded94ef2a1761e7b002eaf0bd1fb2ce85e971e55f9ea6da47

Observation 19febf02-e693-4e25-bf7f-eaebb37e1834 · outbound

This paper cites UTMOS: UTokyo-SaruLab system for voice- mos challenge 2022,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information UTMOS: UTokyo-SaruLab system for voice- mos challenge 2022,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.932575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:45.173094Z digest=sha256:9bbef174fb8fa55ebfe1cd46a2621bcefef137ea578cc85499fdba79f6e576e4

Observation a6a144a5-47dc-4424-ac23-037deb7e233c · outbound

This paper cites A layer-wise analysis of man- darin and english suprasegmentals in ssl speech models,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information A layer-wise analysis of man- darin and english suprasegmentals in ssl speech models,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.774824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:45.232249Z digest=sha256:63aed7cae645aa7a3c933916fb7c0fb993a50ff334208675b577cff10da27afb

Observation ad4fea51-e328-4b6c-84fe-4195775c3f69 · outbound

This paper cites blind speech segmentation: au- tomatic segmentation of speech without linguistic knowledge,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information blind speech segmentation: au- tomatic segmentation of speech without linguistic knowledge,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.662287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:45.290095Z digest=sha256:b0f87319d71820ef587e346eb12548af5bb4707936efd31d386c7179c879e9de

Pith citing papers

Observation 9bbcac71-74b0-4c85-aebf-96dc035cd577 · inbound

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information cites this paper.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:16:45.591375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T15:16:42.871348Z digest=sha256:9de3b31a265dfd71e9e67f19c89d732130d80bb6a13a04b362691bb34ab14dea