Pith. sign in

Paper Citation Record · LEDGER

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information

As of 8 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2505.15667.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15667 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:16:45.290095Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:16:42.871348Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T15:16:45.530174Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact1
  • verified fuzzy26
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5c1a7c5a-f82f-48c0-99d1-dadd05f77c3e · outbound

This paper cites an unresolved cited work.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:50.090889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:42.724908Z digest=sha256:7033b4d9502af1ba43afceba2a71d99a0ec12d1e10c829050efaf4cb5f39c9a2

Observation c65e3999-5a0d-403a-9d95-7cd30b799efe · outbound

This paper cites Are you kidding me?!.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Are you kidding me?!

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:49.909689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:42.799145Z digest=sha256:d0d05ebb8fcc10c2ac705ec998ee44be1240ec5d94a5943523446873aeae44b4

Observation 9bbcac71-74b0-4c85-aebf-96dc035cd577 · outbound

This paper cites Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:16:45.591375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:42.871348Z digest=sha256:dc405899092cc921368f587f2279b76860829c8a5647f5810c8e51b878a86540

Observation 88a1359c-8961-4727-9298-ab4f85eb3114 · outbound

This paper cites Segmentation-Variant Codebooks We encode speech into continuous representations using the frozen HuBERT-large model [17].

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Segmentation-Variant Codebooks We encode speech into continuous representations using the frozen HuBERT-large model [17]

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:49.763315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:42.966193Z digest=sha256:18e1a2285fda05e19627257878afc0e8f23fe87ffcb339a4d43f4dfe34ed0670

Observation 5ae054a2-2fe4-47d5-81a6-1155e9aade0f · outbound

This paper cites Datasets and Alignment Process Naver Prosody Control [21] is a dataset designed for study- ing prosody control in TTS systems.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Datasets and Alignment Process Naver Prosody Control [21] is a dataset designed for study- ing prosody control in TTS systems

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:49.572574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:43.000083Z digest=sha256:26b2bb1e863c24108e064ed25e2722616d0f55d86b6f503e8e8d0ccace1d99b8

Observation 795a7989-4c17-4e90-b2f6-b2e979a45480 · outbound

This paper cites an unresolved cited work.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:49.350000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:43.040761Z digest=sha256:97c44ab356e2f0b7e20064789001cd542c3c23427d547da17702462af2b89d64

Observation ece5708e-0da5-4225-88ec-d0e7b934f5fb · outbound

This paper cites an unresolved cited work.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:49.060304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:43.092557Z digest=sha256:961480741a5053633d3c43a7fe6bf972d86f53ca953bd63f8ccd1d5a8338121b

Observation 0db052a4-8a80-459f-a483-56bdb38dbcff · outbound

This paper cites The higher the score, the better.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information The higher the score, the better

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.907843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:43.197519Z digest=sha256:358975d7467eacfabce53dcbd83c9120a68af12a5d5690d042b408500962b30e

Observation 6e272bc9-7c89-410f-a156-539b597a0daa · outbound

This paper cites This study contributes to ongoing research on the use of DSUs, demonstrating their po- tential in speech representation learning and downstream tasks.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information This study contributes to ongoing research on the use of DSUs, demonstrating their po- tential in speech representation learning and downstream tasks

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.689125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:43.270137Z digest=sha256:77e89cbb111f1b1f7cedbfc690bce39f7505372d8b2524bd6a606975a7dc32d2

Observation 88d2db23-b85e-426a-9398-377d80e93580 · outbound

This paper cites an unresolved cited work.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:48.496560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:43.351918Z digest=sha256:89adfb39a28efbd22bca3d1635214825a6091f4e306d31d2ce63a4110f907d8b

Observation 3f653aa5-3026-4b40-a1f9-f0df0b20afb4 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:43.399696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:43.399696Z digest=sha256:89fb2faef72468cd0bc2739677752be14a15ec7f93fc220c91d8844087b32da2

Observation 57c35fe4-7054-47c1-9511-7031d4396df1 · outbound

This paper cites Exploring speech recognition, translation, and understanding with discrete speech units: A com- parative study,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Exploring speech recognition, translation, and understanding with discrete speech units: A com- parative study,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.391693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:43.437788Z digest=sha256:89b8872dd3e18d706026931d8cf3dd6a750220020032b9b0eddc41a2ef750125

Observation 8b8bae58-1fa5-4470-8ac8-5a7e55fd40a7 · outbound

This paper cites On the use of self-supervised speech representations in spontaneous speech synthesis,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information On the use of self-supervised speech representations in spontaneous speech synthesis,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.247460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:43.481755Z digest=sha256:bf88f7462aa6a8e00d8a9eaba018a4279c852a999984636f1c797e489e8cdb5d

Observation 35b2e03c-73d3-467e-af37-12c1cad7df1d · outbound

This paper cites Selecttts: Syn- thesizing anyone’s voice via discrete unit-based frame selection,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Selecttts: Syn- thesizing anyone’s voice via discrete unit-based frame selection,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.126458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:43.540400Z digest=sha256:39381bc061235c8aadfd0e56d0327e8c30d75e8ea958c8e13a236f6ec63381f6

Observation 72b2e005-e738-4c3b-a928-d33964c18e1f · outbound

This paper cites Exploration of a self- supervised speech model: A study on emotional corpora,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Exploration of a self- supervised speech model: A study on emotional corpora,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.995252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:43.626702Z digest=sha256:7206eea3d5be348de842a5d20e973d1085dc052e61a748edf37dd63c83cd8523

Observation b188d8a3-253b-459b-801c-2ca318becb79 · outbound

This paper cites Exploring Wav2vec 2.0 fine tun- ing for improved speech emotion recognition,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Exploring Wav2vec 2.0 fine tun- ing for improved speech emotion recognition,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.865776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:43.724297Z digest=sha256:e19040a27a83f1d533cbd27bc85873e46afa583f1a9c529d4560959b8b1698a4

Observation d36e856a-ce21-4066-b3b0-096d6b906caa · outbound

This paper cites Layer-wise analysis of self-supervised acoustic word embeddings: A study on speech emotion recognition,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Layer-wise analysis of self-supervised acoustic word embeddings: A study on speech emotion recognition,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.707895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:43.797913Z digest=sha256:5a0992d7f4ff59d8439163a0a47a6cbf5cca743fe9432fde1d2522e87cc32faa

Observation d14d4f7f-73dd-4beb-aead-02aa54898e94 · outbound

This paper cites Crossmodal ASR error correc- tion with discrete speech units,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Crossmodal ASR error correc- tion with discrete speech units,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.557302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:43.897140Z digest=sha256:3b88471a764d8271a3cd32b672d428404e598b1bab9c5a8dd3083ac4d5208697

Observation 9c85d235-3631-4b2a-8c01-d735bd2b04d9 · outbound

This paper cites DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:43.941022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:43.941022Z digest=sha256:dc2de3312d154fb75bec19d43756b9cef8e17f010095ee3ac5c65a97c0e7b4b3

Observation 8c79a01d-fdc2-4364-a74b-257daa0a1297 · outbound

This paper cites Analyzing acoustic word embeddings from pre-trained self-supervised speech models,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Analyzing acoustic word embeddings from pre-trained self-supervised speech models,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.423247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:43.974819Z digest=sha256:f327b7f8ae52520f47df13116a3ae3b212dd95157ee2ab061e487f4e96507e3f

Observation a11d4300-c46c-4f5a-90ea-c8b1a0c6d2a0 · outbound

This paper cites Ex- presso: A benchmark and analysis of discrete expressive speech resynthesis,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Ex- presso: A benchmark and analysis of discrete expressive speech resynthesis,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.278145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:44.049595Z digest=sha256:774a43a51b6256666200820d2bbb28f81f3a4afc09150e88a55d77034bf442eb

Observation 6b5ad4b3-752d-4953-a95f-8457fd82eab2 · outbound

This paper cites Emo-codec: An in-depth look at emotion preservation capacity of legacy and neural codec models with subjective and objective evaluations,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Emo-codec: An in-depth look at emotion preservation capacity of legacy and neural codec models with subjective and objective evaluations,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.171102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:44.115212Z digest=sha256:1c09add18864021fd58e6c44f1143cece8de779bbc57c531b698e60a856f8141

Observation bbf3e34f-2dda-491a-a99d-b68783173e64 · outbound

This paper cites Neural discrete represen- tation learning,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Neural discrete represen- tation learning,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.156454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.156454Z digest=sha256:d2e4480b889d6b4cf6914da433384a691c4c9befff68a923e4545dafbbf4121b

Observation 4c0be852-d92c-423b-88e8-88c50b532cce · outbound

This paper cites vq-wav2vec: Self- supervised learning of discrete speech representations,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information vq-wav2vec: Self- supervised learning of discrete speech representations,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.047351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:44.199984Z digest=sha256:673d6fbfc1008bfbd2521cc798ecb9466a72ef9fdaa8d93857fda95d63a92bf0

Observation 8eafc58b-0597-44a9-ac22-ff1ceba174d8 · outbound

This paper cites High Fidelity Neural Audio Compression.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information High Fidelity Neural Audio Compression

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.240679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.240679Z digest=sha256:dda27fb2c511ecb0666fed70e6e483a6c1cc161641668842495e6482ae9d23e1

Observation 4d15db86-6185-4789-8f19-c1e0798d2143 · outbound

This paper cites Soundstream: An end-to-end neural audio codec,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Soundstream: An end-to-end neural audio codec,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.302788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.302788Z digest=sha256:55182877b2c1b488dc9910d5b61ef69537571b1baedb43ce4c8f3c5c90bcbe25

Observation a408aa65-4d83-4929-b89a-a68074fb7725 · outbound

This paper cites Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.430963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.430963Z digest=sha256:98eaa3513bb04e67a1c54a4280aed898852814ab7ba1e8fa4091624e985d8356

Observation a7b29f47-7768-4684-8867-28bd5bc6838c · outbound

This paper cites Phonetic analysis of self-supervised representations of english speech,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Phonetic analysis of self-supervised representations of english speech,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.901418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:44.504728Z digest=sha256:f562977c74468613875c010044dd896aee156e51192679901667fb585bc9bca6

Observation 7d56b6eb-9f5e-466b-a8d7-2fe03012f0e7 · outbound

This paper cites Analysing discrete self supervised speech representation for spoken language modeling,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Analysing discrete self supervised speech representation for spoken language modeling,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.791664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:44.560689Z digest=sha256:b6b7aaa9eb8deaef4d02e4f5b79d4a2a5bbe1dcc18a1d257b794830e305ed908

Observation 89422f5a-da23-452f-81c1-21090ab9d992 · outbound

This paper cites The Interspeech 2024 Challenge on Speech Processing Using Discrete Units.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information The Interspeech 2024 Challenge on Speech Processing Using Discrete Units

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.632634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.632634Z digest=sha256:8feb7103d313f10c8f15968f87d2d6f1be3ec60ba53e7c30195c6ed7441730c8

Observation 65ef997c-b806-4173-ae93-9404b37028de · outbound

This paper cites Controlling prosody in end-to-end TTS: A case study on contrastive focus generation,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Controlling prosody in end-to-end TTS: A case study on contrastive focus generation,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.672049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:44.708653Z digest=sha256:a9cb88995f39bc45fc5e2934e2027346ee90f792cc62282fff584ac937287741

Observation 11f3a1a4-3e78-4702-b64c-88b285f5efd8 · outbound

This paper cites IEMOCAP: Interactive emotional dyadic motion capture database,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information IEMOCAP: Interactive emotional dyadic motion capture database,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.767109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.767109Z digest=sha256:166652d8132a8007e4da73aae43532522b5a85fba417a9f490710860a9a53064

Observation d0956122-9443-4441-8c7b-08d9fa0bf926 · outbound

This paper cites Montreal forced aligner: Trainable text-speech align- ment using Kaldi,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Montreal forced aligner: Trainable text-speech align- ment using Kaldi,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.487028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:44.798029Z digest=sha256:ffbfd68834e1f422c275f3e5d4ef6c7e0fe7aeb994787ad16ea30caae6fccb72

Observation 1afc9f5f-e468-4347-ba72-1dba800550b7 · outbound

This paper cites The HTK book,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information The HTK book,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.363396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:44.842896Z digest=sha256:5a39fc17ed52ba0100c73e3dc23ee462cf4ba32ab683642bcbdc8f60fb31cddc

Observation c9acfe29-2a41-421d-9171-ed54aad14be2 · outbound

This paper cites k-means++: The advantages of careful seeding,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information k-means++: The advantages of careful seeding,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.900505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.900505Z digest=sha256:c32afe2b4a0be5b81f2d726b4e2637fc3172a6568e54bb9c17dc9dbbda5aad60

Observation e277a6ef-0e09-4456-a720-351c89d10147 · outbound

This paper cites Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.972228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.972228Z digest=sha256:e1e88872fbbea92ab1d6171d29ae822a247a0ed173cd17480ae45b44762e95c3

Observation fa69722a-ed3e-4821-98c7-1861a4afbdca · outbound

This paper cites Su- perb: Speech processing universal performance benchmark,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Su- perb: Speech processing universal performance benchmark,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.199145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:45.013346Z digest=sha256:e5359e35c4eda4cefac0ccacde29d01b5e9c4a1cc171de37bb88db35484773e9

Observation e54207ef-db25-4e0b-9637-4f8ba3a6ea1d · outbound

This paper cites A Fine-tuned Wav2vec 2.0/HuBERT Benchmark For Speech Emotion Recognition, Speaker Verification and Spoken Language Understanding.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information A Fine-tuned Wav2vec 2.0/HuBERT Benchmark For Speech Emotion Recognition, Speaker Verification and Spoken Language Understanding

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:45.044605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:45.044605Z digest=sha256:4134e0d63ae513871b991f6e824d6074af61a6f12912809041f92ee16775eb92

Observation 910719cb-7426-431a-a585-091dc3b4fa4f · outbound

This paper cites Speech emotion di- arization: Which emotion appears when?.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Speech emotion di- arization: Which emotion appears when?

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.101884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:45.083406Z digest=sha256:8ee326edf3c6a8ec461e8e286bc92863b985e054588a0e6158d3bce116d50e9c

Observation 88baefd7-917e-467f-99c3-cda2827ab1bc · outbound

This paper cites EmphAssess : a Prosodic Benchmark on Assessing Emphasis Transfer in Speech-to-Speech Models.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information EmphAssess : a Prosodic Benchmark on Assessing Emphasis Transfer in Speech-to-Speech Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:45.122530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:45.122530Z digest=sha256:83766d3eefa3074994e8e34b4e638f29fe4cb131f943db63a20847f5cf4707c6

Observation 19febf02-e693-4e25-bf7f-eaebb37e1834 · outbound

This paper cites UTMOS: UTokyo-SaruLab system for voice- mos challenge 2022,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information UTMOS: UTokyo-SaruLab system for voice- mos challenge 2022,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.932575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:45.173094Z digest=sha256:c5ca0a3270e5411d7399a211fcfc0a9354d2159646312057a3d0e7b5bb5637ea

Observation a6a144a5-47dc-4424-ac23-037deb7e233c · outbound

This paper cites A layer-wise analysis of man- darin and english suprasegmentals in ssl speech models,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information A layer-wise analysis of man- darin and english suprasegmentals in ssl speech models,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.774824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:45.232249Z digest=sha256:040815bd3c86aac823e8465a712d66db0cfe54db6e84beea516c2e21bc337cfd

Observation ad4fea51-e328-4b6c-84fe-4195775c3f69 · outbound

This paper cites blind speech segmentation: au- tomatic segmentation of speech without linguistic knowledge,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information blind speech segmentation: au- tomatic segmentation of speech without linguistic knowledge,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.662287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:45.290095Z digest=sha256:5cf5713cbd187b3ce3e2df134fd88dae6dc8de2b11fac8188f291528d048ee9d

Pith citing papers

Observation 9bbcac71-74b0-4c85-aebf-96dc035cd577 · inbound

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information cites this paper.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:16:45.591375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:16:42.871348Z digest=sha256:dc405899092cc921368f587f2279b76860829c8a5647f5810c8e51b878a86540