Pith. sign in

Paper Citation Record · LEDGER

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech

As of 20 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 1 inbound Pith citation observation for arXiv:2506.01618.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.01618 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:44:12.195614Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:44:11.728259Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T11:44:12.241972Z

Reference resolution

28 of 28 outbound references displayed

  • verified exact0
  • verified fuzzy25
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ad81ae64-0f7b-4bb5-b196-3f8212a33528 · outbound

This paper cites As a result, Automatic Speech Recognition (ASR) systems trained on typ- ical speech often struggle to process dysarthric speech accu- rately [2].

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech As a result, Automatic Speech Recognition (ASR) systems trained on typ- ical speech often struggle to process dysarthric speech accu- rately [2]

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:14.716260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:11.712020Z digest=sha256:30a8298afdef5b8f36342a965654653ad509198f90d2857cd0a7410a7e91e72c

Observation 45e1977f-c37e-4893-ac5a-e7fcaf0c9bea · outbound

This paper cites Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T11:44:12.260141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:11.728259Z digest=sha256:8d343b0d28ccb629120f5669108c6c664545bf674f3851abe25acdf7af2f1fd4

Observation ea05b2a9-141f-4af2-b352-8260ab674575 · outbound

This paper cites We further train and adapt ASR models on the converted speech to assess more thoroughly whether conversion helps improve recognition performance.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech We further train and adapt ASR models on the converted speech to assess more thoroughly whether conversion helps improve recognition performance

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:14.674799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:11.747659Z digest=sha256:4d71b9b5e8f21c4c65146b79319205ea661548ab904208dabc00ef68dab2b20b

Observation 29d2ca52-509f-4ae2-992e-a48e86f383fd · outbound

This paper cites RnV implementation We implement the framework similarly to [6].

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech RnV implementation We implement the framework similarly to [6]

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:14.632338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:11.768861Z digest=sha256:7e6cc901a58ec22163577dae2cd7b088dc8011ddcc2b6e159badedc10e8babe6

Observation 25ec0a5f-a3f1-456f-82bd-e289f280ab64 · outbound

This paper cites We can observe that speaking rates increase with lower severity levels as expected.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech We can observe that speaking rates increase with lower severity levels as expected

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:14.555941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:11.821575Z digest=sha256:9d6f6f667434b979985ffc0a6facd32f74c4a52b8c43f30315c851c697e38bd2

Observation ec18420e-5145-45f5-8ca4-dad4fa1ffaa2 · outbound

This paper cites The clear correlation between speaking rate and dysarthria severity supports this approach, as speaking rate in- creases with lower severity.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech The clear correlation between speaking rate and dysarthria severity supports this approach, as speaking rate in- creases with lower severity

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:14.513500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:11.840332Z digest=sha256:9aca59b8965dfdd0db3d873d1e94219a8f9aae1a68f3118d96fb3881f5309388

Observation 60897f04-1e08-4895-b8f4-ce1284d9d296 · outbound

This paper cites Pathological Speech Synthesis (PaSS).

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech Pathological Speech Synthesis (PaSS)

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:14.472797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:11.864753Z digest=sha256:1ddf6de3a2a3dfbd0c938de5d7421f3245391db6ff40005b965fe9279344490b

Observation 7f509ffe-d893-4063-ae1b-ae042e573ad5 · outbound

This paper cites Purely Sequence-Trained Neural Networks for ASR Based on Lattice-Free MMI,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech Purely Sequence-Trained Neural Networks for ASR Based on Lattice-Free MMI,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:14.096978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:12.007554Z digest=sha256:38bf560a6c2a58b97ee521b5800ce89886daf437e5c41cf9998339246ffe94e4

Observation e1f7bb4d-453e-4f8f-8adb-7c8d2361607d · outbound

This paper cites an unresolved cited work.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:44:14.427202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:11.884607Z digest=sha256:dfd9bf0fbdf571feb6f7a4309b3f21c788dd4b0ce639c5392ffe1b8bb884fae1

Observation cd796991-b276-4294-8b30-c141bf85c307 · outbound

This paper cites Whistle- blowing ASRs: Evaluating the Need for More Inclusive Speech Recognition Systems,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech Whistle- blowing ASRs: Evaluating the Need for More Inclusive Speech Recognition Systems,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:14.370058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:11.899700Z digest=sha256:7249feb64ff76667501b2b10d5e929f0ae301460671122d50092a7594c95c8af

Observation 1797803c-5ba9-4151-90b2-4ff9e1d629d7 · outbound

This paper cites Synthesizing Dysarthric Speech Using Multi-Speaker TTS For Dysarthric Speech Recognition,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech Synthesizing Dysarthric Speech Using Multi-Speaker TTS For Dysarthric Speech Recognition,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:14.327722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:11.916066Z digest=sha256:cbc66c7a5530930b9c0d64fd919e7d5e51a5fa0f779a286f5967df8a85f004f7

Observation 5b3d3d72-3b33-4e85-96a6-051307373925 · outbound

This paper cites Few-shot Dysarthric Speech Recognition with Text-to-Speech Data Augmentation,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech Few-shot Dysarthric Speech Recognition with Text-to-Speech Data Augmentation,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:14.270207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:11.932724Z digest=sha256:38991c58616862a2fbe0fa19adb7297fc009f921be42d1be602a0c6eeb25440f

Observation 44a58de9-3b59-4a58-b2c9-769681abf7d5 · outbound

This paper cites Training Data Augmentation for Dysarthric Automatic Speech Recognition by Text-to-Dysarthric-Speech Synthesis,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech Training Data Augmentation for Dysarthric Automatic Speech Recognition by Text-to-Dysarthric-Speech Synthesis,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:14.226461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:11.954136Z digest=sha256:ec0d4a582a74a101a9667cdde1551292ae9b9d7bceb3ce9b406058f04b908b41

Observation 1f30b39c-7b1f-4ff1-b597-cfc6bea7c54b · outbound

This paper cites Unsupervised rhythm and voice conversion of dysarthric to healthy speech for ASR,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech Unsupervised rhythm and voice conversion of dysarthric to healthy speech for ASR,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:14.182796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:11.969545Z digest=sha256:30044b9a46c76d160dfe5b214a0e7e6f48d2887fc9277e185bf77dbadaa3f58d

Observation dfe45953-b9de-475d-8eaa-94d2e4636409 · outbound

This paper cites The TORGO database of acoustic and articulatory speech from speakers with dysarthria,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech The TORGO database of acoustic and articulatory speech from speakers with dysarthria,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:14.143463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:11.989189Z digest=sha256:7cff1f682d9e641bba574a5fb621e159d84dfab9d98d8112877da55c2bb9773e

Observation c3c188df-5b1a-4368-bcbc-c22b294cdd25 · outbound

This paper cites HiFi-GAN: Generative adversarial networks for efficient and high fidelity speech synthesis,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech HiFi-GAN: Generative adversarial networks for efficient and high fidelity speech synthesis,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:12.511868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:12.130827Z digest=sha256:834d5abe40ea1ec2f9f3651b0335d4b574a4e22fad7d65d23e3e84a659df90af

Observation be460ca5-a731-455d-b9b6-48d51a510e42 · outbound

This paper cites Robust speech recognition via large-scale weak su- pervision,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech Robust speech recognition via large-scale weak su- pervision,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:44:12.024400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:44:12.024400Z digest=sha256:46da20fc71f35e18861ebd5f0aa569a37330795d0d88d16b0ddf6db4b5bf1f8c

Observation b2f82749-1a2a-45ef-99ec-7ffd58bf77bb · outbound

This paper cites The models are trained in Kaldi [19] using the train- ing recipe from [20], i.e.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech The models are trained in Kaldi [19] using the train- ing recipe from [20], i.e

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:14.597032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:11.793383Z digest=sha256:63bb8bdfb7b40983459c036f9c2e6e24fc3b05af8c647117c2083ad42feecc0c

Observation b78cb52b-4a78-4e2a-89e0-6ea0920e1971 · outbound

This paper cites Rhythm modeling for voice conversion,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech Rhythm modeling for voice conversion,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:14.034203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:12.039615Z digest=sha256:c043f229e2bae6813577d0bcf4fd07d3d8405cb13eb664af7f2927bab5a95d6c

Observation 536d63cc-f246-4139-96a8-669c80ffc76b · outbound

This paper cites V oice conversion with just nearest neighbors,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech V oice conversion with just nearest neighbors,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:13.990482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:12.054024Z digest=sha256:2bba911012668b5214d33a5af459df14a78fbead645e98582f4e737ee721a2ec

Observation 5ed471bc-3c76-4296-acea-1511c628ad97 · outbound

This paper cites Estimating the speaking rate by vowel detection,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech Estimating the speaking rate by vowel detection,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:13.951588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:12.069051Z digest=sha256:0c9f4ac06f3c7e48a0e87c18f2d33c0234af9b145c284d51bd616bbee50e3c66

Observation bb45322b-c199-4384-b063-68eb407324c0 · outbound

This paper cites Syllable level features for Parkinson’s disease detection from speech,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech Syllable level features for Parkinson’s disease detection from speech,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:13.907993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:12.086341Z digest=sha256:bc53645bea61209ba17c363155278198e8d20950d8fdafb1b445b3034510867f

Observation cfa1c791-9509-4cfa-81b6-b1d06417b420 · outbound

This paper cites Pre-linguistic segmen- tation of speech into syllable-like units,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech Pre-linguistic segmen- tation of speech into syllable-like units,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:12.602034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:12.103914Z digest=sha256:0bd8a57c49aa8e709a5549ef76735124ef8bbff4505abeeed3a6a2b52726b99b

Observation 8ddc97db-bcc5-4ad3-9451-635bf6f73fa7 · outbound

This paper cites WavLM: Large- scale self-supervised pre-training for full stack speech process- ing,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech WavLM: Large- scale self-supervised pre-training for full stack speech process- ing,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:12.559326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:12.116656Z digest=sha256:113c4ea411eefefabcc2af35ba552da23c1d945af8347afffaea7dcfa0981ac0

Observation fe6e94c1-1b68-4fd9-96e7-ee6bb9448f1b · outbound

This paper cites The LJ speech dataset,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech The LJ speech dataset,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:12.454431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:12.144264Z digest=sha256:3f41c3d82927f71ee07b5af23eaeaba1f0e07a15c350775fa9d07fac2ce31522

Observation f007e083-ef2e-4127-a262-f6425dcef2bd · outbound

This paper cites Semi-Orthogonal Low-Rank Matrix Factor- ization for Deep Neural Networks,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech Semi-Orthogonal Low-Rank Matrix Factor- ization for Deep Neural Networks,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:12.407372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:12.163114Z digest=sha256:10f1cc75d6417f31896031d2e302cbc044fb37243be685c791348af1747e9869

Observation a99d679a-69e8-4068-a6c3-e39aeb03f551 · outbound

This paper cites The Kaldi Speech Recognition Toolkit,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech The Kaldi Speech Recognition Toolkit,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:12.356703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:12.182663Z digest=sha256:201455fed332dea7226870d96abcb8b9a8baec188d90096853098ff53dd32ff6

Observation d96f659b-562f-4d57-9a9d-5bc446672340 · outbound

This paper cites Dysarthric speech recog- nition with lattice-free MMI,.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech Dysarthric speech recog- nition with lattice-free MMI,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:44:12.295861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:12.195614Z digest=sha256:6d755f0250ca8e6db7becd54ab25357f6b8521b61e0419369ec671210ade2430

Pith citing papers

Observation 45e1977f-c37e-4893-ac5a-e7fcaf0c9bea · inbound

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech cites this paper.

Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T11:44:12.260141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:44:11.728259Z digest=sha256:8d343b0d28ccb629120f5669108c6c664545bf674f3851abe25acdf7af2f1fd4