Pith. sign in

Paper Citation Record · LEDGER

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition

As of 4 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 1 inbound Pith citation observation for arXiv:2606.06211.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.06211 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T01:58:13.871771Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T01:58:13.871771Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-02T12:36:57.080220Z

Reference resolution

28 of 28 outbound references displayed

  • verified exact6
  • verified fuzzy0
  • unresolved20
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cd7c94bd-c407-44a8-91a5-04a49d8a79a2 · outbound

This paper cites Yet these sys- tems continue to struggle with speech produced by individuals with neuromotor disorders such as amyotrophic lateral sclerosis (ALS) or Parkinson’s disease (PD).

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Yet these sys- tems continue to struggle with speech produced by individuals with neuromotor disorders such as amyotrophic lateral sclerosis (ALS) or Parkinson’s disease (PD)

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:4ef373f568eed269ffd99fd38e5a26393bdf2c9bff4b69c9fa7033c568458f58

Observation 3ffaa63b-ca08-487a-b85e-422a68904470 · outbound

This paper cites FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T12:36:57.081596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:a1ffa8c5107329ce509d1b38ac7138f04e9852463f86830b17068aeefa830f76

Observation 44abbcee-3d31-4ed3-b2ff-9c3061b79fd8 · outbound

This paper cites What is the sex of the speaker?.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition What is the sex of the speaker?

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:313996ec5791eb22c1ee62f794c0fa16345c5eaba35ab4337b7acc8a6158bb6b

Observation da710cfc-d670-4881-b010-2af3043d477c · outbound

This paper cites Table 1:WER (%) of Voxtral-Mini and adapted variants on the NeuroVoz and TORGO test sets.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Table 1:WER (%) of Voxtral-Mini and adapted variants on the NeuroVoz and TORGO test sets

Reference 4

Resolution
malformed identifier
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:d4100ea04a0d3c9d907f3c60b1999aba44a1affa5cf47a82756ed97a08715455

Observation 097dd103-e412-4bd1-bfc8-709966e72d2f · outbound

This paper cites However, this requires prior knowl- edge of whether the input speech is pathological, which may not always be available in practice.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition However, this requires prior knowl- edge of whether the input speech is pathological, which may not always be available in practice

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:79adbd1f96d6a95189d4b882f2552d4c706c737a2949458e7240c6efc60d4238

Observation 036a036d-ec36-4b9a-90b5-34456c3f2b4a · outbound

This paper cites It is a lightweight approach to pathological speech adaptation that maintains the base model untouched.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition It is a lightweight approach to pathological speech adaptation that maintains the base model untouched

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:d1157ace2b91c306a298ac53f47d338599c4ba743dca58458d8101b958c883c0

Observation 9457d3f2-a074-41a2-8a06-5d7924d52f64 · outbound

This paper cites 101135916).

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition 101135916)

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:4ded409a9596218c420835edd26f78becc24566d525a6f3cad20790cfc59000c

Observation 96a9a6cf-00e6-43d5-a1d1-76cc01b47d5c · outbound

This paper cites Automatic speech recognition: A survey of deep learning techniques and ap- proaches,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Automatic speech recognition: A survey of deep learning techniques and ap- proaches,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:6fd2ca20f4735e131e92ecc7db9b45a128c2b7e1dd9ba5e81038b95719d04bcf

Observation b1f45bb2-9ac9-4f78-91f1-feb9f507dbc2 · outbound

This paper cites The torgo database of acoustic and articulatory speech from speakers with dysarthria,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition The torgo database of acoustic and articulatory speech from speakers with dysarthria,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:d8e3e49e61665915e552fc840fa8381d6c439db015cc68720ddb128fcfe1d84e

Observation b5aef97d-0fc8-4347-a41f-16d5a7645a26 · outbound

This paper cites Dysarthric speech database for universal access research.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Dysarthric speech database for universal access research

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:6e0da536b88cc1d241248df6908aa612aedd7a0db58f6d0574042bd30536b802

Observation b5698af9-ff93-49eb-adaa-d764956d2f32 · outbound

This paper cites The interspeech 2025 speech accessibility project challenge,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition The interspeech 2025 speech accessibility project challenge,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:b507834d47153f629e677975be403263f40a0798dcfb260f1c33b60199f01faa

Observation cefad5ca-443f-4ed3-813d-9fcdacc0384b · outbound

This paper cites New spanish speech cor- pus database for the analysis of people suffering from parkinson’s disease.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition New spanish speech cor- pus database for the analysis of people suffering from parkinson’s disease

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:f877da13fa375dc452ba347f349caf5d44407e88abbc82c3dd0666afddad4df0

Observation b27da1e1-e2e8-4d44-acd3-5c95a16b67d6 · outbound

This paper cites Neurovoz: a castillian spanish corpus of parkinsonian speech,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Neurovoz: a castillian spanish corpus of parkinsonian speech,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:af1694f03a1a42fb2b616c617e2b4f0f63ee1303bbb52cea1da5285aaebbabee

Observation 6f6f906f-defd-4428-a84a-4f797e5ecf97 · outbound

This paper cites Neurovoz: a castillian spanish corpus of parkinsonian speech,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Neurovoz: a castillian spanish corpus of parkinsonian speech,

Reference 14

Resolution
verified exact
doi, observed 2026-06-28T02:01:28.924627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:d41ff8a300059ee1278e065c8d117ed38a9364e1c35d3803200b05925bceeb2b

Observation 65e55592-1eb3-487a-88ac-2961d7b97795 · outbound

This paper cites Speaker adaptation for Wav2vec2 based dysarthric ASR.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Speaker adaptation for Wav2vec2 based dysarthric ASR

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:36:57.095835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:a227eb94d05a015f0c18565c900212d2316b7b7bcc89ecd021d6487925e8422e

Observation 85a46a3d-6636-41be-82ca-2cf38ff88f6d · outbound

This paper cites Use of speech impairment severity for dysarthric speech recognition,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Use of speech impairment severity for dysarthric speech recognition,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:2e617c29e24ecb7920f06e8355b0b0cca139b498477249ded9a94e5c0bf46b6a

Observation db331eaf-b467-496b-bcd1-16b032982589 · outbound

This paper cites Personalized fine-tuning with controllable synthetic speech from llm-generated transcripts for dysarthric speech recognition,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Personalized fine-tuning with controllable synthetic speech from llm-generated transcripts for dysarthric speech recognition,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:a7a36c3ec00b0df7b36703b4f7b3b8eba1806f6f9cca1d971def357134d53868

Observation 86126588-6e8e-480b-b74a-dffa74f35dda · outbound

This paper cites Qwen3-ASR Technical Report.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Qwen3-ASR Technical Report

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-02T12:36:57.092874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:6d3a90cbcf8a0af78dfa6c6b9449659a97938df5044323ccbab9cb33715c1d6a

Observation 810f0b79-c138-419c-9bec-1d2fc7ba49ea · outbound

This paper cites Kimi-Audio Technical Report.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Kimi-Audio Technical Report

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-07-02T12:36:57.087151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:0bdb1a54f17932f5863034e01ffd8a24b03fe6a70f61b05b1f2f760ec2dbb3e3

Observation 48c11686-9c8e-4be7-87b9-08abc5a2be08 · outbound

This paper cites Audio flamingo 3: Advancing audio intelligence with fully open large audio language models,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Audio flamingo 3: Advancing audio intelligence with fully open large audio language models,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:f4a0ff58b783e2ed84ac2dcd69c27754b2ea45551ba33308ded3790fb88eafbc

Observation b8bec4d3-db49-4042-8c15-1bb682c0d115 · outbound

This paper cites Available: https://openreview.net/forum?id= FjByDpDVIO.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Available: https://openreview.net/forum?id= FjByDpDVIO

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:8d202ed35c561ba6431040adf6d3e8099d6e42ce93f69df465952d08d7f61d0e

Observation d44d6b97-d702-4272-b6db-bac2e0c8b265 · outbound

This paper cites Voxtral.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Voxtral

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:36:57.087582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:7e8aaa89e7f7c6286dfcdca676194e4dfb085ceb09bc6a8b6eb7b1de88413107

Observation c59ebe52-eec0-4879-8f0e-7d099ccf476d · outbound

This paper cites Film: Visual reasoning with a general conditioning layer,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Film: Visual reasoning with a general conditioning layer,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:35cab8d3213cdbf7d68f9a77d12e940f329a0dc2d7407bf0a86f11867da0af2b

Observation 8152422a-8a80-4a55-99e7-e87787e8e814 · outbound

This paper cites Ministral 3.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Ministral 3

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-07-02T12:36:57.090395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:9ec8b17837083d219c17237598c0ab226d59a8780557f2d80438c9fed759c58a

Observation 2d89e8b0-8ad1-4448-946e-e98b3a33b24f · outbound

This paper cites V oxblink2: A 100k+ speaker recognition corpus and the open- set speaker-identification benchmark,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition V oxblink2: A 100k+ speaker recognition corpus and the open- set speaker-identification benchmark,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:f0dcd8b960dc05f66c90c2987e77d46926fd03bd6841ba1e7ac81212e97c0ef3

Observation a748a3f4-633d-48d6-8f92-5ffb3cf8c925 · outbound

This paper cites V oxceleb2: Deep speaker recognition,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition V oxceleb2: Deep speaker recognition,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:816a829eb670aa6a062b0a0ac04acd087196dece29ecd0143e7ac6c23ceba3f1

Observation 3866aa14-931c-4b78-8f70-fbdb815981f5 · outbound

This paper cites Com- mon voice: A massively-multilingual speech corpus,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Com- mon voice: A massively-multilingual speech corpus,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:4a550935ec07b781b3c250306d703d53085b88039bfda3188b4b63a88c0b9ca0

Observation 1378f56a-515f-457d-beec-b8acc83d59c0 · outbound

This paper cites Cba-whisper: Curriculum learning-based adalora fine-tuning on whisper for low-resource dysarthric speech recognition,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Cba-whisper: Curriculum learning-based adalora fine-tuning on whisper for low-resource dysarthric speech recognition,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:8e881656e6db18f33cf4904b95fcc7cb99d959e59958c31cfd100a7e7bd24220

Pith citing papers

Observation 3ffaa63b-ca08-487a-b85e-422a68904470 · inbound

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition cites this paper.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T12:36:57.081596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:a1ffa8c5107329ce509d1b38ac7138f04e9852463f86830b17068aeefa830f76