Pith. sign in

Paper Citation Record · LEDGER

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition

As of 22 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 1 inbound Pith citation observation for arXiv:2606.06211.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.06211 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T01:58:13.871771Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T01:58:13.871771Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-02T12:36:57.080220Z

Reference resolution

28 of 28 outbound references displayed

  • verified exact6
  • verified fuzzy0
  • unresolved20
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cd7c94bd-c407-44a8-91a5-04a49d8a79a2 · outbound

This paper cites Yet these sys- tems continue to struggle with speech produced by individuals with neuromotor disorders such as amyotrophic lateral sclerosis (ALS) or Parkinson’s disease (PD).

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Yet these sys- tems continue to struggle with speech produced by individuals with neuromotor disorders such as amyotrophic lateral sclerosis (ALS) or Parkinson’s disease (PD)

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:6c9e976c200d22d2b5b6d08eb1746ddf36b7a635291aae66f15ee96e1451e03a

Observation 3ffaa63b-ca08-487a-b85e-422a68904470 · outbound

This paper cites FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T12:36:57.081596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:350035facaf9a1f60830f73c59b632be97db5dcacb0166f0c63be66f443a9f0b

Observation 44abbcee-3d31-4ed3-b2ff-9c3061b79fd8 · outbound

This paper cites What is the sex of the speaker?.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition What is the sex of the speaker?

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:641983feab2a5eb97e11c166488407deec65b7ffe242ccacee0726b3f221da1f

Observation da710cfc-d670-4881-b010-2af3043d477c · outbound

This paper cites Table 1:WER (%) of Voxtral-Mini and adapted variants on the NeuroVoz and TORGO test sets.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Table 1:WER (%) of Voxtral-Mini and adapted variants on the NeuroVoz and TORGO test sets

Reference 4

Resolution
malformed identifier
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:63f5ab6e8f62fc6fa70205304a023d8819efdf1c7696964171ca901523219d95

Observation 097dd103-e412-4bd1-bfc8-709966e72d2f · outbound

This paper cites However, this requires prior knowl- edge of whether the input speech is pathological, which may not always be available in practice.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition However, this requires prior knowl- edge of whether the input speech is pathological, which may not always be available in practice

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:ef99c5e4bd98507e9d9c92057424550a81980126365160042b12e2807b75f688

Observation 036a036d-ec36-4b9a-90b5-34456c3f2b4a · outbound

This paper cites It is a lightweight approach to pathological speech adaptation that maintains the base model untouched.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition It is a lightweight approach to pathological speech adaptation that maintains the base model untouched

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:f6fabeb4c22eb5db1b0f08aa734e37d9ba12f0a4dc393aa02357fdd9abc735ae

Observation 9457d3f2-a074-41a2-8a06-5d7924d52f64 · outbound

This paper cites 101135916).

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition 101135916)

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:7e4d643be7c55c1e4cbb8428c11284800923b6d56b7f0ba7c126e813f427eac7

Observation 96a9a6cf-00e6-43d5-a1d1-76cc01b47d5c · outbound

This paper cites Automatic speech recognition: A survey of deep learning techniques and ap- proaches,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Automatic speech recognition: A survey of deep learning techniques and ap- proaches,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:46fe2a91813f25fcf90211765e31dddb845eccff547f23158f8178f5bd10bf58

Observation b1f45bb2-9ac9-4f78-91f1-feb9f507dbc2 · outbound

This paper cites The torgo database of acoustic and articulatory speech from speakers with dysarthria,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition The torgo database of acoustic and articulatory speech from speakers with dysarthria,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:e6fc667043402f95c78076f5a09084100bc132af59e1ce6fdc5048adbca3f9f4

Observation b5aef97d-0fc8-4347-a41f-16d5a7645a26 · outbound

This paper cites Dysarthric speech database for universal access research.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Dysarthric speech database for universal access research

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:f16f954412897c98045a703c14f345aa3a6c9a978eb70c08e4389a8dae224ccf

Observation b5698af9-ff93-49eb-adaa-d764956d2f32 · outbound

This paper cites The interspeech 2025 speech accessibility project challenge,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition The interspeech 2025 speech accessibility project challenge,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:1cbfbf35b104ef82a82a5cd83a2d29a74dcc121034c07e301922119f234ee877

Observation cefad5ca-443f-4ed3-813d-9fcdacc0384b · outbound

This paper cites New spanish speech cor- pus database for the analysis of people suffering from parkinson’s disease.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition New spanish speech cor- pus database for the analysis of people suffering from parkinson’s disease

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:f1da6e7836bbcca80b4525b8d74fe3f2d312da9721d49c5079708ce6ceddb9c9

Observation b27da1e1-e2e8-4d44-acd3-5c95a16b67d6 · outbound

This paper cites Neurovoz: a castillian spanish corpus of parkinsonian speech,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Neurovoz: a castillian spanish corpus of parkinsonian speech,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:0f4e018ab803e4b8d8db70dc7833f53e860680ad44e05f6876b4ad75782e5d1a

Observation 6f6f906f-defd-4428-a84a-4f797e5ecf97 · outbound

This paper cites Neurovoz: a castillian spanish corpus of parkinsonian speech,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Neurovoz: a castillian spanish corpus of parkinsonian speech,

Reference 14

Resolution
verified exact
doi, observed 2026-06-28T02:01:28.924627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:661624e37f26f13f4ee63a8d6da71100c8366b2045350b9baf3bccc12ce51d7e

Observation 65e55592-1eb3-487a-88ac-2961d7b97795 · outbound

This paper cites Speaker adaptation for Wav2vec2 based dysarthric ASR.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Speaker adaptation for Wav2vec2 based dysarthric ASR

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:36:57.095835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:2f61a5422793a4068b961afa0d0f69f88b514f5ecd0560c5af0565c98c34d6bd

Observation 85a46a3d-6636-41be-82ca-2cf38ff88f6d · outbound

This paper cites Use of speech impairment severity for dysarthric speech recognition,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Use of speech impairment severity for dysarthric speech recognition,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:3c9e4e7a81c90ce5bd9f6aa77f9b7d1698aea6e9adc49fd249efc33ff3f8da32

Observation db331eaf-b467-496b-bcd1-16b032982589 · outbound

This paper cites Personalized fine-tuning with controllable synthetic speech from llm-generated transcripts for dysarthric speech recognition,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Personalized fine-tuning with controllable synthetic speech from llm-generated transcripts for dysarthric speech recognition,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:2d77ba5f535ee9e683abef7a5b7626da483f2d34759db24b6b27f871385372df

Observation 86126588-6e8e-480b-b74a-dffa74f35dda · outbound

This paper cites Qwen3-ASR Technical Report.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Qwen3-ASR Technical Report

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-02T12:36:57.092874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:1eec8b86ea6ba3b0203e119b9d9bf85ef688309bdbbff943d00f6e5754ad0ac7

Observation 810f0b79-c138-419c-9bec-1d2fc7ba49ea · outbound

This paper cites Kimi-Audio Technical Report.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Kimi-Audio Technical Report

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-07-02T12:36:57.087151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:424ed9c3faae241ff1fc38779d2024e0167b53ac5a57de0cf9da9db0fedc0d27

Observation 48c11686-9c8e-4be7-87b9-08abc5a2be08 · outbound

This paper cites Audio flamingo 3: Advancing audio intelligence with fully open large audio language models,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Audio flamingo 3: Advancing audio intelligence with fully open large audio language models,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:7dd938922b623d5f7646c81d7a5aba27c978be315d981c6a890e49c5644a5013

Observation b8bec4d3-db49-4042-8c15-1bb682c0d115 · outbound

This paper cites Available: https://openreview.net/forum?id= FjByDpDVIO.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Available: https://openreview.net/forum?id= FjByDpDVIO

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:c06094c73d4eb3ba0ae6c3044a91aa2e42f9bdc1259b9b61f894f2b0044a8db8

Observation d44d6b97-d702-4272-b6db-bac2e0c8b265 · outbound

This paper cites Voxtral.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Voxtral

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:36:57.087582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:a85273b8b9cc0cb3d169caa0b03c7830ae6b1c4ab780b2868cfe12cf628d35d0

Observation c59ebe52-eec0-4879-8f0e-7d099ccf476d · outbound

This paper cites Film: Visual reasoning with a general conditioning layer,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Film: Visual reasoning with a general conditioning layer,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:0da8e6878a2a39003f84f80ab1fe8556bcfc8e8d86e68cb0ef462f27c016ed39

Observation 8152422a-8a80-4a55-99e7-e87787e8e814 · outbound

This paper cites Ministral 3.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Ministral 3

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-07-02T12:36:57.090395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:ad13979e399a7c865616a1d0ceb2d69ca5ce5f239ae4a719a682a7440d3114c9

Observation 2d89e8b0-8ad1-4448-946e-e98b3a33b24f · outbound

This paper cites V oxblink2: A 100k+ speaker recognition corpus and the open- set speaker-identification benchmark,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition V oxblink2: A 100k+ speaker recognition corpus and the open- set speaker-identification benchmark,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:502652ea1999b4c3ad7143885775125aa26f30f2ada5345b5c7197da31318ddf

Observation a748a3f4-633d-48d6-8f92-5ffb3cf8c925 · outbound

This paper cites V oxceleb2: Deep speaker recognition,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition V oxceleb2: Deep speaker recognition,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:f3d85adf23923109120736cb2d50e020400d05e76527b8eaa974bb59dc132c83

Observation 3866aa14-931c-4b78-8f70-fbdb815981f5 · outbound

This paper cites Com- mon voice: A massively-multilingual speech corpus,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Com- mon voice: A massively-multilingual speech corpus,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:d26dabc9f90c89adc8f620d30165fc6f556a33cd6dba614376406d35cf091dac

Observation 1378f56a-515f-457d-beec-b8acc83d59c0 · outbound

This paper cites Cba-whisper: Curriculum learning-based adalora fine-tuning on whisper for low-resource dysarthric speech recognition,.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition Cba-whisper: Curriculum learning-based adalora fine-tuning on whisper for low-resource dysarthric speech recognition,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-28T01:58:13.871771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:fa356cabf20972878e8200019421e7c17479d7489649157c21d1508b975c243c

Pith citing papers

Observation 3ffaa63b-ca08-487a-b85e-422a68904470 · inbound

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition cites this paper.

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T12:36:57.081596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T01:58:13.871771Z digest=sha256:350035facaf9a1f60830f73c59b632be97db5dcacb0166f0c63be66f443a9f0b