Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:37:58.527198Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 1 inbound Pith citation observation for arXiv:2608.11587.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:37:58.527198Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:37:58.415460Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-16T00:37:58.604758Z
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2f19e0b3-41b9-4b0d-966f-c8e930a8b1d8 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Depending on the label taxonomy, the task can be formulated as speaker diarization, vocalization classifi- cation, or a combination of both
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 85931b77-fc6b-41e1-b17f-d272322f871b · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5613fc63-21f2-4f34-86c2-ab752e73caca · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f893d402-5bae-476b-a70f-272e3268979d · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e6dcbee3-c214-474a-a2db-b41e1dee46a5 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 340c13f6-98fb-4592-882b-7fa2a8ae33e1 · outbound
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9ef79544-396e-4c91-b6a9-22f6905adf36 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning In the current design, offsets are learned only for training families, and inference on new families relies solely on the shared tier tokens
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation be981798-466b-42d9-9a9d-cf5eda55c8a2 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning By combining a LoRA-finetuned Whisper encoder with structured speaker conditioning and tier-specific heads, the model supports overlapping speakers and framewise prediction
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 066d5771-9987-4c7c-b5d0-0cfb4e7f08e6 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning All technical content, experimental design, analysis, and scientific contribu- tions are entirely the work of the authors
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d39187cc-5b69-49fa-986e-d7bed726bb28 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning For the experiments presented here, we used the Delta System at the National Center for Supercomputing Applications through AC- CESS allocations CIS240417 and CIS250040
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d43b58f2-4d25-4582-9e8b-d163416162ab · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning An open-source voice type classifier for child-centered daylong recordings,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1f08f8ed-e446-422d-9de2-b712e32e509c · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Analysis of acoustic and voice quality features for the classification of infant and mother vocalizations,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 83c5e8d7-f678-4bad-9126-b99b602b2ce8 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81970ea0-3654-40c7-ab9e-089bdb55826d · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f723355-a9d6-45ef-a69c-6f5b0f6b1486 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Ssast: Self- supervised audio spectrogram transformer,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ca0306a4-6514-4e7f-bd3d-fe7c070f8be9 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Towards ro- bust family-infant audio analysis based on unsupervised pretrain- ing of wav2vec 2.0 on large-scale unlabeled family audio,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d1fc9b1f-8512-4744-945f-4c449bdaf826 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Band- split self-supervised mamba for infant-centered audio analysis,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fcab2698-f3fd-4943-b7ad-08bba330d823 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Robust self supervised speech embeddings for child-adult classification in interactions involving children with autism,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 10861f42-de29-4d87-9f7b-df53877a8d01 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Robust speech recognition via large-scale weak supervision,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44225a6c-1bc8-40df-aeea-ee68f5961914 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Wavlm: Large-scale self- supervised pre-training for full stack speech processing,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f4cbe6a-f1ec-437d-9ef6-05e396f4b027 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Lora: Low-rank adaptation of large language models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3699ebe0-0dce-4014-9e6c-6b81f38bd462 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning LoRA-Whisper: Parameter-Efficient and Extensible Multilingual ASR
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4e021d2-ec4e-49c0-9e39-b22a74218af4 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Towards Rehearsal-Free Multilingual ASR: A LoRA-based Case Study on Whisper
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbbf2178-691b-4d7d-8502-b2cfebee8cde · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Sparsely shared lora on whisper for child speech recognition,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f7800dad-1039-4625-9f92-e59fc25be48c · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Whisper-at: Noise-robust automatic speech recognizers are also strong audio event taggers,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 03a70b6f-972e-482f-b23b-f5fd29698aa4 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Exploring Speech Foundation Models for Speaker Diarization in Child-Adult Dyadic Interactions
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 262b03b5-6715-41b1-b411-9ad2b52a825e · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Data efficient child-adult speaker diarization with simulated conversations,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03d92929-91e8-4f51-8300-d15577639681 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f633839-6a74-45d0-9798-85f6edf2048b · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Preliminary technical validation of LittleBeats™: A multimodal sensing platform to capture cardiac physiology, mo- tion, and vocalizations,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 173baa84-40d6-484a-aede-d591008a6fb9 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Praat: doing phonetics by computer [computer pro- gram],
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8038aa1b-a8bd-48ef-8431-969e39e68630 · outbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Listen, adapt, better wer: Source-free single-utterance test-time adaptation for automatic speech recognition,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 85931b77-fc6b-41e1-b17f-d272322f871b · inbound
Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.