Pith. sign in

Paper Citation Record · LEDGER

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition

As of 16 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 1 inbound Pith citation observation for arXiv:2506.16969.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.16969 v2

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T19:19:59.764201Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T19:19:59.656455Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-15T19:19:59.861302Z

Reference resolution

31 of 31 outbound references displayed

  • verified exact2
  • verified fuzzy17
  • unresolved10
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 29b7e862-1586-40f1-b092-eaa59ece4617 · outbound

This paper cites State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T19:19:59.866003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.656455Z digest=sha256:847359b24ab81748ead798dbdd6f781d5b55d7748d452ec8a4c48478834cda81

Observation e43acfad-951c-4d1e-9cbd-aa0b1b243dfc · outbound

This paper cites The primary distinction lies in how they are produced.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition The primary distinction lies in how they are produced

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:20:00.153712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.661179Z digest=sha256:76761d351e1ccfef231169ada4c9e2ce93e7d2f7cc600c198445ac16e0fe2705

Observation 72731bd3-fea3-495e-881a-1693122076c0 · outbound

This paper cites Specifically, we introduce the wTIMIT and CHAINS.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Specifically, we introduce the wTIMIT and CHAINS

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:20:00.140770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.665829Z digest=sha256:241ace128f4164ca80482666c4c24eea56a40151f21f7177b8ca463fd55036c8

Observation 67f7ec60-44c6-4a08-b8cd-6cc40b262965 · outbound

This paper cites As a base- line, we evaluated the performance of the pre-trained Whisper Large-v2 model on the test set to assess the need for a special- ized system for the proposed challenges.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition As a base- line, we evaluated the performance of the pre-trained Whisper Large-v2 model on the test set to assess the need for a special- ized system for the proposed challenges

Reference 4

Resolution
malformed identifier
raw_fallback, observed 2026-08-15T19:20:00.118240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.673878Z digest=sha256:d509122b36f1e7bede9046ac3a3d53d4172a7941d00264a0d0c17467b752b7ca

Observation 85f063d8-b6b9-4e73-99ed-7ab21a4699e6 · outbound

This paper cites These challenges can severely degrade the performance of traditional systems, highlighting the necessity of developing ASR models specifically tailored to handle whispered speech.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition These challenges can severely degrade the performance of traditional systems, highlighting the necessity of developing ASR models specifically tailored to handle whispered speech

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:20:00.105826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.677498Z digest=sha256:688f4798f0bbdede05c050affa312ca2a2366e96f32a14f989ff73fe3e09407f

Observation affd1cb9-744f-4efd-a73e-75489f3cc4a2 · outbound

This paper cites Gener- ative models for improved naturalness, intelligibility, and voicing of whispered speech,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Gener- ative models for improved naturalness, intelligibility, and voicing of whispered speech,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:20:00.004955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.699284Z digest=sha256:2f64c580c28f7a874f8a058a931f04db02e5c164564003273b50b4283486577f

Observation 738ba118-8258-482e-ad88-17290d9a97ed · outbound

This paper cites Gammatonegram representation for end-to-end dysarthric speech processing tasks: Speech recogni- tion, speaker identification, and intelligibility assessment,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Gammatonegram representation for end-to-end dysarthric speech processing tasks: Speech recogni- tion, speaker identification, and intelligibility assessment,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:20:00.092558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.680878Z digest=sha256:31386bb519f61bda694c72f169aeac5a7ec467063f35e1a0a39c927b4451c9a3

Observation 6ad10499-a51d-4611-9320-2e6c29dc32ca · outbound

This paper cites Dysarthric speaker identification with different degrees of dysarthria severity using deep belief networks,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Dysarthric speaker identification with different degrees of dysarthria severity using deep belief networks,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:20:00.078638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.684390Z digest=sha256:a0540531ec853c95accb47e12ed51b8c3c50e3b0a8903db75382b7a728aba3b9

Observation 01904d7e-5fcf-4b9c-831a-a633d837ca66 · outbound

This paper cites Analysis of deep generative model impact on feature extraction and dimension reduction for short ut- terance text-independent speaker verification,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Analysis of deep generative model impact on feature extraction and dimension reduction for short ut- terance text-independent speaker verification,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:20:00.059860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.687768Z digest=sha256:82009bb61d1ff38a9e0367804e24eca7b7d686b895f68697bad2de5ca9df97ea

Observation 50adb962-8278-4a59-a9de-fef72ff4b763 · outbound

This paper cites Leveraging Self-Supervised Models for Automatic Whispered Speech Recognition.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Leveraging Self-Supervised Models for Automatic Whispered Speech Recognition

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-15T19:19:59.851655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.692042Z digest=sha256:90285960eb8f107ca4db1d6af61a5f65df81b2d00c549812e55c7ed9e4dc31b5

Observation 4943a0d9-d4ad-432e-af54-cb1bbb2372df · outbound

This paper cites Whisper to normal speech conversion using sequence-to-sequence mapping model with auditory attention,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Whisper to normal speech conversion using sequence-to-sequence mapping model with auditory attention,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:20:00.038672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.696042Z digest=sha256:12a6b95ad4d952203abcbc4ae19b8c50b3fa94774e1c732c1aa9610b25e36fc1

Observation 55a6288d-e757-4447-9c5d-d46c9eec9c55 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T19:19:59.720018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:19:59.720018Z digest=sha256:00b409e3daec65840ca039b074fc826b66f3e83b1c8d04e3c951fa5ac3ec457b

Observation 5f77349d-2b6e-4c09-84c3-cd8e6901e049 · outbound

This paper cites End-to- end whispered speech recognition with frequency-weighted ap- proaches and pseudo whisper pre-training,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition End-to- end whispered speech recognition with frequency-weighted ap- proaches and pseudo whisper pre-training,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:19:59.990581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.702683Z digest=sha256:6ea5e28adcefc133e6bb464dc1cf164479b2729b1f606cd02406dc8a5708fc68

Observation a7d48d02-748e-431a-88cf-186b49dd3d7b · outbound

This paper cites Improving whispered speech recognition performance using pseudo-whispered based data augmentation,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Improving whispered speech recognition performance using pseudo-whispered based data augmentation,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:19:59.979546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.705892Z digest=sha256:f221efacac216f179efac564dc5720654b3628ec1b85ad0554008c5eb156d9df

Observation c8dddfea-b597-45e4-af39-88a511962d43 · outbound

This paper cites Whispered speech recogni- tion using deep denoising autoencoder and inverse filtering,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Whispered speech recogni- tion using deep denoising autoencoder and inverse filtering,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:19:59.969448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.709042Z digest=sha256:6c3b0de7a8c3faa7ec8cc26b128f1de14cb6e49102ff8c428c22ac6e2b5ca4fb

Observation 9f526990-9172-46e9-a10f-5ca1c4f7725b · outbound

This paper cites an unresolved cited work.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-15T19:20:00.129015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.669766Z digest=sha256:42070086f1e5e304848c91f7a98093b55c043b262f1498652e8123fbec79b590

Observation b7e8fbd9-361e-489d-8881-e8b372672603 · outbound

This paper cites Multi-dialect speech recognition with a single sequence-to-sequence model,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Multi-dialect speech recognition with a single sequence-to-sequence model,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:19:59.958953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.712964Z digest=sha256:9fc0587436ea20f5a979f9885f7419ddc416a93768f09e211ca79cad7edfd581

Observation 9b0bc379-8d2c-483d-abaf-4f8ad4c63302 · outbound

This paper cites Multi-dialect speech recognition in english using attention on ensemble of experts,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Multi-dialect speech recognition in english using attention on ensemble of experts,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:19:59.948819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.716538Z digest=sha256:c1478512ac2fa8e5fe6f000c95fa1bf16735a66cc633d7a628540b11c1fd9fcb

Observation 8167bd77-c85d-4ebe-a24d-47dc0900bcc8 · outbound

This paper cites Wavlm: Large-scale self- supervised pre-training for full stack speech processing,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Wavlm: Large-scale self- supervised pre-training for full stack speech processing,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T19:19:59.723374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:19:59.723374Z digest=sha256:d7a07b5e3804078c0a8d4036b07552832dcbbd18a4f9f9e991f82254980a4986

Observation 716cfe03-0fb3-442c-9049-245d6aff0cd3 · outbound

This paper cites Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T19:19:59.726713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:19:59.726713Z digest=sha256:da30d9a6019ea53b148e8a2f1685cb2f588aee895332e9ff103c5e9142baca6c

Observation 8be2b83c-d5b6-4c6a-acdd-c7196df6fdeb · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Robust speech recognition via large-scale weak supervision,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T19:19:59.730071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:19:59.730071Z digest=sha256:0f737f7c8d97f93eea4aaf2019d5c6962a169315fb04b2cee4a52e182ea8629c

Observation dc6196fd-7e5b-4638-82c3-9ba663f4d6a4 · outbound

This paper cites wtimit whispered timit dataset,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition wtimit whispered timit dataset,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:19:59.915703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.733128Z digest=sha256:b2dadda877361ad7393ae11dbebdb5ba69d2c4bc83340c162b70516090b558fb

Observation f52d3c34-3268-4ca5-a590-b987f8bc954c · outbound

This paper cites Acoustic analysis of consonants in whispered speech,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Acoustic analysis of consonants in whispered speech,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:19:59.905064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.736135Z digest=sha256:1ba982b0d895ed850571a1f21e0fcc88e37766fe98b4b1c1da3305837e264dba

Observation b776020f-0336-45ca-a896-51d2636d081e · outbound

This paper cites La- ryngeal adjustment in whispering: magnetic resonance imaging study,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition La- ryngeal adjustment in whispering: magnetic resonance imaging study,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:19:59.894509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.739285Z digest=sha256:f9f75142dfa82a8f1b6030461688c06691a143a8e8c79eeaa53b9bb72bae5f64

Observation bd8b54c9-73bb-40e4-a82c-872769e8dc83 · outbound

This paper cites Acoustic differences between voiced and whispered speech in gender diverse speakers,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Acoustic differences between voiced and whispered speech in gender diverse speakers,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:19:59.883849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.742466Z digest=sha256:d54f0e3a1ab24ae6a5f0c65db9acd849b440742487da4bbb5dd19d49dee288d4

Observation f736476c-e39d-4dea-83bf-601666351fed · outbound

This paper cites Lib- rispeech: an asr corpus based on public domain audio books,.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Lib- rispeech: an asr corpus based on public domain audio books,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T19:19:59.745826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:19:59.745826Z digest=sha256:f3670eed818624448a59d5cff8d51abedc5f739d2509703d58c6b8d41649f240

Observation 44223e4c-b682-40b9-8e45-7b50687989da · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T19:19:59.749075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:19:59.749075Z digest=sha256:a4981f9db34dbba58c275b374848bd0ad8199d03e41fca1bb4cff78ddaf54065

Observation 882bca90-7434-4f43-a582-00645041aa4c · outbound

This paper cites Speech Slytherin: Examining the Performance and Efficiency of Mamba for Speech Separation, Recognition, and Synthesis.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Speech Slytherin: Examining the Performance and Efficiency of Mamba for Speech Separation, Recognition, and Synthesis

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T19:19:59.752685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:19:59.752685Z digest=sha256:102acd11b2f4d1e8d693bd2873a2b9c9d8d00a452ca7a520d5dd53344719be51

Observation 72e2873c-942a-42b9-bb99-33a848072dd0 · outbound

This paper cites SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T19:19:59.756520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:19:59.756520Z digest=sha256:abffe144128c2e3cdbdc25c14bc169a91c418a3d2dbc1e01282060dcde9f7054

Observation 05da14a7-142e-4e02-90c1-1a083138f265 · outbound

This paper cites Decoupled Weight Decay Regularization.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition Decoupled Weight Decay Regularization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T19:19:59.760069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:19:59.760069Z digest=sha256:26ea342d28dedd01c085583b823d9f46cf24df7fda45795222f8e073a4704d64

Observation 415652aa-8cc6-4254-9525-79464aa22cee · outbound

This paper cites PaddleSpeech: An Easy-to-Use All-in-One Speech Toolkit.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition PaddleSpeech: An Easy-to-Use All-in-One Speech Toolkit

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-15T19:19:59.799570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.764201Z digest=sha256:dfcd30d9e410682b111b29ae5b9ca6581a1a5fbd484bce203abfc04c0296183d

Pith citing papers

Observation 29b7e862-1586-40f1-b092-eaa59ece4617 · inbound

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition cites this paper.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T19:19:59.866003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:19:59.656455Z digest=sha256:847359b24ab81748ead798dbdd6f781d5b55d7748d452ec8a4c48478834cda81