Pith. sign in

Paper Citation Record · LEDGER

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution

As of 8 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 1 inbound Pith citation observation for arXiv:2506.02181.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02181 v1

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:32:17.610645Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:32:13.765343Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T11:32:17.963370Z

Reference resolution

46 of 46 outbound references displayed

  • verified exact1
  • verified fuzzy36
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 32be8668-a1d0-4b36-9bae-41ad920028b8 · outbound

This paper cites an unresolved cited work.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:32:25.650868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:13.696269Z digest=sha256:6152dd4814aac88d2318c4d5b944d371350d71c527d2bb63c6587ce4d25a6961

Observation 0b075e5a-482f-47e0-ab8a-9b01b6920469 · outbound

This paper cites Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T11:32:18.040121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:13.765343Z digest=sha256:24bec928b579753b4528b1226b8f0d96e638a7c9529b036e2dd47bc928811605

Observation b0cc402e-91c3-4731-bb11-86d9ab1b1021 · outbound

This paper cites Output text is encoded into BPE.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Output text is encoded into BPE

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:25.519045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:13.851341Z digest=sha256:d27c8d091f9216f1b7eea6b5b623fd9bdc9a27c0f1b5a9a10ed0821c458f9307

Observation 55b8654c-edf4-4196-a818-749efbd5c139 · outbound

This paper cites Time Fig.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Time Fig

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:25.090547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:14.077263Z digest=sha256:b639831064508a981a067b603eae8934449de00ef521733669cc4666dde7d1a6

Observation df277d58-ba0c-4348-87c0-c04a91f3b2e3 · outbound

This paper cites The encoder layers use a convolution kernel size of31, an em- bedding size of512, and a linear layer hidden size of 2 048.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution The encoder layers use a convolution kernel size of31, an em- bedding size of512, and a linear layer hidden size of 2 048

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:25.248711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:13.957612Z digest=sha256:1f52ceb16b6d1d18ae79412258313a3e0719d04fd1ef97121caa0129619ef71c

Observation 237cf734-eaba-43a1-a1a0-261ae9493e39 · outbound

This paper cites an unresolved cited work.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:32:24.899323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:14.179728Z digest=sha256:16713a597756475e891a80c5e2f939ae03db3488db123878dc6aca3406124071

Observation 192b2f02-aa97-4c3d-82e6-7deffcdda717 · outbound

This paper cites This is the first in-depth analysis of saliency maps in relation to fine-grained acoustic patterns across three phoneme classes.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution This is the first in-depth analysis of saliency maps in relation to fine-grained acoustic patterns across three phoneme classes

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:25.001122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:14.125900Z digest=sha256:0a25c41b8ffafb42e1fb3016f411bd02356694a4333856fc92d1b006a63f2ccd

Observation bfa9bd22-277f-41f7-99d1-40563c5d85ee · outbound

This paper cites Understanding the representa- tion and computation of multilayer perceptrons: A case study in speech recognition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Understanding the representa- tion and computation of multilayer perceptrons: A case study in speech recognition,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.407824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:14.927096Z digest=sha256:0ff97f8e6fc10a415c8b25f1e714e5399ddd46262c78463c7bc544b006d1b604

Observation 8427c8e7-6bed-4d11-9105-189ad95a3a56 · outbound

This paper cites Analyzing phonetic and graphemic representations in end-to-end automatic speech recog- nition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Analyzing phonetic and graphemic representations in end-to-end automatic speech recog- nition,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:24.714209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:14.217408Z digest=sha256:cc81c9abbec1097c1a6a2df5a9e00df1ad9fee88d8b6a377deb98e102679254e

Observation 29201fdd-5aa4-4714-9f74-777854121a46 · outbound

This paper cites Probing phoneme, language and speaker information in unsupervised speech representations,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Probing phoneme, language and speaker information in unsupervised speech representations,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:24.485514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:14.303707Z digest=sha256:92725cd7105657dc53ad88fbcb76e9718859254ee4b7296eb39d568334bcb851

Observation dbe9d202-7fd3-4c78-aeab-be3dba4606a2 · outbound

This paper cites Domain-Informed Probing of wav2vec 2.0 Embeddings for Pho- netic Features,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Domain-Informed Probing of wav2vec 2.0 Embeddings for Pho- netic Features,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:24.312303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:14.412565Z digest=sha256:965d4665847682685cc0cfd25dfa926733e48a2d4b8fd6a16150d8fa9f180d11

Observation d1e2d11f-e127-44a7-88c1-9123d62fd077 · outbound

This paper cites Probing self- supervised speech models for phonetic and phonemic informa- tion: A case study in aspiration,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Probing self- supervised speech models for phonetic and phonemic informa- tion: A case study in aspiration,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:24.127112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:14.556191Z digest=sha256:6c9dcc00fdffcb8a2ea058f16535897674a82375aa52dc6855c9d2e541d2debb

Observation c469bcd5-9759-4445-90a0-6495cc0b9107 · outbound

This paper cites Self-supervised speech representations are more phonetic than semantic,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Self-supervised speech representations are more phonetic than semantic,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.912683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:14.636029Z digest=sha256:1728b77c8241c7d3c5df017cfc23408c18f86c7cb6f0baee3ad3637f11d23316

Observation a9a232c3-1c1d-478b-b710-8fb9d5be4099 · outbound

This paper cites Encoding of phonol- ogy in a recurrent neural model of grounded speech,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Encoding of phonol- ogy in a recurrent neural model of grounded speech,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.745518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:14.725407Z digest=sha256:9e01ad9796d418b468fc400c3fbb4a43e79e4b153506df0fc5d5a65e79913343

Observation cbaf82b0-f548-412b-b390-a1ef98ef0803 · outbound

This paper cites Learn- ing weakly supervised multimodal phoneme embeddings,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Learn- ing weakly supervised multimodal phoneme embeddings,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.605363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:14.829948Z digest=sha256:aca587f5f830a754cca86eee0e858a3eb2fd65cf5cef30f88211dad45b545bb6

Observation 4da04087-fc0b-499c-a2b8-d111cefd052a · outbound

This paper cites Acous- tic characteristics of American English vowels,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Acous- tic characteristics of American English vowels,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:22.306585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:15.511358Z digest=sha256:49c24973823191996d378fe556d83bd810f46436db42b7c2660af7d915b2e9be

Observation 235f25a8-d890-43e1-8166-70c1b049d5af · outbound

This paper cites Neuron Activation Profiles for Interpreting Convolutional Speech Recognition Models,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Neuron Activation Profiles for Interpreting Convolutional Speech Recognition Models,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.215208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:15.004400Z digest=sha256:d3f8def767f8138c6f68459beca4aff54de7431d6e072dcefc904059c584c75a

Observation af0cc813-9177-45ea-841e-5d938a432c13 · outbound

This paper cites Interpretable Convolutional Filters with SincNet,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Interpretable Convolutional Filters with SincNet,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.047254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:15.077490Z digest=sha256:1333d01c815b565e6ea4a34851d874df8202bfe301959a3174edd6212a7774a8

Observation 1c7f36dc-b57b-4146-8590-64f887c27d60 · outbound

This paper cites Introspection for convolutional automatic speech recognition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Introspection for convolutional automatic speech recognition,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:22.858990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:15.146016Z digest=sha256:e7608b0df5250848232bd32b60e85679c14cb54e029a3f7f34dea8d61a423110

Observation cbad7af6-f918-4f29-b034-15bd8714eaef · outbound

This paper cites End-to-end acoustic modeling using convolutional neural networks for HMM- based automatic speech recognition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution End-to-end acoustic modeling using convolutional neural networks for HMM- based automatic speech recognition,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:22.652588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:15.227371Z digest=sha256:e61006b9635e7bcff1ee0409a356bbf9aea7148b4d5e2882f315641ef89e2ecb

Observation 2c80d404-16f8-4f7f-876e-d7acd6fc5683 · outbound

This paper cites Gradient-Adjusted Neuron Activation Profiles for Comprehensive Introspection of Convolutional Speech Recognition Models.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Gradient-Adjusted Neuron Activation Profiles for Comprehensive Introspection of Convolutional Speech Recognition Models

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:32:17.857757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:15.281363Z digest=sha256:28883aea265a09366aa4dc5c9059f7367f11bde0b6b9c3fd2875bf9fe61244ca

Observation 34779c97-1dc5-4c41-89fe-784ecac7585e · outbound

This paper cites Directly Comparing the Listening Strategies of Humans and Machines,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Directly Comparing the Listening Strategies of Humans and Machines,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:22.489366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:15.360537Z digest=sha256:66888d57b367d23742a30888ff62d4359f70512370e7d92e6e9fe01f142909ae

Observation c0ff9443-cf9f-4f20-b0a5-8365f0174d6d · outbound

This paper cites SPES: Spectrogram Perturbation for Explainable Speech-to-Text Generation.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution SPES: Spectrogram Perturbation for Explainable Speech-to-Text Generation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:32:15.428570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:32:15.428570Z digest=sha256:311f7b73302e3b60beb7f7c03c9f66e9afb07ca77df340f615cd47be9cbac8c5

Observation b03f39f8-479e-4e1c-85ab-9dd34483a66b · outbound

This paper cites an unresolved cited work.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:32:25.364852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:13.923390Z digest=sha256:1b4b4fbca9edc31bb4be9c0a554df0d976a715c90c753d851b31dcaccaaf6f81

Observation 9a5d5fc3-9bff-4c7e-9af2-d9bccee1abff · outbound

This paper cites Burst and Transition Cues to V oicing Perception for Spo- ken Initial Stops by Impaired- and Normal-Hearing Listeners,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Burst and Transition Cues to V oicing Perception for Spo- ken Initial Stops by Impaired- and Normal-Hearing Listeners,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:22.124476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:15.594193Z digest=sha256:635f5ab2607459fef5a0d039b772453bfcbf695f13bf35b1b643353d025dcc3c

Observation 4c5491e7-8eea-45f3-95e2-2a5e777da75e · outbound

This paper cites Acoustic cues of voiced and voiceless plosives for determining place of articulation,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Acoustic cues of voiced and voiceless plosives for determining place of articulation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:21.905743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:15.692957Z digest=sha256:067b8529096b8f6d1b2aceace55a97ccd81caf3e78a0397db1ab9839900ff1e4

Observation 1fe81541-41d6-48f7-95b1-c105c5a05685 · outbound

This paper cites Acoustic characteristics of English fricatives,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Acoustic characteristics of English fricatives,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:21.687005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:15.779334Z digest=sha256:864e87222b001458f65aa1cb053049ddec402929f85489913a6eff5368d043e9

Observation 24f587b8-0313-400c-866b-b6335394924b · outbound

This paper cites Acoustic characteristics of clearly spoken English fricatives,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Acoustic characteristics of clearly spoken English fricatives,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:21.511061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:15.889219Z digest=sha256:1def439b2edbb9618913eb244cbab0f055c328cd2f2f8ca1cf2987b9e8ed3505

Observation a2da6397-2f44-4d45-ad16-4085f91b2d35 · outbound

This paper cites Conformer: Convolution-augmented Transformer for Speech Recognition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Conformer: Convolution-augmented Transformer for Speech Recognition,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:21.258248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:15.968308Z digest=sha256:c8a171c91e4e71707397f9262da15057ff02345621394c6b3ebb387e27fa8c44

Observation 8a09b4e1-738f-41ad-b7aa-3375bf1f9000 · outbound

This paper cites Introducing Parsel- mouth: A Python interface to Praat,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Introducing Parsel- mouth: A Python interface to Praat,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:32:16.075519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:32:16.075519Z digest=sha256:7bdb247a1a9a966e34d13524991d865f4d7c8ed82dac7d34d904f36cc7deba1c

Observation af314bcb-d75a-494b-bde2-ee3b78eecc9f · outbound

This paper cites Pykaldi: A Python Wrapper for Kaldi,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Pykaldi: A Python Wrapper for Kaldi,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:21.020089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:16.143157Z digest=sha256:202e4278f1608ceba2f2a03205c2fe08c84b67bb0ac1955b71469d5f4342ab88

Observation 3effd402-33aa-4488-80dd-0fb1fe60c79c · outbound

This paper cites Neural Machine Transla- tion of Rare Words with Subword Units,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Neural Machine Transla- tion of Rare Words with Subword Units,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:20.786942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:16.218548Z digest=sha256:69bae6faf1216e9c12358eda48ef5943e789b3f50b5ca33a03ed0080ddcf7ff2

Observation 592131ef-154a-4488-bb6f-b1564daa0dd9 · outbound

This paper cites SentencePiece: A simple and lan- guage independent subword tokenizer and detokenizer for Neural Text Processing,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution SentencePiece: A simple and lan- guage independent subword tokenizer and detokenizer for Neural Text Processing,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:20.523379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:16.297689Z digest=sha256:e676816d351e50f93855e5d2b6fe4fff30c74dd3721aa0216cd3e9af420fa7c8

Observation 3f931957-e4ad-4cb7-95f1-c7a131c9a156 · outbound

This paper cites Attention is All you Need,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Attention is All you Need,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:20.287706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:16.359174Z digest=sha256:820a054206e98349e23886a34debcc996717945d1b21fa5e566ce5f520b08c7f

Observation fe1b260f-6a28-4646-9634-947f5a0e244d · outbound

This paper cites Fairseq S2T: Fast Speech-to-Text Modeling with Fairseq,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Fairseq S2T: Fast Speech-to-Text Modeling with Fairseq,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:32:16.455994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:32:16.455994Z digest=sha256:ed9034bf104a264c24713b6e49109d310608609dce566a112db8cda7aae70d51

Observation 6d166b02-a9eb-4aad-b7b8-5a8562bf286a · outbound

This paper cites Com- mon V oice: A Massively-Multilingual Speech Corpus,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Com- mon V oice: A Massively-Multilingual Speech Corpus,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:20.058654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:16.536895Z digest=sha256:d0fb74055cd9db531f830605ee18fdc8fc15d3e34afe1809dc29a90bae13278c

Observation 5b9d54ab-ea59-49dd-a540-7eca7aa6a668 · outbound

This paper cites Lib- rispeech: An ASR corpus based on public domain audio books,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Lib- rispeech: An ASR corpus based on public domain audio books,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:32:16.626379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:32:16.626379Z digest=sha256:f46d8b7714e0605220560d24e2ad1217fdc55eb46cc7b3bb41ce1acefd4f06e9

Observation db3fa82a-0329-4f53-950d-1afe1a917983 · outbound

This paper cites TED-LIUM 3: Twice as Much Data and Corpus Repartition for Experiments on Speaker Adaptation,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution TED-LIUM 3: Twice as Much Data and Corpus Repartition for Experiments on Speaker Adaptation,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:19.771580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:16.724921Z digest=sha256:c040618cce2271c635e5445ed344e0d1658560dbbc73c5aa461e93964deb1dc2

Observation 824f9cd2-f374-4c14-b28d-3409165a5104 · outbound

This paper cites V oxPopuli: A Large- Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution V oxPopuli: A Large- Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:19.558003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:16.814577Z digest=sha256:852cfc5512953a603d434d2867f1bf9c29ab40c707bb967ce43b48d56fe0c49e

Observation 26f4e42a-4f62-481e-9f70-c3a43f6a9d06 · outbound

This paper cites Rethinking the Inception Architecture for Computer Vision,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Rethinking the Inception Architecture for Computer Vision,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:19.372863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:16.949565Z digest=sha256:4b67c52581cf6cdab63bbb1f830d49ef14c8b968b8e0e5a52835c0cbc6a341f7

Observation 800f6bae-2e69-4255-b17b-ccd2ea389988 · outbound

This paper cites Con- nectionist temporal classification: labelling unsegmented se- quence data with recurrent neural networks,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Con- nectionist temporal classification: labelling unsegmented se- quence data with recurrent neural networks,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:19.203475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:17.048361Z digest=sha256:6964ae56a6fb6d3c1abf97b4eacfe09c8d08c7c82342a07eb99d123207dba725

Observation b5794d0f-42b7-49d1-aad3-55220a5f8f8a · outbound

This paper cites Adam: A Method for Stochastic Opti- mization,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Adam: A Method for Stochastic Opti- mization,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:32:17.167561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:32:17.167561Z digest=sha256:fd2fab9272d122ba18c9f505e3acb92f93010de7f27395a176252d117b4a8dd4

Observation 62b93cdb-19e3-43f4-b3ce-ada0bfe6a740 · outbound

This paper cites SpecAugment: A Simple Data Aug- mentation Method for Automatic Speech Recognition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution SpecAugment: A Simple Data Aug- mentation Method for Automatic Speech Recognition,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:18.943826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:17.259642Z digest=sha256:4f41522545c4e328746aa384ad3655d1499ffb9367cb154bc207104932654b6d

Observation 57088e64-c670-4bf6-a24c-a8f1cf64f16e · outbound

This paper cites Darpa timit acoustic-phonetic continuous speech corpus cd-rom TIMIT,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Darpa timit acoustic-phonetic continuous speech corpus cd-rom TIMIT,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:18.661884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:17.396416Z digest=sha256:f9e02b538795e90a9efe9f6856cf8a3794440f4ab842761c3f2960b32e791646

Observation 53152cd0-8d1c-437f-9961-44133f0028fe · outbound

This paper cites Twists, humps, and pebbles: Multilingual speech recognition models exhibit gen- der performance gaps,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Twists, humps, and pebbles: Multilingual speech recognition models exhibit gen- der performance gaps,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:18.424129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:17.502417Z digest=sha256:ec14562444fc1a9c3b6154d6d4d3ada46085e3814c49a956dfc33c25aad3b105

Observation a73cfbae-9702-461d-be97-e71fab6edf86 · outbound

This paper cites Burst spectrum as a cue for the stop voicing contrast in American English,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Burst spectrum as a cue for the stop voicing contrast in American English,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:18.247636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:17.610645Z digest=sha256:3f7af36141998c942a64402e5160ece3de1edf6c90ad10ddcaa72aea46d5043f

Pith citing papers

Observation 0b075e5a-482f-47e0-ab8a-9b01b6920469 · inbound

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution cites this paper.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T11:32:18.040121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:32:13.765343Z digest=sha256:24bec928b579753b4528b1226b8f0d96e638a7c9529b036e2dd47bc928811605