Pith. sign in

Paper Citation Record · LEDGER

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution

As of 8 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 1 inbound Pith citation observation for arXiv:2506.02181.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02181 v1

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:32:17.610645Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:32:13.765343Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T11:32:17.963370Z

Reference resolution

46 of 46 outbound references displayed

  • verified exact1
  • verified fuzzy36
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 32be8668-a1d0-4b36-9bae-41ad920028b8 · outbound

This paper cites an unresolved cited work.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:32:25.650868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:13.696269Z digest=sha256:97a2d05210efc52c41adfa64a499f14ac4f3070ce01684c6bf123a7b10692bf0

Observation 0b075e5a-482f-47e0-ab8a-9b01b6920469 · outbound

This paper cites Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T11:32:18.040121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:13.765343Z digest=sha256:fd278bf44154299419587db4b5c51010276c94ebebafa2f388dc4826712e4590

Observation b0cc402e-91c3-4731-bb11-86d9ab1b1021 · outbound

This paper cites Output text is encoded into BPE.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Output text is encoded into BPE

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:25.519045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:13.851341Z digest=sha256:46592fe358a6e26cd43763b9b3e29eb7826174a6afd68580ba9e8699215ea28e

Observation 55b8654c-edf4-4196-a818-749efbd5c139 · outbound

This paper cites Time Fig.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Time Fig

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:25.090547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:14.077263Z digest=sha256:faa24e1a1f8796f4af9642c743d346fa4569d7ae9d9d1248cf6416951aa62908

Observation df277d58-ba0c-4348-87c0-c04a91f3b2e3 · outbound

This paper cites The encoder layers use a convolution kernel size of31, an em- bedding size of512, and a linear layer hidden size of 2 048.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution The encoder layers use a convolution kernel size of31, an em- bedding size of512, and a linear layer hidden size of 2 048

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:25.248711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:13.957612Z digest=sha256:b7872339a86ff43fb8b5f7966eaaba8bc1feaf96045685130a68dbf3bcab06a4

Observation 237cf734-eaba-43a1-a1a0-261ae9493e39 · outbound

This paper cites an unresolved cited work.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:32:24.899323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:14.179728Z digest=sha256:e177f989ff12b8131312549b0e327186404bf0ffaa2c756eb1b49fd4203f155b

Observation 192b2f02-aa97-4c3d-82e6-7deffcdda717 · outbound

This paper cites This is the first in-depth analysis of saliency maps in relation to fine-grained acoustic patterns across three phoneme classes.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution This is the first in-depth analysis of saliency maps in relation to fine-grained acoustic patterns across three phoneme classes

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:25.001122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:14.125900Z digest=sha256:784696ba9e929db7e988bb6c98234c240367224f9aefcd8b490b6cb45ea59154

Observation bfa9bd22-277f-41f7-99d1-40563c5d85ee · outbound

This paper cites Understanding the representa- tion and computation of multilayer perceptrons: A case study in speech recognition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Understanding the representa- tion and computation of multilayer perceptrons: A case study in speech recognition,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.407824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:14.927096Z digest=sha256:31761fde77b730d0e9f073859918d36b45ac63922f8a34158e9dc269c00797eb

Observation 8427c8e7-6bed-4d11-9105-189ad95a3a56 · outbound

This paper cites Analyzing phonetic and graphemic representations in end-to-end automatic speech recog- nition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Analyzing phonetic and graphemic representations in end-to-end automatic speech recog- nition,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:24.714209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:14.217408Z digest=sha256:c161e17ede6fbe0fa8b02ee82cabaddf59e44faf687e97f4604e76aacd4d3c44

Observation 29201fdd-5aa4-4714-9f74-777854121a46 · outbound

This paper cites Probing phoneme, language and speaker information in unsupervised speech representations,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Probing phoneme, language and speaker information in unsupervised speech representations,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:24.485514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:14.303707Z digest=sha256:597a710757d30093987734d9df78ac6fabd22802c4a67969d2b988e8654ee7bd

Observation dbe9d202-7fd3-4c78-aeab-be3dba4606a2 · outbound

This paper cites Domain-Informed Probing of wav2vec 2.0 Embeddings for Pho- netic Features,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Domain-Informed Probing of wav2vec 2.0 Embeddings for Pho- netic Features,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:24.312303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:14.412565Z digest=sha256:777a653cb6489a53e7e80bbdf72fd6b5832fbf3276fdc76a236088e573250b6a

Observation d1e2d11f-e127-44a7-88c1-9123d62fd077 · outbound

This paper cites Probing self- supervised speech models for phonetic and phonemic informa- tion: A case study in aspiration,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Probing self- supervised speech models for phonetic and phonemic informa- tion: A case study in aspiration,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:24.127112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:14.556191Z digest=sha256:529b9cbc81eb6be67dbbed4ec2afd32e6e56f2bbe5eb055e510937adbb158f6d

Observation c469bcd5-9759-4445-90a0-6495cc0b9107 · outbound

This paper cites Self-supervised speech representations are more phonetic than semantic,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Self-supervised speech representations are more phonetic than semantic,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.912683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:14.636029Z digest=sha256:b9107317eacc120a5301fbd946bce88058679a306685b034b3ae42f5581662d1

Observation a9a232c3-1c1d-478b-b710-8fb9d5be4099 · outbound

This paper cites Encoding of phonol- ogy in a recurrent neural model of grounded speech,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Encoding of phonol- ogy in a recurrent neural model of grounded speech,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.745518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:14.725407Z digest=sha256:01c4ed0ebad7bfe686682c673b9d21c2332ce6d37515187a3a3810b653ab3f18

Observation cbaf82b0-f548-412b-b390-a1ef98ef0803 · outbound

This paper cites Learn- ing weakly supervised multimodal phoneme embeddings,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Learn- ing weakly supervised multimodal phoneme embeddings,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.605363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:14.829948Z digest=sha256:3c6508efe7e0fc6ad81633f16654b36d93b9587db86378ee9e5516a22e46d841

Observation 4da04087-fc0b-499c-a2b8-d111cefd052a · outbound

This paper cites Acous- tic characteristics of American English vowels,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Acous- tic characteristics of American English vowels,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:22.306585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:15.511358Z digest=sha256:0580ef8738c9008191a6f93440a21179437cda78eb5146a36aa8de4d30762486

Observation 235f25a8-d890-43e1-8166-70c1b049d5af · outbound

This paper cites Neuron Activation Profiles for Interpreting Convolutional Speech Recognition Models,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Neuron Activation Profiles for Interpreting Convolutional Speech Recognition Models,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.215208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:15.004400Z digest=sha256:1c082fcd869db5c30391ff4d887d1c0b55f9ff3f1eaa4739891ed5bf4a413c90

Observation af0cc813-9177-45ea-841e-5d938a432c13 · outbound

This paper cites Interpretable Convolutional Filters with SincNet,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Interpretable Convolutional Filters with SincNet,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.047254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:15.077490Z digest=sha256:5ca0a2a58ef31e5bb4420163de1888ce8aa67eee3d04a001c8dd7ff78acbd0c9

Observation 1c7f36dc-b57b-4146-8590-64f887c27d60 · outbound

This paper cites Introspection for convolutional automatic speech recognition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Introspection for convolutional automatic speech recognition,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:22.858990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:15.146016Z digest=sha256:74504de91f5a86e8e89570c4efc14c33355c25215ca3abbd469e71cf168d5dfa

Observation cbad7af6-f918-4f29-b034-15bd8714eaef · outbound

This paper cites End-to-end acoustic modeling using convolutional neural networks for HMM- based automatic speech recognition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution End-to-end acoustic modeling using convolutional neural networks for HMM- based automatic speech recognition,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:22.652588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:15.227371Z digest=sha256:e26405aff268bb40fa190a50dc02e41af3497f7df4fbf28684f3f79e9db506f4

Observation 2c80d404-16f8-4f7f-876e-d7acd6fc5683 · outbound

This paper cites Gradient-Adjusted Neuron Activation Profiles for Comprehensive Introspection of Convolutional Speech Recognition Models.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Gradient-Adjusted Neuron Activation Profiles for Comprehensive Introspection of Convolutional Speech Recognition Models

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:32:17.857757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:15.281363Z digest=sha256:4322cc29308a18f6b6b0ad4e780c5afdb35ef7a60fe893a8e8fb0f12fead9e69

Observation 34779c97-1dc5-4c41-89fe-784ecac7585e · outbound

This paper cites Directly Comparing the Listening Strategies of Humans and Machines,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Directly Comparing the Listening Strategies of Humans and Machines,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:22.489366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:15.360537Z digest=sha256:17500e377d0ab969a22345fbcf432e2f304084a513caa372574492dd08ea5382

Observation c0ff9443-cf9f-4f20-b0a5-8365f0174d6d · outbound

This paper cites SPES: Spectrogram Perturbation for Explainable Speech-to-Text Generation.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution SPES: Spectrogram Perturbation for Explainable Speech-to-Text Generation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:32:15.428570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:32:15.428570Z digest=sha256:3cfbd8e2857241be03abaacdb0328cb473573c63c0c665c3be662990ae1e3fd4

Observation b03f39f8-479e-4e1c-85ab-9dd34483a66b · outbound

This paper cites an unresolved cited work.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:32:25.364852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:13.923390Z digest=sha256:b516e29eafe8418d8ac3d4275f8e3474f401d1786e529b2dc88664c69162ae7f

Observation 9a5d5fc3-9bff-4c7e-9af2-d9bccee1abff · outbound

This paper cites Burst and Transition Cues to V oicing Perception for Spo- ken Initial Stops by Impaired- and Normal-Hearing Listeners,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Burst and Transition Cues to V oicing Perception for Spo- ken Initial Stops by Impaired- and Normal-Hearing Listeners,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:22.124476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:15.594193Z digest=sha256:5c93aa0e09ec37196f3eef66e934141db57018fbd0b07c220ea3e738a0ae15a2

Observation 4c5491e7-8eea-45f3-95e2-2a5e777da75e · outbound

This paper cites Acoustic cues of voiced and voiceless plosives for determining place of articulation,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Acoustic cues of voiced and voiceless plosives for determining place of articulation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:21.905743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:15.692957Z digest=sha256:47674075544d944f7adb1f3757b6a7d6409e41e322508458694981830a03fde0

Observation 1fe81541-41d6-48f7-95b1-c105c5a05685 · outbound

This paper cites Acoustic characteristics of English fricatives,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Acoustic characteristics of English fricatives,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:21.687005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:15.779334Z digest=sha256:18741f6586e6d0608454e0f5b972329be4612e7f34a49384ac6111a592f24b2c

Observation 24f587b8-0313-400c-866b-b6335394924b · outbound

This paper cites Acoustic characteristics of clearly spoken English fricatives,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Acoustic characteristics of clearly spoken English fricatives,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:21.511061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:15.889219Z digest=sha256:42cb1a9d219c8d1d0a9096b9430167c1f2cbe6b342963da1bc53a6589f28c948

Observation a2da6397-2f44-4d45-ad16-4085f91b2d35 · outbound

This paper cites Conformer: Convolution-augmented Transformer for Speech Recognition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Conformer: Convolution-augmented Transformer for Speech Recognition,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:21.258248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:15.968308Z digest=sha256:fe45fbb945724f90ccdf2dd4fc767005b740e9816fed92c330ef99379d9c8ba5

Observation 8a09b4e1-738f-41ad-b7aa-3375bf1f9000 · outbound

This paper cites Introducing Parsel- mouth: A Python interface to Praat,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Introducing Parsel- mouth: A Python interface to Praat,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:32:16.075519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:32:16.075519Z digest=sha256:99ba3ebcd727512b072c565532ebe85235ae90e0cf90c3215189901a38dd53c9

Observation af314bcb-d75a-494b-bde2-ee3b78eecc9f · outbound

This paper cites Pykaldi: A Python Wrapper for Kaldi,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Pykaldi: A Python Wrapper for Kaldi,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:21.020089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:16.143157Z digest=sha256:fe69c71eb5d3d8b8468364ab73e66131cdd08fcc8b1c75e3f1be52cd67554f26

Observation 3effd402-33aa-4488-80dd-0fb1fe60c79c · outbound

This paper cites Neural Machine Transla- tion of Rare Words with Subword Units,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Neural Machine Transla- tion of Rare Words with Subword Units,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:20.786942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:16.218548Z digest=sha256:2c8c28a5930e1cbcdab13ce43b51b7004582d7ec0563e967212ccbeb24691638

Observation 592131ef-154a-4488-bb6f-b1564daa0dd9 · outbound

This paper cites SentencePiece: A simple and lan- guage independent subword tokenizer and detokenizer for Neural Text Processing,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution SentencePiece: A simple and lan- guage independent subword tokenizer and detokenizer for Neural Text Processing,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:20.523379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:16.297689Z digest=sha256:d059bf3af79f85d36811d0615abe6fd62047b756603365af5bce7f0dbccba797

Observation 3f931957-e4ad-4cb7-95f1-c7a131c9a156 · outbound

This paper cites Attention is All you Need,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Attention is All you Need,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:20.287706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:16.359174Z digest=sha256:f5ca2d5eb12a92e35d616da591f80cb4f024f7416b5da42e788491f905186a6e

Observation fe1b260f-6a28-4646-9634-947f5a0e244d · outbound

This paper cites Fairseq S2T: Fast Speech-to-Text Modeling with Fairseq,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Fairseq S2T: Fast Speech-to-Text Modeling with Fairseq,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:32:16.455994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:32:16.455994Z digest=sha256:b5b9cb22694590103d1de769ab697bcb87276f743296736175a587ccbe824f78

Observation 6d166b02-a9eb-4aad-b7b8-5a8562bf286a · outbound

This paper cites Com- mon V oice: A Massively-Multilingual Speech Corpus,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Com- mon V oice: A Massively-Multilingual Speech Corpus,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:20.058654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:16.536895Z digest=sha256:497c76d39c97727bec24f21981a2206c2c5a124d027cf25f0e6e228e9a9a38ee

Observation 5b9d54ab-ea59-49dd-a540-7eca7aa6a668 · outbound

This paper cites Lib- rispeech: An ASR corpus based on public domain audio books,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Lib- rispeech: An ASR corpus based on public domain audio books,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:32:16.626379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:32:16.626379Z digest=sha256:c8cff63ac2a6c048ccec8801592a5a938799d21b68db4fc3326cc56bb9f9c21c

Observation db3fa82a-0329-4f53-950d-1afe1a917983 · outbound

This paper cites TED-LIUM 3: Twice as Much Data and Corpus Repartition for Experiments on Speaker Adaptation,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution TED-LIUM 3: Twice as Much Data and Corpus Repartition for Experiments on Speaker Adaptation,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:19.771580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:16.724921Z digest=sha256:9c5b36a56d57d981a9a02651efe09e18f9a0b7c47ffa5956866ffe541e33ae55

Observation 824f9cd2-f374-4c14-b28d-3409165a5104 · outbound

This paper cites V oxPopuli: A Large- Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution V oxPopuli: A Large- Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:19.558003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:16.814577Z digest=sha256:904257bf11b9407f58264c08f1a04b96d682cc2f311eed8181001b8fce942f32

Observation 26f4e42a-4f62-481e-9f70-c3a43f6a9d06 · outbound

This paper cites Rethinking the Inception Architecture for Computer Vision,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Rethinking the Inception Architecture for Computer Vision,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:19.372863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:16.949565Z digest=sha256:99d7663dec2469fb45b279c55a7807d6e7475c50b8e514281cb374d0837f68b6

Observation 800f6bae-2e69-4255-b17b-ccd2ea389988 · outbound

This paper cites Con- nectionist temporal classification: labelling unsegmented se- quence data with recurrent neural networks,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Con- nectionist temporal classification: labelling unsegmented se- quence data with recurrent neural networks,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:19.203475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:17.048361Z digest=sha256:9a9265a9aee20b189526ed2b1a17499a6ba8b14b7343c8b16d0a80ba58cb63b9

Observation b5794d0f-42b7-49d1-aad3-55220a5f8f8a · outbound

This paper cites Adam: A Method for Stochastic Opti- mization,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Adam: A Method for Stochastic Opti- mization,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:32:17.167561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:32:17.167561Z digest=sha256:182296812eb433f25bf1c288149a97466eedadfe5a1f9d31a3c8957a44d71fa5

Observation 62b93cdb-19e3-43f4-b3ce-ada0bfe6a740 · outbound

This paper cites SpecAugment: A Simple Data Aug- mentation Method for Automatic Speech Recognition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution SpecAugment: A Simple Data Aug- mentation Method for Automatic Speech Recognition,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:18.943826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:17.259642Z digest=sha256:bfc787553e5d04a0305d7c20511c445750afb83ec81dc35038004a38a99eb6d6

Observation 57088e64-c670-4bf6-a24c-a8f1cf64f16e · outbound

This paper cites Darpa timit acoustic-phonetic continuous speech corpus cd-rom TIMIT,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Darpa timit acoustic-phonetic continuous speech corpus cd-rom TIMIT,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:18.661884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:17.396416Z digest=sha256:acc91a7820825e48af381b80d4b716ac83ed6f29161da8807dc4393227584d2e

Observation 53152cd0-8d1c-437f-9961-44133f0028fe · outbound

This paper cites Twists, humps, and pebbles: Multilingual speech recognition models exhibit gen- der performance gaps,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Twists, humps, and pebbles: Multilingual speech recognition models exhibit gen- der performance gaps,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:18.424129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:17.502417Z digest=sha256:abc73cbe04ae76ad3f338c33043975224bc959732d737f9d3ca0522d2130fe8b

Observation a73cfbae-9702-461d-be97-e71fab6edf86 · outbound

This paper cites Burst spectrum as a cue for the stop voicing contrast in American English,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Burst spectrum as a cue for the stop voicing contrast in American English,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:18.247636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:17.610645Z digest=sha256:a16f54f9e673eaa433c0bfeefabc65be4896bf1ababde8218903eb16ba53d8e8

Pith citing papers

Observation 0b075e5a-482f-47e0-ab8a-9b01b6920469 · inbound

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution cites this paper.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T11:32:18.040121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:32:13.765343Z digest=sha256:fd278bf44154299419587db4b5c51010276c94ebebafa2f388dc4826712e4590