Pith. sign in

Paper Citation Record · LEDGER

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends

As of 23 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2502.03260.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.03260 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T05:26:14.104367Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact0
  • verified fuzzy51
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ae385333-7bda-4bdc-bbc3-3000fb6d90a5 · outbound

This paper cites Learning filterbanks from raw speech for phone recogni- tion,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Learning filterbanks from raw speech for phone recogni- tion,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.837636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.879856Z digest=sha256:022141c04a8a4ee53f651a6caa9cff88a8decb483b220640e5279d3cd153eb42

Observation a940fb83-a029-4d50-89fd-67c331b422c3 · outbound

This paper cites Acoustic modelling with cd-ctc-smbr lstm rnns,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Acoustic modelling with cd-ctc-smbr lstm rnns,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.827262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.884229Z digest=sha256:db5b0ace1b761e699ccf7c5d907df357abc81a3f1770a6d1c62cf4f53f59f125

Observation b433856f-f59f-4116-9541-a783e22ab4b7 · outbound

This paper cites Speaker recognition from raw waveform with sincnet,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Speaker recognition from raw waveform with sincnet,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.817135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.887667Z digest=sha256:aa17ba1b439a6c3263051fbcbdac8335fc8fa535fbe6a00e1c23f61af5031bbc

Observation 568ab058-3cde-4018-b704-9df4bb8da0ec · outbound

This paper cites End-to-end spoofing detection with raw waveform CLDNNS,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends End-to-end spoofing detection with raw waveform CLDNNS,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.806780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.891094Z digest=sha256:0a62a802c6695794771235796b0edbe3a89292fa7bc28c3fae1329ecfd4b36a6

Observation 987b28f1-b92f-4cc5-99f8-aaf1b881d4c3 · outbound

This paper cites ESC: Dataset for environmental sound classification,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends ESC: Dataset for environmental sound classification,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.796481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.894684Z digest=sha256:b70da9457533ad28ceaa4de0ec3c56d574f6e928ab5590034151f6328c091404

Observation ffcad32c-bb2f-4883-9c41-351cd4b67183 · outbound

This paper cites Towards learning universal audio representations,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Towards learning universal audio representations,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.786436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.898261Z digest=sha256:7866fe85c503e9521b633105ebd29408342a22ac2a4d1a970b464ad9abd7254a

Observation e2b7b99e-5d1f-45af-8475-c0f596a7a4ad · outbound

This paper cites Filterbank design for end-to-end speech separation,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Filterbank design for end-to-end speech separation,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.777552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.901994Z digest=sha256:cbdb5b362c6cbee292d370ee557516e7eb8608b6d1d89903513ca16f7b2f286d

Observation 7b1da039-a887-4681-a7e8-edace45ee0a6 · outbound

This paper cites Conformer: Convolution-augmented transformer for speech recognition,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Conformer: Convolution-augmented transformer for speech recognition,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.768720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.905152Z digest=sha256:42768d4882adf329b7b08ef2ce56500f4a8d9e3aaa159de2854054de02361553

Observation c13d1567-eb07-4a9e-9bf0-1572fb382d1d · outbound

This paper cites Deep scattering spectrum,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Deep scattering spectrum,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.759439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.908110Z digest=sha256:7ea173761f18366969e78a64716113382a14708ace193c6daab793598dfadaf3

Observation 07170184-b5fb-4077-8755-ccca841376d0 · outbound

This paper cites Pitch-Adaptive Front-End Features for Robust Children’s ASR.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Pitch-Adaptive Front-End Features for Robust Children’s ASR

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.749936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.911056Z digest=sha256:d33244493f2637516ff783ead24ec3aaa59999b59543aa13cd1bbf1a25fd2352

Observation bc707b4f-06f4-4d4d-9196-9cb09ca71960 · outbound

This paper cites Deep residual learning for image recognition,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Deep residual learning for image recognition,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T05:26:13.914144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T05:26:13.914144Z digest=sha256:8e3c833362f4ee09d77ea44a1ca43fd0ad0ad8dfd1e81b8e5a488333c057d711

Observation ec42de38-d2cf-442c-af3b-f8492da4dc89 · outbound

This paper cites Conv-tasnet: Surpassing ideal time– frequency magnitude masking for speech separation,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Conv-tasnet: Surpassing ideal time– frequency magnitude masking for speech separation,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.734147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.917582Z digest=sha256:97f7584c2a716f062b61b1ad33f3973461003684bc02882ec258775cbb7027d3

Observation 188b866e-8f95-4d63-9637-a3d9948b6e5c · outbound

This paper cites Real time speech enhancement in the waveform domain,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Real time speech enhancement in the waveform domain,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.725155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.921110Z digest=sha256:529295239ebae348c4827dcb31221d6b071b9001d1706a712c5777b4b6f8fdba

Observation c3f36b24-9ed4-49d3-8ac2-d58ead9879f6 · outbound

This paper cites Learning the speech front-end with raw waveform cldnns,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Learning the speech front-end with raw waveform cldnns,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.715971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.924582Z digest=sha256:b306f8bf0f367654acf34286e5a6a61b29a64189efdfcd3e25eee0eebcdaba4f

Observation e04c8143-d829-4d61-995b-dd344e649c4c · outbound

This paper cites wav2vec: Unsupervised Pre-training for Speech Recognition.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends wav2vec: Unsupervised Pre-training for Speech Recognition

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T05:26:13.931617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T05:26:13.931617Z digest=sha256:387c8ffeb200665c5c42436f59e9abaf206aeeb00ab8d880db31963dad85fb95

Observation 2fc00c2b-de34-49f3-9328-f756ff147cc3 · outbound

This paper cites A deep neural network integrated with filterbank learning for speech recognition,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends A deep neural network integrated with filterbank learning for speech recognition,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.696742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.935380Z digest=sha256:dcc04d258e481942522b95ad02a272d0b3892610e06d7b4563a04e0bdde79bb2

Observation f9eb6d70-cb69-48e1-b275-8a2feacc7aed · outbound

This paper cites Learning the speech front-end with raw waveform CLDNNs,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Learning the speech front-end with raw waveform CLDNNs,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T05:26:13.938772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T05:26:13.938772Z digest=sha256:4ed05387155ec9fa403d95298a82bec8a37afd46d8b41c9a8ddeed965ab75daf

Observation 96635e46-c47b-491a-a2b6-367606f4e5b6 · outbound

This paper cites LEAF: A Learnable Frontend for Audio Classification,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends LEAF: A Learnable Frontend for Audio Classification,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.682192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.942316Z digest=sha256:cd1f82c7c1832295d284d5411946ab4a0cbc12cd4b269cdf944c3d0c058eac45

Observation 96ed5ad4-f44f-4122-8c3b-77edf07c504d · outbound

This paper cites Train- able frontend for robust and far-field keyword spotting,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Train- able frontend for robust and far-field keyword spotting,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.673484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.946069Z digest=sha256:19662828940b2a84cc3efd8d38ff7f4368e93ade41361de9be0749deb734a31b

Observation 823bf4ce-98f8-4d73-bab6-3b59788ebf0a · outbound

This paper cites Automatic gain control in cochlear mechanics,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Automatic gain control in cochlear mechanics,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.544843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.949703Z digest=sha256:a82003fae7a540fb29fb873817a93d5f8b23220d97c4b53018e9c9b7730d0c46

Observation 3ac8ca3a-36c1-4d2f-926b-d1a21d2bf462 · outbound

This paper cites The cochlea as a smart structure,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends The cochlea as a smart structure,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.536970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.953238Z digest=sha256:1866a5c129cd7837fb2da1a8bac80aaa33c0b834350ac49ec16ccb8fdc5b7432

Observation 30e8efdd-54cd-4e67-86aa-5e978e959651 · outbound

This paper cites Integrating the active process of hair cells with cochlear function,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Integrating the active process of hair cells with cochlear function,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.528781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.956772Z digest=sha256:2205a8c4ef3bb310f8beae59fa237e2eb4963cbbcf68f20ff541f503d90b9eed

Observation 2a437d72-3168-4151-a7bf-25b8d9f34c15 · outbound

This paper cites Auditory processing of speech signals for robust speech recognition in real-world noisy environments,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Auditory processing of speech signals for robust speech recognition in real-world noisy environments,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.520006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.960227Z digest=sha256:18bf4cbef95b1dbb0101c76cc41738acb9f4646cd42a3ffa47b27223f0b3ede1

Observation b2c10c75-0936-4af9-96e9-675fb9a6b305 · outbound

This paper cites Power-normalized cepstral coefficients (pncc) for robust speech recognition,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Power-normalized cepstral coefficients (pncc) for robust speech recognition,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.512342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.963653Z digest=sha256:a9d9406ce9ab3c8233ffde09fac943c234bc146a4e2e74f93550fac2a3f8f001

Observation 49d37a51-9fd5-4a01-b1de-447ae15fa9eb · outbound

This paper cites Dnn controlled adaptive front-end for replay attack detection systems,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Dnn controlled adaptive front-end for replay attack detection systems,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.504058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.967195Z digest=sha256:2c93099c6bc25cbfba5382e69feb1a19e55da87ff0a0788d61f6a1574bee3cca

Observation b4935bf3-18c0-43da-b8e7-eb4805662af2 · outbound

This paper cites Towards learning a universal non- semantic representation of speech,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Towards learning a universal non- semantic representation of speech,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.495664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.970914Z digest=sha256:eb8570aab68d9beb7a90245435a101910714ceccc5da1c05132366692796c903

Observation e335e00c-0181-45e6-af87-8e7fd632a97d · outbound

This paper cites Learning filter banks within a deep neural network framework,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Learning filter banks within a deep neural network framework,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.486754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.974756Z digest=sha256:0660663aab2bbcd01bf1ef1600d2fbdd7b0b137ca664cf9f81c695c4d579e117

Observation 696a8325-e2f6-4544-b5d4-80b415189c30 · outbound

This paper cites Data-driven harmonic filters for audio representation learning,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Data-driven harmonic filters for audio representation learning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.478110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.978505Z digest=sha256:447b9df415cbe027ce165f757b856f11441a49ce6955d4cc9b4f83b53f631203

Observation 5f364cc1-84c6-4f33-b30d-4ae5821f6d8c · outbound

This paper cites Learning a better representation of speech soundwaves using restricted boltzmann machines,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Learning a better representation of speech soundwaves using restricted boltzmann machines,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.469827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.982089Z digest=sha256:31513c5b0a9b68b1d66d470e8e593bfcb02bbc27203ad52e1d164e6099aa2ab1

Observation 34089aae-4b4a-49a9-bd3c-a082aee5dcec · outbound

This paper cites Estimating phoneme class conditional probabilities from raw speech signal using convolutional neural networks,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Estimating phoneme class conditional probabilities from raw speech signal using convolutional neural networks,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.460196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.985562Z digest=sha256:430205d8a4fd26b4cbd0d5022e6f09450744c22a9f401b49be048542ae281bdd

Observation 4f853213-f608-4682-8cf0-ca09f16564dd · outbound

This paper cites Speech acoustic modeling from raw multichannel waveforms,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Speech acoustic modeling from raw multichannel waveforms,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.706115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.989162Z digest=sha256:2c275d98834ef6db75947d312cd79ae0ede143a8b7e8c4da9c52db714288d65e

Observation a46357da-afd4-4e0a-be0d-758f37f03f11 · outbound

This paper cites Learnable frontends that do not learn: Quantifying sensitivity to filterbank initialisation,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Learnable frontends that do not learn: Quantifying sensitivity to filterbank initialisation,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.449801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.992627Z digest=sha256:3e426e6ef0ec5801714e83516c31153210c62b954d46e0b368ab2a5a2166ffc9

Observation abebbd01-ee2c-490c-9a5d-ad8ee5937b04 · outbound

This paper cites What is Learnt by the LEArn- able Front-end (LEAF)? Adapting Per-Channel Energy Normalisation (PCEN) to Noisy Conditions,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends What is Learnt by the LEArn- able Front-end (LEAF)? Adapting Per-Channel Energy Normalisation (PCEN) to Noisy Conditions,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.439497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.996137Z digest=sha256:8ebabff5ddc39c8e015530d5d9c8039952b6474d5d3636d8444d58ec73e78e0e

Observation 26602cf0-42cb-4122-a7f0-560c84a13eaf · outbound

This paper cites Masked autoencoders that listen,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Masked autoencoders that listen,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.429727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:13.999808Z digest=sha256:a1f21754e793b2e0b5887ce2cfab193475707fd5549439056735888e3949d809

Observation cd1c2a4b-0a44-4412-b0fa-4b758bcff402 · outbound

This paper cites Beats: Audio pre-training with acoustic tokenizers,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Beats: Audio pre-training with acoustic tokenizers,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.418791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.003648Z digest=sha256:1a49da2571ea1d1fff45f5f45c30343282b576a67027157d3930809430d97170

Observation 7486415e-e506-4f70-9415-29bed3426fef · outbound

This paper cites Byol for audio: Exploring pre-trained general-purpose audio representations,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Byol for audio: Exploring pre-trained general-purpose audio representations,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.408323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.007243Z digest=sha256:aca3dc93991610246a188fc8bd8b619d9ba44282941b78f6562fc169c4915396

Observation 7ae9b71d-951e-48e3-900d-591b787d0c02 · outbound

This paper cites AST: Audio Spectrogram Trans- former,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends AST: Audio Spectrogram Trans- former,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.397602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.010831Z digest=sha256:a72746ea6a536f371ec2520cfc943fcca4b10c3a841a1632f8cb713bb53f7494

Observation 1768a64a-c332-4c4c-b547-4c5f54d8275a · outbound

This paper cites SSAST: Self-supervised audio spectrogram transformer,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends SSAST: Self-supervised audio spectrogram transformer,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-09T05:26:14.014118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T05:26:14.014118Z digest=sha256:46b09aff07b712f5e9a12a5a2a0c9a7e37b21a7fe41fbb7d226513224ea98ff0

Observation eeebf783-eb90-4f01-846d-66814d587806 · outbound

This paper cites Hubert: Self-supervised speech representation learning by masked prediction of hidden units,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Hubert: Self-supervised speech representation learning by masked prediction of hidden units,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-09T05:26:14.017817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T05:26:14.017817Z digest=sha256:94b661e08b8c8aaf15ad2f3021ace2bd942f2e319ef5d2276e50fa0b24da425d

Observation 7e88e4c0-982b-4334-b898-6f64cb2aec42 · outbound

This paper cites Wavlm: Large-scale self-supervised pre- training for full stack speech processing,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Wavlm: Large-scale self-supervised pre- training for full stack speech processing,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-09T05:26:14.021411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T05:26:14.021411Z digest=sha256:f1c71264f8d0026d5d1c144bd4506ccc6dd1a8274be9abc56860ea729ad4c886

Observation c5f014fb-70ac-4fb9-8735-49df7d7d48a8 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends wav2vec 2.0: A framework for self-supervised learning of speech representations,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T05:26:14.024818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T05:26:14.024818Z digest=sha256:f37423eda3ca9c3f6f6be4662369468daad2f62e15078e64546e2e4bd1feae8b

Observation c3f85a8f-b93b-4cb0-a207-d0a72c9e5688 · outbound

This paper cites Learning neural audio features without supervision,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Learning neural audio features without supervision,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.365196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.028472Z digest=sha256:bf5f7f140bb52c71ba52c436350130eaaffa04e13c6dea403df353cfbfb9a54f

Observation bfbff5b8-db82-459b-8b86-5e33261bbef0 · outbound

This paper cites Biologically in- spired adaptive-q filterbanks for replay spoofing attack detection,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Biologically in- spired adaptive-q filterbanks for replay spoofing attack detection,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.355792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.032393Z digest=sha256:d859186a34839ea42cebc67adc53f22c95b4e4aaf1525dad7d27775d290a0af1

Observation 5f1d3b46-a94b-4de2-bd16-f29dabdd349a · outbound

This paper cites Replay detection in voice biometrics: an investiga- tion of adaptive and non-adaptive front-ends,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Replay detection in voice biometrics: an investiga- tion of adaptive and non-adaptive front-ends,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.347059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.036008Z digest=sha256:72c04d40568b9c477f0d9d17d160a86a50d2136d756d872ce1d173e77ebe38ce

Observation bb2a5e94-9e64-4e1a-830f-a125d7202c21 · outbound

This paper cites Speech nonlinearities, mod- ulations, and energy operators,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Speech nonlinearities, mod- ulations, and energy operators,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.337486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.039066Z digest=sha256:c03b9ed41fa2e57368f524881700f1d117fb241321e609bd8ced3d72a03197b6

Observation 9201f60c-b231-4260-a4ca-5c3bb7e82257 · outbound

This paper cites Encoding frequency modulation to improve cochlear implant performance in noise,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Encoding frequency modulation to improve cochlear implant performance in noise,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.328118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.042172Z digest=sha256:7abd45cb16bc5a26aa79f6862ba31de4ef188f0ac121837f2ca4cd588c59cb48

Observation 8c08a03b-68c8-492e-99fd-53df820cfe74 · outbound

This paper cites Detection of replay-spoofing attacks using frequency modulation features,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Detection of replay-spoofing attacks using frequency modulation features,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.317950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.045326Z digest=sha256:49105c01c4ffe8eb417b9a9d85f5523e8bec9a7acf7467eb2d3694d6259e3fee

Observation c9ac29bd-7e39-41a9-806c-1dd3adb472bd · outbound

This paper cites Auditory inspired spatial differentiation for replay spoofing attack detection,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Auditory inspired spatial differentiation for replay spoofing attack detection,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.308196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.048523Z digest=sha256:0d18c39c264500095aa41329f209ae3835064165a91447cb21be7904e17e1215

Observation 0ab6efef-faf6-4056-8537-be3afd456cda · outbound

This paper cites SUPERB: Speech Processing Universal PERformance Benchmark,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends SUPERB: Speech Processing Universal PERformance Benchmark,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.297393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.051418Z digest=sha256:b5ddf2b9b257bd8124d603789e4cde05a3a87403b8a61eaec998ed67dfc8bbab

Observation 2caf3c72-d1f0-445b-a828-395bdde77a3f · outbound

This paper cites The GTZAN dataset: Its contents, its faults, their effects on evaluation, and its future use.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends The GTZAN dataset: Its contents, its faults, their effects on evaluation, and its future use

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-09T05:26:14.054794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T05:26:14.054794Z digest=sha256:148fbf6d8bd43441f25c1796d19e899b1b14f6efb8f757eb8e6029b379d45f69

Observation 3779dd4e-19ec-4fe5-954c-9cd3fa21b426 · outbound

This paper cites Deep learning and music adversaries,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Deep learning and music adversaries,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.287812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.058810Z digest=sha256:96d2ecff751ac88a5e3d2cbf78f47e88cbb80dbf96295ba1399d0b574c0f1970

Observation 250ce5a4-a1a9-4ba3-9192-eca598d227de · outbound

This paper cites FMA: A dataset for music analysis,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends FMA: A dataset for music analysis,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.275726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.062146Z digest=sha256:5cb1b06515d935d70e4f844a0527fac4c05c9e905cb706a5f157cd1730858fb8

Observation 05b29f96-459b-4640-88ab-fd42e35594ac · outbound

This paper cites CREMA-D: Crowd-sourced emotional multimodal actors dataset,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends CREMA-D: Crowd-sourced emotional multimodal actors dataset,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.264969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.065607Z digest=sha256:15324f12ceb594fbb65f79a3c15581a313b90f8c2668a600ac3faac2738c4af1

Observation e364dc85-aead-457a-812a-15451f7c3fbb · outbound

This paper cites IEMOCAP: Interactive emotional dyadic motion capture database,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends IEMOCAP: Interactive emotional dyadic motion capture database,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.253238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.069586Z digest=sha256:c99dffa44e93c2376a27550ec14c08e9d25d0136d9f09ca55d72719ec529ec14

Observation d553a0da-0f99-4cbc-981a-86d097cd55f3 · outbound

This paper cites Speech Commands: A Dataset for Limited-Vocabulary Speech Recognition.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Speech Commands: A Dataset for Limited-Vocabulary Speech Recognition

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-09T05:26:14.073221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T05:26:14.073221Z digest=sha256:00ff7762e3f46dea8fb7dc69fa53e62873b23f47c0691f31b0f7822b8232c841

Observation 616539e7-81b2-4ce4-b42d-0c78dc2c6876 · outbound

This paper cites V oxceleb: A large-scale speaker identification dataset,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends V oxceleb: A large-scale speaker identification dataset,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.241327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.077714Z digest=sha256:061dc9ff0cc38f42f5dc5d5b950a7802b181fdb9f1f73dad092e1e5acf969fce

Observation 4227709d-dc71-45af-a170-1dfaec37d54c · outbound

This paper cites Efficientnet: Rethinking model scaling for convolutional neural networks,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Efficientnet: Rethinking model scaling for convolutional neural networks,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.230834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.081608Z digest=sha256:cfaa20aad228fe2d45d072d4673d8b8aafd851046b5dd5f5f1da35531cdf31d0

Observation 57cef8dd-893e-4a9a-a016-19520797f8f0 · outbound

This paper cites Squeeze-and-excitation networks,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Squeeze-and-excitation networks,

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-09T05:26:14.085174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T05:26:14.085174Z digest=sha256:4e5004d118c1debd146100629c2c6e362ef79edd33b318e6a0e276382ac9b755

Observation 775b265a-9b47-4a8d-aad4-3f138c61ac15 · outbound

This paper cites MobilenNetv2: Inverted residuals and linear bottlenecks,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends MobilenNetv2: Inverted residuals and linear bottlenecks,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.212700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.089407Z digest=sha256:590b96966047f3cb67655d5c3bc73e04747443c2f9b5ecfb4f390b6025d55082

Observation f53d3120-767c-4e3a-b08b-efdb9fa3a50f · outbound

This paper cites Adam: A method for stochastic optimization,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Adam: A method for stochastic optimization,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-09T05:26:14.093471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T05:26:14.093471Z digest=sha256:a1389e5e45703c69ae64c2637bf82e0b8e4227f215685414b2677f992164757f

Observation 9169a92b-ae62-4e6e-8daf-ac86329aa8f0 · outbound

This paper cites Audio set: An ontology and human- labeled dataset for audio events,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Audio set: An ontology and human- labeled dataset for audio events,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.194668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.097222Z digest=sha256:79bea3a70bcf11782c6252267675c57a4703c4eeb258efc388ff34f51140ebf5

Observation 1f4c2ae0-74c0-48ba-af35-65f75a28ad04 · outbound

This paper cites PANNs: Large-scale pretrained audio neural networks for audio pattern recognition,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends PANNs: Large-scale pretrained audio neural networks for audio pattern recognition,

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-09T05:26:14.100684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T05:26:14.100684Z digest=sha256:613201149714c43841b46b1ff6ec5206a52c6da2026374ebcf7ca22471fd71be

Observation 3873f4f9-cb4f-4cbb-b080-eab6cc314233 · outbound

This paper cites Librispeech: an asr corpus based on public domain audio books,.

Should Audio Front-ends be Adaptive? Comparing Learnable and Adaptive Front-ends Librispeech: an asr corpus based on public domain audio books,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:26:14.173769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-09T05:26:14.104367Z digest=sha256:715ac1b74c5e2d4c8f61680b623a21dc142d942019d440734637352e5409720a

Pith citing papers

No inbound Pith citation observations are available.