Pith. sign in

Paper Citation Record · LEDGER

WaveNet: A Generative Model for Raw Audio

As of 12 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 100 inbound Pith citation observations for arXiv:1609.03499.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1609.03499 v2

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-12T20:27:40.037176Z

measured 160 of 160 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 100 of 206 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:11:21.068085Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

60 of 60 outbound references displayed

  • verified exact8
  • verified fuzzy50
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3618
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation ccefc5c3-fb0c-43fb-8496-8450d7ea5fa0 · outbound

This paper cites Vocaine the vocoder and applications is speech synthesis.

WaveNet: A Generative Model for Raw Audio Vocaine the vocoder and applications is speech synthesis

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.529223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:2074a6e3244d60c17a7a69be5db8927549ac7f1115ae46264a6d2ffd9d75c993

Observation 3c56338c-cffe-408f-a826-5f404fa066a3 · outbound

This paper cites Mixture density networks.

WaveNet: A Generative Model for Raw Audio Mixture density networks

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.568577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:2fab23d3060e54abf883a0cae9bb3953aa638e1ae1d35e0302eae3e036aba5c3

Observation dc8daafa-1b98-4800-9216-6d78cbb8b576 · outbound

This paper cites Semantic Image Segmentation with Deep Convolutional Nets and Fully Connected CRFs.

WaveNet: A Generative Model for Raw Audio Semantic Image Segmentation with Deep Convolutional Nets and Fully Connected CRFs

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-12T20:27:40.277427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:325d698546c10d1813ee887345b5f2e921405fe18e6ae4a2ddafcb487e85122e

Observation 02b36a99-1641-4fd4-b3b6-d62f93e44044 · outbound

This paper cites The Vowel: I ts Nature and Structure.

WaveNet: A Generative Model for Raw Audio The Vowel: I ts Nature and Structure

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.572368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:a4eb370cb4b187594aa3eee59a1c008475bc45fb0d9032453b2905da7610d6a0

Observation 498c99fe-ced4-4d83-84d2-6fe9c2c96b6e · outbound

This paper cites Remaking speech.

WaveNet: A Generative Model for Raw Audio Remaking speech

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.577411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:ccbbb919904ea8e2bca39d0ca7cbfda97bf047851da596b05b4a8614f2b8420b

Observation 2fc67396-4d7e-4b9d-9524-db62705edf79 · outbound

This paper cites An implementation of the ``algorithme \`a trous'' to compute the wavelet transform.

WaveNet: A Generative Model for Raw Audio An implementation of the ``algorithme \`a trous'' to compute the wavelet transform

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.581691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:d802e99371bb2a24ae5466c9c953e6063fbedcdfb8653c8ac2ca9ea9c82a1869

Observation 643d1882-4a33-4a49-b411-6422ba7d6ae1 · outbound

This paper cites TTS synthesis with bidirectional LSTM based recurrent neural networks.

WaveNet: A Generative Model for Raw Audio TTS synthesis with bidirectional LSTM based recurrent neural networks

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.585500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:794f408ec03782812735e9e068497f8b65d4f6d7d4d46500bf903f4fb7495ddb

Observation 9c68de14-19cf-4561-8c33-cd67cf6060f6 · outbound

This paper cites Acoustic Theory of Speech Production.

WaveNet: A Generative Model for Raw Audio Acoustic Theory of Speech Production

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.588978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:ac4a74c67a801ad183b94b2577527c126bc28cad57142398014e94cc16ae2694

Observation c4b46274-f20e-40f1-80c0-c7f089479b83 · outbound

This paper cites DARPA TIMIT acoustic-phonetic continuous speech corpus CD-ROM.

WaveNet: A Generative Model for Raw Audio DARPA TIMIT acoustic-phonetic continuous speech corpus CD-ROM

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.592620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:b0d22bd685dd5ad16b78f295ffd47c9d189ebb2e1022f9f476c38dd83cd7064c

Observation d622104a-1b9c-4530-9f7c-864007b61faf · outbound

This paper cites Recent advances in G oogle real-time HMM -driven unit selection synthesizer.

WaveNet: A Generative Model for Raw Audio Recent advances in G oogle real-time HMM -driven unit selection synthesizer

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.328737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:9af53840442bcbde4d62df488170c203fd6c4fcd5631e6910a2d18b581bf19b7

Observation 226da7d5-06e6-430c-88a3-21c71b16a057 · outbound

This paper cites Deep Residual Learning for Image Recognition.

WaveNet: A Generative Model for Raw Audio Deep Residual Learning for Image Recognition

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-12T20:27:40.287030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:7d623eb0731368f8c8d8d679d0870c181cc861ffd579fecde4264b715b2075be

Observation 5171066e-9614-4189-88c7-5fbebac38d83 · outbound

This paper cites and Schmidhuber, J.

WaveNet: A Generative Model for Raw Audio and Schmidhuber, J

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.332469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:a682baa0132bdc6fea9d0ef0bf157296a0b57df37274e4650facce6a4e47e3e4

Observation 201d3590-a625-4122-bf29-61d421c0c3c1 · outbound

This paper cites A real-time algorithm for signal analysis with the help of the wavelet transform.

WaveNet: A Generative Model for Raw Audio A real-time algorithm for signal analysis with the help of the wavelet transform

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.336603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:d946128dbc1307da3c302c689ecad7928994e44ab9020e5e991c9085b765ac8b

Observation 3ccc7158-8ce0-43ac-b79b-2679eb8c3339 · outbound

This paper cites Speech acoustic modeling from raw multichannel waveforms.

WaveNet: A Generative Model for Raw Audio Speech acoustic modeling from raw multichannel waveforms

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.340844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:23d2291456d2a2efeb43f29fcc8b34972645f923529ca2456a7d983abbeaa9d5

Observation c352245f-e9d2-4cb9-a3bb-4d686c0b9920 · outbound

This paper cites and Black, Alan W.

WaveNet: A Generative Model for Raw Audio and Black, Alan W

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.344793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:40662f48b6c4bc41860027883ed4547f258f842d239a52d67936f0883de3c37d

Observation aaf80b9f-78aa-4e82-a437-55653452b15e · outbound

This paper cites Unbiased estimation of log spectrum.

WaveNet: A Generative Model for Raw Audio Unbiased estimation of log spectrum

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.348377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:a9ec1bb1254d97d7040aff67bd82635ee3e02601b86752d062126b9d50812ec0

Observation 50cd9dbe-3209-4c4e-84fa-7a364ff9a93e · outbound

This paper cites Line spectrum representation of linear predictor coefficients of speech signals.

WaveNet: A Generative Model for Raw Audio Line spectrum representation of linear predictor coefficients of speech signals

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.352332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:4801766fbecc945ec85b348c6f236f884c19eebc1c9ab5e43327effcf6552f59

Observation 8da411f8-8e2e-458b-8174-3f81df1bde1f · outbound

This paper cites A statistical method for estimation of speech spectral density and formant frequencies.

WaveNet: A Generative Model for Raw Audio A statistical method for estimation of speech spectral density and formant frequencies

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.356071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:06ff6a28b08009f5b1351e7e9033c0b831d297539c704f9908d21dbb4a8653e6

Observation ea312d8b-ae77-4880-8dd4-af2651bad9f0 · outbound

This paper cites Recommendation G.

WaveNet: A Generative Model for Raw Audio Recommendation G

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.359637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:a28b6ce9e994d3fe4c41b26b9c7df3068c6fdd549bd82ff14005fd24eb2eddb8

Observation 9bc77de7-7299-4d59-bb6c-3dac25075cb0 · outbound

This paper cites Exploring the Limits of Language Modeling.

WaveNet: A Generative Model for Raw Audio Exploring the Limits of Language Modeling

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-12T20:27:40.300459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:1c790319a3bc5f7518a04cc225684f015006a69f88a50a64da8171de1e7656dd

Observation 62c70682-612b-4e25-bbfd-5c735eb3a0e3 · outbound

This paper cites Mixture autoregressive hidden M arkov models for speech signals.

WaveNet: A Generative Model for Raw Audio Mixture autoregressive hidden M arkov models for speech signals

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.363274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:f1a3a967446248c45a82593d84eaff684d41efd0c12e057c0eefb82e00c19e59

Observation 4c923379-4a2e-4f7a-ba3e-9e63c9038eff · outbound

This paper cites Speech analysis with multi-kernel linear prediction.

WaveNet: A Generative Model for Raw Audio Speech analysis with multi-kernel linear prediction

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.367662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:bf667206de6256db16facd1eb1b69122c6d1d338e9952c559fcad20a898074b4

Observation d85ff923-4e06-48e9-9c83-6eddd4ac7510 · outbound

This paper cites Text-to-speech conversion with neural networks: A recurrent TDNN approach.

WaveNet: A Generative Model for Raw Audio Text-to-speech conversion with neural networks: A recurrent TDNN approach

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.371821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:7607a7b255f128f1e42bc9e0562546ca50043f7aa3db6635e7230a4565b5cd93

Observation 21596be3-1044-49eb-a5fb-4c005305be91 · outbound

This paper cites an unresolved cited work.

WaveNet: A Generative Model for Raw Audio Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-05-12T20:27:40.388742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:9f6dca0980c0bff0f4a146fd374b14bce65d2bd32f0e6215930da226c982d61e

Observation d001d266-47cb-40d1-9255-6145475d4f9a · outbound

This paper cites Aperiodicity extraction and control using mixed mode excitation and group delay manipulation for a high quality speech analysis, modification and synthesis system STRAIGHT.

WaveNet: A Generative Model for Raw Audio Aperiodicity extraction and control using mixed mode excitation and group delay manipulation for a high quality speech analysis, modification and synthesis system STRAIGHT

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.395522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:220704cc575eca8b2c65187579afb6e0ad5c5b25d816163cffdcd7f58372436a

Observation 0b58d73e-ac8a-4e94-b8ca-3d56b6d8f6d5 · outbound

This paper cites Input-agreement: a new mechanism for collecting data using human computation games.

WaveNet: A Generative Model for Raw Audio Input-agreement: a new mechanism for collecting data using human computation games

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.400631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:163964228adfe7785027f8e7dec0a2e83dc005bdfb68952a07a2a630bbf6b8cb

Observation ee4b5b70-cc5c-41a8-80b9-b31e69d981b9 · outbound

This paper cites an unresolved cited work.

WaveNet: A Generative Model for Raw Audio Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-05-12T20:27:40.405003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:b5849ad5386970141d394baf91332f588fe8ca3a5eb451f8b138aff2628cdfe4

Observation 96c927c7-7f1f-40c7-9829-f8016bafe341 · outbound

This paper cites WORLD : A vocoder-based high-quality speech synthesis system for real-time applications.

WaveNet: A Generative Model for Raw Audio WORLD : A vocoder-based high-quality speech synthesis system for real-time applications

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.412000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:0bc69ac39cda3b490465ed6c856d88e1b7a5a09296cf78522b2b1a9ff4087919

Observation 7d8b97b0-d481-45c2-84dc-75948cd482c4 · outbound

This paper cites Pitch synchronous waveform processing techniques for text-to-speech synthesis using diphones.

WaveNet: A Generative Model for Raw Audio Pitch synchronous waveform processing techniques for text-to-speech synthesis using diphones

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.417385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:1bb4b77097587e73e49d54bcf1bd79c2b38c1e0eea6964b955bd38c90cf48ca5

Observation 91e7108d-2fad-4adc-bfa1-baa792248f1e · outbound

This paper cites A Deep Learning Approach to Data-driven Parameterizations for Statistical Parametric Speech Synthesis.

WaveNet: A Generative Model for Raw Audio A Deep Learning Approach to Data-driven Parameterizations for Statistical Parametric Speech Synthesis

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:05:24.719892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:33f82e38ca087e4cdbc60b417311da917f7194e4c94686f1449ec6ff4b02e424

Observation 83aa7022-3b3e-4c89-b04f-e7a1bdb4c19b · outbound

This paper cites Rectified linear units improve restricted B oltzmann machines.

WaveNet: A Generative Model for Raw Audio Rectified linear units improve restricted B oltzmann machines

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.423725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:9609fe09a11263d8d14ba1838249359a71c025bd1a91f617447ec9ffded49237

Observation bca85b06-c248-4720-9ef9-4d141186dbc6 · outbound

This paper cites Integration of spectral feature extraction and modeling for HMM -based speech synthesis.

WaveNet: A Generative Model for Raw Audio Integration of spectral feature extraction and modeling for HMM -based speech synthesis

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.427963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:9b4175f53660e745031f1a4b108b057f8289e8ec85bfe52093f49702bc3c27d3

Observation 87ed2107-5a57-4980-9a1e-712c2e35e4f2 · outbound

This paper cites Estimating phoneme class conditional probabilities from raw speech signal using convolutional neural networks.

WaveNet: A Generative Model for Raw Audio Estimating phoneme class conditional probabilities from raw speech signal using convolutional neural networks

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.432935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:ec7beedb7264f64fb8b0080a1b4c1cfd444d1c9ec5693a1d31d5fb654118e065

Observation e3fcc7c3-4ecf-42cf-9c30-e8e4b92475e6 · outbound

This paper cites Nonlinear filter design: methodologies and challenges.

WaveNet: A Generative Model for Raw Audio Nonlinear filter design: methodologies and challenges

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.438918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:2756ffc154e6d74df400056fc694a2976e107a1f6d535757550e7fb23d94ecbb

Observation 5efe7188-309d-4996-9edc-2d8f852c13b8 · outbound

This paper cites Linear predictive hidden M arkov models and the speech signal.

WaveNet: A Generative Model for Raw Audio Linear predictive hidden M arkov models and the speech signal

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.443179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:0752d95d54a462723f3ff21e9b87d2d531621d67a84e4aecf51ed4fdad30ef4e

Observation e3873c6b-47ba-4fc1-9ee2-9e82b8697182 · outbound

This paper cites Fundamentals of Speech Recognition.

WaveNet: A Generative Model for Raw Audio Fundamentals of Speech Recognition

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.447656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:a856b7229011dc71cc2b1445ec135d51fb9b511b308a8295a0949a9e794b3667

Observation 8c7697ef-9ac9-4d3c-8d1f-152a64c1a7b2 · outbound

This paper cites ATR -talk speech synthesis system.

WaveNet: A Generative Model for Raw Audio ATR -talk speech synthesis system

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.451995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:fc7d99ea2675d5a96557ab5f09695adbe721726ebda03a866b19d3e7f4093515

Observation 95f61bbd-5778-42f8-83e0-2ec8e5ec666e · outbound

This paper cites Learning the speech front-end with raw waveform CLDNN s.

WaveNet: A Generative Model for Raw Audio Learning the speech front-end with raw waveform CLDNN s

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.458834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:476d77b829c93a2956d9afe7e6d052a57e7d4afb3568e052dffd95f67c8cb752

Observation a9b5c17e-e536-4839-bf6c-04c3c29fdeed · outbound

This paper cites A deep auto-encoder based low-dimensional feature extraction from FFT spectral envelopes for statistical parametric speech synthesis.

WaveNet: A Generative Model for Raw Audio A deep auto-encoder based low-dimensional feature extraction from FFT spectral envelopes for statistical parametric speech synthesis

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.463864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:0ded5e88fdcee686bdf92a856f54215f9e33bee0058818baad06e9c0df6d08f4

Observation 486b95c6-502f-42f8-aa93-2a8625e9b21e · outbound

This paper cites Postfilters to modify the modulation spectrum for statistical parametric speech synthesis.

WaveNet: A Generative Model for Raw Audio Postfilters to modify the modulation spectrum for statistical parametric speech synthesis

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.469229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:37f597a2c5ae8946cdcdf4ba865adc6acd7de38d69e39cb27374b4e3130ede48

Observation fa7c4ea9-594c-4f17-bc98-3723de51df3a · outbound

This paper cites Generative image modeling using spatial LSTM s.

WaveNet: A Generative Model for Raw Audio Generative image modeling using spatial LSTM s

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.475340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:3186c692cfacc12bef30c12cfa67663d5d6c877adf72afe2c62256b691040d42

Observation add153d4-5afd-4794-aa86-9458040d9d89 · outbound

This paper cites A speech parameter generation algorithm considering global variance for HMM -based speech synthesis.

WaveNet: A Generative Model for Raw Audio A speech parameter generation algorithm considering global variance for HMM -based speech synthesis

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.480660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:5e2ce0fba0fdef26d91ea04e1b1ec8b88fcf56bc11edb989d07b6b950266ccef

Observation b4005e76-e90a-42ea-8582-a70ea1a12990 · outbound

This paper cites Statistical approach to vocal tract transfer function estimation based on factor analyzed trajectory hmm.

WaveNet: A Generative Model for Raw Audio Statistical approach to vocal tract transfer function estimation based on factor analyzed trajectory hmm

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.485186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:710645d346abd71aaffb5922bc5554312ee54f27cf1e8be306cc40583dd3356a

Observation 60332101-a50c-44f0-b7c5-06452af4cba8 · outbound

This paper cites Speech synthesis as a statistical machine learning problem.

WaveNet: A Generative Model for Raw Audio Speech synthesis as a statistical machine learning problem

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.489858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:d408116202c037d301db2b889512afa980f07bdda60332cb1a51acf33d192fd4

Observation 63efedd9-f713-4b2e-8041-bfe5601b674b · outbound

This paper cites Directly modeling speech waveforms by neural networks for statistical parametric speech synthesis.

WaveNet: A Generative Model for Raw Audio Directly modeling speech waveforms by neural networks for statistical parametric speech synthesis

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.498494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:122c65b71275cccebaec7a2833cd79df3b0d090fa4e8da14d2e5e92cbba7e4ca

Observation da9bd224-14f4-4c02-8f1d-75a675aab7d8 · outbound

This paper cites Directly modeling voiced and unvoiced components in speech waveforms by neural networks.

WaveNet: A Generative Model for Raw Audio Directly modeling voiced and unvoiced components in speech waveforms by neural networks

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.504745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:fd057c86e2ff41fe682b376dc7c7d8d06e2c6a5d48794e95b0de7760a2f31578

Observation db26ac27-4204-4632-a6e5-4cd825954355 · outbound

This paper cites Speech synthesis using artificial neural networks trained on cepstral coefficients.

WaveNet: A Generative Model for Raw Audio Speech synthesis using artificial neural networks trained on cepstral coefficients

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.509602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:1d6822dd6cf88f5408e1a19764624b1634598008a2335f6ab0627ede982cadc5

Observation 93b1e614-7c5c-416a-bad3-f92f2f494149 · outbound

This paper cites u ske, Zolt \'a n, Golik, Pavel, Schl \.

WaveNet: A Generative Model for Raw Audio u ske, Zolt \'a n, Golik, Pavel, Schl \

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.514165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:9ccba4c38d0b68b7f9d19152429802db16bc4a11ff1833ac9e072a2b04b1adf5

Observation f8a8b8af-23c0-4b67-b5e5-d5b01a053c4f · outbound

This paper cites Modelling acoustic feature dependencies with artificial neural networks: T rajectory- RNADE.

WaveNet: A Generative Model for Raw Audio Modelling acoustic feature dependencies with artificial neural networks: T rajectory- RNADE

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.518783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:a16eb493aa0684aa4c72efc17d6d5b5ea0f71ce8254c9b5b5e9f42ed47105f91

Observation 0cd7fa5f-9733-41b2-a61b-b96e0d9953ce · outbound

This paper cites Pixel Recurrent Neural Networks.

WaveNet: A Generative Model for Raw Audio Pixel Recurrent Neural Networks

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-12T20:27:40.306428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:da29202e6a68eb882085988876a82b4b274350f8a5096ceba8b4facabbcf8615

Observation 3f83e73b-f2fd-4783-a030-2fb4d12f8f8b · outbound

This paper cites Conditional Image Generation with PixelCNN Decoders.

WaveNet: A Generative Model for Raw Audio Conditional Image Generation with PixelCNN Decoders

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-12T20:27:40.313149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:00cc26a88ce51343fbb49d175aeded8b58a6e098f178a27efa41f1808e4beca7

Observation 1d714af5-cea9-4927-ae49-26bea11f1cbb · outbound

This paper cites Minimum generation error training with direct log spectral distortion on LSP s for HMM -based speech synthesis.

WaveNet: A Generative Model for Raw Audio Minimum generation error training with direct log spectral distortion on LSP s for HMM -based speech synthesis

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.523722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:84011157eba5d1c72425fd2076eb45a9ed0d52832a50793e4762dc5e75967823

Observation 8147b9a6-db1a-4ae6-bc88-8ddcaad6d7c5 · outbound

This paper cites English multi-speaker corpus for CSTR voice cloning toolkit.

WaveNet: A Generative Model for Raw Audio English multi-speaker corpus for CSTR voice cloning toolkit

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.558962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:06b36d5363dbacbe9670b8e019b6aecd8e5d549451c8f897ef018f3e844a35c6

Observation 22565bf5-63d0-4868-a3b4-65aaeca6a1c9 · outbound

This paper cites Simultaneous modeling of phonetic and prosodic parameters, and characteristic conversion for HMM -based text-to-speech systems.

WaveNet: A Generative Model for Raw Audio Simultaneous modeling of phonetic and prosodic parameters, and characteristic conversion for HMM -based text-to-speech systems

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.533877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:ccedd99e7e92e643bf5ac7e64fe9ad04cc9731b8116c00efc7bde3b36968ae1e

Observation 3507755f-6d5e-4013-8277-6cd722f7b82a · outbound

This paper cites Multi-Scale Context Aggregation by Dilated Convolutions.

WaveNet: A Generative Model for Raw Audio Multi-Scale Context Aggregation by Dilated Convolutions

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-12T20:27:40.293972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:97503debbc1901d98a9c2a7f1d54a4ca948291572e9fba7244759e101b63bb3e

Observation cb5651c2-f9cd-45b5-b062-092173ec0747 · outbound

This paper cites An example of context-dependent label format for HMM -based speech synthesis in E nglish.

WaveNet: A Generative Model for Raw Audio An example of context-dependent label format for HMM -based speech synthesis in E nglish

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.540543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:e58edaeb2f38cb24bde42ca3fabc431ab77d52942c34ec01fa573534ff4027f7

Observation 0aa55e31-6921-45c8-b5ec-96a39fe057e1 · outbound

This paper cites Reformulating the HMM as a trajectory model by imposing explicit relationships between static and dynamic features.

WaveNet: A Generative Model for Raw Audio Reformulating the HMM as a trajectory model by imposing explicit relationships between static and dynamic features

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.546294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:e110fd2a2e26e720b97387c43b81812d5523ca6e2b9a96177d36a8433e1a677a

Observation caf8afbf-f03e-4021-8c25-d59b29a647df · outbound

This paper cites Statistical parametric speech synthesis.

WaveNet: A Generative Model for Raw Audio Statistical parametric speech synthesis

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.551394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:2962dc16a19c91c2d92185e020a7706f7d50b7d9da9a5ac4240474a045116d37

Observation 67f747b9-d820-43c8-8a27-bc9d887b755a · outbound

This paper cites Statistical parametric speech synthesis using deep neural networks.

WaveNet: A Generative Model for Raw Audio Statistical parametric speech synthesis using deep neural networks

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T20:27:40.555537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:ee54603b50388907892d66014039e16b6e2de0e9d1830c80547444610a483ed9

Observation bb484fef-6e4a-44c2-b171-017b7642ac7a · outbound

This paper cites Fast, Compact, and High Quality LSTM-RNN Based Statistical Parametric Speech Synthesizers for Mobile Devices.

WaveNet: A Generative Model for Raw Audio Fast, Compact, and High Quality LSTM-RNN Based Statistical Parametric Speech Synthesizers for Mobile Devices

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-07-04T21:05:47.149369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T20:27:40.037176Z digest=sha256:63669c8240351899fc743eec910eb5ab359f31dde9c2b96522a79b7810f97199

Pith citing papers

Observation 25396cc0-5d9b-40b7-93f6-0fc15a6a6744 · inbound

Progressive Growing of GANs for Improved Quality, Stability, and Variation cites this paper.

Progressive Growing of GANs for Improved Quality, Stability, and Variation WaveNet: A Generative Model for Raw Audio

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T20:27:40.593869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T12:24:03.252331Z digest=sha256:cb7573a843eea3daabf7d34a1d0f98f30991943f0e25c92313677887d31418d1

Observation a5d6c55c-9686-4f81-b3a4-84b1cb2e4f05 · inbound

Generating Long Sequences with Sparse Transformers cites this paper.

Generating Long Sequences with Sparse Transformers WaveNet: A Generative Model for Raw Audio

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-12T20:27:40.593869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T19:50:53.385943Z digest=sha256:924bf4104a344ad6bcec2302faaf3c05d04fd9caaada76d0b125e4fb3c4e32f7

Observation 765ff3db-c420-4625-9767-a57b77ac1f35 · inbound

Singing Voice Synthesis Using Deep Autoregressive Neural Networks for Acoustic Modeling cites this paper.

Singing Voice Synthesis Using Deep Autoregressive Neural Networks for Acoustic Modeling WaveNet: A Generative Model for Raw Audio

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-25T18:41:08.301926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-25T18:37:53.990536Z digest=sha256:502719f216136054b9e822b364bf38d03d7a157092e33d6cd11452a5b371d974

Observation 2ac6a384-dc47-4dac-a84d-53581d0442d0 · inbound

End-to-End Emotional Speech Synthesis Using Style Tokens and Semi-Supervised Training cites this paper.

End-to-End Emotional Speech Synthesis Using Style Tokens and Semi-Supervised Training WaveNet: A Generative Model for Raw Audio

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-25T15:35:58.945560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-25T15:33:09.748398Z digest=sha256:88d2f2f208387bb2cf3bdb3279aa5e09733a979de7846b0fa156341839ea601b

Observation ab9e09ca-474c-4142-8ad3-92023ba43f4c · inbound

RUSLAN: Russian Spoken Language Corpus for Speech Synthesis cites this paper.

RUSLAN: Russian Spoken Language Corpus for Speech Synthesis WaveNet: A Generative Model for Raw Audio

Reference 22

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T15:15:58.489413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-25T15:15:38.672352Z digest=sha256:280d5bf75e0bafe5130ef92eaa11be698bef5ed82808f430e34ab3fdb0183713

Observation 7e0f6a8d-4a40-4ffa-9194-43d57ccbfbc8 · inbound

Analysis by Adversarial Synthesis -- A Novel Approach for Speech Vocoding cites this paper.

Analysis by Adversarial Synthesis -- A Novel Approach for Speech Vocoding WaveNet: A Generative Model for Raw Audio

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-25T11:35:43.705807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-25T11:31:31.494334Z digest=sha256:632980c8ea3c364d976d7b6627b6f1aefdd6a775c3fda092587ed1e88f4afef7

Observation 0506599a-285c-44cd-8699-4f7dd3c54f97 · inbound

Multitasking with Alexa Multitasking with Alexa: How Using Intelligent Personal Assistants Impacts Language-based Primary Task Performance cites this paper.

Multitasking with Alexa Multitasking with Alexa: How Using Intelligent Personal Assistants Impacts Language-based Primary Task Performance WaveNet: A Generative Model for Raw Audio

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-05-25T09:56:51.305971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-25T09:56:13.087146Z digest=sha256:0ff09bcd92742be537acda4a47836b6471626d20043863069083359e6629af82

Observation dad61d25-0bb5-4b23-9366-265181562df0 · inbound

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning cites this paper.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning WaveNet: A Generative Model for Raw Audio

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-25T00:10:07.041425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:77fc8f6b66f17aaad793c192fe5ebf57616528b563031e9ddcd12512b7e85e16

Observation 78ff8076-a283-434b-948e-315857d6ccb0 · inbound

Autoencoding sensory substitution cites this paper.

Autoencoding sensory substitution WaveNet: A Generative Model for Raw Audio

Reference 127

Resolution
verified exact
local_arxiv, observed 2026-05-24T21:24:57.830206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T21:24:11.898508Z digest=sha256:7354a34834fdd592374abc7a171ccd471a614cc5e06baaac03e0250cd6f2f2f1

Observation b6513f91-188c-4fea-91fe-4fe3e210dd58 · inbound

Hierarchical Sequence to Sequence Voice Conversion with Limited Data cites this paper.

Hierarchical Sequence to Sequence Voice Conversion with Limited Data WaveNet: A Generative Model for Raw Audio

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-05-24T21:34:58.851967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T21:30:16.923017Z digest=sha256:8ad270a8072accde01f671107bfbe7aa4cd0338b050a2973517e7360bf947df6

Observation 7cbfa735-d2ad-494c-a29c-69b7ec9d4ec4 · inbound

DNN-based Speaker Embedding Using Subjective Inter-speaker Similarity for Multi-speaker Modeling in Speech Synthesis cites this paper.

DNN-based Speaker Embedding Using Subjective Inter-speaker Similarity for Multi-speaker Modeling in Speech Synthesis WaveNet: A Generative Model for Raw Audio

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-24T19:06:18.897869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T19:05:32.567355Z digest=sha256:a7c69d37ca9f34a59ec8591f8a419e7349c54100a448ec1fe669672ed290f4c8

Observation 47f3f8a2-142b-4b20-a7d9-40c1996d44d2 · inbound

Forward-Backward Decoding for Regularizing End-to-End TTS cites this paper.

Forward-Backward Decoding for Regularizing End-to-End TTS WaveNet: A Generative Model for Raw Audio

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-24T19:29:50.889655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T19:27:54.013474Z digest=sha256:49bb1a9cd06a7d3b5531b3eb2b26e492753d0e0a9c7e1973e63882cc523fdc64

Observation bd8dd9f0-b5fc-4035-98ab-d92bf4eef222 · inbound

Non-Parallel Voice Conversion with Cyclic Variational Autoencoder cites this paper.

Non-Parallel Voice Conversion with Cyclic Variational Autoencoder WaveNet: A Generative Model for Raw Audio

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-24T17:06:16.923305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T17:05:25.448346Z digest=sha256:fa9cda358b3b322b6c0b1b89ef1081a0876aa8667a9c9bd62920fddea9a6d315

Observation 870e7831-3113-4507-a274-02196d43553c · inbound

Compressive Transformers for Long-Range Sequence Modelling cites this paper.

Compressive Transformers for Long-Range Sequence Modelling WaveNet: A Generative Model for Raw Audio

Reference 139

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.671855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:bb235e7681d3b581ba5b9e161d4b2fe3138c57a4a9c96e8a4372d0ad458e23be

Observation d6f94607-a95a-4f29-8a02-543ce382d26d · inbound

Jukebox: A Generative Model for Music cites this paper.

Jukebox: A Generative Model for Music WaveNet: A Generative Model for Raw Audio

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-24T15:19:37.741156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T15:19:06.065102Z digest=sha256:31d7b4051bbbab66abeb7aefd3cadc4e81a517c5f551a838a6cc327532e4b47d

Observation 550dc318-e255-422c-a6fa-42b435e7481f · inbound

Denoising Diffusion Probabilistic Models cites this paper.

Denoising Diffusion Probabilistic Models WaveNet: A Generative Model for Raw Audio

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-12T20:27:40.593869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-11T03:24:24.752572Z digest=sha256:1de36224609ff4347b3d3ee629c9dfe8a4662edc6944e1b84c01af12266d9fee

Observation f1b3c8e9-3841-4fb8-a2be-458d386304f5 · inbound

GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding cites this paper.

GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding WaveNet: A Generative Model for Raw Audio

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-12T20:27:40.593869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-11T02:26:44.624137Z digest=sha256:93500fcdfbf94975898574f33afed0a52a008c63095c0d6861d0176a2c0375a9

Observation c7def291-5593-49de-b441-c9751af99b70 · inbound

DiffWave: A Versatile Diffusion Model for Audio Synthesis cites this paper.

DiffWave: A Versatile Diffusion Model for Audio Synthesis WaveNet: A Generative Model for Raw Audio

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-15T13:13:37.129089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-15T13:13:37.085932Z digest=sha256:379066ce70d4f94072a992e6f0f5f2a0a16d564d9c07f6295b7db9e3180c4f08

Observation 5dd56de1-6972-4ec6-8a42-0e1ab7cb09a2 · inbound

Denoising Diffusion Implicit Models cites this paper.

Denoising Diffusion Implicit Models WaveNet: A Generative Model for Raw Audio

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-24T14:44:36.954327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T14:41:23.935708Z digest=sha256:813a386c8665ff431e24dfc3f8fa398c21f8e1caaa0131deffd29a86d0247242

Observation 14096d3c-3752-4525-96e0-66efeb690dc3 · inbound

VideoGPT: Video Generation using VQ-VAE and Transformers cites this paper.

VideoGPT: Video Generation using VQ-VAE and Transformers WaveNet: A Generative Model for Raw Audio

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-13T17:24:33.845712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-13T17:24:33.725187Z digest=sha256:b023fffd317498de952c73ba46c4f28999242d03e8daf023f8745eb611ecf5f1

Observation c555b75f-c5a6-4e17-b92f-9198820fa97b · inbound

Geometric Deep Learning: Grids, Groups, Graphs, Geodesics, and Gauges cites this paper.

Geometric Deep Learning: Grids, Groups, Graphs, Geodesics, and Gauges WaveNet: A Generative Model for Raw Audio

Reference 95

Resolution
verified exact
local_arxiv, observed 2026-05-13T02:39:30.003876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-13T02:39:29.411021Z digest=sha256:23db10310c1dc2b99e71d617c44207ccada5745ccb06fd72460c187bae003d02

Observation d2b9af6c-f04e-499c-9620-f1bbc75e6ac4 · inbound

Diffusion Models Beat GANs on Image Synthesis cites this paper.

Diffusion Models Beat GANs on Image Synthesis WaveNet: A Generative Model for Raw Audio

Reference 64

Resolution
verified exact
local_arxiv, observed 2026-05-13T11:16:28.487571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-13T11:16:28.445702Z digest=sha256:372c05142ed35a73bf13dc5bcb1fe0c111c3f74c0872bff07e22f713c3dfa58f

Observation 5a14c320-3519-4396-a963-80c10ca18787 · inbound

Efficiently Modeling Long Sequences with Structured State Spaces cites this paper.

Efficiently Modeling Long Sequences with Structured State Spaces WaveNet: A Generative Model for Raw Audio

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-12T20:27:40.593869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-11T10:41:55.618357Z digest=sha256:d8726e61b4cccb0bb6eb6382f4cefcfa678fda52e3f74704b3c9309377c10495

Observation 26a18e11-b72b-4575-b765-b819bc20598a · inbound

Text and Code Embeddings by Contrastive Pre-Training cites this paper.

Text and Code Embeddings by Contrastive Pre-Training WaveNet: A Generative Model for Raw Audio

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-15T19:24:12.083924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-15T19:24:11.907204Z digest=sha256:4844e2895d72fed91979d86bb2540bd448b09c5a95d8d9d5326b64c1a41679ae

Observation 2fdfeb2e-804c-4c68-a0b5-45caef47e4b8 · inbound

A Generalist Agent cites this paper.

A Generalist Agent WaveNet: A Generative Model for Raw Audio

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-13T06:24:49.935627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-13T06:24:49.833638Z digest=sha256:891e84a76b0a4e41aa098d55b0e92d37a671d2bf7b35417624309807e3355c85

Observation 27868269-23e0-4e14-a36a-82fe74e2d2b9 · inbound

Simplified State Space Layers for Sequence Modeling cites this paper.

Simplified State Space Layers for Sequence Modeling WaveNet: A Generative Model for Raw Audio

Reference 133

Resolution
verified exact
local_arxiv, observed 2026-05-16T08:16:11.516486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-16T08:16:11.406774Z digest=sha256:797d3b79a1cd73e0d782d0ea1cc9bde9d257b620404e0619dcd48539b4069fe6

Observation 6433f12a-e595-4a38-a021-ab6f1bb82b4d · inbound

High Fidelity Neural Audio Compression cites this paper.

High Fidelity Neural Audio Compression WaveNet: A Generative Model for Raw Audio

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-13T21:49:52.247674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-13T21:49:52.184932Z digest=sha256:a256a87ef6441c4e3aead6511d1aeb7afac3ba40cf34bc03fd6f7ac4fcb15d96

Observation 3b6fac2a-65bf-4c59-a0bd-0e5e75d09b4f · inbound

Is Conditional Generative Modeling all you need for Decision-Making? cites this paper.

Is Conditional Generative Modeling all you need for Decision-Making? WaveNet: A Generative Model for Raw Audio

Reference 194

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T15:35:10.985464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-15T15:35:10.593969Z digest=sha256:cd4bd59474f1d1c7da63bff44a70cf4e282e71544211f3e4fe3defac14b1970f

Observation f9673700-c241-4c4b-9615-23b9d4d2fde7 · inbound

A decoder-only foundation model for time-series forecasting cites this paper.

A decoder-only foundation model for time-series forecasting WaveNet: A Generative Model for Raw Audio

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-16T18:07:21.278674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-16T18:07:21.246053Z digest=sha256:43a6d81903335875fba2def2071dd3ae7936536bf9bb062e964b1113c0264f7b

Observation 81570b34-cd30-4dd0-bbd5-82a9f85e4768 · inbound

Mamba: Linear-Time Sequence Modeling with Selective State Spaces cites this paper.

Mamba: Linear-Time Sequence Modeling with Selective State Spaces WaveNet: A Generative Model for Raw Audio

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-12T20:27:40.593869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T11:53:06.570968Z digest=sha256:28027dec5eac6ea87f9e3841da7b66e45bd2b1ab809c8d9cff8f8dab8a3c23ba

Observation 90c7731a-54c6-40fe-b20d-fcd8dadc4053 · inbound

Massive Activations in Large Language Models cites this paper.

Massive Activations in Large Language Models WaveNet: A Generative Model for Raw Audio

Reference 85

Resolution
verified exact
local_arxiv, observed 2026-05-16T07:02:53.919878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-16T07:02:53.740597Z digest=sha256:c88178f0c099d5da70d456f70fc9c32015d1d543708731fb47190a8b13472b48

Observation b00f98e0-9052-4415-8ce6-a6b428035075 · inbound

Chronos: Learning the Language of Time Series cites this paper.

Chronos: Learning the Language of Time Series WaveNet: A Generative Model for Raw Audio

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-05-13T08:27:23.386102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-13T08:27:23.298009Z digest=sha256:bfb0b0e9b8f0334bbf42dcf33343857cd2e73918c2bc63b19429ba94aeb16ef2

Observation 1ac39f3f-ebad-4f45-b7e0-d7f259c72f03 · inbound

Deep Time Series Models: A Comprehensive Survey and Benchmark cites this paper.

Deep Time Series Models: A Comprehensive Survey and Benchmark WaveNet: A Generative Model for Raw Audio

Reference 130

Resolution
verified exact
local_arxiv, observed 2026-05-23T23:05:51.461510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-23T23:03:45.096751Z digest=sha256:3879f62105613747b95121094bb4623d18c8e05acc19119664bc1dd870938e55

Observation d8f9c0b7-475c-47ed-b53e-a23d8e08f1bd · inbound

Moshi: a speech-text foundation model for real-time dialogue cites this paper.

Moshi: a speech-text foundation model for real-time dialogue WaveNet: A Generative Model for Raw Audio

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-12T20:27:40.593869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-12T08:13:21.962488Z digest=sha256:57aab9dad6ced7633a0a6eedb2aa51c49908c3e7832adb114945b2f7ca5f532e

Observation 23cf175a-bce8-4654-9620-e39ae0ca35d3 · inbound

Listening for Expert Identified Linguistic Features: Assessment of Audio Deepfake Discernment among Undergraduate Students cites this paper.

Listening for Expert Identified Linguistic Features: Assessment of Audio Deepfake Discernment among Undergraduate Students WaveNet: A Generative Model for Raw Audio

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T15:11:21.068085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:11:21.068085Z digest=sha256:c5e63361dbff76402db7a728ec2cd3ac5cb3e42bd998bec516e169a273da3345

Observation 0b87d536-ad02-4d34-8103-156ec6d0bcc2 · inbound

VQalAttent: a Transparent Speech Generation Pipeline based on Transformer-learned VQ-VAE Latent Space cites this paper.

VQalAttent: a Transparent Speech Generation Pipeline based on Transformer-learned VQ-VAE Latent Space WaveNet: A Generative Model for Raw Audio

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T15:07:25.275090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:07:25.275090Z digest=sha256:ae7b3828030fd85517883141969a3c2f0f95d731bd71528bdad5ab9d7e3a2059

Observation 44302862-b9cc-440a-890c-44b4f516a23f · inbound

Nd-BiMamba2: A Unified Bidirectional Architecture for Multi-Dimensional Data Processing cites this paper.

Nd-BiMamba2: A Unified Bidirectional Architecture for Multi-Dimensional Data Processing WaveNet: A Generative Model for Raw Audio

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T14:25:18.886889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:25:18.886889Z digest=sha256:cebca20fde79606365818b464de84dd46a527e080602ce7c62bc91028f5525d6

Observation bb796e25-1903-400b-a62b-40a5ddf7ad50 · inbound

Scaling Transformers for Low-Bitrate High-Quality Speech Coding cites this paper.

Scaling Transformers for Low-Bitrate High-Quality Speech Coding WaveNet: A Generative Model for Raw Audio

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T05:56:42.839450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T05:56:42.839450Z digest=sha256:05b49cd943ed3e534f9af5d3fed479270f8b8d8dcb4cb3df36086a8ad5cb5639

Observation c4197214-845f-47ba-98db-b03e5b82f365 · inbound

Input-Output Optics as a Causal Time Series Mapping: A Generative Machine Learning Solution cites this paper.

Input-Output Optics as a Causal Time Series Mapping: A Generative Machine Learning Solution WaveNet: A Generative Model for Raw Audio

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:47.673947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:47.673947Z digest=sha256:904b7875bdca393c35ba2d5bf8fa27fb70b8a0cbe37ce1d836966ffa0097bc0b

Observation de80e184-caaf-4486-adbd-cf29eb9c4743 · inbound

Deep Learning-Based Approach for Identification and Compensation of Nonlinear Distortions in Parametric Array Loudspeakers cites this paper.

Deep Learning-Based Approach for Identification and Compensation of Nonlinear Distortions in Parametric Array Loudspeakers WaveNet: A Generative Model for Raw Audio

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T04:47:48.563368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:47:48.563368Z digest=sha256:9ed2abaada167563e28828dc355547a8cdf136847811d10db2eb32dde21d8260

Observation 80e11592-8b73-4872-b0aa-bd56c304a878 · inbound

Schedule On the Fly: Diffusion Time Prediction for Faster and Better Image Generation cites this paper.

Schedule On the Fly: Diffusion Time Prediction for Faster and Better Image Generation WaveNet: A Generative Model for Raw Audio

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T04:37:35.007035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:37:35.007035Z digest=sha256:08f18368d2b7c64b5e368dc608e1e5e71f93aeb0edff55518f056ac0964b615a

Observation 592b121f-6d10-48d0-8d18-df0a35c2b369 · inbound

Machine Learning Analysis of Anomalous Diffusion cites this paper.

Machine Learning Analysis of Anomalous Diffusion WaveNet: A Generative Model for Raw Audio

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-12T04:27:34.797915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:27:34.797915Z digest=sha256:10621b13cc633fa0fa8bc2ba1f25c8d96ea7860475126e4ee7b5a2472f3e9e0a

Observation 4628b374-f65f-4d5c-989b-28b7163e358f · inbound

Deep Learning Modeling Method for RF Devices Based on Uniform Noise Training Set cites this paper.

Deep Learning Modeling Method for RF Devices Based on Uniform Noise Training Set WaveNet: A Generative Model for Raw Audio

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T21:59:28.646021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:59:28.646021Z digest=sha256:7678f7be7ed167418ba2f4d06f9698adc693093ea4aa24a7a3111fb4b79084c2

Observation af09983c-47e3-4522-b4f9-ced1eb201898 · inbound

LMDM:Latent Molecular Diffusion Model For 3D Molecule Generation cites this paper.

LMDM:Latent Molecular Diffusion Model For 3D Molecule Generation WaveNet: A Generative Model for Raw Audio

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T21:41:49.762079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:41:49.762079Z digest=sha256:7e71167122c5cb1493fee67947a82fa1ab26a7cb3fd87ce3a73c67a671d12f8a

Observation b3901e03-157b-463b-bc79-44ce9350a8b0 · inbound

Effective Reward Specification in Deep Reinforcement Learning cites this paper.

Effective Reward Specification in Deep Reinforcement Learning WaveNet: A Generative Model for Raw Audio

Reference 234

Resolution
unresolved
no resolver link, observed 2026-08-11T19:09:55.136086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T19:09:55.136086Z digest=sha256:0be86fc4131e557679fbfed4df994c31affba5a2644c6c8abccca86d512d9026

Observation fa1813f9-3053-4de2-8ed4-920a997a2721 · inbound

QuantFormer: Learning to Quantize for Neural Activity Forecasting in Mouse Visual Cortex cites this paper.

QuantFormer: Learning to Quantize for Neural Activity Forecasting in Mouse Visual Cortex WaveNet: A Generative Model for Raw Audio

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T19:00:03.760512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:00:03.760512Z digest=sha256:aa6e74e3168cdb66b170f9d6a51e3ccda90e60a817287255c3a6c7154d27a65c

Observation 45b8afa9-96e9-4123-a765-1c4e54567888 · inbound

Non-Normal Diffusion Models cites this paper.

Non-Normal Diffusion Models WaveNet: A Generative Model for Raw Audio

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T18:31:59.939362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:31:59.939362Z digest=sha256:3bfbc5afd45b57dfcb815cabd8bdf1e24ac3def7dd97e762f85a6fcf9af1b5f2

Observation 53f5181f-6cc7-4fd9-b5e0-d53044501f8d · inbound

A Unified Model For Voice and Accent Conversion In Speech and Singing using Self-Supervised Learning and Feature Extraction cites this paper.

A Unified Model For Voice and Accent Conversion In Speech and Singing using Self-Supervised Learning and Feature Extraction WaveNet: A Generative Model for Raw Audio

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-11T18:01:14.165761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T18:01:14.165761Z digest=sha256:e0d27771e78d7eef8d373b9c00811e18f8f6b34c4a03e8b9c6f3576c45de1306

Observation db5a8688-9ea4-4e82-b032-7fd09029d8c0 · inbound

CSSinger: End-to-End Chunkwise Streaming Singing Voice Synthesis System Based on Conditional Variational Autoencoder cites this paper.

CSSinger: End-to-End Chunkwise Streaming Singing Voice Synthesis System Based on Conditional Variational Autoencoder WaveNet: A Generative Model for Raw Audio

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T17:31:06.866412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:31:06.866412Z digest=sha256:68776568a598fe2a87a529f6a16b2a84e0b2316767aaae38dc4b4638ee0e11f6

Observation 5aa73031-419c-4f51-a949-4ed69ad28f61 · inbound

Speech-Forensics: Towards Comprehensive Synthetic Speech Dataset Establishment and Analysis cites this paper.

Speech-Forensics: Towards Comprehensive Synthetic Speech Dataset Establishment and Analysis WaveNet: A Generative Model for Raw Audio

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T17:24:22.390076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:24:22.390076Z digest=sha256:3614eb21fb307296c096e5ffb940b9bf682fb443609b3e7397136d3b5eeab24d

Observation 03cd3e37-ab42-4411-97ff-9c31f280a311 · inbound

Learning Latent Spaces for Domain Generalization in Time Series Forecasting cites this paper.

Learning Latent Spaces for Domain Generalization in Time Series Forecasting WaveNet: A Generative Model for Raw Audio

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T15:17:18.781697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:17:18.781697Z digest=sha256:7ab7db598fb0985f38075b6bc71f1463116b8f4732ab86930fb8f45b0aac21f2

Observation d2bfc67e-447f-46ff-a2fb-684b61a9c752 · inbound

Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music cites this paper.

Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music WaveNet: A Generative Model for Raw Audio

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T15:00:02.704954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:00:02.704954Z digest=sha256:476b5886709d2588f1fe8687ea526b9957d24ed1d4785a5bac4150c96bd89d34

Observation 47bddea1-c0f3-4e49-a7c1-e1042d9b1111 · inbound

Phoneme-Level Feature Discrepancies: A Key to Detecting Sophisticated Speech Deepfakes cites this paper.

Phoneme-Level Feature Discrepancies: A Key to Detecting Sophisticated Speech Deepfakes WaveNet: A Generative Model for Raw Audio

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T13:57:20.917865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:57:20.917865Z digest=sha256:d8953fbc82941f92c179f386ad11128f8eb31d684ca339c9c19f118756b60bb9

Observation 15dee101-6225-40c4-b916-5ceafb872b5e · inbound

Falcon: Faster and Parallel Inference of Large Language Models through Enhanced Semi-Autoregressive Drafting and Custom-Designed Decoding Tree cites this paper.

Falcon: Faster and Parallel Inference of Large Language Models through Enhanced Semi-Autoregressive Drafting and Custom-Designed Decoding Tree WaveNet: A Generative Model for Raw Audio

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T13:56:39.734235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:56:39.734235Z digest=sha256:89a0c0ee5a9685a89a328db78eeb446c3ed670ab01b6b8461215c81a2beee81f

Observation 1ecd6d59-2419-45e9-b905-792981723578 · inbound

Cherry-Picking in Time Series Forecasting: How to Select Datasets to Make Your Model Shine cites this paper.

Cherry-Picking in Time Series Forecasting: How to Select Datasets to Make Your Model Shine WaveNet: A Generative Model for Raw Audio

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.623125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T12:18:39.623125Z digest=sha256:adfc4f00701f419d4c0dd0ae162819c2512e80f2301a1c60eb9e67d9ebc33d4a

Observation c3571804-71ec-4ce4-ac19-35a70a84b943 · inbound

Synthetic Time Series Data Generation for Healthcare Applications: A PCG Case Study cites this paper.

Synthetic Time Series Data Generation for Healthcare Applications: A PCG Case Study WaveNet: A Generative Model for Raw Audio

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T13:26:31.598140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:26:31.598140Z digest=sha256:29a6d7dc046b24740e6823486f44481cff56668e4d667872653833003b882ddd

Observation ef7937de-2a79-48ee-9cf0-43cba92739f0 · inbound

SoK: On the Offensive Potential of AI cites this paper.

SoK: On the Offensive Potential of AI WaveNet: A Generative Model for Raw Audio

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T04:47:13.612051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:47:13.612051Z digest=sha256:3ddee70f6a7de00a0edb75d1b65374894c6dc47e17d184aa5d64d8e78f303c39

Observation 19a5df6a-434a-4307-a63c-01b564b267c0 · inbound

BSDB-Net: Band-Split Dual-Branch Network with Selective State Spaces Mechanism for Monaural Speech Enhancement cites this paper.

BSDB-Net: Band-Split Dual-Branch Network with Selective State Spaces Mechanism for Monaural Speech Enhancement WaveNet: A Generative Model for Raw Audio

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T01:04:03.326861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T01:04:03.326861Z digest=sha256:ba1cff5aff057a94faa781b1d4eb28634a9ae5b0234e4b3b0640d0ede71493e3

Observation 8756600b-a2ad-42ba-9761-68db09328321 · inbound

Improving Generalization for AI-Synthesized Voice Detection cites this paper.

Improving Generalization for AI-Synthesized Voice Detection WaveNet: A Generative Model for Raw Audio

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T00:49:27.847846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:49:27.847846Z digest=sha256:32390cda0db2d80640bd42af06b6d84a046a46ff4c2454e98b3e43c6b5383265

Observation 614d5c98-5ce2-4392-8be2-08ff8d25809c · inbound

CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation cites this paper.

CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation WaveNet: A Generative Model for Raw Audio

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T23:41:01.095472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:41:01.095472Z digest=sha256:08845f0675e1f0613be2aed6c210a722198fa199991555b485097f9961ca6c2d

Observation e1b33c7b-fa58-44ca-abbb-8ef64363fde4 · inbound

Analog Alchemy: Neural Computation with In-Memory Inference, Learning and Routing cites this paper.

Analog Alchemy: Neural Computation with In-Memory Inference, Learning and Routing WaveNet: A Generative Model for Raw Audio

Reference 130

Resolution
unresolved
no resolver link, observed 2026-08-10T23:14:07.002483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:14:07.002483Z digest=sha256:196aeedaa523cfa621c51f432e8d26e21ca669e12456470dfc2227231e9bcf15

Observation 7061aed6-cde3-4b39-951a-889f4b70f02f · inbound

Few-Shot Radar Signal Recognition through Self-Supervised Learning and Radio Frequency Domain Adaptation cites this paper.

Few-Shot Radar Signal Recognition through Self-Supervised Learning and Radio Frequency Domain Adaptation WaveNet: A Generative Model for Raw Audio

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T21:55:50.454144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:55:50.454144Z digest=sha256:4299bb89708c8eaa42142e431de7c70ddb398b2463803314daeefd7d2e8d9537

Observation 8b6a760d-85ad-4ff9-82f1-5b7fb3d8b4b9 · inbound

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription cites this paper.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription WaveNet: A Generative Model for Raw Audio

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.653080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.653080Z digest=sha256:9f5dc3a14440da0e35095e87108f9639efeda9ed761178dea76349300e6bf130

Observation 21ae3f99-7ab1-486f-b819-b0095109a3c1 · inbound

Likelihood Training of Cascaded Diffusion Models via Hierarchical Volume-preserving Maps cites this paper.

Likelihood Training of Cascaded Diffusion Models via Hierarchical Volume-preserving Maps WaveNet: A Generative Model for Raw Audio

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-08-10T20:56:10.284004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:56:10.284004Z digest=sha256:31306a4d57641cef57a4ed21af7b1d6a32ae1084ecda6304008c5438f7c40c5f

Observation 74340849-317d-46ea-8d02-3d5e4f01def4 · inbound

Explore the Use of Time Series Foundation Model for Car-Following Behavior Analysis cites this paper.

Explore the Use of Time Series Foundation Model for Car-Following Behavior Analysis WaveNet: A Generative Model for Raw Audio

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T20:53:16.449766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:53:16.449766Z digest=sha256:409b668479918996c70973ffc7c9132541e831b270635b36114c75b0992bb126

Observation 85cab1f6-3833-4292-beb0-99196bf8e6f2 · inbound

STTS-EAD: Improving Spatio-Temporal Learning Based Time Series Prediction via cites this paper.

STTS-EAD: Improving Spatio-Temporal Learning Based Time Series Prediction via WaveNet: A Generative Model for Raw Audio

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T20:38:48.150043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:38:48.150043Z digest=sha256:60cbcbc72cd166fc7a86a5945bfe94f8ddf12d7f359e7f769c5bd1b188dd66ea

Observation dc334105-3d5b-4439-8189-969ae6c280df · inbound

Bridge-SR: Schr\"odinger Bridge for Efficient SR cites this paper.

Bridge-SR: Schr\"odinger Bridge for Efficient SR WaveNet: A Generative Model for Raw Audio

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T20:36:57.404572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:36:57.404572Z digest=sha256:cfb39a0747de518d1a68ce94c8b984e7dc96f83422211220b2ef37a95ffde832

Observation 3124be9c-6d0e-4872-9108-6d9827f33b4d · inbound

Mutual Regression Distance cites this paper.

Mutual Regression Distance WaveNet: A Generative Model for Raw Audio

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-10T19:08:48.862373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T19:08:48.862373Z digest=sha256:88c1bb61c6d1ddb8443b64efef932dcb6cf6a8746d49e41f5db66189d96b7aac

Observation 0f1707ea-3eac-4b01-ba0b-5e0d943caac2 · inbound

Noise-Resilient Point-wise Anomaly Detection in Time Series Using Weak Segment Labels cites this paper.

Noise-Resilient Point-wise Anomaly Detection in Time Series Using Weak Segment Labels WaveNet: A Generative Model for Raw Audio

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-10T17:45:45.995849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:45:45.995849Z digest=sha256:051ef11cb7e87ff128529b73a61cf94a0e5b6a94d263bc687e8761cfbb74ddc1

Observation 640b5ffb-de29-44d9-9f9c-44e07a269046 · inbound

T-Graphormer: Using Transformers for Spatiotemporal Forecasting cites this paper.

T-Graphormer: Using Transformers for Spatiotemporal Forecasting WaveNet: A Generative Model for Raw Audio

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T16:21:57.443951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T16:21:57.443951Z digest=sha256:a38ee6a6f8c337f288ed4d83027567d84a2a61b26c7177cdd206b7931512b8bc

Observation 18582bdb-2f48-44fc-9078-8e3f6501fe95 · inbound

Neural Vocoders as Speech Enhancers cites this paper.

Neural Vocoders as Speech Enhancers WaveNet: A Generative Model for Raw Audio

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T15:59:21.046334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:59:21.046334Z digest=sha256:459c7c5f12106f5f256560d6b59dc91e516626cb4df64097c61b5f4366acf4b7

Observation 276af8c5-afba-497c-bd78-dd83e2c9f782 · inbound

BARNN: A Bayesian Autoregressive and Recurrent Neural Network cites this paper.

BARNN: A Bayesian Autoregressive and Recurrent Neural Network WaveNet: A Generative Model for Raw Audio

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T23:38:10.292452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:38:10.292452Z digest=sha256:3e42be35cd4b7fabd6f1b169eb837e8373788fcc8a14c217988e2de0b173551a

Observation d0f759fd-81aa-46d8-b6c0-b866c42e6977 · inbound

OOD Detection with immature Models cites this paper.

OOD Detection with immature Models WaveNet: A Generative Model for Raw Audio

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T17:43:14.294425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:43:14.294425Z digest=sha256:99987cee9e67eca8d116eefeba23217a8129b3cbd3192caa3b5d1089eae147e7

Observation cba65a0d-e72e-47f2-a4dc-53c734ef45bb · inbound

Multimodal Brain-Computer Interfaces: AI-powered Decoding Methodologies cites this paper.

Multimodal Brain-Computer Interfaces: AI-powered Decoding Methodologies WaveNet: A Generative Model for Raw Audio

Reference 143

Resolution
unresolved
no resolver link, observed 2026-08-09T11:00:16.848844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:00:16.848844Z digest=sha256:002d1bf816edef4781deeb7bad7f18de9c95dbf2f4dccb056a7aab1708333994

Observation 7fa8aef8-8086-4156-8024-43485f67940b · inbound

MobiCLR: Mobility Time Series Contrastive Learning for Urban Region Representations cites this paper.

MobiCLR: Mobility Time Series Contrastive Learning for Urban Region Representations WaveNet: A Generative Model for Raw Audio

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T10:44:27.936799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:44:27.936799Z digest=sha256:1ad350bb14067d9253cc298fb6e2497f8ad3c9f2904fe6a169551ba390b4ed26

Observation b8be1c4b-cadf-4a29-b8fd-1ca592a2eb54 · inbound

Towards Explainable Spoofed Speech Attribution and Detection:a Probabilistic Approach for Characterizing Speech Synthesizer Components cites this paper.

Towards Explainable Spoofed Speech Attribution and Detection:a Probabilistic Approach for Characterizing Speech Synthesizer Components WaveNet: A Generative Model for Raw Audio

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-08T23:48:28.957134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:48:28.957134Z digest=sha256:08dac03f72ecb8dc5a80a258cdbd83cc97e0edb0229eda161b3455365eee01a4

Observation ff9dfb75-04a8-4166-a8ae-d6d1ef3c2093 · inbound

Reusing Embeddings: Reproducible Reward Model Research in Large Language Model Alignment without GPUs cites this paper.

Reusing Embeddings: Reproducible Reward Model Research in Large Language Model Alignment without GPUs WaveNet: A Generative Model for Raw Audio

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T11:32:47.954111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:32:47.954111Z digest=sha256:51bc7754511b541ee3cd82eb554f717116b2b9119c73c68d4c32039289ab4bf1

Observation 49d0fd8b-9410-4c57-86bb-6b09667ed3eb · inbound

What makes a good feedforward computational graph? cites this paper.

What makes a good feedforward computational graph? WaveNet: A Generative Model for Raw Audio

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T14:33:16.029766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:33:16.029766Z digest=sha256:af336ffaa980fe4f4db6531dc21380b9603d86574a17b33aa6d831e87d3b1592

Observation 0b0fad9c-15ba-44af-8c65-7959d2edbdad · inbound

DiffNMR3: Advancing NMR Resolution Beyond Instrumental Limits cites this paper.

DiffNMR3: Advancing NMR Resolution Beyond Instrumental Limits WaveNet: A Generative Model for Raw Audio

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T22:35:47.762112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:35:47.762112Z digest=sha256:ee7ac6cb46c83e07463f740a014957e8a63ecc2157b6249d69b7ff011f2efada

Observation 212fbf2d-35e4-4d71-b131-fae7ed2b1053 · inbound

Hookpad Aria: A Copilot for Songwriters cites this paper.

Hookpad Aria: A Copilot for Songwriters WaveNet: A Generative Model for Raw Audio

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T10:25:17.122248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:25:17.122248Z digest=sha256:ca1b3e5dcc47b60c5a236aa68c4b8bbdd8fb760767ef7e0399ae6ad6821d26c7

Observation daf066ab-d4ab-4c72-af8b-b320338f2213 · inbound

ASVspoof 5: Design, Collection and Validation of Resources for Spoofing, Deepfake, and Adversarial Attack Detection Using Crowdsourced Speech cites this paper.

ASVspoof 5: Design, Collection and Validation of Resources for Spoofing, Deepfake, and Adversarial Attack Detection Using Crowdsourced Speech WaveNet: A Generative Model for Raw Audio

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T23:32:41.933335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:32:41.933335Z digest=sha256:62147df57824193dac882020d6f90ba8ffda0cbd8b543801d3bd312eb2cfd98c

Observation d8eba047-a0fc-4978-adf7-ceb656644f82 · inbound

Harnessing Vision Models for Time Series Analysis: A Survey cites this paper.

Harnessing Vision Models for Time Series Analysis: A Survey WaveNet: A Generative Model for Raw Audio

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T23:27:13.303771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:27:13.303771Z digest=sha256:ac529c5bf007b0f36fbee082583fb185f9b55aead9ee536e23730e57b8268ef8

Observation e56fda20-2609-4af0-96e1-81938a4439bf · inbound

Causal Covariate Shift Correction using Fisher information penalty cites this paper.

Causal Covariate Shift Correction using Fisher information penalty WaveNet: A Generative Model for Raw Audio

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T12:04:37.441316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T12:04:37.441316Z digest=sha256:3879bc723dca57433d021af54f7148b5dafbb58cf9bc7f0137eed0ce7e4c9ad0

Observation b631993f-1a1e-4a52-b85e-ca91850650a3 · inbound

ModRWKV: Transformer Multimodality in Linear Time cites this paper.

ModRWKV: Transformer Multimodality in Linear Time WaveNet: A Generative Model for Raw Audio

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.141592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.141592Z digest=sha256:8114a74f817a5fd0910395bf5e68ed9b6f609ddf157dc362121aa7ac044c9a81

Observation 9bdfb6a8-ab96-431b-8cdc-596432ac9ea4 · inbound

Learning-based Airflow Inertial Odometry for MAVs using Thermal Anemometers in a GPS and vision denied environment cites this paper.

Learning-based Airflow Inertial Odometry for MAVs using Thermal Anemometers in a GPS and vision denied environment WaveNet: A Generative Model for Raw Audio

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:30:59.227529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:30:59.227529Z digest=sha256:0cab72c27a079029fd5f409138b212729776eff50c8f8381cf1bb7312fd65af7

Observation 413e8f29-202c-4454-b323-98c4e88440ab · inbound

Generative AI for Autonomous Driving: A Review cites this paper.

Generative AI for Autonomous Driving: A Review WaveNet: A Generative Model for Raw Audio

Reference 171

Resolution
unresolved
no resolver link, observed 2026-08-07T15:24:06.895949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:24:06.895949Z digest=sha256:01073c0d3ac1e0eda50caa3099b943c7a1e7a98d69ba8ab3b066cbd4d4d9d2d7

Observation 7e5e0f0e-b334-43cc-82f0-28d95c3b394e · inbound

An Exploratory Study on Multi-modal Generative AI in AR Storytelling cites this paper.

An Exploratory Study on Multi-modal Generative AI in AR Storytelling WaveNet: A Generative Model for Raw Audio

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T15:12:02.369628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:12:02.369628Z digest=sha256:16b1fe407e8635f1a8bf1dd9d45ba09d8d7eb66fce07149e6b61efd56cf913a3

Observation 5fc16efa-f3b9-456d-89f8-3a0d26f281b4 · inbound

Wavelet Probabilistic Recurrent Convolutional Network for Multivariate Time Series Classification cites this paper.

Wavelet Probabilistic Recurrent Convolutional Network for Multivariate Time Series Classification WaveNet: A Generative Model for Raw Audio

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:44.630851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:44.630851Z digest=sha256:2962e03acfd0406a14ae1a031f458da3daaf42009211782e97eaeca842843d3a

Observation 0c8522db-a3e9-4706-9c1e-77d85e6baec3 · inbound

Beyond Equilibrium: Non-Equilibrium Foundations Should Underpin Generative Processes in Complex Dynamical Systems cites this paper.

Beyond Equilibrium: Non-Equilibrium Foundations Should Underpin Generative Processes in Complex Dynamical Systems WaveNet: A Generative Model for Raw Audio

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:34.143637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:34.143637Z digest=sha256:0df77f3cf2974e4f88d4adb111d7695c7424fc79a765bfc8861399a78f3dd4b1

Observation 976e46b6-ae17-4def-af59-9111a2c06811 · inbound

Detecting Informative Channels: ActionFormer cites this paper.

Detecting Informative Channels: ActionFormer WaveNet: A Generative Model for Raw Audio

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T13:51:42.248579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:51:42.248579Z digest=sha256:04d71386138f6bb627d7162d1afce536b7503b43a4c3f5cb2b549b198dc51854

Observation d5b1d85e-49f2-4d2d-b2f3-f8b40b86493e · inbound

Versatile Cardiovascular Signal Generation with a Unified Diffusion Transformer cites this paper.

Versatile Cardiovascular Signal Generation with a Unified Diffusion Transformer WaveNet: A Generative Model for Raw Audio

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T13:16:59.636755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:16:59.636755Z digest=sha256:e98d726fc95f88d9697e0534c1aace8ebe168d66c3900668eb70a1d1308fda3d

Observation cee255ca-f30a-4f02-a882-eb659088f941 · inbound

Nonparametric Estimation of Conditional Survival Function with Time-Varying Covariates Using DeepONet cites this paper.

Nonparametric Estimation of Conditional Survival Function with Time-Varying Covariates Using DeepONet WaveNet: A Generative Model for Raw Audio

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:18.344173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:18.344173Z digest=sha256:7007cd6140f347fd7f7753d1cd493bbe1675500f854d685cba406c6c3d898720

Observation 53412ab4-300f-4825-9a76-46584e8aaf77 · inbound

BinauralFlow: A Causal and Streamable Approach for High-Quality Binaural Speech Synthesis with Flow Matching Models cites this paper.

BinauralFlow: A Causal and Streamable Approach for High-Quality Binaural Speech Synthesis with Flow Matching Models WaveNet: A Generative Model for Raw Audio

Reference 2012

Resolution
unresolved
no resolver link, observed 2026-08-07T13:04:13.162866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:04:13.162866Z digest=sha256:f259a3d757de003f3e8406ea937a14097fd2e5eff812c3f45e688fe5f5aef3bc

Observation 29eaab13-5848-44f4-94db-c5597f4da6a3 · inbound

SpeechVerifier: Robust Acoustic Fingerprint against Tampering Attacks via Watermarking cites this paper.

SpeechVerifier: Robust Acoustic Fingerprint against Tampering Attacks via Watermarking WaveNet: A Generative Model for Raw Audio

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:21.394211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:30:21.394211Z digest=sha256:746d0cb3b9709b3fb89d69dac521daddb5a009b5fe8d7cda32b2a28c3918bd8f

Observation cd989631-8db0-4189-84c2-20dd9ef2871f · inbound

SwitchCodec: A High-Fidelity Nerual Audio Codec With Sparse Quantization cites this paper.

SwitchCodec: A High-Fidelity Nerual Audio Codec With Sparse Quantization WaveNet: A Generative Model for Raw Audio

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-19T13:12:18.200370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:10:14.839742Z digest=sha256:cbd3f1252558d96141a7c9b50dc8ca2d166d41238e502dcc0a817319d98bd52c

Observation 28c6cab5-1945-48c2-9fc3-f4162815c91f · inbound

DiffDSR: Dysarthric Speech Reconstruction Using Latent Diffusion Model cites this paper.

DiffDSR: Dysarthric Speech Reconstruction Using Latent Diffusion Model WaveNet: A Generative Model for Raw Audio

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:12:57.357887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:12:57.357887Z digest=sha256:05b7e1b58caba68eebcda340f4b8d02dd572d6546a964d762f915059a4a268a0

Observation c8bfe558-57ec-4f69-82b4-b383d3bf917b · inbound

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models cites this paper.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models WaveNet: A Generative Model for Raw Audio

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:10.368822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:00:10.368822Z digest=sha256:b3d3e36734a7067af53f8ba97c52662fcffb15d82374ab00803a90fe697e0621

Observation 3ed84848-c431-4617-8919-5075d1f70ebe · inbound

The Promise of Spiking Neural Networks for Ubiquitous Computing: A Survey and New Perspectives cites this paper.

The Promise of Spiking Neural Networks for Ubiquitous Computing: A Survey and New Perspectives WaveNet: A Generative Model for Raw Audio

Reference 201

Resolution
unresolved
no resolver link, observed 2026-08-07T11:41:18.397823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:41:18.397823Z digest=sha256:a8e9aa5abf2a7a2d54fc21a920b77c2bbb12f82596d3bd07e7c7e4b2e1221e38

Observation 9d35f1bf-339e-4b51-b565-7b01b396321e · inbound

The cost of ensembling: is it always worth combining? cites this paper.

The cost of ensembling: is it always worth combining? WaveNet: A Generative Model for Raw Audio

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T10:40:49.016291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:40:49.016291Z digest=sha256:1250beddef0d2b608df7bfa34edb701ceb08df28df5bee6ae13748477f64119d

Observation 859c22ac-99a8-47f0-a808-49434e87a16f · inbound

Winner-takes-all for Multivariate Probabilistic Time Series Forecasting cites this paper.

Winner-takes-all for Multivariate Probabilistic Time Series Forecasting WaveNet: A Generative Model for Raw Audio

Reference 2005

Resolution
unresolved
no resolver link, observed 2026-08-07T10:26:15.614683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:26:15.614683Z digest=sha256:b067c019277d135be87cff3599740939e12a538c17a43db0cf18a1d4dfcf136a