Pith. sign in

Paper Citation Record · LEDGER

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration

As of 23 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 2 inbound Pith citation observations for arXiv:2505.04457.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.04457 v4

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:32:42.771927Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:16:41.772452Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T17:16:21.330378Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact1
  • verified fuzzy29
  • unresolved24
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1892013c-9568-464b-80ce-723bf99c7ec3 · outbound

This paper cites Parametric resynthesis with neural vocoders,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Parametric resynthesis with neural vocoders,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.577489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.577489Z digest=sha256:7a924da5d58f501d61887b4e214ad3d71f00c66397461203e09d5d25480c3e23

Observation 778a495f-e658-4be3-ae75-6a7c9bd758ec · outbound

This paper cites SelfRemaster: Self-supervised speech restoration with analysis-by-synthesis approach using channel modeling,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration SelfRemaster: Self-supervised speech restoration with analysis-by-synthesis approach using channel modeling,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.581647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.581647Z digest=sha256:637bce0159954ffc7e7a346d448f7deaafb29fed0d28fd6206c02f9113b0bf98

Observation 1cd3b168-3ff4-41d9-ae23-d2334f44a0e3 · outbound

This paper cites HiFi-GAN-2: Studio-quality speech enhancement via generative adversarial networks conditioned on acoustic features,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration HiFi-GAN-2: Studio-quality speech enhancement via generative adversarial networks conditioned on acoustic features,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.585194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.585194Z digest=sha256:a9b0ade910fcc5581b4bee57351add76fac43126202251869cf8663817c02742

Observation 95ada63f-d7a9-4d8d-bf3c-c71042b6e100 · outbound

This paper cites V oiceFixer: A unified framework for high-fidelity speech restoration,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration V oiceFixer: A unified framework for high-fidelity speech restoration,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.588750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.588750Z digest=sha256:c672fb35cb8c09508c1518f40e7697d30987333a24fd4b66397b4df1f36db521

Observation d0df2acf-bee1-4b3e-9940-dd1ea559941e · outbound

This paper cites Universal Speech Enhancement with Score-based Diffusion.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Universal Speech Enhancement with Score-based Diffusion

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.592156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.592156Z digest=sha256:e8635784a13e5785ccf991a8b456598ee1a3dad285cac9f0fbacedd24e84559b

Observation be161f6f-05f5-4344-8764-1b6fecee3df7 · outbound

This paper cites Miipher: A robust speech restoration model integrating self-supervised speech and text representations,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Miipher: A robust speech restoration model integrating self-supervised speech and text representations,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.524115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.596030Z digest=sha256:b5e0abfd3ce0b917e1e6f7602922c582cf5da2ffa6c184e41be3e653e6833b5f

Observation d2483b2c-fa1a-4264-8060-f215cd0916f1 · outbound

This paper cites Speech enhancement and dereverberation with diffusion-based generative models,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Speech enhancement and dereverberation with diffusion-based generative models,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.513705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.599772Z digest=sha256:40eb7a7390fb6d98208a8b6abc59c55aa1e22dfe6471fa26eb8ed0e2b7a99167

Observation 73686880-e64b-49e2-967e-5ce7a52a83d4 · outbound

This paper cites Diffusion models for audio restoration: A review,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Diffusion models for audio restoration: A review,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.603671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.603671Z digest=sha256:41e42185ce2fac97c05924d1e9232055bea2201b9f14be05ec15f77d499b654c

Observation 615a9790-bf7b-4de2-a0c3-989fca2d99c1 · outbound

This paper cites Universal score-based speech enhancement with high content preservation,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Universal score-based speech enhancement with high content preservation,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.607148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.607148Z digest=sha256:27decac225d982ae546402a9c8a0bb418e108c8f47f054f9bddbd4516b1270f0

Observation 585c329c-d32d-47ab-8fbc-992d7eee9ba9 · outbound

This paper cites LLaSE-G1: Incentivizing Generalization Capability for LLaMA-based Speech Enhancement.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration LLaSE-G1: Incentivizing Generalization Capability for LLaMA-based Speech Enhancement

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.611362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.611362Z digest=sha256:293e52334882d27e5d4db7e2aed45a6b474d9ea82bb073243e607475f939047b

Observation cf7b59b4-66aa-4be7-8a72-56662bc83d58 · outbound

This paper cites Genhancer: High-fidelity speech enhancement via generative modeling on discrete codec tokens,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Genhancer: High-fidelity speech enhancement via generative modeling on discrete codec tokens,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.615300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.615300Z digest=sha256:cc40ca866cec4c66bbc6a61a61190a1d4c964aeef84f64cf8acd1204a95f516e

Observation d27fdb2b-17db-4015-9151-d9a958bf881f · outbound

This paper cites Joint semantic knowledge distillation and masked acoustic modeling for full-band speech restoration with improved intelligibility,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Joint semantic knowledge distillation and masked acoustic modeling for full-band speech restoration with improved intelligibility,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.619218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.619218Z digest=sha256:a8c667fc7ac4d8bc970a457eea1f4f6b414b5b8450b1a0ba5c284ea766ccab4c

Observation d94742fa-34c0-49de-bee9-c52637cf941a · outbound

This paper cites DiTSE: High-fidelity generative speech enhancement via latent diffusion transformers,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration DiTSE: High-fidelity generative speech enhancement via latent diffusion transformers,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.622796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.622796Z digest=sha256:b7ecf78966cfb719140a5e54e9f1d7894c73fdb6f42b2902b2eb6bc312a27c05

Observation 16073fb8-1f62-48f9-a256-5dc406c3b5c6 · outbound

This paper cites LibriTTS-R: A restored multi-speaker text- to-speech corpus,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration LibriTTS-R: A restored multi-speaker text- to-speech corpus,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.478525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.625916Z digest=sha256:f3406e2c38b06e5e53519b649a3872d77519e32dc6a12538ae87f197ce707712

Observation c3a24491-1321-4643-831e-51358c5c4a15 · outbound

This paper cites FLEURS-R: A restored multilingual speech corpus for generation tasks,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration FLEURS-R: A restored multilingual speech corpus for generation tasks,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.467600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.629351Z digest=sha256:a62dd04679e8e16bd41073e7bed5db137d52142a70cb0ba1af86cad1af0e31b7

Observation 84d3122f-ee83-494f-9072-6db649542fc2 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.632918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.632918Z digest=sha256:68a6b0fdd3623acf01766cf93534835e34855ff2b5d192ba7dd0bca945c46ca5

Observation f2d1c267-72e2-4e99-8d3f-866c30c6d5c4 · outbound

This paper cites GPT-4 Technical Report.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration GPT-4 Technical Report

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.636339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.636339Z digest=sha256:300c680594ac3ae55552cf4f5da1968882d24ec3f41910793ab8096266733aa9

Observation d461d7bc-73da-4704-8a08-68406539fc23 · outbound

This paper cites Moshi: a speech-text foundation model for real-time dialogue.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Moshi: a speech-text foundation model for real-time dialogue

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.639502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.639502Z digest=sha256:d787ff338ab1de910bd46a647ecbbafe433fae0d4d91c42126a042e9f3d63c98

Observation 51fdadb0-5a67-494b-a17f-22dff5ed06e7 · outbound

This paper cites Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.643117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.643117Z digest=sha256:17388ed4881cc4cfb9dd2a1a977803f96d5e38045495d3e53456b3cf0c457604

Observation e8467dce-e907-4d8e-a05d-27a56c737e5c · outbound

This paper cites Towards a unified view of parameter- efficient transfer learning,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Towards a unified view of parameter- efficient transfer learning,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.456216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.646893Z digest=sha256:14146508db216c4be41dcb3bace2717eb36c99ea0dffc097ac528157f5b82db4

Observation c61cbd93-9bee-461a-b9ec-f25c3c9507a1 · outbound

This paper cites WaveFit: An iterative and non- autoregressive neural vocoder based on fixed-point iteration,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration WaveFit: An iterative and non- autoregressive neural vocoder based on fixed-point iteration,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.650195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.650195Z digest=sha256:f2edac019721a7d45d06d05bd8a1a75171e120e54f9b1490afad765d485f101d

Observation a9bda9ef-65bb-4b20-a0b9-5fb924c65bf5 · outbound

This paper cites Ten lessons from three generations shaped google’s tpuv4i : Industrial product,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Ten lessons from three generations shaped google’s tpuv4i : Industrial product,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.440054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.653503Z digest=sha256:a5c56560bb90c7e0ce1c34988da17591541571102cc3b6cfbbd9b2e6248b9f3a

Observation 2770856e-7a88-4d33-aa92-bcd5f315043d · outbound

This paper cites Self-supervised learning with random- projection quantizer for speech recognition,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Self-supervised learning with random- projection quantizer for speech recognition,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.429272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.656687Z digest=sha256:481f53262d38ab7ffa075a10245560d7e64ba9aa9c2fda7ea96cb39079762318

Observation 53636106-62be-4ec5-8c96-a6039a9aec64 · outbound

This paper cites wav2vec 2.0: A framework for self- supervised learning of speech representations,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration wav2vec 2.0: A framework for self- supervised learning of speech representations,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.418184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.660180Z digest=sha256:aec5304b2eee5a327a8227dac3fb23b730bc50fe8e8b00edadc53f9896597af7

Observation d079c908-6a74-4505-9183-dff2337e4a70 · outbound

This paper cites w2v-BERT: Combining contrastive learning and masked language modeling for self-supervised speech pre- training,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration w2v-BERT: Combining contrastive learning and masked language modeling for self-supervised speech pre- training,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.407977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.663312Z digest=sha256:48e8181903aa8030e10d18f0f43eebb75a075d0eb15f29a5327e84cdc60e7fa4

Observation 67e8fe65-f6a3-4b73-a393-0e3259c58ce5 · outbound

This paper cites Hubert: Self-supervised speech representation learning by masked prediction of hidden units,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Hubert: Self-supervised speech representation learning by masked prediction of hidden units,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.667269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.667269Z digest=sha256:7a2b17b13b0521a7814a72412afb29d654212613d6c79c0d06bb237698f3737c

Observation 0048fb43-db67-41a6-a2c8-f54870cc487d · outbound

This paper cites Wavlm: Large-scale self-supervised pre-training for full stack speech processing,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Wavlm: Large-scale self-supervised pre-training for full stack speech processing,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.670280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.670280Z digest=sha256:0cd662c2f541d3451258803e048713b68e858abdd1acb3fe1e393f8292160253

Observation cbe5840e-0e8a-4c5d-8779-49d389b9969a · outbound

This paper cites V oice conversion with just nearest neighbors,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration V oice conversion with just nearest neighbors,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.673735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.673735Z digest=sha256:602abd7c98bd273f2faa5286d7cd82d80e71f5a8b867c08611e4e7b0c26becde

Observation 9dd4b508-8b0d-45cb-addf-509425016c4e · outbound

This paper cites Vec-Tok Speech: speech vectorization and tokenization for neural speech generation.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Vec-Tok Speech: speech vectorization and tokenization for neural speech generation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.677294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.677294Z digest=sha256:e03f6e4b18f8950d5b817c575b87f0a777b1d898e9bd9536ebc5b46b8b899c12

Observation f2f635e1-73ec-46ff-9d4a-2650381d9c76 · outbound

This paper cites DF-Conformer: Integrated architecture of Conv-TasNet and Conformer using linear complexity self-attention for speech enhancement,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration DF-Conformer: Integrated architecture of Conv-TasNet and Conformer using linear complexity self-attention for speech enhancement,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.392944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.680939Z digest=sha256:272217b3787b55c07e2929f9fbde9792d376987c80d24c7fab167b6b7440109a

Observation 2d1e8a46-6c55-4459-8e94-c30730d8cce9 · outbound

This paper cites Fast spectrogram inversion using multi-head convolutional neural networks,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Fast spectrogram inversion using multi-head convolutional neural networks,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.382625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.684230Z digest=sha256:38b57b0b0ea38a9f69beafca08112cec599c3d6f437b0b5b45528b7bfbf47519

Observation 871215ce-ed79-44db-9e98-2a8a6e24db45 · outbound

This paper cites FiLM: Visual reasoning with a general conditioning layer,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration FiLM: Visual reasoning with a general conditioning layer,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.373206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.687344Z digest=sha256:8a8f2885ef06c5252d81c73b4f7155d887e3b83994c40c6c3acb4447951de2e4

Observation 2ebaffcf-0a71-4e91-b653-52c6c3ecbe5b · outbound

This paper cites WaveGrad: Estimating gradients for waveform generation,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration WaveGrad: Estimating gradients for waveform generation,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.363030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.690334Z digest=sha256:21dcb452e64c38610115f78ce00c8c2715698ef99e308f6e733cad481ce1a764

Observation 65ba69fd-8587-4349-bc89-a3655b1cb8cf · outbound

This paper cites URGENT challenge 2025 baseline,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration URGENT challenge 2025 baseline,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.353482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.693720Z digest=sha256:51733dad91fb43d775d9c29fc5285732e40f71683914387be66a5dd55d8c3710

Observation 3b08f8d3-3f65-490c-b8c3-6ba4d88883a7 · outbound

This paper cites PromptTTS++: Controlling speaker identity in prompt-based text-to-speech using natural language descrip- tions,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration PromptTTS++: Controlling speaker identity in prompt-based text-to-speech using natural language descrip- tions,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.344013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.696554Z digest=sha256:f1429f9bd2026e372e38d8a07c69dc0aa9292cdd31469dd26f40bd2e63e51c9a

Observation 15c7dead-5fb3-4453-8ee5-ed72b6391fe2 · outbound

This paper cites Xtts: a massively multilingual zero-shot text-to-speech model,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Xtts: a massively multilingual zero-shot text-to-speech model,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.332596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.700307Z digest=sha256:b45f314bd424e021e8fec961d02839d5dd281ae22cfed1bef494ff07c5a96eae

Observation c291ea86-2bb2-4ca4-ab4f-a5d2346c982b · outbound

This paper cites Natural language guidance of high-fidelity text-to-speech with synthetic annotations.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Natural language guidance of high-fidelity text-to-speech with synthetic annotations

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.703745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.703745Z digest=sha256:0de95f3442eaab84f5c296d06d47cde2cd564546d554bf6e727f037b6da1dd82

Observation 1dd4c81b-8604-4235-a1a6-3cbc42f9ae57 · outbound

This paper cites CoVoST: A diverse multilingual speech-to-text translation corpus,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration CoVoST: A diverse multilingual speech-to-text translation corpus,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.322714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.707501Z digest=sha256:de3f4e7082468a8f732bb425ce6b0490e14606b32f96ce753d81684884da286e

Observation ee5e7ca5-5d77-4210-9447-935d32ec8e32 · outbound

This paper cites CVSS corpus and massively multilingual speech-to-speech translation,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration CVSS corpus and massively multilingual speech-to-speech translation,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.313287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.710773Z digest=sha256:afc6c5851ee9adcfaf23ddbfc1006ffe2a16cc3d25965838f58e62e49f3f3e38

Observation bf39bc81-3527-4e8f-839f-5ad6dc60da75 · outbound

This paper cites Mls: A large-scale multilingual dataset for speech research,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Mls: A large-scale multilingual dataset for speech research,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.302906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.713922Z digest=sha256:d2a8bb2f10b11d99349740637e486953047896c02865b0447a08f6e073f1c16b

Observation e4c9d48a-cef2-4e97-b66e-d99e56c7e9a6 · outbound

This paper cites Fleurs: Few-shot learning evaluation of universal representations of speech,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Fleurs: Few-shot learning evaluation of universal representations of speech,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.293112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.717154Z digest=sha256:a36e89f03e980f3c3286a1bfa649a49785e868b779ea9c03d00f403983b0129e

Observation fd8c7d3d-9924-4d7e-9f71-fafb58022c4e · outbound

This paper cites Image method for efficiently simulating small-room acoustics,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Image method for efficiently simulating small-room acoustics,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.720338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.720338Z digest=sha256:3319707f77df2323bee19fb41e5374d3c27540b20638a4f4504e730be1f90ad0

Observation 9ae9d80c-0921-4630-96ce-4f332333bc93 · outbound

This paper cites HiFi-GAN: High-fidelity denoising and dereverberation based on speech deep features in adversarial networks,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration HiFi-GAN: High-fidelity denoising and dereverberation based on speech deep features in adversarial networks,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.277235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.723658Z digest=sha256:c39a6750cef301189299c5d7e15388d1816d4d9ab67a3b05c58f91417053a5c1

Observation 5679b8be-58e7-4077-ac28-39ead45a4358 · outbound

This paper cites Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,

Reference 44

Resolution
malformed identifier
no resolver link, observed 2026-08-15T23:32:42.727294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.727294Z digest=sha256:2284956f37f4db1a551a928662cc36350c775fda8f196b5e2e4361e062d89a27

Observation ca40632e-e652-42b8-9eac-909d4dcdc76e · outbound

This paper cites Transfer learning from speaker verification to multispeaker text-to-speech synthesis,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Transfer learning from speaker verification to multispeaker text-to-speech synthesis,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.730687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.730687Z digest=sha256:83f7e94723a4c7df405952b3712cd6270e53b39ea997942c123277ab330e1dc7

Observation 04b3664d-ceb5-48fb-80d4-41e322ac1adb · outbound

This paper cites Sample efficient adaptive text-to-speech,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Sample efficient adaptive text-to-speech,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:42.738754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:42.738754Z digest=sha256:4f1f9096a424c81bce02a64d666c50f1239b6a23de2053dd3d13fb96240258c5

Observation 0c57575c-e60e-4e88-bd6f-cfd191fedb24 · outbound

This paper cites DNSMOS: A non-intrusive perceptual objective speech quality metric to evaluate noise suppressors,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration DNSMOS: A non-intrusive perceptual objective speech quality metric to evaluate noise suppressors,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.255856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.742652Z digest=sha256:d41041bbced5a905edc9d07c6c0520d883760fa1063916cc9d1d2b8c2adb2cc6

Observation aa719d13-24d4-4cd8-9a89-27f858593ef9 · outbound

This paper cites SQuId: Measuring Speech Naturalness in Many Languages.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration SQuId: Measuring Speech Naturalness in Many Languages

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:32:42.810919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.745784Z digest=sha256:bc0bfa48448a3dc82095499a47acfb4cb1f208c1935602eaa25f8e56c6e033cd

Observation c001b40f-7490-4293-bead-05baeefdbca3 · outbound

This paper cites Tf-gridnet: Making time-frequency domain models great again for monaural speaker separation,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Tf-gridnet: Making time-frequency domain models great again for monaural speaker separation,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.245389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.749404Z digest=sha256:eb0ac0ad5563de8b0284ceb360c4eba052e1853e7170c2600dc7ee366e596e74

Observation de58b87a-7c84-4eda-8fa4-cd8f287a196b · outbound

This paper cites The design for the Wall Street Journal-based CSR corpus,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration The design for the Wall Street Journal-based CSR corpus,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.234101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.752771Z digest=sha256:9f0c1ca31a34d0d798342698109d42a316d0afc78608cedbb0b27ec77d3ff989

Observation 4cd4ef79-a611-490d-98b4-b348c5c6a614 · outbound

This paper cites Common V oice: A massively-multilingual speech corpus,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Common V oice: A massively-multilingual speech corpus,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.222412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.756115Z digest=sha256:f4067efff6a5a47b2b89f848dd01cad554aaa8d881c4cd1d194da91d5413e4d2

Observation 536c3c56-7a8b-458a-a2aa-494a2af2134f · outbound

This paper cites mHuBERT-147: A Compact Multilingual HuBERT Model,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration mHuBERT-147: A Compact Multilingual HuBERT Model,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.212033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.760033Z digest=sha256:9eac530a8443fb7fd8750fc29e90b336c3326c680ee11c3950444561f3681a8c

Observation ad065ed2-98ad-4d8a-9582-eb659a01789f · outbound

This paper cites Towards robust speech representation learning for thousands of languages,.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Towards robust speech representation learning for thousands of languages,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.201845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.763985Z digest=sha256:289951b3c3137250b9b9572425a85fe164bd6281fcdf1e020f32dd9573dfb7d2

Observation 0d47ae0f-f161-404f-80d1-33fd0b826017 · outbound

This paper cites Nakata, https://github.com/Wataru-Nakata/miipher.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Nakata, https://github.com/Wataru-Nakata/miipher

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.192008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.768431Z digest=sha256:3de9d789a0ffd7a7c7f1e6f6f41d390798d386c15efbd99402ffb3a1bcd637d4

Observation 7a4c618e-962f-4008-b88c-bcbd163b7e1f · outbound

This paper cites Ikemiya, https://github.com/yukara-ikemiya/wavefit-pytorch.

Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration Ikemiya, https://github.com/yukara-ikemiya/wavefit-pytorch

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:32:43.181306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T23:32:42.771927Z digest=sha256:04d1edf158df5a1d9d6cd5a32113027a881a02dfac17c22cd90a4787f99646dc

Pith citing papers

Observation 2f81d4fb-c484-47a4-a834-a2affe547eb0 · inbound

ReverbMiipher: Generative Speech Restoration meets Reverberation Characteristics Controllability cites this paper.

ReverbMiipher: Generative Speech Restoration meets Reverberation Characteristics Controllability Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T23:16:41.772452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:16:41.772452Z digest=sha256:13c8f293f64bc7dbebd113d3f06ec2836da6fb3d721506ce605db35d7c87f059

Observation 08e637f5-1af5-4aa6-9c46-10db19b147a8 · inbound

Rethinking Training Targets, Architectures and Data Quality for Universal Speech Enhancement cites this paper.

Rethinking Training Targets, Architectures and Data Quality for Universal Speech Enhancement Miipher-2: A Universal Speech Restoration Model for Million-Hour Scale Data Restoration

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:16:21.332886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-15T17:14:55.602201Z digest=sha256:a48c97b9c1e91f98982d04b147cfc5f6b8e57feb4cdaf165d36b33f8dc6813d9