Pith. sign in

Paper Citation Record · LEDGER

HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2010.05646.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2010.05646 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:55:55.933535Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

743
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 968f3035-3656-43ed-aa5d-cb8f91ee8ba5 · inbound

From Tens of Hours to Tens of Thousands: Scaling Back-Translation for Speech Recognition cites this paper.

From Tens of Hours to Tens of Thousands: Scaling Back-Translation for Speech Recognition HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:55.933535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:55:55.933535Z digest=sha256:5f1cf7dae90ecd9a87a1a44f74b5509a8c3444d31d860a86ae6033cd6c100657

Observation ead1e5f4-0116-4f3a-9204-a73e4d863fef · inbound

MuteSwap: Visual-informed Silent Video Identity Conversion cites this paper.

MuteSwap: Visual-informed Silent Video Identity Conversion HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:51.800306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:19:51.800306Z digest=sha256:0d700110addca5be72ca3a09fb645055c571df5ad486a5a4b71de9fe10c0440c

Observation 45171cba-3a9e-437e-8d8f-363ea2e819ae · inbound

Quantize More, Lose Less: Autoregressive Generation from Residually Quantized Speech Representations cites this paper.

Quantize More, Lose Less: Autoregressive Generation from Residually Quantized Speech Representations HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T16:55:50.034113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:55:50.034113Z digest=sha256:7083296fba9023ae097920700d4685dc1f93483232dec021be838b620709dfd4

Observation dca4162f-ba8b-4bd4-9128-75eaac40980d · inbound

Step-Audio 2 Technical Report cites this paper.

Step-Audio 2 Technical Report HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:59:51.164871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T05:59:50.900436Z digest=sha256:ceec501d810be8f7c36e73f1316c1c23fb45261a21c6dae467303a4a56a25c69

Observation 724417c4-7bab-481d-90ce-930880d48b18 · inbound

Technical report: Impact of Duration Prediction on Speaker-specific TTS for Indian Languages cites this paper.

Technical report: Impact of Duration Prediction on Speaker-specific TTS for Indian Languages HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:29.951696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:14:29.951696Z digest=sha256:e6c5b7390863ddb967d214b51925087b0b99d5b68a5f364aa275d5a4f84496fd

Observation bba85d5b-0f43-4f82-8e37-6464b8284f01 · inbound

WaveLLDM: Design and Development of a Lightweight Latent Diffusion Model for Speech Enhancement and Restoration cites this paper.

WaveLLDM: Design and Development of a Lightweight Latent Diffusion Model for Speech Enhancement and Restoration HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T14:36:37.600739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:36:37.600739Z digest=sha256:ebe293ea0c662698f05cc8173380fb9169208b9b28e390f53306f87cf93730b6

Observation 36c2ed4e-c239-424e-a24f-596773e1d546 · inbound

Entropy-based Coarse and Compressed Semantic Speech Representation Learning cites this paper.

Entropy-based Coarse and Compressed Semantic Speech Representation Learning HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T13:36:05.134630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:36:05.134630Z digest=sha256:c7cd93bd451c07afaef443dda388aa514d758fa5a3d792c1b3d0be72ecd3edc6

Observation 1c146714-4cef-43c3-8b37-bc6b139b9d95 · inbound

LTX-2: Efficient Joint Audio-Visual Foundation Model cites this paper.

LTX-2: Efficient Joint Audio-Visual Foundation Model HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:06:20.644035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:06:20.470686Z digest=sha256:7c3ddef3900e32559a0b7e0577b7f2311103318de687e0ecfb6c4cb13a0f804c

Observation d992a0f8-e8f1-4232-a319-f8b823dfd20b · inbound

Woosh: A Sound Effects Foundation Model cites this paper.

Woosh: A Sound Effects Foundation Model HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:53:15.808236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T20:51:08.144573Z digest=sha256:128fd6e6488388386c8fc1a6162c82c9f8e594edec78823c2d73dab84195afd3

Observation df2cb341-f2bc-4c59-9ecf-44c7a2e6e304 · inbound

Few-Shot Synthetic Accented Speech for ASR Fine-Tuning: What Helps and When? cites this paper.

Few-Shot Synthetic Accented Speech for ASR Fine-Tuning: What Helps and When? HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:51:26.738614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T09:07:41.143902Z digest=sha256:f27076b0dedb2de4d150dc9ec271068a10878aa6647cc1455283954c1816e274

Observation a130a1d8-0b7c-4809-87ad-13474a9a4969 · inbound

Tibetan-TTS:Low-Resource Tibetan Speech Synthesis with Large Model Adaptation cites this paper.

Tibetan-TTS:Low-Resource Tibetan Speech Synthesis with Large Model Adaptation HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:17:09.717259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T02:59:39.588729Z digest=sha256:92e0ada2236784e8996ca419ebb7cf56dcca470fd8f987a15eaabfa205e3dc89

Observation 3bd21313-8051-483a-a0d5-7911639f44f3 · inbound

Audio Deepfake Detection with Half-Truth Localisation Using Cross-Attentive Feature Fusion cites this paper.

Audio Deepfake Detection with Half-Truth Localisation Using Cross-Attentive Feature Fusion HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-29T06:03:07.645923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T06:02:59.488895Z digest=sha256:3a1d610f16282cb0961183cb0d0a610aca8a84167b1ac4e1b84ea13b04361404