Pith. sign in

Paper Citation Record · LEDGER

HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2010.05646.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2010.05646 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T06:02:14.705173Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

743
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1cd59664-8011-4b3b-abdd-476fb2bea821 · inbound

HiFi-SR: A Unified Generative Transformer-Convolutional Adversarial Network for High-Fidelity Speech Super-Resolution cites this paper.

HiFi-SR: A Unified Generative Transformer-Convolutional Adversarial Network for High-Fidelity Speech Super-Resolution HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:35.334307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:35.334307Z digest=sha256:c1015a554ee93a01820d7e6410d009d6c39a5640c666a191b19d8ac1ce640433

Observation bc343b69-0b10-4179-9c48-d180ef718c45 · inbound

Generative Adversarial Network based Voice Conversion: Techniques, Challenges, and Recent Advancements cites this paper.

Generative Adversarial Network based Voice Conversion: Techniques, Challenges, and Recent Advancements HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T06:02:14.705173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:02:14.705173Z digest=sha256:62144d2cc03fb142e4b6205dd5ad6f2de0d00927afd467c33d679ddbcce0b3cc

Observation 968f3035-3656-43ed-aa5d-cb8f91ee8ba5 · inbound

From Tens of Hours to Tens of Thousands: Scaling Back-Translation for Speech Recognition cites this paper.

From Tens of Hours to Tens of Thousands: Scaling Back-Translation for Speech Recognition HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:55.933535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:55:55.933535Z digest=sha256:7251778d34441085878275047c78d92b0f28309afba2c2a4e27cef22bd35eae7

Observation ead1e5f4-0116-4f3a-9204-a73e4d863fef · inbound

MuteSwap: Visual-informed Silent Video Identity Conversion cites this paper.

MuteSwap: Visual-informed Silent Video Identity Conversion HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:51.800306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:19:51.800306Z digest=sha256:fc54105bd02f7350830a99aa26d4c992907c0eda629c45f4d0e5694926a05a37

Observation 45171cba-3a9e-437e-8d8f-363ea2e819ae · inbound

Quantize More, Lose Less: Autoregressive Generation from Residually Quantized Speech Representations cites this paper.

Quantize More, Lose Less: Autoregressive Generation from Residually Quantized Speech Representations HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T16:55:50.034113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:55:50.034113Z digest=sha256:e16e2c3fbeed78d7c89740b6dec637fb32696b40d8cae38cc69c51744cfde9ff

Observation dca4162f-ba8b-4bd4-9128-75eaac40980d · inbound

Step-Audio 2 Technical Report cites this paper.

Step-Audio 2 Technical Report HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:59:51.164871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-16T05:59:50.900436Z digest=sha256:1b9fd17c0c1036e0f1c5a872bc51f5b99ab6afbf07c70ed54ad2714f05cfad99

Observation 724417c4-7bab-481d-90ce-930880d48b18 · inbound

Technical report: Impact of Duration Prediction on Speaker-specific TTS for Indian Languages cites this paper.

Technical report: Impact of Duration Prediction on Speaker-specific TTS for Indian Languages HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:29.951696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:14:29.951696Z digest=sha256:37d19a5652eebc9c8c651eb32c9b0fc6f773c762d35ddc99a95e4f746fd00b9c

Observation bba85d5b-0f43-4f82-8e37-6464b8284f01 · inbound

WaveLLDM: Design and Development of a Lightweight Latent Diffusion Model for Speech Enhancement and Restoration cites this paper.

WaveLLDM: Design and Development of a Lightweight Latent Diffusion Model for Speech Enhancement and Restoration HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T14:36:37.600739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:36:37.600739Z digest=sha256:221bf1faa5d5a23ae3eb5a12907e1ff6916a96108c9a0de61e8d389a69438ff8

Observation 36c2ed4e-c239-424e-a24f-596773e1d546 · inbound

Entropy-based Coarse and Compressed Semantic Speech Representation Learning cites this paper.

Entropy-based Coarse and Compressed Semantic Speech Representation Learning HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T13:36:05.134630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:36:05.134630Z digest=sha256:94cb1260a03baf0f0436c78b2854baaf59e56fbffbdc3068a7bd0f86998b8a35

Observation 1c146714-4cef-43c3-8b37-bc6b139b9d95 · inbound

LTX-2: Efficient Joint Audio-Visual Foundation Model cites this paper.

LTX-2: Efficient Joint Audio-Visual Foundation Model HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:06:20.644035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-13T07:06:20.470686Z digest=sha256:07583e28673bc6b6edec6608dd7ba29ae2a1ae45df3963bd3e0b714535ed0ab9

Observation d992a0f8-e8f1-4232-a319-f8b823dfd20b · inbound

Woosh: A Sound Effects Foundation Model cites this paper.

Woosh: A Sound Effects Foundation Model HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:53:15.808236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-13T20:51:08.144573Z digest=sha256:e889a353d3964b4b5c6cd6c8c007415ece8180b044e6d6aaac2d033fe48147f4

Observation df2cb341-f2bc-4c59-9ecf-44c7a2e6e304 · inbound

Few-Shot Synthetic Accented Speech for ASR Fine-Tuning: What Helps and When? cites this paper.

Few-Shot Synthetic Accented Speech for ASR Fine-Tuning: What Helps and When? HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:51:26.738614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-07T09:07:41.143902Z digest=sha256:12ad472e56c1221a9a8510cbaf9d560cabbd00ab35c4a89454d8c3b9b0170470

Observation a130a1d8-0b7c-4809-87ad-13474a9a4969 · inbound

Tibetan-TTS:Low-Resource Tibetan Speech Synthesis with Large Model Adaptation cites this paper.

Tibetan-TTS:Low-Resource Tibetan Speech Synthesis with Large Model Adaptation HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:17:09.717259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-08T02:59:39.588729Z digest=sha256:279062b77f275035dddd5ff3f0bc55e449aba4c4c7c85260c8dff266c4ded08d

Observation 3bd21313-8051-483a-a0d5-7911639f44f3 · inbound

Audio Deepfake Detection with Half-Truth Localisation Using Cross-Attentive Feature Fusion cites this paper.

Audio Deepfake Detection with Half-Truth Localisation Using Cross-Attentive Feature Fusion HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-29T06:03:07.645923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T06:02:59.488895Z digest=sha256:03c627ab8c514d3e8413745db3f9858d703f7f9d70d7ae8c4eea5cd924de3e51

Observation fedec146-001e-4fd3-a522-7f5592eb1cef · inbound

A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies cites this paper.

A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:48.098151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:48.098151Z digest=sha256:c0ee488d689df4918b582a1b8a3aa1817fc6d759eaea8a88ddc863a19d261da3