Pith. sign in

Paper Citation Record · LEDGER

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models

As of 21 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 1 inbound Pith citation observation for arXiv:2506.00832.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00832 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:00:12.519114Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:00:09.374265Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T12:00:13.204804Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact2
  • verified fuzzy21
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 02f6f1f8-c1ca-42de-870b-2c079d7d87b2 · outbound

This paper cites What would the internal representation look like if the model aimed for a higher pitch rather than the current, lower one?.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models What would the internal representation look like if the model aimed for a higher pitch rather than the current, lower one?

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:18.318653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:09.318584Z digest=sha256:e940273ac0ae0abae160b8393a92af7dd43bd8c1c00a9bbadc9a6eea99b5507a

Observation 5638f2c9-2732-4a60-a3ac-02461578072a · outbound

This paper cites Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T12:00:13.248560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:09.374265Z digest=sha256:109dfb0efba06b4a14520b07c1559b07c150cbeabf21afb5a69b97f08c29b374

Observation 7241112a-e882-4420-8835-0c7fb9c7aefb · outbound

This paper cites What would the internal representation look like if the synthesized speech had a higher pitch rather than a lower pitch?.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models What would the internal representation look like if the synthesized speech had a higher pitch rather than a lower pitch?

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:18.062228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:09.480496Z digest=sha256:8bf8c0390fc14c1c3faca222b32cce1e9ab381a61810640a902d34031e0f6168

Observation 3590fff4-d059-4644-a4e7-015aed35fa08 · outbound

This paper cites an unresolved cited work.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:00:17.902086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:09.563895Z digest=sha256:89ab0f89a4ea76200af19359285419c44192759a1600bde2213a51921d350f26

Observation 37349320-eb49-4298-81b1-7a3994aa0f7b · outbound

This paper cites In planning its data processing techniques.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models In planning its data processing techniques

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:17.722710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:09.654733Z digest=sha256:ccfefb32193e36279da660816dda0469c62cebd55834ee56e55c3e91fe745652

Observation 44d11211-003e-41f0-8d42-4ce76193197d · outbound

This paper cites an unresolved cited work.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:00:17.633955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:09.732784Z digest=sha256:0785fb63069443f857d54977d7a9740548c4cff37d2bac8714e08906b2d4044b

Observation 1003b833-5259-4b16-bc68-dfe11d55928d · outbound

This paper cites 2022-0-00984, Devel- opment of Artificial Intelligence Technology for Personalized Plug-and-Play Explanation and Verification of Explanation; No.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models 2022-0-00984, Devel- opment of Artificial Intelligence Technology for Personalized Plug-and-Play Explanation and Verification of Explanation; No

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:17.417830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:09.840699Z digest=sha256:3ba696b94fe368a28877e6608e120947fbf250cbddde6d63e0829305af5222b3

Observation 2b339be2-3a3e-42ca-a3fb-9dd0a901e141 · outbound

This paper cites Tacotron: Towards End-to-End Speech Synthesis.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Tacotron: Towards End-to-End Speech Synthesis

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:09.959311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:00:09.959311Z digest=sha256:3111fe61cb56d72d531bc527365b093ae396a92b87530a8cbf13e1d429d4dca0

Observation d7419863-81a5-4109-a059-61c5e407d01b · outbound

This paper cites Natural tts synthesis by conditioning wavenet on mel spectrogram predic- tions,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Natural tts synthesis by conditioning wavenet on mel spectrogram predic- tions,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:17.294685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:10.014611Z digest=sha256:09a56ac0668c219e4449ad267504bd67ffe0471b593e9f35db0ba38f1b1f0f36

Observation 24f6eeb6-59af-4911-b21b-dd9f13ab3f56 · outbound

This paper cites Neural speech synthe- sis with transformer network,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Neural speech synthe- sis with transformer network,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:17.049983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:10.124986Z digest=sha256:910161ed3d2645792bc420aad6c6d744efa67b83a72c980a2dec96b3c04ec5d7

Observation 508dffc0-9d53-459f-ba21-77ad334033e4 · outbound

This paper cites Fastspeech: Fast, robust and controllable text to speech,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Fastspeech: Fast, robust and controllable text to speech,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:10.220550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:00:10.220550Z digest=sha256:057a7e86cae5c60b2251bee780c16b94c7719db481f865e4100a5d9087f0b79b

Observation fc3de9b6-9a65-4842-a3ef-4433dcddfc69 · outbound

This paper cites FastSpeech 2: Fast and High-Quality End-to-End Text to Speech.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models FastSpeech 2: Fast and High-Quality End-to-End Text to Speech

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:10.317617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:00:10.317617Z digest=sha256:dd08c7499093fc9e340aa1044ccecca75e1816f8612902f0a4ccdc7cb09ec43e

Observation c8bfe558-57ec-4f69-82b4-b383d3bf917b · outbound

This paper cites WaveNet: A Generative Model for Raw Audio.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models WaveNet: A Generative Model for Raw Audio

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:10.368822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:00:10.368822Z digest=sha256:995350a8c83b808d39370ddb48170efa7282ad68d47a753d6e9d244f394ce5e1

Observation a44eca22-eeaf-46c1-9a54-1f7da6dacf01 · outbound

This paper cites Waveglow: A flow-based generative network for speech synthesis,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Waveglow: A flow-based generative network for speech synthesis,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:16.909764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:10.424096Z digest=sha256:e3514b27fe1fb71f8a66074835815720049dc382f6234d24289de017d20b92ea

Observation 3bff5da7-4ba2-473e-86e7-8a80941dfa07 · outbound

This paper cites Mel- gan: Generative adversarial networks for conditional waveform synthesis,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Mel- gan: Generative adversarial networks for conditional waveform synthesis,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:16.716059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:10.536694Z digest=sha256:6bc696dca28e288f420eba8198344067504ac16250ebd619a56c3e0e45ec45c5

Observation 4337127c-beb1-442c-8cd1-27d4b13c3992 · outbound

This paper cites Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:16.447390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:10.614594Z digest=sha256:cd764193d8d2ee10ef1f8b5f17ed70683a424b47e4e19f101a40eedd50b1e6ad

Observation cb462432-b81c-4d94-a303-910b17317d71 · outbound

This paper cites Style tokens: Un- supervised style modeling, control and transfer in end-to-end speech synthesis,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Style tokens: Un- supervised style modeling, control and transfer in end-to-end speech synthesis,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:10.682549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:00:10.682549Z digest=sha256:ceddbc2942760b3921c33d1d1aae8a9e32f0b710c1629946b20180885c160a17

Observation 3b54d476-2d76-49c6-aa5d-b571d2f94059 · outbound

This paper cites To- wards end-to-end prosody transfer for expressive speech synthesis with tacotron,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models To- wards end-to-end prosody transfer for expressive speech synthesis with tacotron,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:10.754568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:00:10.754568Z digest=sha256:79854f6fd4a9908edef9eeb36a4ddaab6029be80477dad8b31a6abaa782d5964

Observation 0fd643ee-44fa-46ab-a53d-8e9ccab5cb3a · outbound

This paper cites Hierarchical Generative Modeling for Controllable Speech Synthesis.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Hierarchical Generative Modeling for Controllable Speech Synthesis

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:10.808080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:00:10.808080Z digest=sha256:b730f7b964789fc34b4acd694c10c06169c4f00b02cd6909731b398ff5914ffb

Observation 004539f0-2355-42fe-9072-d05e53e7391b · outbound

This paper cites Prosody under con- trol: Controlling prosody in text-to-speech synthesis by adjust- ments in latent reference space,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Prosody under con- trol: Controlling prosody in text-to-speech synthesis by adjust- ments in latent reference space,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:15.506225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:10.916035Z digest=sha256:c11a663f41ac0bbecb7907b236662f4c7ee11cf8dcbb4e85904aa6b22a33ee06

Observation 86511840-7bf2-40cd-b188-343e069a30e6 · outbound

This paper cites Ctrl-P: Temporal Control of Prosodic Variation for Speech Synthesis.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Ctrl-P: Temporal Control of Prosodic Variation for Speech Synthesis

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:00:13.054479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:11.005095Z digest=sha256:04689fac001eb9febf792caf7aeb6b88baaa63ff091659007ab1125643a46e37

Observation f3586622-df59-4260-90dd-f93722557864 · outbound

This paper cites Hierarchical prosody modeling and control in non-autoregressive parallel neural tts,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Hierarchical prosody modeling and control in non-autoregressive parallel neural tts,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:15.025126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:11.095853Z digest=sha256:0729578c01183286adc0e90a02bc3557b43e6124b8ca624b9f9bb3950d70dde3

Observation 1d4052e0-cd1c-4850-990a-bdb2b00d5acb · outbound

This paper cites Analysis of pronunciation learning in end-to-end speech synthesis.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Analysis of pronunciation learning in end-to-end speech synthesis

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:14.866577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:11.170821Z digest=sha256:bccdeaff31f639220b7fad5d1e5a83c2abb3eaf842323a5340dcaec3fb27a9cf

Observation a4613d82-41a0-40cc-ab7b-439523ce1418 · outbound

This paper cites Expressive prosody for unit- selection speech synthesis.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Expressive prosody for unit- selection speech synthesis

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:14.657076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:11.217114Z digest=sha256:6094251ec456e867e12ff7f09031429ba6538653a2e27e191986310ccebdb7eb

Observation 68cf2e50-5554-4ae8-9655-447e62ecce8d · outbound

This paper cites Speaking rate attention-based duration prediction for speed control TTS.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Speaking rate attention-based duration prediction for speed control TTS

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:00:12.835344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:11.312003Z digest=sha256:0c35b2361af922da1526b53ccaa042c75af034e246ff777af426bd095aadc4de

Observation 445b59ed-41b1-4406-9923-d6d1812161ea · outbound

This paper cites Speaking rate control of end-to-end tts models by direct manipulation of the encoder’s out- put embeddings,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Speaking rate control of end-to-end tts models by direct manipulation of the encoder’s out- put embeddings,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:14.461220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:11.374501Z digest=sha256:0a20f5f30d9a9394b0fd2f20f771021974e66f1d31031da9355524c8a0e0d932

Observation 0b83458a-f700-4ac0-ad80-36057ab40fb3 · outbound

This paper cites Unified Mandarin TTS Front-end Based on Distilled BERT Model.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Unified Mandarin TTS Front-end Based on Distilled BERT Model

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:11.440776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:00:11.440776Z digest=sha256:15b3213859c9b81bf2bb88c093b16ee359bbde0841d4281e757d517fcc626db5

Observation d1e7c6ee-5379-450d-b595-528859d2db11 · outbound

This paper cites Speech audio corrector: using speech from non-target speakers for one- off correction of mispronunciations in grapheme-input text-to- speech,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Speech audio corrector: using speech from non-target speakers for one- off correction of mispronunciations in grapheme-input text-to- speech,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:14.256621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:11.517200Z digest=sha256:47cad829a6cd1c5a0e195b0a1a9c8d1e98e5bc13e085f914a6bf8cf3ac0c8d0b

Observation e376def8-8962-4098-8b46-8ebfe1c77cba · outbound

This paper cites Exact prosody cloning in zero- shot multispeaker text-to-speech,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Exact prosody cloning in zero- shot multispeaker text-to-speech,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:14.115182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:11.587352Z digest=sha256:d0209610230df555ca9ed5c73a5bf2678c7ff11fc914feaec1891e9e5ba386bd

Observation f68dbea8-097f-45fe-9527-0df732a749dd · outbound

This paper cites Hubert: Self-supervised speech rep- resentation learning by masked prediction of hidden units,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Hubert: Self-supervised speech rep- resentation learning by masked prediction of hidden units,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:11.667468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:00:11.667468Z digest=sha256:5e6d765215b4e1193205d6c10c6d3314573768979b3eebcc81084b49f7970386

Observation 5d2ce58f-ac88-44b7-bbdd-7403a9ea9228 · outbound

This paper cites Predicting within and across language phoneme recognition performance of self-supervised learning speech pre-trained models.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Predicting within and across language phoneme recognition performance of self-supervised learning speech pre-trained models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:11.722250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:00:11.722250Z digest=sha256:c7ec6387b1d15ddaa890e500775a692372b3e835b34f72e74ef9da4a34c0c71e

Observation b1e1493a-f485-411c-97f1-fc0529edeae0 · outbound

This paper cites What do Neural Machine Translation Models Learn about Morphology?.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models What do Neural Machine Translation Models Learn about Morphology?

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:11.782293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:00:11.782293Z digest=sha256:382f021093aa35ec8fdd5324007ded35da6f901117c926182749a64a7eeaa9bb

Observation 4b48f0f9-6544-478c-b1d1-0d2c492e11bc · outbound

This paper cites What is one grain of sand in the desert? analyzing individual neu- rons in deep nlp models,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models What is one grain of sand in the desert? analyzing individual neu- rons in deep nlp models,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:14.001445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:11.876789Z digest=sha256:ecd0a74e549959736171aee4496159f9621e3d008a0c2405a70921061395563d

Observation 7e3f52dc-1e90-4081-86ad-cb68208c4814 · outbound

This paper cites Towards Realistic Individual Recourse and Actionable Explanations in Black-Box Decision Making Systems.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Towards Realistic Individual Recourse and Actionable Explanations in Black-Box Decision Making Systems

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:11.931116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:00:11.931116Z digest=sha256:b37684cfc9236cce0fd36ef04fe544a451495b839768d4178c79b9d65afcecf4

Observation 1952a89f-b8e6-4015-b9a0-e0f5dc6cdf65 · outbound

This paper cites Diffeomorphic counterfactuals with generative models,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Diffeomorphic counterfactuals with generative models,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:13.854760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:12.023797Z digest=sha256:7de28cee734e4e3db0a5fb9bbe205eb09c26e0d22982231489487685bfde5630

Observation 1db4c802-4710-465f-bea5-b579d3e189c3 · outbound

This paper cites beta-vae: Learning basic visual concepts with a constrained variational framework,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models beta-vae: Learning basic visual concepts with a constrained variational framework,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:13.742857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:12.065057Z digest=sha256:cea6a4f4fc4a9be905e4a391a00a74b932e1747b8d38433783ad79562a04c09f

Observation fd3bb7e5-c28d-4611-ad1d-542c4bbf2e2f · outbound

This paper cites Neural discrete represen- tation learning,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Neural discrete represen- tation learning,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:13.586825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:12.131190Z digest=sha256:c25c56d3bc68717fdbad05c45cb72bb7fc4b4798a5ca8ded6fa85407fd0c3b60

Observation 861b9cab-a2b4-4017-8617-33324a2c61ab · outbound

This paper cites The lj speech dataset,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models The lj speech dataset,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:12.209634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:00:12.209634Z digest=sha256:cebab05cdc8dfd0ab44308bfc1cdb5a702798362fc7edebf19448b7d9191a8d5

Observation f8e062e2-2d29-4f38-bb32-edd9b8a6d929 · outbound

This paper cites Large Scale GAN Training for High Fidelity Natural Image Synthesis.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Large Scale GAN Training for High Fidelity Natural Image Synthesis

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:12.264666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:00:12.264666Z digest=sha256:8564befac92bbade29407520dad6ff65637d85423d75dde1c74c6c7aa2c314aa

Observation c22909a4-cdd4-41c5-a5c2-13d369506f8e · outbound

This paper cites Robust speech recognition via large-scale weak su- pervision,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Robust speech recognition via large-scale weak su- pervision,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:12.370607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:00:12.370607Z digest=sha256:26f447b28205c6bd14038f28e3d16c88c30ea7d6358da5f711e1cca4b735b859

Observation a972daaa-7f52-4136-8cb9-cfa7279a8627 · outbound

This paper cites Language models are unsupervised multitask learners,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Language models are unsupervised multitask learners,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:12.455556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:00:12.455556Z digest=sha256:3eb40e8d47fe933f487fdbac6dad72ac32f63f91dd0ee1a588c7922676e82349

Observation 3363bde7-799e-4e73-8c61-0ef606ae6fa3 · outbound

This paper cites Speech quality assessment,.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Speech quality assessment,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:00:13.409116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:12.519114Z digest=sha256:3c26bbdee510a0c83523716c0048f155b08c39f8f1422350b63da1357b5bf85f

Pith citing papers

Observation 5638f2c9-2732-4a60-a3ac-02461578072a · inbound

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models cites this paper.

Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T12:00:13.248560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:00:09.374265Z digest=sha256:109dfb0efba06b4a14520b07c1559b07c150cbeabf21afb5a69b97f08c29b374