Pith. sign in

Paper Citation Record · LEDGER

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

As of 20 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 8 inbound Pith citation observations for arXiv:2506.11130.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.11130 v2

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:02:44.434638Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T08:34:04.507414Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T07:56:57.825825Z

Reference resolution

40 of 40 outbound references displayed

  • verified exact1
  • verified fuzzy29
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 09713c24-144e-4d48-ac34-f892f444cc1a · outbound

This paper cites Canary-1b,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Canary-1b,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.860662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:42.793832Z digest=sha256:5380213bdd1da05803a3f02d658b90568d6e91ce9e08c66d9c0c7caedb8efad6

Observation 662533db-9e42-442e-af6d-a128959cbbbe · outbound

This paper cites Robust speech recognition via large- scale weak supervision,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Robust speech recognition via large- scale weak supervision,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.852209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:42.848259Z digest=sha256:93b053afca60df577af4f062a420f521aba93fe6952f87060e7c10c45ed66593

Observation 104cc75b-bc4f-4529-8055-cebd11a13c5d · outbound

This paper cites Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:42.960494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:42.960494Z digest=sha256:0e42f8b58c0307489b490c8e19561a2549cee0da1d5a72e7552bbf4d28e1dbd0

Observation 15f02759-05b4-497e-9bda-c790d208db1e · outbound

This paper cites Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:43.054744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:43.054744Z digest=sha256:591c0ebf5a3e89c2d5e03d5743c7601c6c01b8f53058aa831c7a03389b271243

Observation 0f229948-93ad-4221-a704-877f2e96ec7b · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data LLaMA: Open and Efficient Foundation Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:43.165097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:43.165097Z digest=sha256:70a6021577088efe3d6c97789d3157c63ba9cd9c72067e86a56d98e151c7ed5b

Observation 0bda8398-d3db-4345-a7c1-161a1d882905 · outbound

This paper cites Olmo: Accelerating the science of language models,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Olmo: Accelerating the science of language models,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.843500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:43.255563Z digest=sha256:b8c7acfb339fcbf5062c6e1be07784734cb29e66cc5d5ce893d657dd0e3f30ed

Observation 4e756a88-b01c-4fcf-83d7-f6ffcb8777cc · outbound

This paper cites Using synthetic audio to improve the recognition of out-of-vocabulary words in end-to-end asr systems,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Using synthetic audio to improve the recognition of out-of-vocabulary words in end-to-end asr systems,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.834919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:43.319915Z digest=sha256:fac60415e2b2b82681cb5580301cc103594eee714a572dd5e7bff723a82fe7f6

Observation c3209f4f-15b8-43d3-98ec-6b8185379d5c · outbound

This paper cites Text-only domain adaptation for end-to-end asr using integrated text-to-mel-spectrogram generator,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Text-only domain adaptation for end-to-end asr using integrated text-to-mel-spectrogram generator,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.825752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:43.432373Z digest=sha256:34626f55871184a6209ff00084b80a63755ef81551d7bbf0e80399a14cb82908

Observation 919869c1-63a6-4b95-87a9-eb9ac0d2469f · outbound

This paper cites Text is all you need: Personalizing asr models using controllable speech synthesis,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Text is all you need: Personalizing asr models using controllable speech synthesis,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.816617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:43.458564Z digest=sha256:192a29e72930f062482a5ae2fc5b1d0b18d061bbeb51b16733cc528d317626ad

Observation 638f1136-88cc-470e-bc51-b5a45698d866 · outbound

This paper cites Corpus synthesis for zero-shot asr domain adaptation using large language models,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Corpus synthesis for zero-shot asr domain adaptation using large language models,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.807632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:43.572469Z digest=sha256:17980294556d553b9db842d856b06811df88e1064cfc59cc40c5b2e204a2fa5e

Observation 87178acd-1693-4f7f-b19f-03122e216dfb · outbound

This paper cites Task arithmetic can mitigate synthetic-to-real gap in automatic speech recognition,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Task arithmetic can mitigate synthetic-to-real gap in automatic speech recognition,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.798599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:43.695199Z digest=sha256:ab85c2ae1cd4574bb8d4e8929c37901a208028c9b9e00a11e5401e2a3cd9b406

Observation f653ef1d-cd9b-4cca-bdd8-d6849a7ca75f · outbound

This paper cites Enhancing low-resource asr through versatile tts: Bridging the data gap,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Enhancing low-resource asr through versatile tts: Bridging the data gap,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.789636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:43.828322Z digest=sha256:7c7983331c57b6b472508117228d7bb617e57da79e7f179d86e681149d72e8ec

Observation 1f28d616-630e-49f7-be2b-d40a4e527f94 · outbound

This paper cites Coqui tts,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Coqui tts,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.780191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:43.984291Z digest=sha256:98dc277f040caffd0fb5b2b83173c294b07fff3f016e7bdd34a0fe8c0fb9d151

Observation ac740d26-9413-4b1f-b83f-82037dfb0eff · outbound

This paper cites Matcha-tts: A fast tts architecture with conditional flow matching,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Matcha-tts: A fast tts architecture with conditional flow matching,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.770859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.138447Z digest=sha256:e62f711347ef8ad0cec4cf1dc17a373d5182e53f0eff1b5c97fd31ca2f9df511

Observation 32169c15-5d06-4017-9e15-760990f98dc8 · outbound

This paper cites CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:44.246789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:44.246789Z digest=sha256:40a30cdbcc1d9d403747ff7ee3eb9b5cce85f98b0269aa41c732a7e9a6b0643f

Observation ba4d3c8e-de90-4309-8294-03b2b6fa800a · outbound

This paper cites BreezyVoice: Adapting TTS for Taiwanese Mandarin with Enhanced Polyphone Disambiguation -- Challenges and Insights.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data BreezyVoice: Adapting TTS for Taiwanese Mandarin with Enhanced Polyphone Disambiguation -- Challenges and Insights

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:44.355741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:44.355741Z digest=sha256:4751d93c1ac7b0de2a8d4785d68082964b3924036365d5a0351afaafd7d04b21

Observation d0e7526e-f689-4e97-81cc-bdf2b866b07b · outbound

This paper cites Leave no knowledge behind during knowledge distillation: Towards practical and effective knowledge distillation for code-switching asr using realistic data,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Leave no knowledge behind during knowledge distillation: Towards practical and effective knowledge distillation for code-switching asr using realistic data,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.760966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.359603Z digest=sha256:5b77befd7c1d78753b972d0d442ec28aa646060f54bd70fd3535afb224095d22

Observation 4e6f13ba-272b-4c1a-bdee-4f22176ceb35 · outbound

This paper cites Whispering in Amharic: Fine-tuning Whisper for Low-resource Language.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Whispering in Amharic: Fine-tuning Whisper for Low-resource Language

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:44.363082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:44.363082Z digest=sha256:2a138fa73040b559c477a8d7ff941c8e7e78797e4bbd1bdc0eff9adb3ecbedda

Observation 6d857015-18fb-448d-b0ba-4585d63eb3f2 · outbound

This paper cites Whispering in norwegian: Navigating orthographic and dialectic challenges,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Whispering in norwegian: Navigating orthographic and dialectic challenges,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.750751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.366290Z digest=sha256:d63980f12c6c6fb863b41d083445419f1633a93170ecfe45ec72ddcac10d3f9e

Observation 980ba760-e2d5-4d5d-bad9-30617b419090 · outbound

This paper cites Improving the Inclusivity of Dutch Speech Recognition by Fine-tuning Whisper on the JASMIN-CGN Corpus.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Improving the Inclusivity of Dutch Speech Recognition by Fine-tuning Whisper on the JASMIN-CGN Corpus

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:44.369784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:44.369784Z digest=sha256:b719a09e7afd87ce9c8939505701f152d662f5da2b762931dbeed5d245c2fbc7

Observation 934b101f-821e-41f5-bafa-13a4274c834b · outbound

This paper cites Whisper Finetuning on Nepali Language.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Whisper Finetuning on Nepali Language

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:44.373293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:44.373293Z digest=sha256:03143567a70b3a0aa9666cd6d31948e69c60d1f0cb6b6de4fb1b1ebbd2497af6

Observation 2e399dec-e9d1-410c-9f25-1cdf3804fc0b · outbound

This paper cites Efficient Adaptation of Multilingual Models for Japanese ASR.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Efficient Adaptation of Multilingual Models for Japanese ASR

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:02:44.481540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.376241Z digest=sha256:6d7d0c67939c15508f6a7dfff3f8d8865719d970904839d88cb6055246228c03

Observation 7d06113f-401f-4f05-87f1-0ed9bc90cad3 · outbound

This paper cites Fine-tuning Whisper on Low-Resource Languages for Real-World Applications.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Fine-tuning Whisper on Low-Resource Languages for Real-World Applications

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:44.379377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:44.379377Z digest=sha256:72cc6cfb27d80beabc8c790469ad2c7d29357954dc8132f7c5db2e05f4528b92

Observation 7f81d8e4-e523-4ca3-9c4c-e8179e59ff0d · outbound

This paper cites The NTNU ASR system for Formosa speech recognition challenge 2023,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data The NTNU ASR system for Formosa speech recognition challenge 2023,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.741425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.382508Z digest=sha256:9fa1c62f4e2ad0f1dfd120b5083975afb38ed48612c6a4d26f1120284a3facc0

Observation c062ea13-214c-407f-9d7b-e43ccb36cdb3 · outbound

This paper cites Zero resource code-switched speech benchmark using speech utterance pairs for multiple spoken languages,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Zero resource code-switched speech benchmark using speech utterance pairs for multiple spoken languages,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.730942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.385418Z digest=sha256:464d5b78550134a95d0e4b826a78749d38f1ad78f8bef8be4d501d62d125752c

Observation 4a444104-473b-4f62-93c5-834d5a2416d5 · outbound

This paper cites IEEE, 2024, pp.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data IEEE, 2024, pp

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.719164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.388771Z digest=sha256:e047c0b1b4edcb4ec05e1f2ce559e285c14ec8ee68d13aec042a763267b92760

Observation cca4f21b-4b3f-470d-89e1-d445227237ff · outbound

This paper cites Zero-shot domain-sensitive speech recognition with prompt-conditioning fine-tuning,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Zero-shot domain-sensitive speech recognition with prompt-conditioning fine-tuning,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.709818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.392048Z digest=sha256:e1037b24d44e3a318ea015e270e258cc0899c410ada7559f9a2d731e9187efff

Observation 72df09f8-47ad-44b5-ada4-b9e1644f1054 · outbound

This paper cites Prompting the hidden talent of web-scale speech models for zero-shot task generalization,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Prompting the hidden talent of web-scale speech models for zero-shot task generalization,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.700018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.395736Z digest=sha256:a75f47494acbe1c7687e4722b287fc6dfde861954d25cfea6b33d5c0f9838a9e

Observation b80664d1-98f3-40ad-8896-11761ffe46ea · outbound

This paper cites Can whisper perform speech-based in-context learning?,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Can whisper perform speech-based in-context learning?,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.689371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.399184Z digest=sha256:d059fdd41174078bac95f92c302354183cd52637cc1acf46e3130c8a8b138cdc

Observation c4ad91b3-e85a-4ecb-afac-f6d26d9cd3db · outbound

This paper cites High fidelity neural audio compression,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data High fidelity neural audio compression,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.678618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.402693Z digest=sha256:0dca7a306d4b02ad75749038dcaf876467ed2db6da2299934cadb874e6fd66b2

Observation 23141108-be84-401c-9425-b0ea07c0f062 · outbound

This paper cites Funcodec: A fundamental, reproducible and integrable open-source toolkit for neural speech codec,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Funcodec: A fundamental, reproducible and integrable open-source toolkit for neural speech codec,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.667779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.406127Z digest=sha256:010f69b6c8235978c7451de60ecfdf9270f112c3e07d0a9d446fb361a74cae17

Observation 7a163650-412f-4999-b35a-f3a40cb377bf · outbound

This paper cites Funaudiollm: V oice understanding and generation foundation models for natural interaction between humans and llms,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Funaudiollm: V oice understanding and generation foundation models for natural interaction between humans and llms,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.657772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.409058Z digest=sha256:64109f20b2db27359cda548970be8776d6ccc98b507148f807ece3721f3aa7af

Observation bd0590be-51b2-47d1-91e2-d6935e2c57c8 · outbound

This paper cites Denes and E.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Denes and E

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.647512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.412025Z digest=sha256:aefbf7b30d0028dcf75f9dadcf0749f004f75382fd1d587ef820920d4380e993

Observation 97e7a478-4632-45aa-9323-fc8b3108d053 · outbound

This paper cites Listening while speaking: Speech chain by deep learning,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Listening while speaking: Speech chain by deep learning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.636777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.415140Z digest=sha256:0bea2505abc68c4e0b7380698efe96ca58215c465318806adecc1d36fe1967e2

Observation 89f55cb9-41a1-419c-ade4-b08e5efe1a43 · outbound

This paper cites Machine speech chain with one-shot speaker adaptation,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Machine speech chain with one-shot speaker adaptation,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.626338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.418056Z digest=sha256:012b7e560eff4d53567ec921848efdf6c704b4596c442144f01e2c03ba7b20de

Observation 49e5124b-0bfc-4865-82c1-0145612caaf9 · outbound

This paper cites Speech chain for semi-supervised learning of japanese-english code-switching asr and tts,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Speech chain for semi-supervised learning of japanese-english code-switching asr and tts,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.616372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.421031Z digest=sha256:0c6493e124055e6ddf9c6bac890d93c070d26d49bdafb08f11043a49908b5b6d

Observation 8030fb06-d709-4e3f-8a2d-ce8a2a503d84 · outbound

This paper cites Montreal forced aligner [computer program],.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Montreal forced aligner [computer program],

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.605435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.424377Z digest=sha256:2f3642ee757efea249d7c4118e9bc880fe5aa9a2a1678de50193eaca7509c32b

Observation fb1a391c-924d-4876-884e-39b2ab741760 · outbound

This paper cites Fineweb2: A sparkling update with 1000s of languages,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Fineweb2: A sparkling update with 1000s of languages,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.595146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.427779Z digest=sha256:21a4511c848a99f73c9c10acf5691d2fd68d4d9b34a8bf3c2bc091ad652b20ab

Observation 53f5a8a8-d7a1-4338-9232-31c84f2c0b5c · outbound

This paper cites Common voice: A massively-multilingual speech corpus,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Common voice: A massively-multilingual speech corpus,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:44.431369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:44.431369Z digest=sha256:a393042fa5b5271047d53d882b84e7ccd84312784f652836ab6c80b003f7567f

Observation 3317f119-64c7-471b-8ada-0066f2c58191 · outbound

This paper cites Ascend: A spontaneous chinese-english dataset for code- switching in multi-turn conversation,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Ascend: A spontaneous chinese-english dataset for code- switching in multi-turn conversation,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.576759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:02:44.434638Z digest=sha256:9854738feec3744313c06ff1c731aa37b7b8db682f6fa02435b5e7853f682dfe

Pith citing papers

Observation 66bc31a7-5e68-4c2b-ad52-848df8d5018e · inbound

ASPIRin: Action Space Projection for Interactivity-Optimized Reinforcement Learning in Full-Duplex Speech Language Models cites this paper.

ASPIRin: Action Space Projection for Interactivity-Optimized Reinforcement Learning in Full-Duplex Speech Language Models A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:35:58.880514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T17:05:45.214298Z digest=sha256:edbb55753c05607ee86796c527de3d76e88f7a33e4a0266b34d90e01c65fcd77

Observation 7b470afe-9f7a-4760-9ff1-a46b8f602637 · inbound

How to Leverage Synthetic Speech for LLM-Based ASR Systems? cites this paper.

How to Leverage Synthetic Speech for LLM-Based ASR Systems? A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-06-30T12:54:40.707154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T09:35:03.660671Z digest=sha256:fd559e66a7ace2ccd7f4b67565ed1e8de22c929b734f370394003f9db89411f2

Observation 3d6112c8-506e-4003-9006-343e34a6122f · inbound

How to Leverage Synthetic Speech for LLM-Based ASR Systems? cites this paper.

How to Leverage Synthetic Speech for LLM-Based ASR Systems? A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-12T11:11:08.878850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:11:08.878850Z digest=sha256:8aabb7f93c372396e623fb06fefb788d2f6bd5a4e319bae96eab12e53270fd3f

Observation 44fc9157-acf4-47d7-b5b9-92517e912a3e · inbound

Context-Aware ASR for Mandarin Technical Lectures cites this paper.

Context-Aware ASR for Mandarin Technical Lectures A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-11T09:26:54.286790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T09:26:54.286790Z digest=sha256:5e3f55dcd6a0b5d2e8ef86e3e922b20f6e77174b2009f78fb14cdbaa77465de8

Observation ac312fe2-935e-4db9-a082-1052100c4ef5 · inbound

REDDIT: Correcting Model-Generated Timestamp Drift in ASR without Forgetting via Replay-Based Distribution Editing cites this paper.

REDDIT: Correcting Model-Generated Timestamp Drift in ASR without Forgetting via Replay-Based Distribution Editing A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-07-07T14:53:55.858603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-07T14:52:35.127533Z digest=sha256:d2f17a02780055738541adea5a389650950da2d0bb6322e43bde1a048d25e060

Observation 219c8161-cbe9-43e2-a9b9-94e4ba10c391 · inbound

REDDIT: Correcting Model-Generated Timestamp Drift in ASR without Forgetting via Replay-Based Distribution Editing cites this paper.

REDDIT: Correcting Model-Generated Timestamp Drift in ASR without Forgetting via Replay-Based Distribution Editing A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T08:34:04.507414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:34:04.507414Z digest=sha256:25e66b1ff1f9e7c72853e782861b4f92179a1012ed13f018ad9aebe652d72730

Observation 40bfac51-d6fb-42a3-9f54-3faa9e183317 · inbound

BlueMagpie-TTS: A Token-Efficient Tokenizer, Language Model, and TTS for Taiwanese-Accent Code-Switching Speech cites this paper.

BlueMagpie-TTS: A Token-Efficient Tokenizer, Language Model, and TTS for Taiwanese-Accent Code-Switching Speech A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

Reference 29

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T18:15:21.432691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-08T18:09:22.207379Z digest=sha256:5f296c6c45a1bc91e7846122e11a56c50b12483c0f7ce0994d23cdecbf2e8411

Observation f4c1c24d-f962-49e2-a2e3-79a370005f00 · inbound

When Synthetic Speech Is All You Have: Better Call GRPO cites this paper.

When Synthetic Speech Is All You Have: Better Call GRPO A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-10T07:56:57.826984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-10T07:53:41.580597Z digest=sha256:23cd4774fbc11a1eef000035db14feba1c538fd2fc61aae7d431746ef96fab3b