Pith. sign in

Paper Citation Record · LEDGER

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

As of 20 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 8 inbound Pith citation observations for arXiv:2506.11130.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.11130 v2

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:02:44.434638Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T08:34:04.507414Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T07:56:57.825825Z

Reference resolution

40 of 40 outbound references displayed

  • verified exact1
  • verified fuzzy29
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 09713c24-144e-4d48-ac34-f892f444cc1a · outbound

This paper cites Canary-1b,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Canary-1b,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.860662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:42.793832Z digest=sha256:af13bd28f53deac060dc84c29c7b333c82f49d217e2f17f40d79d272c49be50f

Observation 662533db-9e42-442e-af6d-a128959cbbbe · outbound

This paper cites Robust speech recognition via large- scale weak supervision,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Robust speech recognition via large- scale weak supervision,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.852209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:42.848259Z digest=sha256:3a31b30624e437be4b7d06ff9bc05675b54ae48641093f82a4f7bbf308fe043a

Observation 104cc75b-bc4f-4529-8055-cebd11a13c5d · outbound

This paper cites Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:42.960494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:42.960494Z digest=sha256:0e42f8b58c0307489b490c8e19561a2549cee0da1d5a72e7552bbf4d28e1dbd0

Observation 15f02759-05b4-497e-9bda-c790d208db1e · outbound

This paper cites Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:43.054744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:43.054744Z digest=sha256:591c0ebf5a3e89c2d5e03d5743c7601c6c01b8f53058aa831c7a03389b271243

Observation 0f229948-93ad-4221-a704-877f2e96ec7b · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data LLaMA: Open and Efficient Foundation Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:43.165097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:43.165097Z digest=sha256:70a6021577088efe3d6c97789d3157c63ba9cd9c72067e86a56d98e151c7ed5b

Observation 0bda8398-d3db-4345-a7c1-161a1d882905 · outbound

This paper cites Olmo: Accelerating the science of language models,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Olmo: Accelerating the science of language models,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.843500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:43.255563Z digest=sha256:a9c5d2494794f2a26ca264dbead19fd5d24a0b6e9d604edb9c9cc34911680112

Observation 4e756a88-b01c-4fcf-83d7-f6ffcb8777cc · outbound

This paper cites Using synthetic audio to improve the recognition of out-of-vocabulary words in end-to-end asr systems,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Using synthetic audio to improve the recognition of out-of-vocabulary words in end-to-end asr systems,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.834919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:43.319915Z digest=sha256:8b54433f0668530474a2af48f7341fcbe0a89a6712ec891aebeb5f35f63e0569

Observation c3209f4f-15b8-43d3-98ec-6b8185379d5c · outbound

This paper cites Text-only domain adaptation for end-to-end asr using integrated text-to-mel-spectrogram generator,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Text-only domain adaptation for end-to-end asr using integrated text-to-mel-spectrogram generator,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.825752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:43.432373Z digest=sha256:215ee8a030b6511bc1414e13f54f735cb4622997f89d5efc72c5b654a3cab4b2

Observation 919869c1-63a6-4b95-87a9-eb9ac0d2469f · outbound

This paper cites Text is all you need: Personalizing asr models using controllable speech synthesis,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Text is all you need: Personalizing asr models using controllable speech synthesis,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.816617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:43.458564Z digest=sha256:5338500e9dbd886a6d423a4dbd9725aa740ca231b92b3f0fe40465b5964bfe0f

Observation 638f1136-88cc-470e-bc51-b5a45698d866 · outbound

This paper cites Corpus synthesis for zero-shot asr domain adaptation using large language models,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Corpus synthesis for zero-shot asr domain adaptation using large language models,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.807632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:43.572469Z digest=sha256:818facc90255740e4267a2a778867f08b824357f1ea4c63832200f07037a7fba

Observation 87178acd-1693-4f7f-b19f-03122e216dfb · outbound

This paper cites Task arithmetic can mitigate synthetic-to-real gap in automatic speech recognition,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Task arithmetic can mitigate synthetic-to-real gap in automatic speech recognition,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.798599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:43.695199Z digest=sha256:086f7ec485bebbe21521cbaf8fbb00129f493562de9a2affc8f1ca1615cc482b

Observation f653ef1d-cd9b-4cca-bdd8-d6849a7ca75f · outbound

This paper cites Enhancing low-resource asr through versatile tts: Bridging the data gap,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Enhancing low-resource asr through versatile tts: Bridging the data gap,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.789636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:43.828322Z digest=sha256:82307cbddc4ea60eb1b832d32346767157838081a01efc797696bfc546987340

Observation 1f28d616-630e-49f7-be2b-d40a4e527f94 · outbound

This paper cites Coqui tts,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Coqui tts,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.780191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:43.984291Z digest=sha256:4d8094f35c5dcbd0ca9ed92b4ee0b4b2a35ff3ff741a4668dbffc35fb2327665

Observation ac740d26-9413-4b1f-b83f-82037dfb0eff · outbound

This paper cites Matcha-tts: A fast tts architecture with conditional flow matching,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Matcha-tts: A fast tts architecture with conditional flow matching,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.770859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.138447Z digest=sha256:ed9620390c30f61ff57fcbf8160cd8e9b81785b44954a092a95dd9d622f71456

Observation 32169c15-5d06-4017-9e15-760990f98dc8 · outbound

This paper cites CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:44.246789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:44.246789Z digest=sha256:40a30cdbcc1d9d403747ff7ee3eb9b5cce85f98b0269aa41c732a7e9a6b0643f

Observation ba4d3c8e-de90-4309-8294-03b2b6fa800a · outbound

This paper cites BreezyVoice: Adapting TTS for Taiwanese Mandarin with Enhanced Polyphone Disambiguation -- Challenges and Insights.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data BreezyVoice: Adapting TTS for Taiwanese Mandarin with Enhanced Polyphone Disambiguation -- Challenges and Insights

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:44.355741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:44.355741Z digest=sha256:4751d93c1ac7b0de2a8d4785d68082964b3924036365d5a0351afaafd7d04b21

Observation d0e7526e-f689-4e97-81cc-bdf2b866b07b · outbound

This paper cites Leave no knowledge behind during knowledge distillation: Towards practical and effective knowledge distillation for code-switching asr using realistic data,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Leave no knowledge behind during knowledge distillation: Towards practical and effective knowledge distillation for code-switching asr using realistic data,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.760966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.359603Z digest=sha256:04d6e3f28f1c0265703eeb953c428db2d0310de1ba90abc6d868721c302f8547

Observation 4e6f13ba-272b-4c1a-bdee-4f22176ceb35 · outbound

This paper cites Whispering in Amharic: Fine-tuning Whisper for Low-resource Language.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Whispering in Amharic: Fine-tuning Whisper for Low-resource Language

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:44.363082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:44.363082Z digest=sha256:2a138fa73040b559c477a8d7ff941c8e7e78797e4bbd1bdc0eff9adb3ecbedda

Observation 6d857015-18fb-448d-b0ba-4585d63eb3f2 · outbound

This paper cites Whispering in norwegian: Navigating orthographic and dialectic challenges,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Whispering in norwegian: Navigating orthographic and dialectic challenges,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.750751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.366290Z digest=sha256:5370336ddb43a75089f5970778d665bf4c0b8b06fec5fe71ddba4a9b8c70b7b5

Observation 980ba760-e2d5-4d5d-bad9-30617b419090 · outbound

This paper cites Improving the Inclusivity of Dutch Speech Recognition by Fine-tuning Whisper on the JASMIN-CGN Corpus.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Improving the Inclusivity of Dutch Speech Recognition by Fine-tuning Whisper on the JASMIN-CGN Corpus

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:44.369784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:44.369784Z digest=sha256:b719a09e7afd87ce9c8939505701f152d662f5da2b762931dbeed5d245c2fbc7

Observation 934b101f-821e-41f5-bafa-13a4274c834b · outbound

This paper cites Whisper Finetuning on Nepali Language.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Whisper Finetuning on Nepali Language

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:44.373293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:44.373293Z digest=sha256:03143567a70b3a0aa9666cd6d31948e69c60d1f0cb6b6de4fb1b1ebbd2497af6

Observation 2e399dec-e9d1-410c-9f25-1cdf3804fc0b · outbound

This paper cites Efficient Adaptation of Multilingual Models for Japanese ASR.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Efficient Adaptation of Multilingual Models for Japanese ASR

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:02:44.481540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.376241Z digest=sha256:5c12eb90616d2aaee72b4b735f240f801ae1a9abd41d78f5370f6b40675b2113

Observation 7d06113f-401f-4f05-87f1-0ed9bc90cad3 · outbound

This paper cites Fine-tuning Whisper on Low-Resource Languages for Real-World Applications.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Fine-tuning Whisper on Low-Resource Languages for Real-World Applications

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:44.379377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:44.379377Z digest=sha256:72cc6cfb27d80beabc8c790469ad2c7d29357954dc8132f7c5db2e05f4528b92

Observation 7f81d8e4-e523-4ca3-9c4c-e8179e59ff0d · outbound

This paper cites The NTNU ASR system for Formosa speech recognition challenge 2023,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data The NTNU ASR system for Formosa speech recognition challenge 2023,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.741425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.382508Z digest=sha256:03851f6d57e41c1f63931ab4829ddd9f9c2061f4e12f44f42b24a50af8d30e74

Observation c062ea13-214c-407f-9d7b-e43ccb36cdb3 · outbound

This paper cites Zero resource code-switched speech benchmark using speech utterance pairs for multiple spoken languages,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Zero resource code-switched speech benchmark using speech utterance pairs for multiple spoken languages,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.730942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.385418Z digest=sha256:247b440a4b1b846ed77e002114922b67234acc10475581abadbca7f519350bcc

Observation 4a444104-473b-4f62-93c5-834d5a2416d5 · outbound

This paper cites IEEE, 2024, pp.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data IEEE, 2024, pp

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.719164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.388771Z digest=sha256:4941ada0fd1ae1f727195d87d1ff144f453f433d26b00f9af909a3e51de377cc

Observation cca4f21b-4b3f-470d-89e1-d445227237ff · outbound

This paper cites Zero-shot domain-sensitive speech recognition with prompt-conditioning fine-tuning,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Zero-shot domain-sensitive speech recognition with prompt-conditioning fine-tuning,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.709818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.392048Z digest=sha256:d6df3f044f7f44a878dacf1b162535da2141a20a489e6d34e246d735f68fe5da

Observation 72df09f8-47ad-44b5-ada4-b9e1644f1054 · outbound

This paper cites Prompting the hidden talent of web-scale speech models for zero-shot task generalization,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Prompting the hidden talent of web-scale speech models for zero-shot task generalization,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.700018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.395736Z digest=sha256:d09b96ed2d93c5df9146adc23a9e7098c9441461caf6e6e556f17aacafc485f0

Observation b80664d1-98f3-40ad-8896-11761ffe46ea · outbound

This paper cites Can whisper perform speech-based in-context learning?,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Can whisper perform speech-based in-context learning?,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.689371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.399184Z digest=sha256:5d79aca465d25c1ec340b85e1338a8b6d914542e7bf32899481b8b292d55e6d7

Observation c4ad91b3-e85a-4ecb-afac-f6d26d9cd3db · outbound

This paper cites High fidelity neural audio compression,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data High fidelity neural audio compression,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.678618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.402693Z digest=sha256:28f374634b33ff3122d984591d94d5f03e8d8e27372e9581324cd99ef5d62805

Observation 23141108-be84-401c-9425-b0ea07c0f062 · outbound

This paper cites Funcodec: A fundamental, reproducible and integrable open-source toolkit for neural speech codec,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Funcodec: A fundamental, reproducible and integrable open-source toolkit for neural speech codec,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.667779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.406127Z digest=sha256:3f158798e4faff4c6ade563fbd427c72e42ee034184ab3499451ee15cd226d9b

Observation 7a163650-412f-4999-b35a-f3a40cb377bf · outbound

This paper cites Funaudiollm: V oice understanding and generation foundation models for natural interaction between humans and llms,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Funaudiollm: V oice understanding and generation foundation models for natural interaction between humans and llms,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.657772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.409058Z digest=sha256:c19154758ccfd3bf74df828fbb4fe2de92cbf7f119060703fa16a7f8260aeefa

Observation bd0590be-51b2-47d1-91e2-d6935e2c57c8 · outbound

This paper cites Denes and E.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Denes and E

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.647512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.412025Z digest=sha256:fb497b8070368241d157624ff6781e94917fd7bdd191447203224f7175bee46a

Observation 97e7a478-4632-45aa-9323-fc8b3108d053 · outbound

This paper cites Listening while speaking: Speech chain by deep learning,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Listening while speaking: Speech chain by deep learning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.636777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.415140Z digest=sha256:699fb6e6b0f4528b9ecd5a8f72a8ff456e0ef1b35f74674ed4172ee734cc5013

Observation 89f55cb9-41a1-419c-ade4-b08e5efe1a43 · outbound

This paper cites Machine speech chain with one-shot speaker adaptation,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Machine speech chain with one-shot speaker adaptation,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.626338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.418056Z digest=sha256:01003455a541ec3a46f90f373485086df041c8daf66407d9a58d483673b4ca41

Observation 49e5124b-0bfc-4865-82c1-0145612caaf9 · outbound

This paper cites Speech chain for semi-supervised learning of japanese-english code-switching asr and tts,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Speech chain for semi-supervised learning of japanese-english code-switching asr and tts,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.616372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.421031Z digest=sha256:5dc7eab4734eedaa3baa767697315cdf94d242272a740d6809c4ea2208b3111a

Observation 8030fb06-d709-4e3f-8a2d-ce8a2a503d84 · outbound

This paper cites Montreal forced aligner [computer program],.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Montreal forced aligner [computer program],

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.605435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.424377Z digest=sha256:50fa4139d9f9a5fb5bd68469237adf078916800c8a128d9d7fa3da3122611089

Observation fb1a391c-924d-4876-884e-39b2ab741760 · outbound

This paper cites Fineweb2: A sparkling update with 1000s of languages,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Fineweb2: A sparkling update with 1000s of languages,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.595146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.427779Z digest=sha256:2a56f09107fa100160d7acf0169ce5bfc51e78b51ba2062e6b50fe17b7c7c61d

Observation 53f5a8a8-d7a1-4338-9232-31c84f2c0b5c · outbound

This paper cites Common voice: A massively-multilingual speech corpus,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Common voice: A massively-multilingual speech corpus,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:44.431369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:44.431369Z digest=sha256:a393042fa5b5271047d53d882b84e7ccd84312784f652836ab6c80b003f7567f

Observation 3317f119-64c7-471b-8ada-0066f2c58191 · outbound

This paper cites Ascend: A spontaneous chinese-english dataset for code- switching in multi-turn conversation,.

A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data Ascend: A spontaneous chinese-english dataset for code- switching in multi-turn conversation,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:44.576759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:02:44.434638Z digest=sha256:52116b02dd7120581d782dbdc10bf5edbf5c36aca801adb3d3db3747a892dd1a

Pith citing papers

Observation 66bc31a7-5e68-4c2b-ad52-848df8d5018e · inbound

ASPIRin: Action Space Projection for Interactivity-Optimized Reinforcement Learning in Full-Duplex Speech Language Models cites this paper.

ASPIRin: Action Space Projection for Interactivity-Optimized Reinforcement Learning in Full-Duplex Speech Language Models A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:35:58.880514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T17:05:45.214298Z digest=sha256:b3daa914576f11a28002ec8a887953b725b61c6d0a4ad652520ddd39376716a2

Observation 7b470afe-9f7a-4760-9ff1-a46b8f602637 · inbound

How to Leverage Synthetic Speech for LLM-Based ASR Systems? cites this paper.

How to Leverage Synthetic Speech for LLM-Based ASR Systems? A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-06-30T12:54:40.707154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T09:35:03.660671Z digest=sha256:85b9224c776d9f8d6f8bce2e594f775976ea9cb6334cca08dc50e282316288e8

Observation 3d6112c8-506e-4003-9006-343e34a6122f · inbound

How to Leverage Synthetic Speech for LLM-Based ASR Systems? cites this paper.

How to Leverage Synthetic Speech for LLM-Based ASR Systems? A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-12T11:11:08.878850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:11:08.878850Z digest=sha256:8aabb7f93c372396e623fb06fefb788d2f6bd5a4e319bae96eab12e53270fd3f

Observation 44fc9157-acf4-47d7-b5b9-92517e912a3e · inbound

Context-Aware ASR for Mandarin Technical Lectures cites this paper.

Context-Aware ASR for Mandarin Technical Lectures A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-11T09:26:54.286790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T09:26:54.286790Z digest=sha256:5e3f55dcd6a0b5d2e8ef86e3e922b20f6e77174b2009f78fb14cdbaa77465de8

Observation ac312fe2-935e-4db9-a082-1052100c4ef5 · inbound

REDDIT: Correcting Model-Generated Timestamp Drift in ASR without Forgetting via Replay-Based Distribution Editing cites this paper.

REDDIT: Correcting Model-Generated Timestamp Drift in ASR without Forgetting via Replay-Based Distribution Editing A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-07-07T14:53:55.858603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-07T14:52:35.127533Z digest=sha256:0ca6bcb7e1d25a18f75b673cbe1869c5c4caf2dd1b00017448efd5106dd72074

Observation 219c8161-cbe9-43e2-a9b9-94e4ba10c391 · inbound

REDDIT: Correcting Model-Generated Timestamp Drift in ASR without Forgetting via Replay-Based Distribution Editing cites this paper.

REDDIT: Correcting Model-Generated Timestamp Drift in ASR without Forgetting via Replay-Based Distribution Editing A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T08:34:04.507414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:34:04.507414Z digest=sha256:25e66b1ff1f9e7c72853e782861b4f92179a1012ed13f018ad9aebe652d72730

Observation 40bfac51-d6fb-42a3-9f54-3faa9e183317 · inbound

BlueMagpie-TTS: A Token-Efficient Tokenizer, Language Model, and TTS for Taiwanese-Accent Code-Switching Speech cites this paper.

BlueMagpie-TTS: A Token-Efficient Tokenizer, Language Model, and TTS for Taiwanese-Accent Code-Switching Speech A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

Reference 29

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T18:15:21.432691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-08T18:09:22.207379Z digest=sha256:05545b2ce9f518e1b6349c150c1d06d71927a5ced7f4c14d7ad75eb9610f7e48

Observation f4c1c24d-f962-49e2-a2e3-79a370005f00 · inbound

When Synthetic Speech Is All You Have: Better Call GRPO cites this paper.

When Synthetic Speech Is All You Have: Better Call GRPO A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-10T07:56:57.826984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-10T07:53:41.580597Z digest=sha256:f9a28c8fb9e478fcb5c4744a11db73c5ab4e7c26227b2e7bb4d71b7862957ee1