Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:33:53.920089Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 1 inbound Pith citation observation for arXiv:2506.07646.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:33:53.920089Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:33:50.951862Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T05:33:54.309530Z
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 253b81c9-f476-4391-b006-ed4cf90f5099 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 00e469c6-d85e-457b-8842-2a142c8aa630 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 91318715-bb77-4e99-b88e-79028646b3a3 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation #”. For each accent phrase, we divide it into graphemes and TTS labels using the delimiter “|
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 63f3318d-2eeb-4add-bb3c-dd0d4fbbe38b · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation ブィ ” and “ビ
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 87001ae8-3a89-4560-8df0-9357527be1b2 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation To address the phonemic labeling issue associated with Kanji characters, we adopt a decoding strategy that incorporates dic- tionary prior knowledge
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 53b808c4-7044-48ae-b726-aa238163ed11 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation A unified front-end framework for english text-to-speech syn- thesis,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 76430cf7-6798-4e45-8f26-f2dd8121b684 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation E2 TTS: Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18ec3f15-32de-44f4-b89d-3f97b8fafb02 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47f8a2ca-2413-402c-a719-cf2fe2236001 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Fish-Speech: Leveraging Large Language Models for Advanced Multilingual Text-to-Speech Synthesis
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43a7426a-5dad-4650-b1bc-20c9987a7f33 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Impacts of input linguistic feature representation on japanese end-to-end speech synthesis,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 438178da-5627-4800-a118-739148fca03a · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation However, these methods either only focus on prosody annotation or rely on complicated pipelines
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8996299f-79b6-4f74-9010-1b9ae9acf37d · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Prosodic features con- trol by symbols as input of sequence-to-sequence acoustic mod- eling for neural tts,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c45148ba-cc31-46e1-856a-3b8fe0aaa918 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Unified Mandarin TTS Front-end Based on Distilled BERT Model
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2c70ab7f-6038-41df-857c-cc1e8240032b · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation A unified accent esti- mation method based on multi-task learning for japanese text-to- speech,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6ad4b55a-7782-415b-a446-d0522e5e5863 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation End-to-end asr to jointly predict transcriptions and linguistic annotations,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f8e3027f-d9ef-43ec-8bcd-d5f057c689de · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Audio- conditioned phonemic and prosodic annotation for building text- to-speech models from unlabeled speech data,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bdc8a950-e117-4484-9467-848140de031a · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Automatic prosody annotation with pre-trained text- speech model,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation baab77ce-9c1e-4174-ab31-64838d04a974 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation G2pa: G2p with aligned audio for mandarin chinese,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 477f577b-d313-459d-96ab-a70ded181cc7 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Robust speech recognition via large-scale weak supervision,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b7841fa-cec0-4104-b24d-fb822222db2a · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Attention is all you need,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd6e10cc-7013-4121-9300-c5939c1ff6bb · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Zero-shot domain-sensitive speech recognition with prompt- conditioning fine-tuning,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40855888-ffb2-4ddb-a055-da5ee2067c56 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Prompt tuning for speech recognition on unknown spoken name entities,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ae99b0b4-7761-4ff9-aea2-0db55d4b7418 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Extending whisper with prompt tuning to target-speaker asr,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c330a1a8-3056-4d41-9b4c-cca0e7c68aa3 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Perceiver-prompt: Flexible speaker adaptation in whisper for chinese disordered speech recognition,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 712e21e5-a28a-429e-bfd1-d6ab197e3da3 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Applying conditional random fields to japanese morphological analysis,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d758b06d-870d-4a4e-bd7e-c93fb762fde8 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation A proper approach to japanese morphological analysis: Dictionary, model, and eval- uation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b87c52ac-3c7c-42f0-b555-bede0aadabf0 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation JSUT corpus: free large-scale Japanese speech corpus for end-to-end speech synthesis
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f22a6537-8921-4163-b814-8d3cd93b2ad9 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation GPT-4 Technical Report
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66ba85f3-dd33-48a6-a3a2-306299ae381d · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation JVS corpus: free Japanese multi-speaker voice corpus
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6e6ce3c-1240-4e25-b3ea-bdf2c7fbe2dd · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Transformers: State-of-the-art natural language processing,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 96ed9fe9-0f09-466c-bbc4-47643c0ea502 · outbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00e469c6-d85e-457b-8842-2a142c8aa630 · inbound
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.