Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:21:04.007502Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 3 inbound Pith citation observations for arXiv:2507.05911.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:21:04.007502Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:21:03.907098Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T03:27:34.881668Z
34 of 34 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation af6f0ba1-e714-474d-85b0-3c6f96591e36 · outbound
Differentiable Reward Optimization for LLM based TTS system Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a8c3ba41-ce94-48ce-9281-c1dc4798c9f2 · outbound
Differentiable Reward Optimization for LLM based TTS system Differentiable Reward Optimization for LLM based TTS system
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d759b234-f61b-440c-a3a8-0414f5eaaf81 · outbound
Differentiable Reward Optimization for LLM based TTS system Figure 1 shows the difference between the DiffRO and the existing RL method like DPO
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3143c8d2-3fda-4e26-be7a-c317c3a78512 · outbound
Differentiable Reward Optimization for LLM based TTS system Experimental Setup
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e98d9252-e536-4fa8-936f-e4197d2273fe · outbound
Differentiable Reward Optimization for LLM based TTS system Compared to other reinforcement learning methods, DiffRO is capable of directly predicting reward scores from speech tokens rather than from synthesized audio
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fa37fdba-494b-4665-b948-5dbc05bc8b34 · outbound
Differentiable Reward Optimization for LLM based TTS system Denoising diffusion probabilistic models,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation db4d6f1f-e8d5-424b-bbba-0311e1c1fef5 · outbound
Differentiable Reward Optimization for LLM based TTS system Neural codec language models are zero-shot text to speech synthesizers,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b87f9175-72f9-4052-ba6c-4c17668cba53 · outbound
Differentiable Reward Optimization for LLM based TTS system Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3419c96d-06ef-4023-9f1f-7aa98617f27a · outbound
Differentiable Reward Optimization for LLM based TTS system Fine-Tuning Language Models from Human Preferences
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ca2a934-9b4e-4e91-b437-ec4e5b423d9a · outbound
Differentiable Reward Optimization for LLM based TTS system BATON: Aligning Text-to-Audio Model with Human Preference Feedback
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ae14a623-b5bc-4baa-b660-2e36555c476f · outbound
Differentiable Reward Optimization for LLM based TTS system Enhancing Zero-shot Text-to-Speech Synthesis with Human Feedback
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad003da6-0d82-4a6f-ae68-dc40dcb6d347 · outbound
Differentiable Reward Optimization for LLM based TTS system Robust zero- shot text-to-speech synthesis with reverse inference optimization,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 24b90341-4162-4efd-b4e5-185aefe70570 · outbound
Differentiable Reward Optimization for LLM based TTS system Cosyvoice 2: Scalable streaming speech synthesis with large language models,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0b63ce8a-c9bf-4023-b1e3-db6397667678 · outbound
Differentiable Reward Optimization for LLM based TTS system Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d68e9d5-2e5e-40cd-92bd-32082180c84d · outbound
Differentiable Reward Optimization for LLM based TTS system Emo-DPO: Controllable Emotional Speech Synthesis through Direct Preference Optimization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c62705d-c203-4ba0-a71a-92f75d48aa4b · outbound
Differentiable Reward Optimization for LLM based TTS system Proximal Policy Optimization Algorithms
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c36fff5-d410-4ec5-8f65-8e207496bd0e · outbound
Differentiable Reward Optimization for LLM based TTS system Direct Preference Optimization: Your Language Model is Secretly a Reward Model
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39a0bd8d-5ab3-46b2-920b-5f6e1f5668db · outbound
Differentiable Reward Optimization for LLM based TTS system Asq: An ultra-low bit rate asr-oriented speech quantization method,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 82b8e870-57a9-4240-8ada-7202285ac39e · outbound
Differentiable Reward Optimization for LLM based TTS system AIR-Bench: Automated Heterogeneous Information Retrieval Benchmark
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dafd7893-1b29-45a3-af82-fb00420edb9a · outbound
Differentiable Reward Optimization for LLM based TTS system CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b9ce675-5e3d-4150-8ae2-dc62834e45d7 · outbound
Differentiable Reward Optimization for LLM based TTS system Torchaudio-squim: Reference-less speech quality and intelligibility measures in torchaudio,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9cf9a8d5-836e-4f85-a99a-afb1b4a17279 · outbound
Differentiable Reward Optimization for LLM based TTS system Panns: Large-scale pretrained audio neural networks for audio pattern recognition,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 433206e1-8cd7-4efd-9f85-458541ea8b18 · outbound
Differentiable Reward Optimization for LLM based TTS system Common voice: A massively-multilingual speech corpus,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07aae18e-a3e3-4cc6-811e-2c7ab260765d · outbound
Differentiable Reward Optimization for LLM based TTS system The voicemos challenge 2022,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7ab4b5bd-b48f-4fee-9fda-fd0ec86dea26 · outbound
Differentiable Reward Optimization for LLM based TTS system Iemocap: Interactive emotional dyadic motion capture database,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f89e426-80c3-4f7b-889d-7f0670e5d05c · outbound
Differentiable Reward Optimization for LLM based TTS system Robust Speech Recognition via Large-Scale Weak Supervision
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a397a2aa-8407-4328-8ff1-11f688955eb5 · outbound
Differentiable Reward Optimization for LLM based TTS system FunAudioLLM: Voice Understanding and Generation Foundation Models for Natural Interaction Between Humans and LLMs
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 036ae5c2-1986-4023-a598-1804bcc8cea9 · outbound
Differentiable Reward Optimization for LLM based TTS system Dnsmos p.835: A non- intrusive perceptual objective speech quality metric to evaluate noise suppressors,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a913a405-6aa8-4709-99b5-ba60796d5be2 · outbound
Differentiable Reward Optimization for LLM based TTS system EmoBox: Multilingual Multi-corpus Speech Emotion Recognition Toolkit and Benchmark
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19dc18b0-34b9-4b54-9d08-4ceb984b145e · outbound
Differentiable Reward Optimization for LLM based TTS system Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9690060e-f981-4df2-969d-edbf34cee12c · outbound
Differentiable Reward Optimization for LLM based TTS system MinMo: A Multimodal Large Language Model for Seamless Voice Interaction
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cddfe44-4ecd-431a-a4b1-c23c24355f13 · outbound
Differentiable Reward Optimization for LLM based TTS system Paraformer: Fast and accurate parallel transformer for non-autoregressive end-to- end speech recognition,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1db4b377-e5fd-4c35-b9b2-d12bf1683ce2 · outbound
Differentiable Reward Optimization for LLM based TTS system F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f700fc2-640f-4295-a63f-959c6af2aac1 · outbound
Differentiable Reward Optimization for LLM based TTS system Robust Zero-Shot Text-to-Speech Synthesis with Reverse Inference Optimization
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8c3ba41-ce94-48ce-9281-c1dc4798c9f2 · inbound
Differentiable Reward Optimization for LLM based TTS system Differentiable Reward Optimization for LLM based TTS system
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5669cd96-e95c-4b59-a58c-f0a29b7d705e · inbound
Evaluating and Rewarding LALMs for Expressive Role-Play TTS via Mean Continuation Log-Probability Differentiable Reward Optimization for LLM based TTS system
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58bb72f7-cbef-49ea-b7a7-b0098ae6e966 · inbound
End-to-End Training for Discrete Token LLM based TTS System Differentiable Reward Optimization for LLM based TTS system
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.