Pith. sign in

Paper Citation Record · LEDGER

Differentiable Reward Optimization for LLM based TTS system

As of 7 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 3 inbound Pith citation observations for arXiv:2507.05911.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.05911 v1

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:21:04.007502Z

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:21:03.907098Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T03:27:34.881668Z

Reference resolution

34 of 34 outbound references displayed

  • verified exact1
  • verified fuzzy12
  • unresolved20
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation af6f0ba1-e714-474d-85b0-3c6f96591e36 · outbound

This paper cites an unresolved cited work.

Differentiable Reward Optimization for LLM based TTS system Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:21:04.287642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:21:03.903629Z digest=sha256:7e154deaa38184a2d420844e04370581496f4a5fcb3eec0bd7952b65c358e06e

Observation a8c3ba41-ce94-48ce-9281-c1dc4798c9f2 · outbound

This paper cites Differentiable Reward Optimization for LLM based TTS system.

Differentiable Reward Optimization for LLM based TTS system Differentiable Reward Optimization for LLM based TTS system

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.907098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.907098Z digest=sha256:977b0b195dac69f99ae8fde83d34cbcf1837facaa03583df7d33f30c2760c7e5

Observation d759b234-f61b-440c-a3a8-0414f5eaaf81 · outbound

This paper cites Figure 1 shows the difference between the DiffRO and the existing RL method like DPO.

Differentiable Reward Optimization for LLM based TTS system Figure 1 shows the difference between the DiffRO and the existing RL method like DPO

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:21:04.279359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:21:03.910576Z digest=sha256:1de742915d06779dbed49ad6643a9dccaff1949cbf4d966435e30b4ca4965a20

Observation 3143c8d2-3fda-4e26-be7a-c317c3a78512 · outbound

This paper cites Experimental Setup.

Differentiable Reward Optimization for LLM based TTS system Experimental Setup

Reference 4

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T19:21:04.270240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:21:03.913744Z digest=sha256:58de656727ad4960ffbe5c446a715ec44cd26f00325c3b41883f98fa4a187de7

Observation e98d9252-e536-4fa8-936f-e4197d2273fe · outbound

This paper cites Compared to other reinforcement learning methods, DiffRO is capable of directly predicting reward scores from speech tokens rather than from synthesized audio.

Differentiable Reward Optimization for LLM based TTS system Compared to other reinforcement learning methods, DiffRO is capable of directly predicting reward scores from speech tokens rather than from synthesized audio

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:21:04.261574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:21:03.917226Z digest=sha256:6e6ab20a79777c9a719045136ce100ca721397c4dbc17694be703b1c1fb6c0ae

Observation fa37fdba-494b-4665-b948-5dbc05bc8b34 · outbound

This paper cites Denoising diffusion probabilistic models,.

Differentiable Reward Optimization for LLM based TTS system Denoising diffusion probabilistic models,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:21:04.252926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:21:03.920367Z digest=sha256:db98ec230c1765445111ad6a70049ef8d9eed01d43a55cc1725593c5a4e6f268

Observation db4d6f1f-e8d5-424b-bbba-0311e1c1fef5 · outbound

This paper cites Neural codec language models are zero-shot text to speech synthesizers,.

Differentiable Reward Optimization for LLM based TTS system Neural codec language models are zero-shot text to speech synthesizers,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:21:04.243318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:21:03.923447Z digest=sha256:052858ea34095c65edac737541e6d6d484f6e30b98b5888aca88c0569295ff90

Observation b87f9175-72f9-4052-ba6c-4c17668cba53 · outbound

This paper cites Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling.

Differentiable Reward Optimization for LLM based TTS system Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.927154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.927154Z digest=sha256:b93e3304b61f809c431db918b0ced63a970eb06c3b6b369096302671719580c9

Observation 3419c96d-06ef-4023-9f1f-7aa98617f27a · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

Differentiable Reward Optimization for LLM based TTS system Fine-Tuning Language Models from Human Preferences

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.930674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.930674Z digest=sha256:bb085f626ab97bb698fbd37704b84406918b7d7036837cd151e9525cf60e2763

Observation 0ca2a934-9b4e-4e91-b437-ec4e5b423d9a · outbound

This paper cites BATON: Aligning Text-to-Audio Model with Human Preference Feedback.

Differentiable Reward Optimization for LLM based TTS system BATON: Aligning Text-to-Audio Model with Human Preference Feedback

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-06T19:21:04.134137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:21:03.933807Z digest=sha256:333bb8f37fde8a8aa9ca69ab7829ad91b5e98cb78a4b9b8566b2521fc941cdcf

Observation ae14a623-b5bc-4baa-b660-2e36555c476f · outbound

This paper cites Enhancing Zero-shot Text-to-Speech Synthesis with Human Feedback.

Differentiable Reward Optimization for LLM based TTS system Enhancing Zero-shot Text-to-Speech Synthesis with Human Feedback

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.937201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.937201Z digest=sha256:6296bf6cd8f74899b0e189f256f4d320b4f3e9e78a2bf2da49c5b3d2d893f94f

Observation ad003da6-0d82-4a6f-ae68-dc40dcb6d347 · outbound

This paper cites Robust zero- shot text-to-speech synthesis with reverse inference optimization,.

Differentiable Reward Optimization for LLM based TTS system Robust zero- shot text-to-speech synthesis with reverse inference optimization,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:21:04.234285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:21:03.940391Z digest=sha256:4a15badeaafcc87a6d67efbf6c5ad54996000c424248c694db2e1ec1feb95815

Observation 24b90341-4162-4efd-b4e5-185aefe70570 · outbound

This paper cites Cosyvoice 2: Scalable streaming speech synthesis with large language models,.

Differentiable Reward Optimization for LLM based TTS system Cosyvoice 2: Scalable streaming speech synthesis with large language models,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:21:04.216021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:21:03.961084Z digest=sha256:451fa95e3d9842b0e7152a7b1027f2612e55c41f703debf18a60a021e588d22c

Observation 0b63ce8a-c9bf-4023-b1e3-db6397667678 · outbound

This paper cites Seed-TTS: A Family of High-Quality Versatile Speech Generation Models.

Differentiable Reward Optimization for LLM based TTS system Seed-TTS: A Family of High-Quality Versatile Speech Generation Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.946996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.946996Z digest=sha256:870cf86bab53ee5827d23d35950499eb765434b264ff276cdc1e51fbdcbc0fcb

Observation 6d68e9d5-2e5e-40cd-92bd-32082180c84d · outbound

This paper cites Emo-DPO: Controllable Emotional Speech Synthesis through Direct Preference Optimization.

Differentiable Reward Optimization for LLM based TTS system Emo-DPO: Controllable Emotional Speech Synthesis through Direct Preference Optimization

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.949976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.949976Z digest=sha256:3594ad4f05dc84bb2c83467ddae0d316d7212a6483fd96d6eb7849e5f529da30

Observation 0c62705d-c203-4ba0-a71a-92f75d48aa4b · outbound

This paper cites Proximal Policy Optimization Algorithms.

Differentiable Reward Optimization for LLM based TTS system Proximal Policy Optimization Algorithms

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.952970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.952970Z digest=sha256:c7199357e50f0630528ebe19a77b74d5a2f80602dfb35ca5c0bcc5b0dddc9ac8

Observation 0c36fff5-d410-4ec5-8f65-8e207496bd0e · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

Differentiable Reward Optimization for LLM based TTS system Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.955658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.955658Z digest=sha256:73325a35bc53e0e815780b99ba8776dd57353e2470e84c8920aab641f5daa1c9

Observation 39a0bd8d-5ab3-46b2-920b-5f6e1f5668db · outbound

This paper cites Asq: An ultra-low bit rate asr-oriented speech quantization method,.

Differentiable Reward Optimization for LLM based TTS system Asq: An ultra-low bit rate asr-oriented speech quantization method,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:21:04.225241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:21:03.958481Z digest=sha256:4c91d655f7e1f2bd0336ccf1735d4f2d18c38059f122df9eda6ed9e035569c8c

Observation 82b8e870-57a9-4240-8ada-7202285ac39e · outbound

This paper cites AIR-Bench: Automated Heterogeneous Information Retrieval Benchmark.

Differentiable Reward Optimization for LLM based TTS system AIR-Bench: Automated Heterogeneous Information Retrieval Benchmark

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.983187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.983187Z digest=sha256:c8af986f078c4558c05a743420618ffe03f380548f198cf8381eee123749dc9f

Observation dafd7893-1b29-45a3-af82-fb00420edb9a · outbound

This paper cites CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models.

Differentiable Reward Optimization for LLM based TTS system CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.964399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.964399Z digest=sha256:9055783421c7f9a92ff1010dc1379d6364553e145d3fd7db2100933dea3a7672

Observation 3b9ce675-5e3d-4150-8ae2-dc62834e45d7 · outbound

This paper cites Torchaudio-squim: Reference-less speech quality and intelligibility measures in torchaudio,.

Differentiable Reward Optimization for LLM based TTS system Torchaudio-squim: Reference-less speech quality and intelligibility measures in torchaudio,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:21:04.207339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:21:03.967235Z digest=sha256:f300aa13c64e3be6228e79df8fb01d276b47ca80bfdd5d142af9b78f98c42970

Observation 9cf9a8d5-836e-4f85-a99a-afb1b4a17279 · outbound

This paper cites Panns: Large-scale pretrained audio neural networks for audio pattern recognition,.

Differentiable Reward Optimization for LLM based TTS system Panns: Large-scale pretrained audio neural networks for audio pattern recognition,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:21:04.198504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:21:03.969868Z digest=sha256:1a1c3204c53f83956bc5b645f79e6da9068fc828180496a06e6aae884c2b4dad

Observation 433206e1-8cd7-4efd-9f85-458541ea8b18 · outbound

This paper cites Common voice: A massively-multilingual speech corpus,.

Differentiable Reward Optimization for LLM based TTS system Common voice: A massively-multilingual speech corpus,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.972402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.972402Z digest=sha256:e416825eb87ba4782b2b8846232eb5e3c1c922264f4a184c20df416b0c1510d8

Observation 07aae18e-a3e3-4cc6-811e-2c7ab260765d · outbound

This paper cites The voicemos challenge 2022,.

Differentiable Reward Optimization for LLM based TTS system The voicemos challenge 2022,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:21:04.184689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:21:03.975942Z digest=sha256:a7f33d7da98e708c1eb6efe1d7da3d4857b27112ff8b5c0bf2ec44bc77360c34

Observation 7ab4b5bd-b48f-4fee-9fda-fd0ec86dea26 · outbound

This paper cites Iemocap: Interactive emotional dyadic motion capture database,.

Differentiable Reward Optimization for LLM based TTS system Iemocap: Interactive emotional dyadic motion capture database,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.979488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.979488Z digest=sha256:8a3560b696201fedf2c721120367af189173a8e514719384239bb3ca4035f4a3

Observation 5f89e426-80c3-4f7b-889d-7f0670e5d05c · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

Differentiable Reward Optimization for LLM based TTS system Robust Speech Recognition via Large-Scale Weak Supervision

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:04.004618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:04.004618Z digest=sha256:db7370d767e69c46baee6955344f946b30fef833705db4ae51e13ef3d3df0b71

Observation a397a2aa-8407-4328-8ff1-11f688955eb5 · outbound

This paper cites FunAudioLLM: Voice Understanding and Generation Foundation Models for Natural Interaction Between Humans and LLMs.

Differentiable Reward Optimization for LLM based TTS system FunAudioLLM: Voice Understanding and Generation Foundation Models for Natural Interaction Between Humans and LLMs

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.987118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.987118Z digest=sha256:8eca020c61e9c6a704266b3e53c1dc973a26aadae3ab03185f5f7d7c5344a577

Observation 036ae5c2-1986-4023-a598-1804bcc8cea9 · outbound

This paper cites Dnsmos p.835: A non- intrusive perceptual objective speech quality metric to evaluate noise suppressors,.

Differentiable Reward Optimization for LLM based TTS system Dnsmos p.835: A non- intrusive perceptual objective speech quality metric to evaluate noise suppressors,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:21:04.172299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:21:03.989942Z digest=sha256:5610cdcb9489ec5a8d716c4b085255467fe727ab50105bf34cd1933697d93c8e

Observation a913a405-6aa8-4709-99b5-ba60796d5be2 · outbound

This paper cites EmoBox: Multilingual Multi-corpus Speech Emotion Recognition Toolkit and Benchmark.

Differentiable Reward Optimization for LLM based TTS system EmoBox: Multilingual Multi-corpus Speech Emotion Recognition Toolkit and Benchmark

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.992535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.992535Z digest=sha256:bdd967352d7a64ced106f1cd1a2bfc93300c1822a685ee517c327db89d848798

Observation 19dc18b0-34b9-4b54-9d08-4ceb984b145e · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

Differentiable Reward Optimization for LLM based TTS system Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.995596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.995596Z digest=sha256:7f220cc40f724dcff0082464fa2a74b29284004cade749a2ad5855899909a7be

Observation 9690060e-f981-4df2-969d-edbf34cee12c · outbound

This paper cites MinMo: A Multimodal Large Language Model for Seamless Voice Interaction.

Differentiable Reward Optimization for LLM based TTS system MinMo: A Multimodal Large Language Model for Seamless Voice Interaction

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.998457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.998457Z digest=sha256:8d5efcf0570a6672413a8e02981140e33ff3ff2199827202b3f7d7ad246942cc

Observation 2cddfe44-4ecd-431a-a4b1-c23c24355f13 · outbound

This paper cites Paraformer: Fast and accurate parallel transformer for non-autoregressive end-to- end speech recognition,.

Differentiable Reward Optimization for LLM based TTS system Paraformer: Fast and accurate parallel transformer for non-autoregressive end-to- end speech recognition,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:21:04.163819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:21:04.002036Z digest=sha256:90ec8907d5b3f59e76738697bfbf1928f67d804b31698473c1c96c5a1ceeac54

Observation 1db4b377-e5fd-4c35-b9b2-d12bf1683ce2 · outbound

This paper cites F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching.

Differentiable Reward Optimization for LLM based TTS system F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:04.007502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:04.007502Z digest=sha256:86ba5d8ab030c1b2c68e98046fa87044afa2a7997acaf2165508b6a3145ed94e

Observation 9f700fc2-640f-4295-a63f-959c6af2aac1 · outbound

This paper cites Robust Zero-Shot Text-to-Speech Synthesis with Reverse Inference Optimization.

Differentiable Reward Optimization for LLM based TTS system Robust Zero-Shot Text-to-Speech Synthesis with Reverse Inference Optimization

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.943678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.943678Z digest=sha256:2b52b0b6a98d6fa38506afae5655bc0e05b362c5915aa6817edecd98d8dc08bd

Pith citing papers

Observation a8c3ba41-ce94-48ce-9281-c1dc4798c9f2 · inbound

Differentiable Reward Optimization for LLM based TTS system cites this paper.

Differentiable Reward Optimization for LLM based TTS system Differentiable Reward Optimization for LLM based TTS system

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:03.907098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:03.907098Z digest=sha256:977b0b195dac69f99ae8fde83d34cbcf1837facaa03583df7d33f30c2760c7e5

Observation 5669cd96-e95c-4b59-a58c-f0a29b7d705e · inbound

Evaluating and Rewarding LALMs for Expressive Role-Play TTS via Mean Continuation Log-Probability cites this paper.

Evaluating and Rewarding LALMs for Expressive Role-Play TTS via Mean Continuation Log-Probability Differentiable Reward Optimization for LLM based TTS system

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T06:36:32.240560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:36:32.240560Z digest=sha256:f49e1dc177e3a77570836147a5a2604c9bc784b436c271c73df880efeff545ac

Observation 58bb72f7-cbef-49ea-b7a7-b0098ae6e966 · inbound

End-to-End Training for Discrete Token LLM based TTS System cites this paper.

End-to-End Training for Discrete Token LLM based TTS System Differentiable Reward Optimization for LLM based TTS system

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:34.883195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T15:22:06.893507Z digest=sha256:04ce7ca94976b4c6d94a5cf76d7801090320c1156819e459d27345ba50f9492a