Pith. sign in

Paper Citation Record · LEDGER

EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2504.12867.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.12867 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:04:48.933028Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T02:49:24.567639Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e8030ec9-5c4e-4c8b-b9de-b51f1d09328a · inbound

MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts cites this paper.

MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T20:04:48.933028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:04:48.933028Z digest=sha256:147bd375b47fda964fd3e98d49b0921af5b64f5561c636469991e87ed5904a39

Observation 8e58e325-e656-4e23-9637-e05c7bfe76a8 · inbound

Semantic-Aware Ship Detection with Vision-Language Integration cites this paper.

Semantic-Aware Ship Detection with Vision-Language Integration EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T17:41:58.565216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:41:58.565216Z digest=sha256:759e005ce56d7492360f4dadce04cffb991177e03545b7f64a83b803ec491881

Observation dd4eee83-4785-4c3c-8e77-e5edad12565d · inbound

Enhancing Conversational TTS with Cascaded Prompting and ICL-Based Online Reinforcement Learning cites this paper.

Enhancing Conversational TTS with Cascaded Prompting and ICL-Based Online Reinforcement Learning EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:01:01.853953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T16:50:59.950646Z digest=sha256:ec0a7524db152a7c7ec262c80eb89838ec4a7a248ece0e359536048bc5a040e0

Observation 4d172644-6192-4f82-b2db-ff95e80a089b · inbound

UniVocal: Unified Speech-Singing Code-Switching Synthesis cites this paper.

UniVocal: Unified Speech-Singing Code-Switching Synthesis EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

Reference 77

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T00:46:24.542484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-28T13:17:13.510587Z digest=sha256:3008f5e0cc7f3ff804ea2de2cc9ab9bd71ea5f81d5afcb2c4239033d8242cda1

Observation ea96c1bd-869f-4af7-8758-c490e8ab2cf9 · inbound

FineCombo-TTS: Collaborative and Precise Controllable Speech Synthesis Using Text Descriptions and Reference Speech cites this paper.

FineCombo-TTS: Collaborative and Precise Controllable Speech Synthesis Using Text Descriptions and Reference Speech EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-04T02:49:24.569112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T19:09:33.605232Z digest=sha256:8d500b485c494ee2da10db35edee858deb8113e41de531c5b6369f6cf2f00e48

Observation 7ffc3279-1555-42dc-b45b-541efe36508d · inbound

EmoInstruct-TTS: Dual-Path Instruction-Guided Emotional Speech Synthesis cites this paper.

EmoInstruct-TTS: Dual-Path Instruction-Guided Emotional Speech Synthesis EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:47:30.568695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T17:01:13.972071Z digest=sha256:729c5aa35f1723024a4237679fd38950fdd3965fb1f1ea5de8c608c6151fcae1