Pith. sign in

Paper Citation Record · LEDGER

F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2504.02407.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.02407 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:48:22.347437Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T12:19:49.614306Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 39b4e942-e3a7-4d56-bf27-a493a78b1194 · inbound

Flow-GRPO: Training Flow Matching Models via Online RL cites this paper.

Flow-GRPO: Training Flow Matching Models via Online RL F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:45:16.759002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T18:45:16.641012Z digest=sha256:a19e98f89253fea3f2e0a54304fd1170314e9970b79dbf694ce4c5e72449214b

Observation 48b43947-6642-48d3-a731-46443b4040d0 · inbound

CosyVoice 3: Towards In-the-wild Speech Generation via Scaling-up and Post-training cites this paper.

CosyVoice 3: Towards In-the-wild Speech Generation via Scaling-up and Post-training F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:27:25.581153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T05:27:25.425188Z digest=sha256:416392aa846cf9eac716bbb8eb0f290a4e5bf14438aa7ee0c48fdae6778b9bd6

Observation 3de44fb7-35b5-40b1-8310-b586fe374102 · inbound

DMOSpeech 2: Reinforcement Learning for Duration Prediction in Metric-Optimized Speech Synthesis cites this paper.

DMOSpeech 2: Reinforcement Learning for Duration Prediction in Metric-Optimized Speech Synthesis F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T15:48:22.347437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:48:22.347437Z digest=sha256:c03d59833a087829d19df3385ef3aef98f6aaff0636c38e40692b9e07ce6bc10

Observation c4e63a3a-d6be-4b11-9a66-f6f111ac4947 · inbound

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance cites this paper.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.405894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.405894Z digest=sha256:f57bb85723ff2e069e93b49efca479041f786a2240b4f842699700e8495e5f3f

Observation ebc77e3f-6077-444c-97b9-5e50df4152f5 · inbound

Group Relative Policy Optimization for Speech Recognition cites this paper.

Group Relative Policy Optimization for Speech Recognition F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T12:07:21.397960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:07:21.397960Z digest=sha256:cf81a5dea0568991a2c4421f04b1eb1cc0cf524c6ab79b3a5e7a6d6247700151

Observation 6535acbc-3aa8-4ff5-ab7e-32a92a12a424 · inbound

TIGFlow-GRPO: Trajectory Forecasting via Interaction-Aware Flow Matching and Reward-Guided Optimization cites this paper.

TIGFlow-GRPO: Trajectory Forecasting via Interaction-Aware Flow Matching and Reward-Guided Optimization F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:58:26.119603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T00:55:45.389178Z digest=sha256:b08c7d13b586548b834fe10689a67503fb0262e96e20595b9f374d2aff7cc9b4

Observation 47596943-c5bb-40eb-a453-69a0a968f8d3 · inbound

WavTTS: Towards High-Quality Zero-Shot TTS via Direct Raw Waveform Modeling cites this paper.

WavTTS: Towards High-Quality Zero-Shot TTS via Direct Raw Waveform Modeling F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-07-02T05:16:39.779566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T08:18:42.002083Z digest=sha256:9dc979ba8f29185d52e74c25e2b27b40e4e4c214b99b4d5fdc574cf0eebe00c3

Observation b97b0ef7-a09c-45d8-83b2-b982174e8a87 · inbound

GLASS: GRPO-Trained LoRA for Acoustic Style Steering in Zero-Shot Text-to-Speech cites this paper.

GLASS: GRPO-Trained LoRA for Acoustic Style Steering in Zero-Shot Text-to-Speech F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T15:27:04.865913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T23:53:55.385445Z digest=sha256:f79a83bfad9196eeb0a9e3cd0675d517a7d8b6f40c3746c5743111944f574a8f

Observation 56412acd-4f70-437b-abdb-91d0744fa439 · inbound

Reinforcement Learning for Flow-Matching Policies with Density Transport cites this paper.

Reinforcement Learning for Flow-Matching Policies with Density Transport F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:27:25.863031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T18:55:02.040180Z digest=sha256:62956c7fce056343078d17ee77221772d2ce6952dc4131e47ad2e2d5990e21f0

Observation 5d01e01a-226b-48d6-8a96-5d36ccdbeded · inbound

End-to-End Training for Discrete Token LLM based TTS System cites this paper.

End-to-End Training for Discrete Token LLM based TTS System F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:34.873088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T15:22:06.893507Z digest=sha256:a8ad3518b58af58e6516c5109b70c34b3521f0db25f1f4b0f90dad174f3aa2f8

Observation 2fcaf6a1-8afa-4b64-bede-153e95a9c8a6 · inbound

FlowTTS-GRPO: Online Reinforcement Learning with Multi-Objective Reward Optimization for Flow-Matching Based Text-to-Speech cites this paper.

FlowTTS-GRPO: Online Reinforcement Learning with Multi-Objective Reward Optimization for Flow-Matching Based Text-to-Speech F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:19:49.615854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T07:02:36.499424Z digest=sha256:8de239525235ea50c0c3bf585f259a86b2971afc96ee5ca9899fd97f3195ca9b

Observation 2923c7d3-0a40-42e7-9697-04dcf657ed6f · inbound

FlowTTS-GRPO: Online Reinforcement Learning with Multi-Objective Reward Optimization for Flow-Matching Based Text-to-Speech cites this paper.

FlowTTS-GRPO: Online Reinforcement Learning with Multi-Objective Reward Optimization for Flow-Matching Based Text-to-Speech F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-12T12:44:20.831164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:44:20.831164Z digest=sha256:17aa9cad2412468db3dbf08b3e920c9e75ec29fe278d3aab7ffd50eaf38c3a63

Observation 0a9329e8-6c1d-43ee-8698-2833321dac54 · inbound

From Objectives to Applications: Aligning Architectural Biases in Audio Self-Supervised Learning cites this paper.

From Objectives to Applications: Aligning Architectural Biases in Audio Self-Supervised Learning F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 139

Resolution
verified exact
arxiv_id, observed 2026-07-02T05:56:39.809035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-02T05:52:55.818877Z digest=sha256:0ee80542ac9e9807a0216125e02b177a5c5284abda5976c7373135a74d5b757d

Observation 4919c597-e41a-4cf4-9b97-aa152b580de7 · inbound

Qwen-Audio-3.0-TTS: Freely Controllable and Highly Robust Speech Synthesis with Multi-Stage Training Paradigm cites this paper.

Qwen-Audio-3.0-TTS: Freely Controllable and Highly Robust Speech Synthesis with Multi-Stage Training Paradigm F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-31T23:35:24.514221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:35:24.514221Z digest=sha256:0fe98f103eab054943e9c62e2d76227c41a0a638b20a5db8324c784e19f1a043