Pith. sign in

Paper Citation Record · LEDGER

F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2504.02407.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.02407 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:48:22.347437Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T12:19:49.614306Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 39b4e942-e3a7-4d56-bf27-a493a78b1194 · inbound

Flow-GRPO: Training Flow Matching Models via Online RL cites this paper.

Flow-GRPO: Training Flow Matching Models via Online RL F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:45:16.759002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T18:45:16.641012Z digest=sha256:aaf2664ea29d97402f98db72dc90b36672847561d98b07a2a13ffed4e340d2bc

Observation 48b43947-6642-48d3-a731-46443b4040d0 · inbound

CosyVoice 3: Towards In-the-wild Speech Generation via Scaling-up and Post-training cites this paper.

CosyVoice 3: Towards In-the-wild Speech Generation via Scaling-up and Post-training F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:27:25.581153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T05:27:25.425188Z digest=sha256:aa772044ccf482106d3fbc01e645c4b9ffbdb95511890873a3ec2f3aa82f84e4

Observation 3de44fb7-35b5-40b1-8310-b586fe374102 · inbound

DMOSpeech 2: Reinforcement Learning for Duration Prediction in Metric-Optimized Speech Synthesis cites this paper.

DMOSpeech 2: Reinforcement Learning for Duration Prediction in Metric-Optimized Speech Synthesis F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T15:48:22.347437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:48:22.347437Z digest=sha256:c03d59833a087829d19df3385ef3aef98f6aaff0636c38e40692b9e07ce6bc10

Observation c4e63a3a-d6be-4b11-9a66-f6f111ac4947 · inbound

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance cites this paper.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.405894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.405894Z digest=sha256:f57bb85723ff2e069e93b49efca479041f786a2240b4f842699700e8495e5f3f

Observation ebc77e3f-6077-444c-97b9-5e50df4152f5 · inbound

Group Relative Policy Optimization for Speech Recognition cites this paper.

Group Relative Policy Optimization for Speech Recognition F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T12:07:21.397960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:07:21.397960Z digest=sha256:cf81a5dea0568991a2c4421f04b1eb1cc0cf524c6ab79b3a5e7a6d6247700151

Observation 6535acbc-3aa8-4ff5-ab7e-32a92a12a424 · inbound

TIGFlow-GRPO: Trajectory Forecasting via Interaction-Aware Flow Matching and Reward-Guided Optimization cites this paper.

TIGFlow-GRPO: Trajectory Forecasting via Interaction-Aware Flow Matching and Reward-Guided Optimization F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:58:26.119603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T00:55:45.389178Z digest=sha256:874874ba6b6562d97903b7d065882451a165ee9c734251116c774136ecdf6f8c

Observation 47596943-c5bb-40eb-a453-69a0a968f8d3 · inbound

WavTTS: Towards High-Quality Zero-Shot TTS via Direct Raw Waveform Modeling cites this paper.

WavTTS: Towards High-Quality Zero-Shot TTS via Direct Raw Waveform Modeling F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-07-02T05:16:39.779566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T08:18:42.002083Z digest=sha256:9d868c28be7139fe7b18466a0a5964f7e0094f6d505cbf57ab5cfa10f21e6211

Observation b97b0ef7-a09c-45d8-83b2-b982174e8a87 · inbound

GLASS: GRPO-Trained LoRA for Acoustic Style Steering in Zero-Shot Text-to-Speech cites this paper.

GLASS: GRPO-Trained LoRA for Acoustic Style Steering in Zero-Shot Text-to-Speech F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T15:27:04.865913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T23:53:55.385445Z digest=sha256:35ae5c5f257b35f35f106a2eb984ff380f8aa282c539366fb722c127eb2c155c

Observation 56412acd-4f70-437b-abdb-91d0744fa439 · inbound

Reinforcement Learning for Flow-Matching Policies with Density Transport cites this paper.

Reinforcement Learning for Flow-Matching Policies with Density Transport F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:27:25.863031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T18:55:02.040180Z digest=sha256:5d91ecbcd944c0780dacefe3b2da2468129f6a0343b5e510775d1b4dbcf4514e

Observation 5d01e01a-226b-48d6-8a96-5d36ccdbeded · inbound

End-to-End Training for Discrete Token LLM based TTS System cites this paper.

End-to-End Training for Discrete Token LLM based TTS System F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:34.873088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T15:22:06.893507Z digest=sha256:858b39f262d5645fe70b21005066e09e11b34002001e93127bccdab04b5cb908

Observation 2fcaf6a1-8afa-4b64-bede-153e95a9c8a6 · inbound

FlowTTS-GRPO: Online Reinforcement Learning with Multi-Objective Reward Optimization for Flow-Matching Based Text-to-Speech cites this paper.

FlowTTS-GRPO: Online Reinforcement Learning with Multi-Objective Reward Optimization for Flow-Matching Based Text-to-Speech F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:19:49.615854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T07:02:36.499424Z digest=sha256:5418092e0f0ea4ba3821492d8a2656404b29a59fdd92a941fba74198fee79e8b

Observation 2923c7d3-0a40-42e7-9697-04dcf657ed6f · inbound

FlowTTS-GRPO: Online Reinforcement Learning with Multi-Objective Reward Optimization for Flow-Matching Based Text-to-Speech cites this paper.

FlowTTS-GRPO: Online Reinforcement Learning with Multi-Objective Reward Optimization for Flow-Matching Based Text-to-Speech F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-12T12:44:20.831164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:44:20.831164Z digest=sha256:17aa9cad2412468db3dbf08b3e920c9e75ec29fe278d3aab7ffd50eaf38c3a63

Observation 0a9329e8-6c1d-43ee-8698-2833321dac54 · inbound

From Objectives to Applications: Aligning Architectural Biases in Audio Self-Supervised Learning cites this paper.

From Objectives to Applications: Aligning Architectural Biases in Audio Self-Supervised Learning F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 139

Resolution
verified exact
arxiv_id, observed 2026-07-02T05:56:39.809035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-02T05:52:55.818877Z digest=sha256:cda35283cecef6e52be27322388a0b7e66fb513cd61a81355101baecec93ca34

Observation 4919c597-e41a-4cf4-9b97-aa152b580de7 · inbound

Qwen-Audio-3.0-TTS: Freely Controllable and Highly Robust Speech Synthesis with Multi-Stage Training Paradigm cites this paper.

Qwen-Audio-3.0-TTS: Freely Controllable and Highly Robust Speech Synthesis with Multi-Stage Training Paradigm F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-31T23:35:24.514221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:35:24.514221Z digest=sha256:0fe98f103eab054943e9c62e2d76227c41a0a638b20a5db8324c784e19f1a043