Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T21:32:17.120091Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 4 inbound Pith citation observations for arXiv:2501.04644.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T21:32:17.120091Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:25:11.587593Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T02:49:24.559099Z
15 of 15 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b2b1e617-6d4d-4a45-aa9f-191bda070c2d · outbound
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts BASE TTS: Lessons from building a billion-parameter Text-to-Speech model on 100K hours of data
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49384579-9107-4709-b7a5-bcd2bcc0c23b · outbound
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts emotion2vec: Self-Supervised Pre-Training for Speech Emotion Representation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b20f6d86-58f1-4b03-b427-5fe86fac6322 · outbound
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87d9f7e2-56b7-46f9-80fe-ff1f47f62436 · outbound
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts Audiobox: Unified Audio Generation with Natural Language Prompts
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 579e16a7-2a35-4fc1-845f-3364f0e92b3b · outbound
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts Kazuki Yamauchi, Yusuke Ijima, and Yuki Saito
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 555b47d4-9723-4506-aee1-1e1c24e5b16e · outbound
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c433296-5a26-40d5-a1b9-1a81bfce6c5b · outbound
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts In ACM Multimedia, pages 7513–7522
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6c7ba811-c705-41a7-9356-61b26b9b4c27 · outbound
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts fast speaking rate
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c54bc1df-086a-417b-84d4-0d471703f3ce · outbound
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts Emotional End-to-End Neural Speech Synthesizer
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8a02476-6325-48d8-8033-4140a9d3bf73 · outbound
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts Unresolved cited work
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 480d6631-a213-4098-a07e-240c853555d6 · outbound
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2d2c83e-cb37-490d-9f91-737d913f5b85 · outbound
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts In ICASSP 2023-2023 IEEE Inter- national Conference on Acoustics, Speech and Signal Processing (ICASSP), pages 1–5
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3e5a84b6-e6bc-49b4-8250-4be61db26bdc · outbound
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts ID-Animator: Zero-Shot Identity-Preserving Human Video Generation
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0632dcb1-5a7c-42aa-82bc-85664d44cfd6 · outbound
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 215e5912-a577-4fab-9ca8-de53817efa26 · outbound
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts Natural language guidance of high-fidelity text-to-speech with synthetic annotations
Reference 7771
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa975243-39c4-4a5b-8925-b279af81baf4 · inbound
MultiActor-Audiobook: Zero-Shot Audiobook Generation with Faces and Voices of Multiple Speakers FleSpeech: Flexibly Controllable Speech Generation with Various Prompts
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea35b21e-fc12-4662-bab0-88f46af95e88 · inbound
JIS: A Speech Corpus of Japanese Idol Speakers with Various Speaking Styles FleSpeech: Flexibly Controllable Speech Generation with Various Prompts
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec264d95-654f-440a-9c15-2b2f25920f8a · inbound
IndexTTS2: A Breakthrough in Emotionally Expressive and Duration-Controlled Auto-Regressive Zero-Shot Text-to-Speech FleSpeech: Flexibly Controllable Speech Generation with Various Prompts
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a402457-476e-405b-bc98-8b35a98c947a · inbound
FineCombo-TTS: Collaborative and Precise Controllable Speech Synthesis Using Text Descriptions and Reference Speech FleSpeech: Flexibly Controllable Speech Generation with Various Prompts
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.