Pith. sign in

Paper Citation Record · LEDGER

FELLE: Autoregressive Speech Synthesis with Token-Wise Coarse-to-Fine Flow Matching

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2502.11128.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.11128 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T10:50:49.029983Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T11:41:42.733852Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2e3767e5-b00b-43a4-8dc8-8e904d53f762 · inbound

Interleaved Speech-Text Language Models for Simple Streaming Text-to-Speech Synthesis cites this paper.

Interleaved Speech-Text Language Models for Simple Streaming Text-to-Speech Synthesis FELLE: Autoregressive Speech Synthesis with Token-Wise Coarse-to-Fine Flow Matching

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T10:50:49.029983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T10:50:49.029983Z digest=sha256:7ac28055066ac7b1d4c6c8f7ca10bb632458a2acc9a09c74dbf2672f9f50da40

Observation c778f8fa-b4e8-4720-a550-e1c19f05d7d0 · inbound

DualDub: Video-to-Soundtrack Generation via Joint Speech and Background Audio Synthesis cites this paper.

DualDub: Video-to-Soundtrack Generation via Joint Speech and Background Audio Synthesis FELLE: Autoregressive Speech Synthesis with Token-Wise Coarse-to-Fine Flow Matching

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T17:46:17.295613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:46:17.295613Z digest=sha256:b576bd2533c7d832c7d520c5f4550793751359dbca2a8dc5d5a75d87e32f4f4d

Observation efbba303-682e-4bdb-9678-fe4e6a456f05 · inbound

Next Tokens Denoising for Speech Synthesis cites this paper.

Next Tokens Denoising for Speech Synthesis FELLE: Autoregressive Speech Synthesis with Token-Wise Coarse-to-Fine Flow Matching

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:27.493156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:27.493156Z digest=sha256:a5f4b1c872b7c88a9ba5a81dd16137f676da3f8bfb948b6986f5a52b2cf80b06

Observation c5b0d8c7-9cc3-4359-8b41-ce886d522b74 · inbound

CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis cites this paper.

CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis FELLE: Autoregressive Speech Synthesis with Token-Wise Coarse-to-Fine Flow Matching

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T16:01:52.818331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:01:52.818331Z digest=sha256:8ad9af39420f2b94e44d8d40fc1b35f65e9f37b8aff3b7ac877843deaaa34a1c

Observation 3882d03c-8f3b-4384-b0b7-7ac957202b0f · inbound

TTA-Bench: A Comprehensive Benchmark for Evaluating Text-to-Audio Models cites this paper.

TTA-Bench: A Comprehensive Benchmark for Evaluating Text-to-Audio Models FELLE: Autoregressive Speech Synthesis with Token-Wise Coarse-to-Fine Flow Matching

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-05T11:41:42.890538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-05T11:41:41.566443Z digest=sha256:f65ae0261198d738c75459649fc7f65119305944eb000aae3af42cc3f7a58ab1