Pith. sign in

Paper Citation Record · LEDGER

CLaM-TTS: Improving Neural Codec Language Model for Zero-Shot Text-to-Speech

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2404.02781.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.02781 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T19:54:58.579073Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T00:04:22.476503Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 61224e10-6883-4e89-a553-dd944a8b051c · inbound

Representing Speech Through Autoregressive Prediction of Cochlear Tokens cites this paper.

Representing Speech Through Autoregressive Prediction of Cochlear Tokens CLaM-TTS: Improving Neural Codec Language Model for Zero-Shot Text-to-Speech

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T19:54:58.579073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:54:58.579073Z digest=sha256:93b633ef40e185f0d25b2ef2a836fa5dd2db05c95948edbaf210679c974f650c

Observation bc03816e-8704-4a65-a3bf-6d7772f6b3be · inbound

CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis cites this paper.

CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis CLaM-TTS: Improving Neural Codec Language Model for Zero-Shot Text-to-Speech

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T16:01:52.654964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:01:52.654964Z digest=sha256:8e5057272ed9d51bb25e05ea4280469d7aec375211274410a2c4cb7a81fdbaae

Observation 27fb6ae3-5b2b-4e9e-831a-be692cc5504b · inbound

FlexiSLM: A Dynamic and Controllable Frame Rate Spoken Language Model cites this paper.

FlexiSLM: A Dynamic and Controllable Frame Rate Spoken Language Model CLaM-TTS: Improving Neural Codec Language Model for Zero-Shot Text-to-Speech

Reference 125

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T11:55:42.231853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-07-01T03:50:26.873406Z digest=sha256:3e540134e2885c5d79d055a00dae7e5372aba4474cd3727db8f7d8ab662c7a45

Observation d7204a67-1ae4-421e-bb6d-28bd395d5ceb · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence CLaM-TTS: Improving Neural Codec Language Model for Zero-Shot Text-to-Speech

Reference 242

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T00:04:22.477746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-07-07T23:59:38.702609Z digest=sha256:63ac37b69e6d051edee7a347b24ae833cc3d44004d93efef15d8853d1488ec6c

Observation 27536e25-fa58-4fb3-b0af-1ee74b11c232 · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence CLaM-TTS: Improving Neural Codec Language Model for Zero-Shot Text-to-Speech

Reference 242

Resolution
unresolved
no resolver link, observed 2026-07-11T07:46:49.059192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T07:46:49.059192Z digest=sha256:6020331f6c5cb4d5d82205c37036cea67324b2ea7528f5eb6569917511010388