Pith. sign in

Paper Citation Record · LEDGER

Hi-Fi Multi-Speaker English TTS Dataset

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2104.01497.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2104.01497 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T17:10:04.049487Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T10:44:07.682416Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 04d6df58-72a5-42d1-b5e3-e1d6e6e605aa · inbound

Length-Aware Rotary Position Embedding for Text-Speech Alignment cites this paper.

Length-Aware Rotary Position Embedding for Text-Speech Alignment Hi-Fi Multi-Speaker English TTS Dataset

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T17:10:04.049487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:10:04.049487Z digest=sha256:8648dc59f52d960fbc67b30e6b3017ac307198f30c57b35d449bf2adef962808

Observation 7d15d4d7-08fe-4d0f-b947-d0bf53fb9d28 · inbound

StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs cites this paper.

StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs Hi-Fi Multi-Speaker English TTS Dataset

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:01:24.417213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-18T12:57:04.450462Z digest=sha256:95db6d44e0e1931ff2b3847439e186a5678e16e2cb588f5448b244e5fd73799e

Observation 2e0b8a9f-af81-4ecb-b30a-e3aca0b1a9d7 · inbound

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning cites this paper.

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning Hi-Fi Multi-Speaker English TTS Dataset

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:35:18.340737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-15T08:34:56.898815Z digest=sha256:d688a3a6c3dd999b8a6170fc5eddda987c30d48d47516859d949465884891b8c

Observation bbcace2b-3698-4b78-9ce9-e17121cfad71 · inbound

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning cites this paper.

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning Hi-Fi Multi-Speaker English TTS Dataset

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-21T10:44:07.685113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-21T10:43:27.176535Z digest=sha256:cac0de986287740fbd065ba435b0cbe289a590222f348520db3364661003489a

Observation 9d510468-0d3f-47c4-8646-d3a90a8337c2 · inbound

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning cites this paper.

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning Hi-Fi Multi-Speaker English TTS Dataset

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-13T22:57:11.059962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:57:11.059962Z digest=sha256:190f582f2f618af8634731fadaf3b1a7eb7045ac2094dd7f7a182cb63fd37118

Observation 0c217e5c-1c13-4177-b03e-0e6e0c8a3616 · inbound

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey cites this paper.

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey Hi-Fi Multi-Speaker English TTS Dataset

Reference 198

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:30:56.907217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-10T16:36:33.264166Z digest=sha256:b2a5d6f9fc2d2954df212d6f10d968c083d0376a2956a6c36be6b2d77725047f

Observation b2aba4bf-c4f9-47c9-bea6-b8ff609a08d8 · inbound

A Survey of Advancing Audio Super-Resolution and Bandwidth Extension from Discriminative to Generative Models cites this paper.

A Survey of Advancing Audio Super-Resolution and Bandwidth Extension from Discriminative to Generative Models Hi-Fi Multi-Speaker English TTS Dataset

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:27:53.861667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-19T20:26:59.049472Z digest=sha256:c98b91442782f1e0638937b819b58773470817ee4d1aa91987620978448c6f68

Observation 0c4ec660-1768-463d-8887-34e4cb624f23 · inbound

Designed Vocalizations Dataset: Sound-Designed Human and Animal Voices for Non-human Voice Conversion cites this paper.

Designed Vocalizations Dataset: Sound-Designed Human and Animal Voices for Non-human Voice Conversion Hi-Fi Multi-Speaker English TTS Dataset

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T08:57:40.034274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:57:40.034274Z digest=sha256:6d268cb0311ea8b39536b2a1438ff6b6f5f3678ab26f657782eddafb396f9ba1