Pith. sign in

Paper Citation Record · LEDGER

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024

As of 15 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 1 inbound Pith citation observation for arXiv:2412.01100.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.01100 v2

Coverage vector

measured 33 of 33 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T04:46:23.404901Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T04:46:23.292385Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-12T04:46:23.543943Z

Reference resolution

33 of 33 outbound references displayed

  • verified exact1
  • verified fuzzy17
  • unresolved13
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1190a25e-e7b2-4c45-889b-8d3fc6a7410d · outbound

This paper cites On the one hand, TTS systems need to accurately replicate the target speakers’ voice, including their timbre, pitch, and prosody, using voice cloning techniques.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 On the one hand, TTS systems need to accurately replicate the target speakers’ voice, including their timbre, pitch, and prosody, using voice cloning techniques

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.787300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.285394Z digest=sha256:d8773e9a19e53bfdf35cff3c3d7b51a20af94a21dd10ec2d0ea19e70788ce627

Observation a5953255-227d-4271-9788-83a2cc98c164 · outbound

This paper cites The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-12T04:46:23.547969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.292385Z digest=sha256:9bfa60955099f7caff485b25a22fa849d5103d14c172238e6790a16b8f2d4809

Observation 592e28c6-cd5f-4b2a-98ac-9cdff6417d90 · outbound

This paper cites Firstly, we will overview the text represen- tations and the discrete speech tokens, and then introduce the LLaMA-based codec language model with a delay pattern.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Firstly, we will overview the text represen- tations and the discrete speech tokens, and then introduce the LLaMA-based codec language model with a delay pattern

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.775974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.296336Z digest=sha256:19402838f3771950876541942c85a2bc2146bd193dcc7770eb6c43bc1a9ca645

Observation 6f26b3b5-fbf7-41de-9a15-3bd8a680e28b · outbound

This paper cites Model Configurations We use the open-source models and parameters of MT5-base, DAC and HuBERT, with both DAC and HuBERT configured for a sampling rate of 16 kHz.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Model Configurations We use the open-source models and parameters of MT5-base, DAC and HuBERT, with both DAC and HuBERT configured for a sampling rate of 16 kHz

Reference 4

Resolution
malformed identifier
raw_fallback, observed 2026-08-12T04:46:23.764622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.300465Z digest=sha256:2d365b2d9c435c4e27bd016658fbcd4bb49e257d7ceb24bae5525999e555d0fc

Observation 01c61244-b2b2-41d0-90ba-e3291cb72429 · outbound

This paper cites 嗯” (En- glish translation: “um.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 嗯” (En- glish translation: “um

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.752707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.304430Z digest=sha256:20a9a89a0de3ca5d266a7b520d270a259ebc663a3489d65794f4cb446130684d

Observation b7fa6797-065c-45fa-8d5d-8d8bf1dff17b · outbound

This paper cites We propose a LLaMA-based codec language model with a delay pattern for spontaneous style voice cloning.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 We propose a LLaMA-based codec language model with a delay pattern for spontaneous style voice cloning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.740832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.308086Z digest=sha256:c9818ffee0b4c81fb7a2002e5a80632c19f03d9590b83f9efe22b1a54f88d801

Observation 542d5b6a-0118-423b-9794-2f06bfeb2511 · outbound

This paper cites an unresolved cited work.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-12T04:46:23.727825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.311926Z digest=sha256:5ea2a04945abdcc78bf79bf4dc2633df6a633cc2445a0311be23718c63a5e73e

Observation 139d5379-45c2-4cf5-bd1c-874b58db1a35 · outbound

This paper cites Deep voice 3: Scaling text- to-speech with convolutional sequence learning,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Deep voice 3: Scaling text- to-speech with convolutional sequence learning,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.716321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.315475Z digest=sha256:42e83f261bd479a4a89ebb933f97a1f639fb42722f21a448681dbd32635a06ee

Observation b710d84c-195f-4657-98c3-a383d04cd232 · outbound

This paper cites Neural voice cloning with a few samples,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Neural voice cloning with a few samples,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.705048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.319118Z digest=sha256:09e2da137a99d09ef811cd94577497ce8588b4c81cef5c7e32175db8450c2db4

Observation ea60877d-f449-4e08-bac4-d82e9ed42ed0 · outbound

This paper cites Adaspeech: Adaptive text to speech for custom voice,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Adaspeech: Adaptive text to speech for custom voice,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.693828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.322596Z digest=sha256:3647f50a9c778f1609a4cff415d8884c8374117120739e82ce7e29b7f00e2ad6

Observation 2be492a3-b375-428d-9fd9-c5743916fb77 · outbound

This paper cites Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.326163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.326163Z digest=sha256:0deafe7d4f2fdfa63dcbcba2e27ff90fe4559320c0e72079aa745ce710c907d4

Observation d446060d-6a77-46c0-81dd-6940d5873c3c · outbound

This paper cites Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.330505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.330505Z digest=sha256:618822ca2178e8bb591c30f66c37ef64aaf32546ce06646fb34b0900f197426d

Observation d293bd2d-c7d7-4380-9a81-fddc37b9b8b6 · outbound

This paper cites Speechx: Neural codec language model as a versatile speech transformer,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Speechx: Neural codec language model as a versatile speech transformer,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.682223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.334229Z digest=sha256:f5b0c43defbdbd2700b79574ed9a34e74abfb8be7bb9d4a121187b4100cee993

Observation 52a4f702-c67b-4136-bb99-3b2fed255e00 · outbound

This paper cites ELLA-V: Stable Neural Codec Language Modeling with Alignment-guided Sequence Reordering.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 ELLA-V: Stable Neural Codec Language Modeling with Alignment-guided Sequence Reordering

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.337892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.337892Z digest=sha256:25aea1665597ff65154ee8e74c210a89b76365c08973dadc9134f875732850c0

Observation 9536cd50-92d2-420f-b1a8-48d92887b299 · outbound

This paper cites High Fidelity Neural Audio Compression.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 High Fidelity Neural Audio Compression

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.341898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.341898Z digest=sha256:f8f890c0ff6ead436cc5f55f673c69d0e1ac8eccb165e9fed1cf63e935770c1a

Observation 56d4e3c2-ab15-4282-8e69-e1e06b5951f9 · outbound

This paper cites Soundstream: An end-to-end neural audio codec,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Soundstream: An end-to-end neural audio codec,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.345998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.345998Z digest=sha256:d8e506b71371c69a4a6a891221d84609a888e710b3e9a9c48018d1e739f03224

Observation 2336ffaf-b83f-4884-8c33-73fefbbd7960 · outbound

This paper cites Conversational end-to-end tts for voice agents,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Conversational end-to-end tts for voice agents,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.662670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.349799Z digest=sha256:993b4de442f866f742b3eb7298719748478b3b0525398a135c6768277459aef1

Observation 9ace13ad-7a94-4da0-b69b-b61de73813ed · outbound

This paper cites End-to-end text-to-speech based on latent representa- tion of speaking styles using spontaneous dialogue,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 End-to-end text-to-speech based on latent representa- tion of speaking styles using spontaneous dialogue,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.651480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.353254Z digest=sha256:66a227c65934cc25ae21bcb6ab68f73178f9f83b11d5492a74da713e1c75e36d

Observation e941ab01-b81a-4b56-82c3-6611513a1fc5 · outbound

This paper cites Spontts: modeling and transferring spontaneous style for tts,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Spontts: modeling and transferring spontaneous style for tts,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.640293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.356699Z digest=sha256:6b43e9d7922a5eaa4c37239dc32c0e938d71fdba215c44372b0b39ae95c846cc

Observation ab184b98-a3ca-4c99-96f8-ebee51f4f082 · outbound

This paper cites Controllable Context-aware Conversational Speech Synthesis.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Controllable Context-aware Conversational Speech Synthesis

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-12T04:46:23.491867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.360982Z digest=sha256:1307310e6e2479e8b24b41a6dcd67455d49a65de3b57a27183c470f6fb4d5a4f

Observation eef021a3-549c-473f-8d2f-0f8ff685f326 · outbound

This paper cites Spontaneous Style Text-to-Speech Synthesis with Controllable Spontaneous Behaviors Based on Language Models.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Spontaneous Style Text-to-Speech Synthesis with Controllable Spontaneous Behaviors Based on Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.364935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.364935Z digest=sha256:54509f71c9ab3b9fa4a86660e86498fc282b591a87f0d4638a14bc5168af726b

Observation 8198aee0-864b-4a99-aae9-39ad1aece243 · outbound

This paper cites WenetSpeech4TTS: A 12,800-hour Mandarin TTS Corpus for Large Speech Generation Model Benchmark.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 WenetSpeech4TTS: A 12,800-hour Mandarin TTS Corpus for Large Speech Generation Model Benchmark

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.368773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.368773Z digest=sha256:e33375213f41d3962acb2a840404e1eff50b340958d52a689207fd43fa4f38b2

Observation 713e2b56-1c94-4773-bb5a-85dd25ea7cc5 · outbound

This paper cites Wenetspeech: A 10000+ hours multi- domain mandarin corpus for speech recognition,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Wenetspeech: A 10000+ hours multi- domain mandarin corpus for speech recognition,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.629094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.371902Z digest=sha256:69ffd287f6d57e700e9e694c73e659260d9873c88d20abe02795a396f0ef9027

Observation d7bf9a78-9f93-4ed1-baf8-69e0a9c4cdcc · outbound

This paper cites mt5: A massively multilingual pre-trained text-to-text transformer,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 mt5: A massively multilingual pre-trained text-to-text transformer,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.617503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.374799Z digest=sha256:d6d14dce7b463cd77490e5013fa2b8db3c9c7f405492ef7652f3166b21123290

Observation a8c3f4cf-12b0-42e2-b03f-74b31a86af06 · outbound

This paper cites HuBERT: Self-Supervised Speech Rep- resentation Learning by Masked Prediction of Hidden Units,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 HuBERT: Self-Supervised Speech Rep- resentation Learning by Masked Prediction of Hidden Units,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.377892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.377892Z digest=sha256:a4240cec51c1f2e7854ed6766a2bb276763abdf240e1e99fb5fb4b345cae2084

Observation f66521be-d789-4b3c-8d13-68af8feaa5a6 · outbound

This paper cites High-fidelity audio compression with improved rvqgan,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 High-fidelity audio compression with improved rvqgan,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.599558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.381246Z digest=sha256:053a10f32eabc03c8720e996f064b49173bdcffb62beb44050968e86eb0f51bc

Observation 0f5ec975-e862-44f8-8c43-08b91f42b660 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 LLaMA: Open and Efficient Foundation Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.384154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.384154Z digest=sha256:cefb3e2b5ca38c52873a1c129d76ba1b47c81f4c47ce707b4581c245319a9649

Observation d9cbc3d8-0b36-4985-b679-2ddd1768a059 · outbound

This paper cites Flashattention: Fast and memory-efficient exact attention with io-awareness,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Flashattention: Fast and memory-efficient exact attention with io-awareness,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.589572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.387495Z digest=sha256:b26675894e27734a7614f87436daca58554fd45502f08ad70eae1d6f0486ca3d

Observation 30144310-756d-41df-ba16-52faf1b7a0e8 · outbound

This paper cites Speak, read and prompt: High-fidelity text-to-speech with min- imal supervision,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Speak, read and prompt: High-fidelity text-to-speech with min- imal supervision,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.390561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.390561Z digest=sha256:ec1b469638f277e93f1c80c12f95da4dda33e9b25c88f70d7d14c5b642cd0686

Observation b902a4c1-e7f8-48ac-8890-11e9b59feea2 · outbound

This paper cites Simple and controllable music gen- eration,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Simple and controllable music gen- eration,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.570739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.394037Z digest=sha256:728afdbf2623281dac2bbc52b27b6cac58b0d4bd7f8065cfbef803a7968caf00

Observation d6439b49-a38e-4d9c-82ea-8e7aacd6f301 · outbound

This paper cites Natural language guidance of high-fidelity text-to-speech with synthetic annotations.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Natural language guidance of high-fidelity text-to-speech with synthetic annotations

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.397620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.397620Z digest=sha256:4c87f36886dafdc368ddf5eda24eb64d30021f0fe10e4955e7432bdf8b576e44

Observation 39d733d5-23be-4d16-a63b-ccb8ab33b86a · outbound

This paper cites Stay on topic with Classifier-Free Guidance.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Stay on topic with Classifier-Free Guidance

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.401316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.401316Z digest=sha256:07917fc9ed34157a4422dde69cb3828fe550ab19bdafe200235f3f6a769a482c

Observation 724b1e42-e0be-4518-9a48-49440cf836ad · outbound

This paper cites V oxinstruct: Expressive human instruction-to-speech generation with unified multilingual codec language modelling,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 V oxinstruct: Expressive human instruction-to-speech generation with unified multilingual codec language modelling,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.559034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.404901Z digest=sha256:99224cf91cb1793ea7be20d3b89afec3c58ebc7c68c56cb01099cb9eecdc3c1a

Pith citing papers

Observation a5953255-227d-4271-9788-83a2cc98c164 · inbound

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 cites this paper.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-12T04:46:23.547969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T04:46:23.292385Z digest=sha256:9bfa60955099f7caff485b25a22fa849d5103d14c172238e6790a16b8f2d4809