Pith. sign in

Paper Citation Record · LEDGER

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024

As of 16 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 1 inbound Pith citation observation for arXiv:2412.01100.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.01100 v2

Coverage vector

measured 33 of 33 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T04:46:23.404901Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T04:46:23.292385Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-12T04:46:23.543943Z

Reference resolution

33 of 33 outbound references displayed

  • verified exact1
  • verified fuzzy17
  • unresolved13
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1190a25e-e7b2-4c45-889b-8d3fc6a7410d · outbound

This paper cites On the one hand, TTS systems need to accurately replicate the target speakers’ voice, including their timbre, pitch, and prosody, using voice cloning techniques.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 On the one hand, TTS systems need to accurately replicate the target speakers’ voice, including their timbre, pitch, and prosody, using voice cloning techniques

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.787300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.285394Z digest=sha256:2a33b1c1ec1e0ed09552f255a4a535618350a88b559ed4a3ad05c35da171720e

Observation a5953255-227d-4271-9788-83a2cc98c164 · outbound

This paper cites The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-12T04:46:23.547969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.292385Z digest=sha256:6b031ddb3f25805284a01e20f898def7591b8fec56406a7c0a0b650188e14785

Observation 592e28c6-cd5f-4b2a-98ac-9cdff6417d90 · outbound

This paper cites Firstly, we will overview the text represen- tations and the discrete speech tokens, and then introduce the LLaMA-based codec language model with a delay pattern.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Firstly, we will overview the text represen- tations and the discrete speech tokens, and then introduce the LLaMA-based codec language model with a delay pattern

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.775974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.296336Z digest=sha256:7bc73c3495afc8c02eb5ec147929e926c34d39af617108f86c29b8888045ee6c

Observation 6f26b3b5-fbf7-41de-9a15-3bd8a680e28b · outbound

This paper cites Model Configurations We use the open-source models and parameters of MT5-base, DAC and HuBERT, with both DAC and HuBERT configured for a sampling rate of 16 kHz.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Model Configurations We use the open-source models and parameters of MT5-base, DAC and HuBERT, with both DAC and HuBERT configured for a sampling rate of 16 kHz

Reference 4

Resolution
malformed identifier
raw_fallback, observed 2026-08-12T04:46:23.764622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.300465Z digest=sha256:0b810ed241c92ce851a1d6e565e7833ed919b11e3a21e2c29f8bcd4c23f74c7e

Observation 01c61244-b2b2-41d0-90ba-e3291cb72429 · outbound

This paper cites 嗯” (En- glish translation: “um.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 嗯” (En- glish translation: “um

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.752707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.304430Z digest=sha256:ce80825a55361ff170cecb671a305c903f4c2e125189857dc4412c2acc528d08

Observation b7fa6797-065c-45fa-8d5d-8d8bf1dff17b · outbound

This paper cites We propose a LLaMA-based codec language model with a delay pattern for spontaneous style voice cloning.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 We propose a LLaMA-based codec language model with a delay pattern for spontaneous style voice cloning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.740832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.308086Z digest=sha256:f43eef1088af1e3d17e7b6d4170a6b88b34dd091f62c2f6bc3890afd348a7cad

Observation 542d5b6a-0118-423b-9794-2f06bfeb2511 · outbound

This paper cites an unresolved cited work.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-12T04:46:23.727825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.311926Z digest=sha256:09762209b685a4a44861e9fa333a52d9b75554141d9eb5d3f566f17a78f62353

Observation 139d5379-45c2-4cf5-bd1c-874b58db1a35 · outbound

This paper cites Deep voice 3: Scaling text- to-speech with convolutional sequence learning,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Deep voice 3: Scaling text- to-speech with convolutional sequence learning,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.716321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.315475Z digest=sha256:232a81e096a9e519025eda188499fb2de8d42318355d893fca7504ff29f1378c

Observation b710d84c-195f-4657-98c3-a383d04cd232 · outbound

This paper cites Neural voice cloning with a few samples,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Neural voice cloning with a few samples,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.705048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.319118Z digest=sha256:3c9aeaee0b821d2420d89805ccf559f9bf5a32a1c42db95c74293770703832b7

Observation ea60877d-f449-4e08-bac4-d82e9ed42ed0 · outbound

This paper cites Adaspeech: Adaptive text to speech for custom voice,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Adaspeech: Adaptive text to speech for custom voice,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.693828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.322596Z digest=sha256:98b407f77e19dadc964134b23c70a438e21833ed0a97e34e615aaec9762a1488

Observation 2be492a3-b375-428d-9fd9-c5743916fb77 · outbound

This paper cites Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.326163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.326163Z digest=sha256:0deafe7d4f2fdfa63dcbcba2e27ff90fe4559320c0e72079aa745ce710c907d4

Observation d446060d-6a77-46c0-81dd-6940d5873c3c · outbound

This paper cites Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.330505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.330505Z digest=sha256:618822ca2178e8bb591c30f66c37ef64aaf32546ce06646fb34b0900f197426d

Observation d293bd2d-c7d7-4380-9a81-fddc37b9b8b6 · outbound

This paper cites Speechx: Neural codec language model as a versatile speech transformer,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Speechx: Neural codec language model as a versatile speech transformer,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.682223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.334229Z digest=sha256:7ba739e4e32c4325245a16958ae8fcf54bc3a07808de852916a92a042f13dd79

Observation 52a4f702-c67b-4136-bb99-3b2fed255e00 · outbound

This paper cites ELLA-V: Stable Neural Codec Language Modeling with Alignment-guided Sequence Reordering.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 ELLA-V: Stable Neural Codec Language Modeling with Alignment-guided Sequence Reordering

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.337892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.337892Z digest=sha256:25aea1665597ff65154ee8e74c210a89b76365c08973dadc9134f875732850c0

Observation 9536cd50-92d2-420f-b1a8-48d92887b299 · outbound

This paper cites High Fidelity Neural Audio Compression.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 High Fidelity Neural Audio Compression

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.341898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.341898Z digest=sha256:0eb6469bebee121e5b4999d19d9b8ce2eead0d9b25efd032b720d9d0692210ed

Observation 56d4e3c2-ab15-4282-8e69-e1e06b5951f9 · outbound

This paper cites Soundstream: An end-to-end neural audio codec,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Soundstream: An end-to-end neural audio codec,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.345998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.345998Z digest=sha256:d8e506b71371c69a4a6a891221d84609a888e710b3e9a9c48018d1e739f03224

Observation 2336ffaf-b83f-4884-8c33-73fefbbd7960 · outbound

This paper cites Conversational end-to-end tts for voice agents,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Conversational end-to-end tts for voice agents,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.662670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.349799Z digest=sha256:fe42837acd5ccae109505e3af49f7d6e32cc751c0d17fac3d8503043eb9d7798

Observation 9ace13ad-7a94-4da0-b69b-b61de73813ed · outbound

This paper cites End-to-end text-to-speech based on latent representa- tion of speaking styles using spontaneous dialogue,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 End-to-end text-to-speech based on latent representa- tion of speaking styles using spontaneous dialogue,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.651480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.353254Z digest=sha256:8e0088394ca3bcdc756977b53d4f9ddcb62a78c376eddba0ede26a75c8820afc

Observation e941ab01-b81a-4b56-82c3-6611513a1fc5 · outbound

This paper cites Spontts: modeling and transferring spontaneous style for tts,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Spontts: modeling and transferring spontaneous style for tts,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.640293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.356699Z digest=sha256:7f63d4405de466ed97611db1f20d46a55068f51147b988deb35d0893855554c5

Observation ab184b98-a3ca-4c99-96f8-ebee51f4f082 · outbound

This paper cites Controllable Context-aware Conversational Speech Synthesis.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Controllable Context-aware Conversational Speech Synthesis

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-12T04:46:23.491867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.360982Z digest=sha256:3d8e50939a9e56dc02fb2f106f79b76c2d3ca141edc78b401ea6cfe43e5a0e53

Observation eef021a3-549c-473f-8d2f-0f8ff685f326 · outbound

This paper cites Spontaneous Style Text-to-Speech Synthesis with Controllable Spontaneous Behaviors Based on Language Models.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Spontaneous Style Text-to-Speech Synthesis with Controllable Spontaneous Behaviors Based on Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.364935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.364935Z digest=sha256:c4378c0db480efd05b1b229a59bd2b4265461ca44e9a93f3618b2f8e0d81b499

Observation 8198aee0-864b-4a99-aae9-39ad1aece243 · outbound

This paper cites WenetSpeech4TTS: A 12,800-hour Mandarin TTS Corpus for Large Speech Generation Model Benchmark.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 WenetSpeech4TTS: A 12,800-hour Mandarin TTS Corpus for Large Speech Generation Model Benchmark

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.368773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.368773Z digest=sha256:e33375213f41d3962acb2a840404e1eff50b340958d52a689207fd43fa4f38b2

Observation 713e2b56-1c94-4773-bb5a-85dd25ea7cc5 · outbound

This paper cites Wenetspeech: A 10000+ hours multi- domain mandarin corpus for speech recognition,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Wenetspeech: A 10000+ hours multi- domain mandarin corpus for speech recognition,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.629094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.371902Z digest=sha256:f37cf6d8d26cf36533f8f3c8d7a191b13e1c11aa4fa4aea3d3d0aec4f4c1fc89

Observation d7bf9a78-9f93-4ed1-baf8-69e0a9c4cdcc · outbound

This paper cites mt5: A massively multilingual pre-trained text-to-text transformer,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 mt5: A massively multilingual pre-trained text-to-text transformer,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.617503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.374799Z digest=sha256:7be777d6aa75a640644e77785395b988a06f7096eb0377f6146c7f4e5814f7c5

Observation a8c3f4cf-12b0-42e2-b03f-74b31a86af06 · outbound

This paper cites HuBERT: Self-Supervised Speech Rep- resentation Learning by Masked Prediction of Hidden Units,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 HuBERT: Self-Supervised Speech Rep- resentation Learning by Masked Prediction of Hidden Units,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.377892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.377892Z digest=sha256:a4240cec51c1f2e7854ed6766a2bb276763abdf240e1e99fb5fb4b345cae2084

Observation f66521be-d789-4b3c-8d13-68af8feaa5a6 · outbound

This paper cites High-fidelity audio compression with improved rvqgan,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 High-fidelity audio compression with improved rvqgan,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.599558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.381246Z digest=sha256:0e4119ff1f257cb1abc48211090bd55f0a9045d57a5ad2e3f4035a979c26a871

Observation 0f5ec975-e862-44f8-8c43-08b91f42b660 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 LLaMA: Open and Efficient Foundation Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.384154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.384154Z digest=sha256:cefb3e2b5ca38c52873a1c129d76ba1b47c81f4c47ce707b4581c245319a9649

Observation d9cbc3d8-0b36-4985-b679-2ddd1768a059 · outbound

This paper cites Flashattention: Fast and memory-efficient exact attention with io-awareness,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Flashattention: Fast and memory-efficient exact attention with io-awareness,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.589572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.387495Z digest=sha256:b192b39b5206ab5e82385ddf73ce912c7e20b6e4b33ad208336deeafea4e9b38

Observation 30144310-756d-41df-ba16-52faf1b7a0e8 · outbound

This paper cites Speak, read and prompt: High-fidelity text-to-speech with min- imal supervision,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Speak, read and prompt: High-fidelity text-to-speech with min- imal supervision,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.390561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.390561Z digest=sha256:ec1b469638f277e93f1c80c12f95da4dda33e9b25c88f70d7d14c5b642cd0686

Observation b902a4c1-e7f8-48ac-8890-11e9b59feea2 · outbound

This paper cites Simple and controllable music gen- eration,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Simple and controllable music gen- eration,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.570739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.394037Z digest=sha256:f334e8d62c318c9abb592a8ca20dbdbd77b89ffb771e98d347f5befc65b9ae6c

Observation d6439b49-a38e-4d9c-82ea-8e7aacd6f301 · outbound

This paper cites Natural language guidance of high-fidelity text-to-speech with synthetic annotations.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Natural language guidance of high-fidelity text-to-speech with synthetic annotations

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.397620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.397620Z digest=sha256:4c87f36886dafdc368ddf5eda24eb64d30021f0fe10e4955e7432bdf8b576e44

Observation 39d733d5-23be-4d16-a63b-ccb8ab33b86a · outbound

This paper cites Stay on topic with Classifier-Free Guidance.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Stay on topic with Classifier-Free Guidance

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T04:46:23.401316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:46:23.401316Z digest=sha256:07917fc9ed34157a4422dde69cb3828fe550ab19bdafe200235f3f6a769a482c

Observation 724b1e42-e0be-4518-9a48-49440cf836ad · outbound

This paper cites V oxinstruct: Expressive human instruction-to-speech generation with unified multilingual codec language modelling,.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 V oxinstruct: Expressive human instruction-to-speech generation with unified multilingual codec language modelling,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:46:23.559034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.404901Z digest=sha256:ab2ff13026108cc2c880462154c7d94015e524deb31d7fe53b90f90442ae574b

Pith citing papers

Observation a5953255-227d-4271-9788-83a2cc98c164 · inbound

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 cites this paper.

The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-12T04:46:23.547969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T04:46:23.292385Z digest=sha256:6b031ddb3f25805284a01e20f898def7591b8fec56406a7c0a0b650188e14785