Pith. sign in

Paper Citation Record · LEDGER

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information

As of 19 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2505.15667.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15667 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:16:45.290095Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:16:42.871348Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T15:16:45.530174Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact1
  • verified fuzzy26
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5c1a7c5a-f82f-48c0-99d1-dadd05f77c3e · outbound

This paper cites an unresolved cited work.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:50.090889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:42.724908Z digest=sha256:08c4561722d019bdd732be0f4cca865870db014c51fcf9f0f82c067ed879976a

Observation c65e3999-5a0d-403a-9d95-7cd30b799efe · outbound

This paper cites Are you kidding me?!.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Are you kidding me?!

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:49.909689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:42.799145Z digest=sha256:6d4a5d24c4133a99b814b6a8a5a1481c741bc4b31d8173761a847b7ad4cc3bfa

Observation 9bbcac71-74b0-4c85-aebf-96dc035cd577 · outbound

This paper cites Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:16:45.591375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:42.871348Z digest=sha256:320ccadf5fb415e2cf64919460a09e08b5437359db81678f8b656096f5a3a8e4

Observation 88a1359c-8961-4727-9298-ab4f85eb3114 · outbound

This paper cites Segmentation-Variant Codebooks We encode speech into continuous representations using the frozen HuBERT-large model [17].

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Segmentation-Variant Codebooks We encode speech into continuous representations using the frozen HuBERT-large model [17]

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:49.763315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:42.966193Z digest=sha256:8016d290f08b886e8acc30e2dd453239038fb26ded7ba40c86edfe4591853d25

Observation 5ae054a2-2fe4-47d5-81a6-1155e9aade0f · outbound

This paper cites Datasets and Alignment Process Naver Prosody Control [21] is a dataset designed for study- ing prosody control in TTS systems.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Datasets and Alignment Process Naver Prosody Control [21] is a dataset designed for study- ing prosody control in TTS systems

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:49.572574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:43.000083Z digest=sha256:0db4e47f13efd5798633b5a1bf118f7138163e7fc3197e9ebad6d1a636f21285

Observation 795a7989-4c17-4e90-b2f6-b2e979a45480 · outbound

This paper cites an unresolved cited work.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:49.350000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:43.040761Z digest=sha256:a2669add8e330e301a4684cd2f4f931d803c661d8b73f12cad5fc29cf5f16944

Observation ece5708e-0da5-4225-88ec-d0e7b934f5fb · outbound

This paper cites an unresolved cited work.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:49.060304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:43.092557Z digest=sha256:4c9ae416e1efb6c8ca38343246ac4a69ae86d7ddad087f69f448de49d2958d60

Observation 0db052a4-8a80-459f-a483-56bdb38dbcff · outbound

This paper cites The higher the score, the better.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information The higher the score, the better

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.907843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:43.197519Z digest=sha256:463a35e950881de1d07c9b01b6f0bd6ecb358fe0f2d42ff6c2a979f9b47d1c6c

Observation 6e272bc9-7c89-410f-a156-539b597a0daa · outbound

This paper cites This study contributes to ongoing research on the use of DSUs, demonstrating their po- tential in speech representation learning and downstream tasks.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information This study contributes to ongoing research on the use of DSUs, demonstrating their po- tential in speech representation learning and downstream tasks

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.689125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:43.270137Z digest=sha256:187de9b2f0290b766f244d87abfca684bded9e7b70d46009a6c59d75fbca2a31

Observation 88d2db23-b85e-426a-9398-377d80e93580 · outbound

This paper cites an unresolved cited work.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:48.496560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:43.351918Z digest=sha256:16ac88d8738e7f50b0165abebff88c176081af959c8fa27e013c994a7ce24e4f

Observation 3f653aa5-3026-4b40-a1f9-f0df0b20afb4 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:43.399696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:43.399696Z digest=sha256:c167189b7221a87f6f24214283dcdad41d707cdc0acf43c64d762a2f01b8282f

Observation 57c35fe4-7054-47c1-9511-7031d4396df1 · outbound

This paper cites Exploring speech recognition, translation, and understanding with discrete speech units: A com- parative study,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Exploring speech recognition, translation, and understanding with discrete speech units: A com- parative study,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.391693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:43.437788Z digest=sha256:634af16bd1067d4996dd63d9465e878c9bef40a6e21f7436502069868075182c

Observation 8b8bae58-1fa5-4470-8ac8-5a7e55fd40a7 · outbound

This paper cites On the use of self-supervised speech representations in spontaneous speech synthesis,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information On the use of self-supervised speech representations in spontaneous speech synthesis,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.247460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:43.481755Z digest=sha256:6e22e56377ddba67e1d2c2176638c65109de1dc66fd75308d2c4163a94957cc0

Observation 35b2e03c-73d3-467e-af37-12c1cad7df1d · outbound

This paper cites Selecttts: Syn- thesizing anyone’s voice via discrete unit-based frame selection,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Selecttts: Syn- thesizing anyone’s voice via discrete unit-based frame selection,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.126458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:43.540400Z digest=sha256:a0117f8a8994717c80a3beac78f91704d7a5aacc3993a6e5c89d2bae7af9cb7c

Observation 72b2e005-e738-4c3b-a928-d33964c18e1f · outbound

This paper cites Exploration of a self- supervised speech model: A study on emotional corpora,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Exploration of a self- supervised speech model: A study on emotional corpora,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.995252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:43.626702Z digest=sha256:9727285789214b2a11b7c75c85303c6adc703dfb2d1df6fdb22824329d34313f

Observation b188d8a3-253b-459b-801c-2ca318becb79 · outbound

This paper cites Exploring Wav2vec 2.0 fine tun- ing for improved speech emotion recognition,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Exploring Wav2vec 2.0 fine tun- ing for improved speech emotion recognition,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.865776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:43.724297Z digest=sha256:d6f1d830765adf76e0cec000993a58dc51238eff2cfc686e82c47a44e2ceb503

Observation d36e856a-ce21-4066-b3b0-096d6b906caa · outbound

This paper cites Layer-wise analysis of self-supervised acoustic word embeddings: A study on speech emotion recognition,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Layer-wise analysis of self-supervised acoustic word embeddings: A study on speech emotion recognition,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.707895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:43.797913Z digest=sha256:d874ffe1d1aad58b70bc76f14f4ed3c0d2174676a8c6ac21cf8ac4f05dd00ae6

Observation d14d4f7f-73dd-4beb-aead-02aa54898e94 · outbound

This paper cites Crossmodal ASR error correc- tion with discrete speech units,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Crossmodal ASR error correc- tion with discrete speech units,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.557302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:43.897140Z digest=sha256:d5243c6e3d3a934c844d339f3bba2db825f16173dc6e7bff87009dd2424ec29b

Observation 9c85d235-3631-4b2a-8c01-d735bd2b04d9 · outbound

This paper cites DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:43.941022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:43.941022Z digest=sha256:b618a864e0ec7c8058f0ed34a5c32b279d6f145b5f91b4fb908afd4133bfd2ef

Observation 8c79a01d-fdc2-4364-a74b-257daa0a1297 · outbound

This paper cites Analyzing acoustic word embeddings from pre-trained self-supervised speech models,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Analyzing acoustic word embeddings from pre-trained self-supervised speech models,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.423247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:43.974819Z digest=sha256:c6815637753d5df192f6e6b2a9707b101aaf54d33e8c6c97870def49bd08a0c5

Observation a11d4300-c46c-4f5a-90ea-c8b1a0c6d2a0 · outbound

This paper cites Ex- presso: A benchmark and analysis of discrete expressive speech resynthesis,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Ex- presso: A benchmark and analysis of discrete expressive speech resynthesis,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.278145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:44.049595Z digest=sha256:cb7886a498f681e03bf058c2837f2b3d4c9f627cea30810b8848a24251342c6a

Observation 6b5ad4b3-752d-4953-a95f-8457fd82eab2 · outbound

This paper cites Emo-codec: An in-depth look at emotion preservation capacity of legacy and neural codec models with subjective and objective evaluations,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Emo-codec: An in-depth look at emotion preservation capacity of legacy and neural codec models with subjective and objective evaluations,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.171102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:44.115212Z digest=sha256:e2808d3b33dca28f42f914d174e5773a084b314c1076ff0713fe1dbb6ce9c934

Observation bbf3e34f-2dda-491a-a99d-b68783173e64 · outbound

This paper cites Neural discrete represen- tation learning,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Neural discrete represen- tation learning,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.156454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.156454Z digest=sha256:78b99664426bfe955cab834ed2ec39f5514a594d31a413b4eacb224195579d21

Observation 4c0be852-d92c-423b-88e8-88c50b532cce · outbound

This paper cites vq-wav2vec: Self- supervised learning of discrete speech representations,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information vq-wav2vec: Self- supervised learning of discrete speech representations,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.047351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:44.199984Z digest=sha256:35786aaf567ea0173684bf65ef902a815cf4798356aa4e1deffe84fc95d23265

Observation 8eafc58b-0597-44a9-ac22-ff1ceba174d8 · outbound

This paper cites High Fidelity Neural Audio Compression.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information High Fidelity Neural Audio Compression

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.240679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.240679Z digest=sha256:5b3e629d2aac6b797aaeeeb3c71bf1bb50183e2d766c778d003a3f1dba7f7dc3

Observation 4d15db86-6185-4789-8f19-c1e0798d2143 · outbound

This paper cites Soundstream: An end-to-end neural audio codec,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Soundstream: An end-to-end neural audio codec,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.302788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.302788Z digest=sha256:525e6e25c128140186a9c97495d76ff471bc396524d2ebfdc6a1f5a174d25530

Observation a408aa65-4d83-4929-b89a-a68074fb7725 · outbound

This paper cites Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.430963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.430963Z digest=sha256:f331db663be2b67deb1d95371a113f2ecb04ab233ae335f9c9e76c882c4fcda9

Observation a7b29f47-7768-4684-8867-28bd5bc6838c · outbound

This paper cites Phonetic analysis of self-supervised representations of english speech,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Phonetic analysis of self-supervised representations of english speech,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.901418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:44.504728Z digest=sha256:794a4c8c69cf3abbbfe20083a7387f933d4e7ec56235a6562e765a04df1967ef

Observation 7d56b6eb-9f5e-466b-a8d7-2fe03012f0e7 · outbound

This paper cites Analysing discrete self supervised speech representation for spoken language modeling,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Analysing discrete self supervised speech representation for spoken language modeling,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.791664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:44.560689Z digest=sha256:45a897b103774e2da5b5e20dc9aa8a2174eb37c23294cfc643695562c6ca918a

Observation 89422f5a-da23-452f-81c1-21090ab9d992 · outbound

This paper cites The Interspeech 2024 Challenge on Speech Processing Using Discrete Units.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information The Interspeech 2024 Challenge on Speech Processing Using Discrete Units

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.632634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.632634Z digest=sha256:bb7c301204fb410702dbacf7a667ce3eb20fb630f63cc4f9da74252543f05958

Observation 65ef997c-b806-4173-ae93-9404b37028de · outbound

This paper cites Controlling prosody in end-to-end TTS: A case study on contrastive focus generation,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Controlling prosody in end-to-end TTS: A case study on contrastive focus generation,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.672049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:44.708653Z digest=sha256:e708680dc81a64de0d2b2177638425084696ddad4aaf840a7429725b82a0cd21

Observation 11f3a1a4-3e78-4702-b64c-88b285f5efd8 · outbound

This paper cites IEMOCAP: Interactive emotional dyadic motion capture database,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information IEMOCAP: Interactive emotional dyadic motion capture database,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.767109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.767109Z digest=sha256:3f81e370990b318eb5b8072b554b69342688a4a756a306922c24d26321cf456a

Observation d0956122-9443-4441-8c7b-08d9fa0bf926 · outbound

This paper cites Montreal forced aligner: Trainable text-speech align- ment using Kaldi,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Montreal forced aligner: Trainable text-speech align- ment using Kaldi,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.487028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:44.798029Z digest=sha256:a9e2f1d403ab049cb8dfa47c79e7d080adedecc8cde0102c064e5c3e85795166

Observation 1afc9f5f-e468-4347-ba72-1dba800550b7 · outbound

This paper cites The HTK book,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information The HTK book,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.363396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:44.842896Z digest=sha256:d690f7d2aa632c3ffe1809cdad932f2fc62ac69ea42ed6773300062a988ea97c

Observation c9acfe29-2a41-421d-9171-ed54aad14be2 · outbound

This paper cites k-means++: The advantages of careful seeding,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information k-means++: The advantages of careful seeding,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.900505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.900505Z digest=sha256:02fb97b94181fa23519147ee4f9cbb8481a0cb97641ee9fac369d33ef1063106

Observation e277a6ef-0e09-4456-a720-351c89d10147 · outbound

This paper cites Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:44.972228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:44.972228Z digest=sha256:2bf7c19afb9e9411ecc26db505309239b4fa41cecb5fb409ea64a76d6eec9362

Observation fa69722a-ed3e-4821-98c7-1861a4afbdca · outbound

This paper cites Su- perb: Speech processing universal performance benchmark,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Su- perb: Speech processing universal performance benchmark,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.199145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:45.013346Z digest=sha256:44307e4bf04b9b536983a384040e76de847a5b1a3575522b0d937ba0fbff3458

Observation e54207ef-db25-4e0b-9637-4f8ba3a6ea1d · outbound

This paper cites A Fine-tuned Wav2vec 2.0/HuBERT Benchmark For Speech Emotion Recognition, Speaker Verification and Spoken Language Understanding.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information A Fine-tuned Wav2vec 2.0/HuBERT Benchmark For Speech Emotion Recognition, Speaker Verification and Spoken Language Understanding

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:45.044605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:45.044605Z digest=sha256:98748f221ddfc66ff7a4d735241c63561c2d42d601dd766496314560f4e812d9

Observation 910719cb-7426-431a-a585-091dc3b4fa4f · outbound

This paper cites Speech emotion di- arization: Which emotion appears when?.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Speech emotion di- arization: Which emotion appears when?

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.101884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:45.083406Z digest=sha256:1b0e98977d3d1da75a926073a9e4719602b789ba6ba9832700dfcb9819664c06

Observation 88baefd7-917e-467f-99c3-cda2827ab1bc · outbound

This paper cites EmphAssess : a Prosodic Benchmark on Assessing Emphasis Transfer in Speech-to-Speech Models.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information EmphAssess : a Prosodic Benchmark on Assessing Emphasis Transfer in Speech-to-Speech Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:45.122530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:45.122530Z digest=sha256:8cb68ba1d357adc03ab0e35afc34117dc27d91e4d73e315dc8b3769eb2bad5c4

Observation 19febf02-e693-4e25-bf7f-eaebb37e1834 · outbound

This paper cites UTMOS: UTokyo-SaruLab system for voice- mos challenge 2022,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information UTMOS: UTokyo-SaruLab system for voice- mos challenge 2022,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.932575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:45.173094Z digest=sha256:483c9f6d7751108e76e19543d95b67c4d4a56d09e0599d9bc20138f3d996439d

Observation a6a144a5-47dc-4424-ac23-037deb7e233c · outbound

This paper cites A layer-wise analysis of man- darin and english suprasegmentals in ssl speech models,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information A layer-wise analysis of man- darin and english suprasegmentals in ssl speech models,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.774824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:45.232249Z digest=sha256:858620f3115bb65b7d5e15b3e27166ec2a637f1a30949afca9d67f24566a8e37

Observation ad4fea51-e328-4b6c-84fe-4195775c3f69 · outbound

This paper cites blind speech segmentation: au- tomatic segmentation of speech without linguistic knowledge,.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information blind speech segmentation: au- tomatic segmentation of speech without linguistic knowledge,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.662287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:45.290095Z digest=sha256:c21f74ed203f3e7965fbeb911678abfe6af6451334689c61414e6e9c388df7d0

Pith citing papers

Observation 9bbcac71-74b0-4c85-aebf-96dc035cd577 · inbound

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information cites this paper.

Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:16:45.591375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:16:42.871348Z digest=sha256:320ccadf5fb415e2cf64919460a09e08b5437359db81678f8b656096f5a3a8e4