Pith. sign in

Paper Citation Record · LEDGER

Rhythm Features for Speaker Identification

As of 19 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 2 inbound Pith citation observations for arXiv:2506.06834.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06834 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:54:07.435997Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:54:04.008475Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T05:54:08.361418Z

Reference resolution

32 of 32 outbound references displayed

  • verified exact5
  • verified fuzzy20
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c47e5e7b-a891-4f4e-80c3-872e70f9e8fe · outbound

This paper cites an unresolved cited work.

Rhythm Features for Speaker Identification Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:54:13.271527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:03.435043Z digest=sha256:27f4c56b1f088e5689018a0cddcd8e371b8edc103b5440bea0afc4f8a4e10a64

Observation a3f4dd10-0d06-474a-bb20-03f7b3ec3f38 · outbound

This paper cites For example, [11] demonstrated that human listeners were able to identify familiar speakers based on only a sinusoidal encoding of the prosodic informa- tion in speech.

Rhythm Features for Speaker Identification For example, [11] demonstrated that human listeners were able to identify familiar speakers based on only a sinusoidal encoding of the prosodic informa- tion in speech

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:13.054283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:03.623277Z digest=sha256:55e9100a584f1f0a525eeb0c206b5512fb7140e83106e5b17ad923b4082d6c22

Observation aa410efc-82a1-4641-9fff-3b368902c93f · outbound

This paper cites Rhythm Features for Speaker Identification.

Rhythm Features for Speaker Identification Rhythm Features for Speaker Identification

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T05:54:08.482430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:04.008475Z digest=sha256:2fd4d69c07e93289d6fcd9b6f00688c55dab9d0380cafc364d6f313844368a12

Observation c6d38783-3ffe-4835-ba9f-9eb975717d90 · outbound

This paper cites Datasets We evaluate our approach on a closed-set SI task using two pop- ular speech datasets.

Rhythm Features for Speaker Identification Datasets We evaluate our approach on a closed-set SI task using two pop- ular speech datasets

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:12.691984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:04.217453Z digest=sha256:65d4d130964e6b34d776682a8f9be093b245f5da12b6e2b53c1633ea2b6e3015

Observation c25e25d8-041c-4101-8ecc-61c15eef1b71 · outbound

This paper cites Additionally, it is better suited to our closed-set identi- fication task than other commonly used measures, such as equal error rate (EER).

Rhythm Features for Speaker Identification Additionally, it is better suited to our closed-set identi- fication task than other commonly used measures, such as equal error rate (EER)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:12.366008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:04.600304Z digest=sha256:29cef61adbc0b8f6157bbd78e049ad962af38d4c96206b04700850c40efb5f23

Observation 406f3c22-f2a8-4119-bd03-8df349f8a56a · outbound

This paper cites Discussion Our results are consistent with prior works [2, 3] that have suggested rhythm features do convey useful information about a speaker’s identity.

Rhythm Features for Speaker Identification Discussion Our results are consistent with prior works [2, 3] that have suggested rhythm features do convey useful information about a speaker’s identity

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:12.175705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:04.771338Z digest=sha256:703a028717b262686e7452e8372a4c1f0551574228cddcafc68ee25b0f7c74f7

Observation 53377cd6-1458-4229-96e6-f00d78470ba2 · outbound

This paper cites Durations of context- dependent phonemes: A new feature in speaker verification,.

Rhythm Features for Speaker Identification Durations of context- dependent phonemes: A new feature in speaker verification,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:10.666295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:05.885515Z digest=sha256:2a5121ffc6708f91f7063d2c93e44bfa9f2fb5b4df3be3767c330896c504b577

Observation 028b0074-465d-45ce-b6bd-e94900cf3c93 · outbound

This paper cites Prosodic pa- rameter for speaker identification,.

Rhythm Features for Speaker Identification Prosodic pa- rameter for speaker identification,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:10.441207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:05.937445Z digest=sha256:e390a9ad3c2909321f4bb7ececfe1e2e8308eb4a9d44c68dbc0930fe4a2a186f

Observation 40c9c841-afce-47e3-88d1-a9483ff02afd · outbound

This paper cites Speaker recognition by machines and humans: A tutorial review,.

Rhythm Features for Speaker Identification Speaker recognition by machines and humans: A tutorial review,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:11.971180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:04.981905Z digest=sha256:ae0e59060175f19c6c9494631a257ad918abf94c50e7d47c8c3770f235803eac

Observation ab08149a-b143-4b5d-a9dc-5e658b131205 · outbound

This paper cites We used the pre-existing identifica- tion split, which uses roughly95%of the utterances for training and the remaining5%for testing.

Rhythm Features for Speaker Identification We used the pre-existing identifica- tion split, which uses roughly95%of the utterances for training and the remaining5%for testing

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:12.514529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:04.381391Z digest=sha256:e22b712fe224a06c232f4d09b6bf647a884cb3de1481c6e0556b13e03dbda1fd

Observation e6651ab3-e85a-430b-beea-90175a08c77b · outbound

This paper cites Is phoneme length and phoneme energy useful in automatic speaker recognition?.

Rhythm Features for Speaker Identification Is phoneme length and phoneme energy useful in automatic speaker recognition?

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:11.788813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:05.128715Z digest=sha256:ae0c39b77b528cb69c72a00a02fa47a969e9f2c341cf44a52461fb0e78f318f9

Observation 79f82602-70e2-48e3-8f0e-de41d3af3c31 · outbound

This paper cites Other works, however, such as [13] have found that content can have a substantial influence on certain aspects of the speaker’s prosody.

Rhythm Features for Speaker Identification Other works, however, such as [13] have found that content can have a substantial influence on certain aspects of the speaker’s prosody

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:12.873829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:03.842347Z digest=sha256:4c668fd1780d6b3d15c6082ce5c8899968b648cf23078e27b010fabfe1f3506c

Observation a3fef487-3d24-482d-bc30-51e7fc883489 · outbound

This paper cites Intrinsic phone durations are speaker-specific,.

Rhythm Features for Speaker Identification Intrinsic phone durations are speaker-specific,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:11.491929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:05.294335Z digest=sha256:0aa2cfccb7e458f783fb017526e5aad5e34fdb0be8a99a1ab432730cb6fa6608

Observation 878ef969-b1df-4b60-8b17-4f754decae99 · outbound

This paper cites What determines duration-based rhythm measures: text or speaker?.

Rhythm Features for Speaker Identification What determines duration-based rhythm measures: text or speaker?

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:11.212587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:05.428798Z digest=sha256:29d2205a6e46de0d24257f623aa9dc0001e0a65ced134893a3f040e96394a79c

Observation 3d853395-be5f-418c-93af-4b809e211051 · outbound

This paper cites Extraction and representation of prosodic features for language and speaker recognition,.

Rhythm Features for Speaker Identification Extraction and representation of prosodic features for language and speaker recognition,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:10.964542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:05.739504Z digest=sha256:b79b84c9a4df909ebfc97e08222431f627ba17c903b6cd7ac2ce9b0901045ace

Observation ff9843ba-ad83-49a2-92a3-5a0cc28e648a · outbound

This paper cites Analysis of Speech Temporal Dynamics in the Context of Speaker Verification and Voice Anonymization.

Rhythm Features for Speaker Identification Analysis of Speech Temporal Dynamics in the Context of Speaker Verification and Voice Anonymization

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:54:08.210687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:05.801906Z digest=sha256:89411b16ba5f2188c47395403e12d8f26e349c3ff41c86c6b4b236ce3df0b232

Observation 95e1751b-1236-4cd7-9a5b-858ef3a85bc0 · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

Rhythm Features for Speaker Identification Robust Speech Recognition via Large-Scale Weak Supervision

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:06.944151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:54:06.944151Z digest=sha256:d408c0eda8af79b65e7b7adde4ff3d465ef9e379dba487ca286dbff35cb0a8d7

Observation 8cd7a73f-c587-42bf-b65f-ba31151625be · outbound

This paper cites Lib- rispeech: an asr corpus based on public domain audio books,.

Rhythm Features for Speaker Identification Lib- rispeech: an asr corpus based on public domain audio books,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:10.169887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:06.047270Z digest=sha256:512d22002ccd155f487234d33c70c8da4ce908a09712d14fc365c5eddcd66ec4

Observation 2cc2a9a4-e143-4516-ac91-7356b31f31f9 · outbound

This paper cites V oxceleb: a large- scale speaker identification dataset,.

Rhythm Features for Speaker Identification V oxceleb: a large- scale speaker identification dataset,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:09.869062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:06.126513Z digest=sha256:4d1bbbab55acec9881ec82435fc08f161b445213fd4b8b8d438db130d5c470d9

Observation 22708e30-6233-4650-af6a-ecf4cd757192 · outbound

This paper cites On the importance of pure prosody in the perception of speaker identity,.

Rhythm Features for Speaker Identification On the importance of pure prosody in the perception of speaker identity,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:09.615160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:06.230664Z digest=sha256:8f954e7fc6fc9a8f29967911e68f72ebe656111c3218eff4fdd1785a0bd1d857

Observation 1fc3e0fc-4607-4b20-82b0-67155870bfbb · outbound

This paper cites Speaker idiosyncratic rhythmic features in the speech signal,.

Rhythm Features for Speaker Identification Speaker idiosyncratic rhythmic features in the speech signal,

Reference 21

Resolution
verified exact
doi, observed 2026-08-07T05:54:07.585336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:06.362057Z digest=sha256:a5a081fcd786ab94998b5097e81bbd981ad9d9cfdeffcb016d33bce22e606527

Observation a67c0dcb-b08a-4288-aac9-bb840b46e271 · outbound

This paper cites How stable are acoustic metrics of contrastive speech rhythm?.

Rhythm Features for Speaker Identification How stable are acoustic metrics of contrastive speech rhythm?

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:09.406348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:06.451841Z digest=sha256:3ee3c19a620214799dcdb2fe9a2f636538242572beb124e7c4fbbac2afd93fc5

Observation 701d603e-08e5-4652-96c2-2827259db652 · outbound

This paper cites Speech rate normalization used to improve speaker verification,.

Rhythm Features for Speaker Identification Speech rate normalization used to improve speaker verification,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:09.247335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:06.559650Z digest=sha256:21825351f42097c7de348bf06668d1cbf199c61e58b4cf1caa13e286cce58474

Observation a31148a8-0d21-46c6-8e8a-c7ef37354a60 · outbound

This paper cites Phoneme duration modeling using speech rhythm-based speaker embeddings for multi-speaker speech synthesis.

Rhythm Features for Speaker Identification Phoneme duration modeling using speech rhythm-based speaker embeddings for multi-speaker speech synthesis

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:08.995644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:06.735474Z digest=sha256:82cc280ebc3f954f48a5dc6391ef8f6bf66d2d6c788965155cfd4e4c49816a11

Observation e3af9745-c40e-4454-8921-92433606a7d6 · outbound

This paper cites Whisperx: Time-accurate speech transcription of long-form audio,.

Rhythm Features for Speaker Identification Whisperx: Time-accurate speech transcription of long-form audio,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:06.859629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:54:06.859629Z digest=sha256:bd58a53843b353a86e2158a198cdf69017d662f9caf24c4b585950ad366516d2

Observation a4b76104-1cff-4184-9d31-d430f5668a06 · outbound

This paper cites Wavlm: Large-scale self- supervised pre-training for full stack speech processing,.

Rhythm Features for Speaker Identification Wavlm: Large-scale self- supervised pre-training for full stack speech processing,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:07.029788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:54:07.029788Z digest=sha256:c6a2373169fbd90694fef5ccb94a9605a91331e21656f03d96d94070713250fe

Observation 8fa06a12-12ce-44d2-a4d7-055ac1174e82 · outbound

This paper cites SpeechBrain: A General-Purpose Speech Toolkit.

Rhythm Features for Speaker Identification SpeechBrain: A General-Purpose Speech Toolkit

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:07.103872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:54:07.103872Z digest=sha256:6d134bc6eef8f6d1fd5575262656c7aa99e37bbed984e0b7af3d4a637576f72f

Observation 8d74c528-c8ac-4ed6-af2f-cc29a99eabe4 · outbound

This paper cites Metrics for Multi-Class Classification: an Overview.

Rhythm Features for Speaker Identification Metrics for Multi-Class Classification: an Overview

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:07.182515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:54:07.182515Z digest=sha256:b856689c2b33e550bd3da8404f5232e5a4c9c9fcbc2903ce4d703cb44f21cf82

Observation 421f1a32-f7e0-47a4-9dc3-69e2765aab3e · outbound

This paper cites Improving Automatic Emotion Recognition from speech using Rhythm and Temporal feature.

Rhythm Features for Speaker Identification Improving Automatic Emotion Recognition from speech using Rhythm and Temporal feature

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:54:08.006889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:07.275427Z digest=sha256:b79803a2d97395392c2d054a79650ffeab258c552db5cc9469cdfe42a3e0d431

Observation d64348cd-c582-481b-8c7d-222d99e010e1 · outbound

This paper cites Probing the in- formation encoded in x-vectors,.

Rhythm Features for Speaker Identification Probing the in- formation encoded in x-vectors,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:08.741765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:07.324574Z digest=sha256:359e1102875c85606f90744f0cd95e3dd6b478cb63d3e16074dd74c019dde0fd

Observation 8ba2e869-2962-4fd8-b72e-65056bfdd04d · outbound

This paper cites An empirical analysis of information encoded in disentangled neural speaker representations.

Rhythm Features for Speaker Identification An empirical analysis of information encoded in disentangled neural speaker representations

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:54:07.850121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:07.435997Z digest=sha256:4ed0a115430d8b72942c1a0121f83c9dd4fa175aaaeafaac6bb7e9d4afbd0946

Observation 2e035620-68bf-4465-95bf-3bfac055ce88 · outbound

This paper cites Available: https://doi.org/10.1515/lp-2013-0012.

Rhythm Features for Speaker Identification Available: https://doi.org/10.1515/lp-2013-0012

Reference 2013

Resolution
verified exact
doi, observed 2026-08-07T05:54:07.708316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:05.581631Z digest=sha256:d0fe19d19ac4d1a097aec9cb4bae6238012d2f58ebee3d4db1534a3327b1b354

Pith citing papers

Observation aa410efc-82a1-4641-9fff-3b368902c93f · inbound

Rhythm Features for Speaker Identification cites this paper.

Rhythm Features for Speaker Identification Rhythm Features for Speaker Identification

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T05:54:08.482430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:54:04.008475Z digest=sha256:2fd4d69c07e93289d6fcd9b6f00688c55dab9d0380cafc364d6f313844368a12

Observation 0e985fe9-4c6d-4ea1-a3fc-8ccaa30980b0 · inbound

Multimodal Speaker Verification as a Threat to Speaker Anonymization cites this paper.

Multimodal Speaker Verification as a Threat to Speaker Anonymization Rhythm Features for Speaker Identification

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T12:12:42.171456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:12:42.171456Z digest=sha256:4334e4f92117da4825847073b3fc4227f0ef0a1b7ca7a5f7b03b1ff4e9445210