Pith. sign in

Paper Citation Record · LEDGER

Rhythm Features for Speaker Identification

As of 9 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 2 inbound Pith citation observations for arXiv:2506.06834.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06834 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:54:07.435997Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:54:04.008475Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T05:54:08.361418Z

Reference resolution

32 of 32 outbound references displayed

  • verified exact5
  • verified fuzzy20
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c47e5e7b-a891-4f4e-80c3-872e70f9e8fe · outbound

This paper cites an unresolved cited work.

Rhythm Features for Speaker Identification Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:54:13.271527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:03.435043Z digest=sha256:dad6ab5d02361b435166036b867a5fb73b7626b6e04b06b0649cb9fad6ea57c3

Observation a3f4dd10-0d06-474a-bb20-03f7b3ec3f38 · outbound

This paper cites For example, [11] demonstrated that human listeners were able to identify familiar speakers based on only a sinusoidal encoding of the prosodic informa- tion in speech.

Rhythm Features for Speaker Identification For example, [11] demonstrated that human listeners were able to identify familiar speakers based on only a sinusoidal encoding of the prosodic informa- tion in speech

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:13.054283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:03.623277Z digest=sha256:b7fb3f9825e4c4e9b631189928774c5e3f4e74126f338740239b4b92ddf161d3

Observation aa410efc-82a1-4641-9fff-3b368902c93f · outbound

This paper cites Rhythm Features for Speaker Identification.

Rhythm Features for Speaker Identification Rhythm Features for Speaker Identification

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T05:54:08.482430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:04.008475Z digest=sha256:3d429f62c9cd926a85284da5d096bcc7cd048b090e5c2d40ccca5e5b127e80e4

Observation c6d38783-3ffe-4835-ba9f-9eb975717d90 · outbound

This paper cites Datasets We evaluate our approach on a closed-set SI task using two pop- ular speech datasets.

Rhythm Features for Speaker Identification Datasets We evaluate our approach on a closed-set SI task using two pop- ular speech datasets

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:12.691984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:04.217453Z digest=sha256:602c98b72d12e99a2198720e358ee1923e9508469c108f7bef65f9052d02e180

Observation c25e25d8-041c-4101-8ecc-61c15eef1b71 · outbound

This paper cites Additionally, it is better suited to our closed-set identi- fication task than other commonly used measures, such as equal error rate (EER).

Rhythm Features for Speaker Identification Additionally, it is better suited to our closed-set identi- fication task than other commonly used measures, such as equal error rate (EER)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:12.366008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:04.600304Z digest=sha256:8638231aee96199f00fb8636dd34a7fd375d42adb523b7f845ad6e76947222fc

Observation 406f3c22-f2a8-4119-bd03-8df349f8a56a · outbound

This paper cites Discussion Our results are consistent with prior works [2, 3] that have suggested rhythm features do convey useful information about a speaker’s identity.

Rhythm Features for Speaker Identification Discussion Our results are consistent with prior works [2, 3] that have suggested rhythm features do convey useful information about a speaker’s identity

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:12.175705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:04.771338Z digest=sha256:70453ce8de2e7c84570234a8652f043d8ca88d109c12c2547b405b37240f9385

Observation 53377cd6-1458-4229-96e6-f00d78470ba2 · outbound

This paper cites Durations of context- dependent phonemes: A new feature in speaker verification,.

Rhythm Features for Speaker Identification Durations of context- dependent phonemes: A new feature in speaker verification,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:10.666295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:05.885515Z digest=sha256:a41dae85ebb9706296eaafe287a328c7180b4c7a3816c0c0a3782a17ca0c5b73

Observation 028b0074-465d-45ce-b6bd-e94900cf3c93 · outbound

This paper cites Prosodic pa- rameter for speaker identification,.

Rhythm Features for Speaker Identification Prosodic pa- rameter for speaker identification,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:10.441207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:05.937445Z digest=sha256:50a0770abb551c0ec2dad21fffcb1720da3e1e18fe09e94506d2894e8f2112ae

Observation 40c9c841-afce-47e3-88d1-a9483ff02afd · outbound

This paper cites Speaker recognition by machines and humans: A tutorial review,.

Rhythm Features for Speaker Identification Speaker recognition by machines and humans: A tutorial review,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:11.971180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:04.981905Z digest=sha256:d2351850e3291a8c4c2f454716618913af5bbc5d88ddf96c50b7abd927f280b6

Observation ab08149a-b143-4b5d-a9dc-5e658b131205 · outbound

This paper cites We used the pre-existing identifica- tion split, which uses roughly95%of the utterances for training and the remaining5%for testing.

Rhythm Features for Speaker Identification We used the pre-existing identifica- tion split, which uses roughly95%of the utterances for training and the remaining5%for testing

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:12.514529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:04.381391Z digest=sha256:0d202c2dd7717f6100edd80fb1749587955c888c4521443186a883fd18026b5a

Observation e6651ab3-e85a-430b-beea-90175a08c77b · outbound

This paper cites Is phoneme length and phoneme energy useful in automatic speaker recognition?.

Rhythm Features for Speaker Identification Is phoneme length and phoneme energy useful in automatic speaker recognition?

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:11.788813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:05.128715Z digest=sha256:d6478d9833a8dd6daab14ca8899ae96273b55eb555e1ad109131e7adef8247d7

Observation 79f82602-70e2-48e3-8f0e-de41d3af3c31 · outbound

This paper cites Other works, however, such as [13] have found that content can have a substantial influence on certain aspects of the speaker’s prosody.

Rhythm Features for Speaker Identification Other works, however, such as [13] have found that content can have a substantial influence on certain aspects of the speaker’s prosody

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:12.873829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:03.842347Z digest=sha256:59ee4575fcb2f7676db68942b85f959e38bbc0cd4b0bf68d3bcbfc06f977dfdf

Observation a3fef487-3d24-482d-bc30-51e7fc883489 · outbound

This paper cites Intrinsic phone durations are speaker-specific,.

Rhythm Features for Speaker Identification Intrinsic phone durations are speaker-specific,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:11.491929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:05.294335Z digest=sha256:9094c90a76b741d477268bee42bc64d17bce6aa1c65af95fe5575d41d161f9e8

Observation 878ef969-b1df-4b60-8b17-4f754decae99 · outbound

This paper cites What determines duration-based rhythm measures: text or speaker?.

Rhythm Features for Speaker Identification What determines duration-based rhythm measures: text or speaker?

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:11.212587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:05.428798Z digest=sha256:83d189a083c1635911929ad99a101c7880d64db5d8ff5f96d091c6ff67251729

Observation 3d853395-be5f-418c-93af-4b809e211051 · outbound

This paper cites Extraction and representation of prosodic features for language and speaker recognition,.

Rhythm Features for Speaker Identification Extraction and representation of prosodic features for language and speaker recognition,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:10.964542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:05.739504Z digest=sha256:8368b8816aa18a65f51140f7b06cdd9f3475d4d31cd325360980ea58044ce74c

Observation ff9843ba-ad83-49a2-92a3-5a0cc28e648a · outbound

This paper cites Analysis of Speech Temporal Dynamics in the Context of Speaker Verification and Voice Anonymization.

Rhythm Features for Speaker Identification Analysis of Speech Temporal Dynamics in the Context of Speaker Verification and Voice Anonymization

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:54:08.210687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:05.801906Z digest=sha256:758ed2b4276a88f68a013eeb15daabf6f486c953ed2866014cdd92e30723bc04

Observation 95e1751b-1236-4cd7-9a5b-858ef3a85bc0 · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

Rhythm Features for Speaker Identification Robust Speech Recognition via Large-Scale Weak Supervision

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:06.944151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:54:06.944151Z digest=sha256:b2acb3d7acf4b24406bd1f837f02470769204f1b6a7873c487311981a1c182e3

Observation 8cd7a73f-c587-42bf-b65f-ba31151625be · outbound

This paper cites Lib- rispeech: an asr corpus based on public domain audio books,.

Rhythm Features for Speaker Identification Lib- rispeech: an asr corpus based on public domain audio books,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:10.169887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:06.047270Z digest=sha256:c80839a988f0cc3616d3c4b568967c32ac71e71c7ffa5bdf96b207061b22c5e1

Observation 2cc2a9a4-e143-4516-ac91-7356b31f31f9 · outbound

This paper cites V oxceleb: a large- scale speaker identification dataset,.

Rhythm Features for Speaker Identification V oxceleb: a large- scale speaker identification dataset,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:09.869062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:06.126513Z digest=sha256:f82e83d9e630ff3b6ab7ea52c1e2b1c3ac3ba51eb1fa69a3cc3557d492205f17

Observation 22708e30-6233-4650-af6a-ecf4cd757192 · outbound

This paper cites On the importance of pure prosody in the perception of speaker identity,.

Rhythm Features for Speaker Identification On the importance of pure prosody in the perception of speaker identity,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:09.615160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:06.230664Z digest=sha256:f823171aec874137f61313c535789321c08cceccbd832e2125a2de0084605551

Observation 1fc3e0fc-4607-4b20-82b0-67155870bfbb · outbound

This paper cites Speaker idiosyncratic rhythmic features in the speech signal,.

Rhythm Features for Speaker Identification Speaker idiosyncratic rhythmic features in the speech signal,

Reference 21

Resolution
verified exact
doi, observed 2026-08-07T05:54:07.585336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:06.362057Z digest=sha256:be26f34f81c8e226f32fff7df179063d6dc8892af020385b3c1c59e8cb23faf1

Observation a67c0dcb-b08a-4288-aac9-bb840b46e271 · outbound

This paper cites How stable are acoustic metrics of contrastive speech rhythm?.

Rhythm Features for Speaker Identification How stable are acoustic metrics of contrastive speech rhythm?

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:09.406348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:06.451841Z digest=sha256:2337861c63e277820671ec1d346dbb97683d94b19a164e42fb962eaebb23ede4

Observation 701d603e-08e5-4652-96c2-2827259db652 · outbound

This paper cites Speech rate normalization used to improve speaker verification,.

Rhythm Features for Speaker Identification Speech rate normalization used to improve speaker verification,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:09.247335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:06.559650Z digest=sha256:a8d60bac1d43410ca93cfeddd3c33f4eb75fdedee1be059cc3c04c43281b7050

Observation a31148a8-0d21-46c6-8e8a-c7ef37354a60 · outbound

This paper cites Phoneme duration modeling using speech rhythm-based speaker embeddings for multi-speaker speech synthesis.

Rhythm Features for Speaker Identification Phoneme duration modeling using speech rhythm-based speaker embeddings for multi-speaker speech synthesis

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:08.995644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:06.735474Z digest=sha256:b2f8b26076aa1742a26a6200654c370441a3b74ba4bbb14e5451f4965ab8a4e2

Observation e3af9745-c40e-4454-8921-92433606a7d6 · outbound

This paper cites Whisperx: Time-accurate speech transcription of long-form audio,.

Rhythm Features for Speaker Identification Whisperx: Time-accurate speech transcription of long-form audio,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:06.859629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:54:06.859629Z digest=sha256:79d62caa10cd1760ff17e197c70c5334564ea5e693efcc5223a48a1f10348806

Observation a4b76104-1cff-4184-9d31-d430f5668a06 · outbound

This paper cites Wavlm: Large-scale self- supervised pre-training for full stack speech processing,.

Rhythm Features for Speaker Identification Wavlm: Large-scale self- supervised pre-training for full stack speech processing,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:07.029788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:54:07.029788Z digest=sha256:00e772c97dcdf1a6f48f59bbd41b280e6baa5e0132ad2a3a077f6dd3811d75f5

Observation 8fa06a12-12ce-44d2-a4d7-055ac1174e82 · outbound

This paper cites SpeechBrain: A General-Purpose Speech Toolkit.

Rhythm Features for Speaker Identification SpeechBrain: A General-Purpose Speech Toolkit

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:07.103872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:54:07.103872Z digest=sha256:e41e2797bba9a959d3295ae798cafad4f991340e47ae2447a105cd3e6241a21e

Observation 8d74c528-c8ac-4ed6-af2f-cc29a99eabe4 · outbound

This paper cites Metrics for Multi-Class Classification: an Overview.

Rhythm Features for Speaker Identification Metrics for Multi-Class Classification: an Overview

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:07.182515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:54:07.182515Z digest=sha256:6badb4f0fb53279f1270f64b7db7149dd72db4dd89e01be131424b0290269cd7

Observation 421f1a32-f7e0-47a4-9dc3-69e2765aab3e · outbound

This paper cites Improving Automatic Emotion Recognition from speech using Rhythm and Temporal feature.

Rhythm Features for Speaker Identification Improving Automatic Emotion Recognition from speech using Rhythm and Temporal feature

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:54:08.006889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:07.275427Z digest=sha256:082d9ab9223626b97fc9cc38025d80be3c3be5ae913c3365581f41b7ca7f9957

Observation d64348cd-c582-481b-8c7d-222d99e010e1 · outbound

This paper cites Probing the in- formation encoded in x-vectors,.

Rhythm Features for Speaker Identification Probing the in- formation encoded in x-vectors,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:08.741765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:07.324574Z digest=sha256:c15097aee9b2fc54cf739984dd9dc15cf14d038377eb6261f9abbaa2d32642b7

Observation 8ba2e869-2962-4fd8-b72e-65056bfdd04d · outbound

This paper cites An empirical analysis of information encoded in disentangled neural speaker representations.

Rhythm Features for Speaker Identification An empirical analysis of information encoded in disentangled neural speaker representations

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:54:07.850121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:07.435997Z digest=sha256:d4e18e9264544aea5d57b3aa3a4340dcae919f549bb0d9a40cdd1c7acacdf921

Observation 2e035620-68bf-4465-95bf-3bfac055ce88 · outbound

This paper cites Available: https://doi.org/10.1515/lp-2013-0012.

Rhythm Features for Speaker Identification Available: https://doi.org/10.1515/lp-2013-0012

Reference 2013

Resolution
verified exact
doi, observed 2026-08-07T05:54:07.708316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:05.581631Z digest=sha256:75df08f40e821d8bab5198f76691bf13bee003197a7f424ec1e1b18fbda89340

Pith citing papers

Observation aa410efc-82a1-4641-9fff-3b368902c93f · inbound

Rhythm Features for Speaker Identification cites this paper.

Rhythm Features for Speaker Identification Rhythm Features for Speaker Identification

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T05:54:08.482430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:54:04.008475Z digest=sha256:3d429f62c9cd926a85284da5d096bcc7cd048b090e5c2d40ccca5e5b127e80e4

Observation 0e985fe9-4c6d-4ea1-a3fc-8ccaa30980b0 · inbound

Multimodal Speaker Verification as a Threat to Speaker Anonymization cites this paper.

Multimodal Speaker Verification as a Threat to Speaker Anonymization Rhythm Features for Speaker Identification

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T12:12:42.171456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:12:42.171456Z digest=sha256:adfdb52fb96420b0aa372494367ae31981a8966a67ca3d1b140a6be762f98836