Pith. sign in

Paper Citation Record · LEDGER

ESPnet-SPK: full pipeline speaker embedding toolkit with reproducible recipes, self-supervised front-ends, and off-the-shelf models

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2401.17230.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.17230 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T00:39:51.411070Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T07:57:44.880446Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 82b3c130-1de6-46c1-aa51-3fb9b5a1f987 · inbound

A Non-autoregressive Model for Joint STT and TTS cites this paper.

A Non-autoregressive Model for Joint STT and TTS ESPnet-SPK: full pipeline speaker embedding toolkit with reproducible recipes, self-supervised front-ends, and off-the-shelf models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:28.455695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:28.455695Z digest=sha256:e4b59fd9945c9086694a91080b37d108768bb42daf6e3f45106a506f44279815

Observation 08a45b31-41f8-4493-87f1-11b526282bea · inbound

Geolocation-Aware Robust Spoken Language Identification cites this paper.

Geolocation-Aware Robust Spoken Language Identification ESPnet-SPK: full pipeline speaker embedding toolkit with reproducible recipes, self-supervised front-ends, and off-the-shelf models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T17:06:17.024757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:06:17.024757Z digest=sha256:3ab374d6bdf958615d579fd27c08a1cf6bea11630af5ed7ca692ed8f3de2a8fa

Observation 24ec7006-2eb0-41ff-8bfc-f8e3284a43aa · inbound

CS-YODAS: A Mined Dataset of In-the-Wild Code-Switched Speech cites this paper.

CS-YODAS: A Mined Dataset of In-the-Wild Code-Switched Speech ESPnet-SPK: full pipeline speaker embedding toolkit with reproducible recipes, self-supervised front-ends, and off-the-shelf models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-03T07:57:44.881976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-27T11:27:11.357245Z digest=sha256:3d1fceadc24652f294772f8da594e8a0e4109d5215a6bc1051ecd71561c9786d

Observation ffd2280f-f4ed-4f6d-a23d-e68325c13327 · inbound

Leveraging Gradient Reversal Loss and Multitask Learning for Datasets-Aware Audio Deepfake Detection cites this paper.

Leveraging Gradient Reversal Loss and Multitask Learning for Datasets-Aware Audio Deepfake Detection ESPnet-SPK: full pipeline speaker embedding toolkit with reproducible recipes, self-supervised front-ends, and off-the-shelf models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-31T23:27:00.427163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:27:00.427163Z digest=sha256:2f0b1f3e732112a1a9566e481e2b3249dc8db0a8988f47731992177c367fa4ab

Observation 10434a84-dbbb-49f2-8569-12e66b2f9777 · inbound

The Voiceprint Fallacy: Why Voices Are Not Unique Biometric Imprints cites this paper.

The Voiceprint Fallacy: Why Voices Are Not Unique Biometric Imprints ESPnet-SPK: full pipeline speaker embedding toolkit with reproducible recipes, self-supervised front-ends, and off-the-shelf models

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-12T00:39:51.411070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:39:51.411070Z digest=sha256:162cd84232df4e6662f8972f20970f119024858eb2381c85d0b872e4650b67fb