Pith. sign in

Paper Citation Record · LEDGER

3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2306.15354.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.15354 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T16:58:02.446598Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:29:44.199281Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5ad194ca-b72e-4f4d-9468-e213687d7f26 · inbound

The Multimodal Information Based Speech Processing (MISP) 2025 Challenge: Audio-Visual Diarization and Recognition cites this paper.

The Multimodal Information Based Speech Processing (MISP) 2025 Challenge: Audio-Visual Diarization and Recognition 3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:04.029122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:43:04.029122Z digest=sha256:fba77e68ff43b8a0b1c196daff4a20a0403b957b3d4a8ecbfaf42f1ca648628d

Observation c2722b4e-f906-42c7-a0ce-6e054e013cac · inbound

Multi-Channel Sequence-to-Sequence Neural Diarization: Experimental Results for The MISP 2025 Challenge cites this paper.

Multi-Channel Sequence-to-Sequence Neural Diarization: Experimental Results for The MISP 2025 Challenge 3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:50.995609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:50.995609Z digest=sha256:f9ee935304313112d62dc434c7eae4149c1576adbf3c1fae1bccd84d9ed92d16

Observation 05674538-aa95-487b-adc8-d195f47d14ff · inbound

VoxAging: Continuously Tracking Speaker Aging with a Large-Scale Longitudinal Dataset in English and Mandarin cites this paper.

VoxAging: Continuously Tracking Speaker Aging with a Large-Scale Longitudinal Dataset in English and Mandarin 3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:32:32.813536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:32:32.813536Z digest=sha256:13d50c85fdbbb9549f6e8a5b06c579d717a3d7e67b069d8daed1940b04a5c7c2

Observation eb907e42-812a-4029-a0a9-f9c868980d34 · inbound

Interpolating Speaker Identities in Embedding Space for Data Expansion cites this paper.

Interpolating Speaker Identities in Embedding Space for Data Expansion 3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T16:58:02.446598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:58:02.446598Z digest=sha256:c7805e32fe90b172a201d80527dcd25f68343754f4970ed66d7e28b8f885c6d1

Observation 60f6bdce-a975-4e39-94d4-1a574ca3b838 · inbound

Audio2Tool: Speak, Call, Act -- A Dataset for Benchmarking Speech Tool Use cites this paper.

Audio2Tool: Speak, Call, Act -- A Dataset for Benchmarking Speech Tool Use 3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:47:13.060741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T07:32:06.306966Z digest=sha256:24a6552dffdfe0ecff40be5c5a1d2bc0a263631719bd6806f68e38d53eebf496

Observation 37562708-62ef-40ff-a7c4-569a13b094c1 · inbound

SoulX-Transcriber: A Robust End-to-End Framework for Multi-Speaker Speech Transcription cites this paper.

SoulX-Transcriber: A Robust End-to-End Framework for Multi-Speaker Speech Transcription 3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-02T01:06:24.520029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T12:37:52.271987Z digest=sha256:3e054f657c54d08a428229751fcf3849009f780866b28bf8b967ece9e3065ff8

Observation a2b513cf-b347-4a4f-8746-82d1435e9877 · inbound

Kiwano: A Cutting-Edge Open-Source Toolkit for Speaker Verification cites this paper.

Kiwano: A Cutting-Edge Open-Source Toolkit for Speaker Verification 3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-04T09:29:44.201695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T09:52:58.661879Z digest=sha256:e7588202acbeee4d7b1a2fbba44dc9e1727b40a9adb62908ed583b48ea774a1d

Observation 505248d2-3a49-4223-b368-65e25321864d · inbound

StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis cites this paper.

StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis 3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T11:33:14.026639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:33:14.026639Z digest=sha256:88ec8e02129e94e697cdbc56ed403472a07529f549c62db46f8a0f49ca07747e