Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:21:58.204834Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 1 inbound Pith citation observation for arXiv:2506.14427.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:21:58.204834Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-09T15:07:52.715880Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-09T15:16:18.188094Z
52 of 52 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d3f82b19-44d6-435f-9c02-be0f261f38bb · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset A review of speaker diarization: Recent advances with deep learning,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e13f9fe6-6fd1-4b5a-961a-4aeb8dd64713 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Speaker diarization: A review of recent research,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ec32f24c-0d18-4624-b2e7-4f1c8a84a765 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Speaker diarization with lstm,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 73ef4eec-b169-4a19-99fd-25a05bff166e · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Front- end factor analysis for speaker verification,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c5e0ead-22dd-4f87-8f96-979565dfbea6 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset X- vectors: Robust dnn embeddings for speaker recognition,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 56007a15-7a95-4dad-9254-627bcce8099a · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Developing on-line speaker diarization system
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1efab697-e996-43c9-b157-0884261e392a · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset A study of the cosine distance-based mean shift for telephone speech diarization,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 09e744c1-aa83-4cfa-bc3a-41559e4f6828 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Speaker diarization with lstm,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fde14acd-eef8-4b76-a3dd-e02491bab023 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset A robust stopping criterion for agglomerative hierarchical clustering in a speaker diarization system
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ad39687d-62ea-4546-9eb8-9db7c3a523f3 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Characterizing performance of speaker diarization systems on far-field speech using standard methods,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f87ca2c2-33fb-4503-8ebe-fe82f85839e9 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Discriminative neural clustering for speaker diarisation,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1799a67a-a6c6-4938-a63d-5af4012ce448 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Bayesian hmm clustering of x-vector sequences (vbx) in speaker diarization: theory, implemen- tation and analysis on standard tasks,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6c09ff5f-6367-4f97-b3af-add1a7bd70f9 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset End-to-End Neural Speaker Diarization with Permutation-Free Objec- tives,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6962df6d-2f1d-46a9-b7ff-7e1488b08cd9 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset End-to-end neural speaker diarization with self-attention,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a66dfec-ffab-4768-bbe7-c47f87b803ec · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Auxiliary loss of transformer with residual connection for end-to-end speaker diarization,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 91d2595b-6fe6-4cd8-ac5e-ddf332add929 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Incorporating end-to-end framework into target- speaker voice activity detection,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 48b1ea25-89ef-42ed-8019-a486ea98782c · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Target-Speaker V oice Activity Detection: A Novel Approach for Multi-Speaker Diarization in a Dinner Party Scenario,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e72957de-844b-4af2-b005-b6d83bdb7240 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Ansd-ma-mse: Adaptive neural speaker diarization using memory-aware multi-speaker embed- ding,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2716b444-1e63-4fc5-9248-731d85914f91 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset The CHiME-7 DASR Challenge: Distant Meeting Transcription with Multiple Devices in Diverse Scenarios
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ff547e6-29f0-42fd-a60f-6f86231dddc2 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset The ustc-nercslip systems for chime-7 challenge,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bca27e2b-c9db-4d31-a9a3-d0314d2ed985 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Audio-visual speaker diarization based on spatiotemporal bayesian fusion,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c773421d-5f15-4cde-850b-8864ec10fbf1 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Quantitative association of vocal-tract and facial behavior,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e43a8855-a018-4c74-aa81-c275a07f6cf7 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Multimodal speaker diarization,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ce9ae35e-a165-40e7-985d-40c08a73cb89 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Who said that?: Audio-visual speaker diarisation of real-world meetings,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2e95f8fa-c44d-44f1-87d8-08440b71e7b2 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Spot the conversation: Speaker diarisation in the wild,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 428b31c9-85f8-4f23-9c29-9dac8e0da984 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset End-to-end audio-visual neural speaker diarization,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c140ea81-7ae9-40be-b600-f63c29793861 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Multi-Input Multi-Output Target-Speaker Voice Activity Detection For Unified, Flexible, and Robust Audio-Visual Speaker Diarization
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0fa53158-fd14-451c-907c-60b3a82249c7 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Quality-Aware End-to-End Audio-Visual Neural Speaker Diarization
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2ffe891a-1a73-45a7-907e-004876e43b21 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Semi-supervised multi-channel speaker diarization with cross- channel attention,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e1c36b09-1a9e-4a1b-8abf-6cfe9b2aafed · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Aishell-4: An open source dataset for speech enhancement, separation, recognition and speaker diarization in conference scenario,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 026db9a8-134e-4b29-b845-835258614c2c · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Ava-avd: Audio-visual speaker diarization in the wild,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 05fbe6af-a073-45db-bb76-3551a48e1ab3 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Semi-supervised training with pseudo-labeling for end- to-end neural diarization,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f4fd928c-abbd-4ec0-8ea1-5b27ae7b4a26 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Chime-6 challenge: Tackling multispeaker speech recognition for unsegmented recordings,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ac8da1ae-9ff5-4ccc-b653-e77e25873ae7 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset The multimodal information based speech processing (misp) 2022 challenge: Audio- visual diarization and recognition,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 284f3f73-a188-42a9-86b4-331cfea318fa · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Notsofar-1 challenge: New datasets, baseline, and tasks for distant meeting transcription,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cedac64d-e72b-4cea-976f-faf267d75d7e · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset The nist speaker recognition evaluations: 1996-2001
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6dad5101-8ed9-4e39-b33c-813a367f287b · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset First dihard challenge evaluation plan,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 68cf3cf7-1053-4eba-8bcb-860721340f5e · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset The Second DIHARD Diarization Challenge: Dataset, task, and baselines
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7b1e1bb5-efde-4eee-9410-07ad6e9434eb · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset The Third DIHARD Diarization Challenge
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5daffbd3-c75b-4dee-adf6-aea506583848 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset M2met: The icassp 2022 multi- channel multi-party meeting transcription challenge,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f88caf64-7715-4f24-94b6-681e862c35be · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset The ami meeting corpus,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dca6c443-3407-4b0b-92fb-977eb884cbde · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset The Multimodal Information Based Speech Processing (MISP) 2025 Challenge: Audio-Visual Diarization and Recognition
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5438a602-d4a0-4270-b8f1-af5dabe50ec5 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Msdwild: Multi-modal speaker diarization dataset in the wild
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 17e3e495-23e7-4b84-b33e-6598d3733f73 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset An approach to scene change detection,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d947b473-73bc-4c55-b50d-545945fa1c29 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Dnsmos p. 835: A non- intrusive perceptual objective speech quality metric to evaluate noise suppressors,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b6875ee2-91e5-4828-9401-ffd2f0c4e60b · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Md-vqa: Multi-dimensional quality assessment for ugc live videos,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c6a24fa6-3ee0-42bd-a628-9d00325f3f8b · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Out of time: automated lip sync in the wild,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8a3154c0-925e-43c5-bf15-519b2d0f01ca · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Retinaface: Single-shot multi-level face localisation in the wild,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6c5c5a94-fc8e-4f58-be3c-5990295ce602 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Simple online and realtime tracking with a deep association metric,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a6efce1f-5493-45e2-acdb-574c78d25f91 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset MediaPipe: A Framework for Building Perception Pipelines
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e4dfd2a-1d75-4bae-a506-09f2c1823712 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset 3d-speaker-toolkit: An open-source toolkit for multimodal speaker verification and diarization,
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 10058e43-5b10-47f4-b6a0-c3e06c096d82 · outbound
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset Dover-lap: A method for combining overlap- aware diarization outputs,
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 627f9364-7480-480c-a8e3-e381cc071047 · inbound
Multimodal Voice Activity Projection for Turn-Taking in Social Robots with Voice-Activity-Related Pretrained Encoders M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.