Pith. sign in

Paper Citation Record · LEDGER

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR

As of 8 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2506.22646.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.22646 v2

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:07:48.267842Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:07:44.192140Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T22:07:49.263030Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact5
  • verified fuzzy30
  • unresolved7
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a4cc5d57-b0cb-43ca-9daf-ea7b1bdfb564 · outbound

This paper cites Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:07:49.354943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:44.192140Z digest=sha256:d02a72844fb73d2f0e2285f910874030f6032a6bc30e6a3779bf7c0bc6f5a983

Observation e7894ad5-706a-4fa4-be77-15675d7f7b94 · outbound

This paper cites However, the performance of such systems is highly dependent on the qual- ity of the provided queries.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR However, the performance of such systems is highly dependent on the qual- ity of the provided queries

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:55.395872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:44.310471Z digest=sha256:a1cec508cffd05307a378b8927516d0b923fccff6b26fee723bb95e337cbb87c

Observation 657f274a-3481-4526-aa89-a3d8cea78902 · outbound

This paper cites there are the characters you can see.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR there are the characters you can see

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:55.215436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:44.386032Z digest=sha256:09412bcef6fc0c29028d476456c14631bba7760d2e315c2276eae5a086cf83c1

Observation 3d5c0200-a639-42dd-88dc-6bb3bee1ec8b · outbound

This paper cites an unresolved cited work.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:07:54.858078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:44.566702Z digest=sha256:61f23e763beb3f677ec02d8a17cc2600057690e6b07c22d39d82c4ff5189ef79

Observation f82c0efb-db0a-4284-ad13-779cf237084a · outbound

This paper cites Learning hidden unit contri- butions for unsupervised speaker adaptation of neural network acoustic models,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Learning hidden unit contri- butions for unsupervised speaker adaptation of neural network acoustic models,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:54.089504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:45.033171Z digest=sha256:dd95096d913b201acf77f073224cd76e02cd69c9713b01cff9f80fee93a1eb89

Observation b1f1f8a4-b051-435c-a206-6168c5831889 · outbound

This paper cites Speaker adap- tation of neural network acoustic models using i-vectors,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Speaker adap- tation of neural network acoustic models using i-vectors,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:54.698967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:44.663080Z digest=sha256:9364d6e70450637e0c72565dea02b8ca7092558d0f54f2e859cb0a63861c2357

Observation 254af8b5-ec3d-4305-81d7-83740cf41295 · outbound

This paper cites Front-end factor analysis for speaker verification,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Front-end factor analysis for speaker verification,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:54.558929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:44.759713Z digest=sha256:96647fa318bfec60c2c83ed6687e69feec992024776c1dfd983d9e4bc17163ba

Observation cff07ccf-4fc5-4fa1-b960-cdf3841d21cc · outbound

This paper cites Fast speaker adaptation of hy- brid nn/hmm model for speech recognition based on discrimina- tive learning of speaker code,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Fast speaker adaptation of hy- brid nn/hmm model for speech recognition based on discrimina- tive learning of speaker code,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:54.415568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:44.859203Z digest=sha256:d605b6a88efedb5c1af9e25ac6cb04d4da185cdd8a554d05060abc3825ae7a7e

Observation d90a4b5d-d156-45c4-b9b7-a80d664dbbbd · outbound

This paper cites Speaker normalization using efficient fre- quency warping procedures,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Speaker normalization using efficient fre- quency warping procedures,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:54.255499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:44.934237Z digest=sha256:c976c66871fb6b21249847aed161eff2eeb501095fc8bd6cad6249ba35bbd8c3

Observation 2ba42655-1d87-4a09-9d49-fed0efeefad8 · outbound

This paper cites The STC System for the CHiME-6 Challenge,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR The STC System for the CHiME-6 Challenge,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:53.457041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:45.450138Z digest=sha256:15364831cd2a50c653c8e862bc69e4998fb1de90eed72a6b9dffed3476c25265

Observation d7459d1d-93c9-403c-8cc2-88ca0d618584 · outbound

This paper cites Parameterised sigmoid and relu hidden activation functions for dnn acoustic modelling,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Parameterised sigmoid and relu hidden activation functions for dnn acoustic modelling,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:53.951772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:45.120263Z digest=sha256:5f7bddf520985c0d2bd3da62fea8ea3aebb338a8c09661cf3b41dfba5e6a6385

Observation 09ea5ed9-9274-4ac8-89f5-e6d21e126ee8 · outbound

This paper cites Low-rank plus diagonal adapta- tion for deep neural networks,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Low-rank plus diagonal adapta- tion for deep neural networks,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:53.757809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:45.204118Z digest=sha256:e6684767d94b95acc5219e7793dbc7f41ef05039aea82cc678b29e58737ef494

Observation c52eaa62-d4a3-465b-96c3-bb92d5553116 · outbound

This paper cites Speaker adaptation for Wav2vec2 based dysarthric ASR.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Speaker adaptation for Wav2vec2 based dysarthric ASR

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T22:07:45.299812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:07:45.299812Z digest=sha256:fcf629b3725eafba8c49107822b1826f7ae3ed083276c98a979dfa49174f42a6

Observation 2323e8dd-936e-4e77-9a8a-3b1e314c4f6f · outbound

This paper cites Confidence score based con- former speaker adaptation for speech recognition,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Confidence score based con- former speaker adaptation for speech recognition,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:53.582236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:45.313330Z digest=sha256:f5066961641006de62513c10efc30dc26a74f3f6db0243216c37867eca0edb38

Observation a7c944ef-cbe6-4249-97c8-9980c380313f · outbound

This paper cites Attention Is All You Need.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Attention Is All You Need

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T22:07:46.100528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:07:46.100528Z digest=sha256:03cd9ff721b36c87e380ea09556e591d449cbfdd1124b880c0e64081cc3191f1

Observation 54351c53-45a3-4cc2-adc8-5451d5b73767 · outbound

This paper cites Front- End Processing for the CHiME-5 Dinner Party Scenario,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Front- End Processing for the CHiME-5 Dinner Party Scenario,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:53.292204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:45.636591Z digest=sha256:9865b27d6c4f9aef254b7aaae20ba282fcb656bae54a96f857c9b9ad2cd67871

Observation 91b9dc7c-8282-423a-8255-6a0741453cb0 · outbound

This paper cites Serialized Output Training for End-to- End Overlapped Speech Recognition,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Serialized Output Training for End-to- End Overlapped Speech Recognition,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:53.162733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:45.799776Z digest=sha256:d2beebc8cce2977590baa466a5d59b7e168a564779b605b7dcb6fb795072ac18

Observation e1850345-326b-4ed5-8812-868e8c00024e · outbound

This paper cites End-to-End Monaural Multi-Speaker ASR System without Pretraining,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR End-to-End Monaural Multi-Speaker ASR System without Pretraining,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:52.958274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:45.921770Z digest=sha256:1e7a40e8672306e981ad1cb3a3a03ba79f2adb20cc142824eaf9df05ed627e68

Observation 9c8e1428-563e-4f17-aee3-45512159a1bd · outbound

This paper cites MIMO-SPEECH: End-to-end Multi- Channel Multi-Speaker Speech Recognition,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR MIMO-SPEECH: End-to-end Multi- Channel Multi-Speaker Speech Recognition,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:52.808889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:46.002063Z digest=sha256:528b4452272bd1c3096ac4fcaf3eb950d96c92dc58f6581937f193dc060116f5

Observation 4b4464c3-864f-47fc-8beb-62d371584eed · outbound

This paper cites Alignment-Free Training for Transducer-based Multi-Talker ASR.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Alignment-Free Training for Transducer-based Multi-Talker ASR

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:07:48.741731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:46.598768Z digest=sha256:9eec6bd829a04dc597b38adab0931a1cb5820974604fc11f2659bee249c9046c

Observation 1d465741-e94e-42de-bf7b-161997c99f33 · outbound

This paper cites Joint Speaker Counting, Speech Recognition, and Speaker Identification for Overlapped Speech of Any Number of Speakers,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Joint Speaker Counting, Speech Recognition, and Speaker Identification for Overlapped Speech of Any Number of Speakers,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:52.531087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:46.168226Z digest=sha256:ef35ae68fc9b1338dacb87f35b34f81a5bd2e907dacc85ee98b52485a25ca5fd

Observation 59f3b787-2415-400b-a260-41f4e1a23e58 · outbound

This paper cites Serialized Output Training by Learned Dominance.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Serialized Output Training by Learned Dominance

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:07:49.088166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:46.277354Z digest=sha256:55bbe1243193d653e8ed6397746caccc323d72be4faf3ed2165cfa18ec002659

Observation 786d5b6c-960a-4add-bb5f-b3bb1138de3e · outbound

This paper cites Sa-sot: Speaker- aware serialized output training for multi-talker asr,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Sa-sot: Speaker- aware serialized output training for multi-talker asr,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:52.352124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:46.397084Z digest=sha256:1fced4e2cc490ccb0dcded616ba270f2bd5308f1263e1137c9d75efa9359ace5

Observation d9f826b6-3b8e-4791-836e-6e9e9f76ff12 · outbound

This paper cites End-to-End Speaker-Attributed ASR with Transformer.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR End-to-End Speaker-Attributed ASR with Transformer

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:07:48.928968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:46.485730Z digest=sha256:07821afd1709a3b924e26c03fd3a994b3dd8c4c1bb1520613d1035bf664fcaf5

Observation c6996e68-cce0-45d3-b528-35210bfe662e · outbound

This paper cites Streaming Sortformer: Speaker Cache-Based Online Speaker Diarization with Arrival-Time Or- dering,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Streaming Sortformer: Speaker Cache-Based Online Speaker Diarization with Arrival-Time Or- dering,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:51.247647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:47.116575Z digest=sha256:737d7d6814f623a52a815b450cbb9e8654e8541735e68a0c1b72dc048c47cda0

Observation be70a8d0-d8fa-41bd-b315-6f2bee44b997 · outbound

This paper cites Di- cow: Diarization-conditioned whisper for target speaker au- tomatic speech recognition,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Di- cow: Diarization-conditioned whisper for target speaker au- tomatic speech recognition,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:52.125698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:46.722266Z digest=sha256:0e13a7f4849d4cbc9a78de98ca320e33d6d473a52fe3a104fad58b2317354911

Observation e513ee35-cc41-4b6f-a5b2-d683ac84762b · outbound

This paper cites Fast conformer with linearly scalable attention for efficient speech recognition,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Fast conformer with linearly scalable attention for efficient speech recognition,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:51.909567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:46.828184Z digest=sha256:5f6d54e275807af55da98c846f48a93bbe3406602823445f5e011e74e96a46f4

Observation b6a02b00-b849-4a83-b397-98809b7a18c5 · outbound

This paper cites Personal vad: Speaker-conditioned voice activity detection,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Personal vad: Speaker-conditioned voice activity detection,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:51.692437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:46.937342Z digest=sha256:4be67d0a18a66d3d58457b58f640b2e8c2f78b7016405887b8e47d830c347987

Observation a1fd8bfa-f722-4e63-b8c8-cccd3fb74e78 · outbound

This paper cites Stateful conformer with cache-based inference for streaming au- tomatic speech recognition,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Stateful conformer with cache-based inference for streaming au- tomatic speech recognition,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:51.413675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:47.027005Z digest=sha256:8fe1f6ad856e329f79ad40592ab9bd450553b7f871cfc0437bf07cd18e8d04e8

Observation eff046bd-78f9-43c0-8793-d0fe98986aed · outbound

This paper cites an unresolved cited work.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Unresolved cited work

Reference 30

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T22:07:55.039997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:44.481041Z digest=sha256:643a5e20958b1afc2aeb31832f2cca7c96b4bcb9341254b150725b6d6a4256de

Observation 44c85ab2-f0a5-4e82-bb92-76de185c0ed9 · outbound

This paper cites Lib- rispeech: an asr corpus based on public domain audio books,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Lib- rispeech: an asr corpus based on public domain audio books,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:51.023499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:47.184773Z digest=sha256:d2721f761f2667c23b0bd9f90f2cee8f80b2d739d9e8969616a7804408023e0d

Observation 8b654938-2300-419d-a166-6ab0cdb419d6 · outbound

This paper cites The Fisher Corpus: A Resource for the Next Generations of Speech-to-text,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR The Fisher Corpus: A Resource for the Next Generations of Speech-to-text,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:50.750431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:47.265546Z digest=sha256:cca815a21d50d3ecbc9386230d9a9b1d1f19274dea8bfd872285fa56aaf7f5d1

Observation eb8aeb34-e50f-4c90-9427-9b9715b47b90 · outbound

This paper cites Callhome american english speech,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Callhome american english speech,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T22:07:47.345797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:07:47.345797Z digest=sha256:21b0e5ecfd19ad84da869740a15d3407310faea3582eb35eaf743e6ae0477827

Observation f385c3cc-70f2-4888-9820-c62456d4e843 · outbound

This paper cites CHiME-6 Chal- lenge: Tackling Multispeaker Speech Recognition for Unseg- mented Recordings,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR CHiME-6 Chal- lenge: Tackling Multispeaker Speech Recognition for Unseg- mented Recordings,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:50.560820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:47.497977Z digest=sha256:69eebd908e948169885dc89bf4c55b498d3571e4248a03ed0bc9039e380ef61a

Observation d29b06c6-a457-4291-9d69-02ce65b9c23a · outbound

This paper cites Sortformer: A Novel Approach for Permutation-Resolved Speaker Supervision in Speech-to-Text Systems.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Sortformer: A Novel Approach for Permutation-Resolved Speaker Supervision in Speech-to-Text Systems

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T22:07:47.575403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:07:47.575403Z digest=sha256:8a72a765fa116fa4be63fd033cac7cf001fa85cdacce349c7c12495ed73c4c83

Observation 8204684c-4ced-4259-a74f-0e420cc4b7c8 · outbound

This paper cites SentencePiece: A Simple and Lan- guage Independent Subword Tokenizer and Detokenizer for Neu- ral Text Processing,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR SentencePiece: A Simple and Lan- guage Independent Subword Tokenizer and Detokenizer for Neu- ral Text Processing,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:50.372984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:47.660954Z digest=sha256:0ce41f229ea7f3b887ea2178a95f7dd5d77e1f92d9e487deb8cb9f6d11ee82d7

Observation de571186-0418-41c9-84b5-65540ff5994f · outbound

This paper cites Decoupled Weight Decay Regularization.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Decoupled Weight Decay Regularization

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T22:07:47.756399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:07:47.756399Z digest=sha256:654ca68567f4276c0636d720842a061cb83eb06b04766f5a22649c5c06c20c3c

Observation 8a4637b6-b75f-40f0-b5ec-4bc4d87df361 · outbound

This paper cites A sidecar separator can convert a single-talker speech recognition system to a multi-talker one,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR A sidecar separator can convert a single-talker speech recognition system to a multi-talker one,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T22:07:47.818489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:07:47.818489Z digest=sha256:96ea4062e2367722267ffe4aed912d3f1e3f76278a421b2192be145ea247574f

Observation 4024ca9e-0a38-4c69-bc39-8e6548e256f3 · outbound

This paper cites Empowering whisper as a joint multi-talker and target-talker speech recognition system,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Empowering whisper as a joint multi-talker and target-talker speech recognition system,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:50.198384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:47.919372Z digest=sha256:3e031174ae4be2d566c73256c491783631364c3aef32aae63f3448a14b1ad20d

Observation 9eafb0da-c064-4bb4-9043-77cfcda37df1 · outbound

This paper cites Large Language Model Can Transcribe Speech in Multi-Talker Scenarios with Versatile Instructions.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Large Language Model Can Transcribe Speech in Multi-Talker Scenarios with Versatile Instructions

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:07:48.514083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:47.977602Z digest=sha256:91c9650a59c5b9249e5e90b9a63854b767b3fc465f0ae78fc35a005829fadaa2

Observation 18d4b92f-a343-4a6c-b593-7105ea9541f5 · outbound

This paper cites Streaming speaker-attributed asr with token-level speaker embeddings,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Streaming speaker-attributed asr with token-level speaker embeddings,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:49.959413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:48.052907Z digest=sha256:38367b00223137f279b06c7d0d553c3c79243c8d93203433170b8740554a942a

Observation 04347f6a-c426-4031-a461-4c894e0ad486 · outbound

This paper cites Self-supervised learning with bi- label masked speech prediction for streaming multi-talker speech recognition,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Self-supervised learning with bi- label masked speech prediction for streaming multi-talker speech recognition,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:49.761065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:48.135079Z digest=sha256:a5b49ff81676c1db951fd751f5bbca5716697de867533d7ec459d71eb2ef48ed

Observation dcc7a032-fd04-4053-96c8-470bfba1d33f · outbound

This paper cites Enhancing Speaker Diarization with Large Language Models: A Contextual Beam Search Approach,.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Enhancing Speaker Diarization with Large Language Models: A Contextual Beam Search Approach,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:07:49.571916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:48.267842Z digest=sha256:554c7404417185f425d4b8526de6e422dc015fd7f21f8a1746ac3d726f0956be

Pith citing papers

Observation a4cc5d57-b0cb-43ca-9daf-ea7b1bdfb564 · inbound

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR cites this paper.

Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:07:49.354943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:07:44.192140Z digest=sha256:d02a72844fb73d2f0e2285f910874030f6032a6bc30e6a3779bf7c0bc6f5a983