Pith. sign in

Paper Citation Record · LEDGER

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling

As of 10 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 1 inbound Pith citation observation for arXiv:2507.09510.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.09510 v3

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:58:41.063893Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:58:34.654304Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T17:58:42.865534Z

Reference resolution

51 of 51 outbound references displayed

  • verified exact6
  • verified fuzzy38
  • unresolved5
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 870e7914-99ff-4e70-b38b-98f092d64188 · outbound

This paper cites The chal- lenge of isolating the target speech while ignoring other inter- ferences is known as the cocktail party problem [1].

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling The chal- lenge of isolating the target speech while ignoring other inter- ferences is known as the cocktail party problem [1]

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:53.038346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:34.470611Z digest=sha256:5fba5260b7369d0e632b8c84c58d4bab4a284ae5b9fbffa526fbb4ed95083c35

Observation 460796ea-b9b2-49bc-a537-08fdb67af52d · outbound

This paper cites Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:58:42.987162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:34.654304Z digest=sha256:855612bc51a3d1c996344552c26330adc83faede16a8951f4795cc47a0f12c03

Observation 5eb2f227-2ea8-410d-b8b2-2201a0991e5a · outbound

This paper cites The proposed methods enhance overall performance, achieving state-of-the-art results.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling The proposed methods enhance overall performance, achieving state-of-the-art results

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:52.741434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:34.551888Z digest=sha256:45c9c5dd1615f0c7831791ade56893f0d46e2fb46ba3080171729076862b24b1

Observation cab9f041-ad09-4942-af0c-63b64cb3edf0 · outbound

This paper cites Dataset We use the clean Libri2Mix [26] dataset with two-speaker mix- tures.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Dataset We use the clean Libri2Mix [26] dataset with two-speaker mix- tures

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:52.458594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:34.793706Z digest=sha256:012bf3c0fd9a2396dec95ea041e80abff5867269bfc71d79be9537ca05138ca7

Observation c075f8a8-3d53-4810-bb23-37fc9b388278 · outbound

This paper cites an unresolved cited work.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:58:50.898381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:35.711552Z digest=sha256:259e8341561544567dac400310ea9e9fd93147e6a9eba3858f9074dd828a010c

Observation 337e1f65-c57a-4676-86b4-feffcd802435 · outbound

This paper cites Comparative studies with proposed methods Table 1 presents the performance of BSRNN with proposed methods on Libri2Mix.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Comparative studies with proposed methods Table 1 presents the performance of BSRNN with proposed methods on Libri2Mix

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:51.893049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:35.033551Z digest=sha256:420461a93e853a025abb46c5af2dcca4f05715f0b47cec4e711d3ef5ea94e38d

Observation db43320f-9074-446f-bf5f-854a9124fb86 · outbound

This paper cites and other metrics in all scenarios.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling and other metrics in all scenarios

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:51.611443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:35.199613Z digest=sha256:1f6e84abfc65d6b09c8b23d1412dbd4e5f6c7fdd6f9213105fd0fafa6f4b2a53

Observation 822d6794-86b9-46b5-8b5b-4d94271d6249 · outbound

This paper cites For ResNet34 in pretrained mode, SI-SDR improved by 0.43 dB, accuracy by 1.56%, and Sim.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling For ResNet34 in pretrained mode, SI-SDR improved by 0.43 dB, accuracy by 1.56%, and Sim

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:51.397031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:35.317717Z digest=sha256:00dc727f1d53f33f54edf6da38201fe55e7dd2f0330ab4d244682da5c81cb356

Observation 28f2b8ea-0ebd-44eb-9479-22d2da22a519 · outbound

This paper cites However, a slight decrease in the Sim.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling However, a slight decrease in the Sim

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:51.118859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:35.445487Z digest=sha256:7cc0afa001ae1bdd12bb53d560026c31a60efdaf52cb617d0cce1d89e246373f

Observation f050ae84-64f6-4b44-bd57-609f7e050a6b · outbound

This paper cites an unresolved cited work.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Unresolved cited work

Reference 10

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T17:58:42.746251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:35.564318Z digest=sha256:567d152374b071c48a6c2747cf461f0d4df967742d2ef9077e3cab4b0c8c0cd2

Observation 57c6b5e7-099a-4b66-b1c0-2b9c376c295c · outbound

This paper cites Spex+: A complete time domain speaker extraction network,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Spex+: A complete time domain speaker extraction network,

Reference 11

Resolution
verified exact
doi, observed 2026-08-06T17:58:41.640664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:37.093240Z digest=sha256:b0a8bf18d86b9f5d67de059e7cacf744260f68f5ea38180eedf37ac5b9cd2cfe

Observation 4e26e3b0-37c9-44a3-aacb-3604a1ade830 · outbound

This paper cites Some further experiments upon the recognition of speech, with one and with two ears,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Some further experiments upon the recognition of speech, with one and with two ears,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:50.695497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:35.827574Z digest=sha256:e443a3302deadc8289a99240420ceafe93172d8c53cc540370860031086955f3

Observation 861f1f0d-d5cb-473c-8324-3503423679c5 · outbound

This paper cites The cocktail-party problem revisited: early process- ing and selection of multi-talker speech,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling The cocktail-party problem revisited: early process- ing and selection of multi-talker speech,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:50.464346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:35.953626Z digest=sha256:598864592f289921cf33f2232dd6721f98c430d6cd9ab8bd581fa278f57f40ed

Observation 7d027329-0848-4d74-a6a2-05b65b530330 · outbound

This paper cites Speakerbeam: Speaker aware neural network for target speaker extraction in speech mixtures,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Speakerbeam: Speaker aware neural network for target speaker extraction in speech mixtures,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:50.159472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:36.061089Z digest=sha256:edef7957ffbb13be8601bd4bc3640e7c111bf5a5c61249a7e7da0d744e5df39b

Observation 22189dbc-c058-4ea1-abee-5f7698904856 · outbound

This paper cites Adaptive blind audio source extraction supervised by dom- inant speaker identification using x-vectors,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Adaptive blind audio source extraction supervised by dom- inant speaker identification using x-vectors,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:49.916815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:36.196820Z digest=sha256:bd9edfd67bf36d9efec6cdf2791cbfa3b85b2246512e30a236417fff8a1d6295

Observation 79c0e0c6-74b9-4def-8d3a-9f331e82fed9 · outbound

This paper cites Looking to listen at the cocktail party: A speaker-independent audio-visual model for speech separation,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Looking to listen at the cocktail party: A speaker-independent audio-visual model for speech separation,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T17:58:36.334856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:58:36.334856Z digest=sha256:103b960ce0a7d31d56842dfce4c5c749311171174bd1dfb0341f25c35ae45549

Observation f6b3d310-16e3-40f0-a688-f132e61e3109 · outbound

This paper cites Audio-visual sound separation via hidden markov models,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Audio-visual sound separation via hidden markov models,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:49.657251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:36.449562Z digest=sha256:f8596d639d43fce502ee61a7a78618eaafc98c9ea73a186f147e86f13020f0b1

Observation af15a8ea-4823-4eef-a39e-c93bd86e7267 · outbound

This paper cites Neural spatial filter: Target speaker speech separation assisted with directional information,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Neural spatial filter: Target speaker speech separation assisted with directional information,

Reference 18

Resolution
verified exact
doi, observed 2026-08-06T17:58:41.908685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:36.579776Z digest=sha256:8ecb908d6f283312c1ec37547a5f368478a2396fa74ecc08b910c9c93b251f60

Observation 0f880b43-d2f2-41c2-8626-efe03a73eef1 · outbound

This paper cites V oice filter: Few- shot text-to-speech speaker adaptation using voice conversion as a post- processing module,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling V oice filter: Few- shot text-to-speech speaker adaptation using voice conversion as a post- processing module,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:49.408598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:36.727248Z digest=sha256:4f7336319388375778c1862a9a81c7fa5060d7ea48e4352148d64dc9833e4347

Observation 5c1b3112-7e5e-434f-95f4-fe3f2eedf8d1 · outbound

This paper cites Single-channel speech extraction using speaker inventory and attention network,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Single-channel speech extraction using speaker inventory and attention network,

Reference 20

Resolution
malformed identifier
no resolver link, observed 2026-08-06T17:58:36.849729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:58:36.849729Z digest=sha256:80ce0ac58be22d457319e0bbf99088cac276d5a23d8f2083b2e6cd8865b042e0

Observation 0565405d-e347-4ded-91e0-f928a88263ed · outbound

This paper cites Target confusion in end- to-end speaker extraction: Analysis and approaches,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Target confusion in end- to-end speaker extraction: Analysis and approaches,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:49.075349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:36.966592Z digest=sha256:54a6cff42ed283c502b0f982dd2cef4a41425095a42ff6794aa2e82c5e20d5f3

Observation f9a99aaa-3293-48a6-ae46-d0c273b5b740 · outbound

This paper cites Sc-glowtts: an efficient zero-shot multi-speaker text-to-speech model,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Sc-glowtts: an efficient zero-shot multi-speaker text-to-speech model,

Reference 22

Resolution
verified exact
doi, observed 2026-08-06T17:58:41.332281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:38.645691Z digest=sha256:89cef1842351dc4657f85043e4dd2140a1ff52467042830df805f60fd4601cdb

Observation 6c39300c-79b4-4bd9-81f3-510e8762c787 · outbound

This paper cites X-sepformer: End-to-end speaker extraction network with explicit optimization on speaker confusion,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling X-sepformer: End-to-end speaker extraction network with explicit optimization on speaker confusion,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:48.848806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:37.191473Z digest=sha256:594361321de235b4da369da2cd2bd488d8afbf3f8913594780abd784c867c3ac

Observation 5b01aa03-ecc7-4709-9135-d444296e4cfd · outbound

This paper cites Speaker ex- traction with detection of presence and absence of target speakers,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Speaker ex- traction with detection of presence and absence of target speakers,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:48.588117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:37.334445Z digest=sha256:e5d9dacba4fafc40a0ba4d467f53343e711047b1738785b4ac6d2875fcd4214c

Observation a27b6499-477a-4aa7-9b33-2cb633290a4c · outbound

This paper cites On the effectiveness of enrollment speech augmentation for target speaker extraction,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling On the effectiveness of enrollment speech augmentation for target speaker extraction,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:48.321486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:37.493824Z digest=sha256:5fddc9d5a4d00749ee4282fb465774033c2c00eba256765ce965e29853556ad5

Observation 65129ac3-5d07-47d4-a9ac-157856c1a130 · outbound

This paper cites Selective hubert: Self- supervised pre-training for target speaker in clean and mixture speech,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Selective hubert: Self- supervised pre-training for target speaker in clean and mixture speech,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:47.974497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:37.612121Z digest=sha256:f4abe4e8134086d11d063bc66240673ddc762469fee340e9513f867a0c970df5

Observation 7ec856e6-7044-4601-8c3e-eae9808aedd0 · outbound

This paper cites Contrastive learning for target speaker extraction with attention-based fusion,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Contrastive learning for target speaker extraction with attention-based fusion,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:47.452568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:37.916014Z digest=sha256:f41d3a35862e009d1b76679a01686435210925840f7604c834ebde424ce6d7aa

Observation 5ec61414-d81e-4b3f-9672-5da76eb3834c · outbound

This paper cites Real-time personalised speech enhancement transformers with dynamic cross-attended speaker representations,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Real-time personalised speech enhancement transformers with dynamic cross-attended speaker representations,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:47.148722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:38.081178Z digest=sha256:67735ba1d57304bcaa2dace79219ba457e553aa3bda90e1c1ca810c3a98844ab

Observation dc186a4c-5358-4844-a292-4e3437f233aa · outbound

This paper cites Smma-net: An audio clue- based target speaker extraction network with spectrogram matching and mu- tual attention,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Smma-net: An audio clue- based target speaker extraction network with spectrogram matching and mu- tual attention,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:46.806035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:38.291163Z digest=sha256:60d6e472ec5d8e651392e1e55b698c3a4896775ea70882fc8dff1ada9b6faa6e

Observation fbb47c09-a1a4-4457-b8f6-55b72c625c02 · outbound

This paper cites Target speaker extraction by di- rectly exploiting contextual information in the time-frequency domain,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Target speaker extraction by di- rectly exploiting contextual information in the time-frequency domain,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:46.475563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:38.420322Z digest=sha256:8581a5ca636f96628438953777f23e6ffb4e273a5dd187d717b7ed6fd7815754

Observation 4d180dee-aad7-483b-85c5-dc9bd8c3c1cb · outbound

This paper cites Music source separation with band-split rnn,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Music source separation with band-split rnn,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:46.199355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:38.536836Z digest=sha256:ec75b2cccaa4ff1e1776a1dfa5e96beba575ebf15276f6f411f41074761a8410

Observation c76a0f6f-7617-463d-b028-365e09664def · outbound

This paper cites An algorithm for intelligibility prediction of time–frequency weighted noisy speech,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling An algorithm for intelligibility prediction of time–frequency weighted noisy speech,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:44.836454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:40.037370Z digest=sha256:cc8c3c13b50c6725c424960bfbcddcc859bd3879b5fb0697c49dd4f1a65198a3

Observation f235ec62-893c-47f2-8c67-67648fa1cd8b · outbound

This paper cites In defence of metric learning for speaker recognition,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling In defence of metric learning for speaker recognition,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.995819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:38.754873Z digest=sha256:85aae337b0f7a85815ff2e4e8aefe39726bbe1d20507050183d3824d716057c0

Observation ec33ac18-6273-44d0-a7e4-c91995132715 · outbound

This paper cites Centroid estimation with transformer-based speaker embedder for robust target speaker extraction,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Centroid estimation with transformer-based speaker embedder for robust target speaker extraction,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.802159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:38.870085Z digest=sha256:e3b37329b0b5c74c3469c88e532d8ad24e89ae0a2b355630ce9d5fe07355caf5

Observation 7553ccb9-467f-4907-8590-877601927c26 · outbound

This paper cites Sdr–half-baked or well done?.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Sdr–half-baked or well done?

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.638351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:39.010501Z digest=sha256:c432c083265034684815d5f4a3c44c16086d48e193675b9ff7662b25106475e9

Observation 31d26293-98a0-43d4-9c27-cc536da3f3d7 · outbound

This paper cites Lib- rimix: An open-source dataset for generalizable speech separation,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Lib- rimix: An open-source dataset for generalizable speech separation,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.473589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:39.151681Z digest=sha256:c9f907af0ffca463216de4497d2400ebca1d2798ac0b41ec35868be0af4268cf

Observation 49af23a8-a315-4150-a6f8-c19823114eaf · outbound

This paper cites But system description to voxceleb speaker recognition challenge 2019,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling But system description to voxceleb speaker recognition challenge 2019,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.304615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:39.253114Z digest=sha256:e3647e1d73c92328eaf8757903015ed3f7837a9a63da05733af2cf9fb74a1a73

Observation f58f258f-bf1c-4927-9cfe-01972083417b · outbound

This paper cites Ecapa-tdnn : Emphasized channel attention, propagation and aggregation in tdnn based speaker verification,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Ecapa-tdnn : Emphasized channel attention, propagation and aggregation in tdnn based speaker verification,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T17:58:39.400949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:58:39.400949Z digest=sha256:6f36f1f196790ad0228df8d6cb97a582322c12e6f10123b603f2129aaaed23a0

Observation b8624695-9fa7-4885-92e4-b63d42c05b9f · outbound

This paper cites Wespeaker: A research and production oriented speaker embed- ding learning toolkit,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Wespeaker: A research and production oriented speaker embed- ding learning toolkit,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.136229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:39.509318Z digest=sha256:3700e01b6e820aaea50364ffe6d125bfee37eb14fafc63ef2ac9a45f753cb274

Observation d861da22-9724-4e98-9b00-3e6cf57d6b19 · outbound

This paper cites V oxceleb2: Deep speaker recognition,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling V oxceleb2: Deep speaker recognition,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T17:58:39.616004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:58:39.616004Z digest=sha256:fc09ba5746827c9bd220f68f1e5a2e244f4d093bcfe036dcf45e6de66cc031c0

Observation 9beb7c1c-b899-43c8-a306-d54f488fcbd1 · outbound

This paper cites WeSep: A Scalable and Flexible Toolkit Towards Generalizable Target Speaker Extraction.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling WeSep: A Scalable and Flexible Toolkit Towards Generalizable Target Speaker Extraction

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:58:42.352929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:39.732096Z digest=sha256:459e7a49c2913605c1d762b92b7cf49dfd1dc0738e71bc200049af5adfef8377

Observation 19611a75-f91d-49d7-82f2-05ece02de0c5 · outbound

This paper cites Perceptual evaluation of speech quality (pesq): An objective method for end-to-end speech quality assessment of narrow-band telephone networks and speech codecs,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Perceptual evaluation of speech quality (pesq): An objective method for end-to-end speech quality assessment of narrow-band telephone networks and speech codecs,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:44.962435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:39.867589Z digest=sha256:3a02d868f2c85cd0380fde2f71932e6432248c1b28c629cf9b2e2e1ec340bdf6

Observation 2b029e47-5f9f-4c78-a18c-4ff84e909f84 · outbound

This paper cites Xtts: a massively mul- tilingual zero-shot text-to-speech model,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Xtts: a massively mul- tilingual zero-shot text-to-speech model,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:44.620467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:40.186751Z digest=sha256:f3e122960ea495bbb1ec5887f74f29634da7f1a0d47447fcad4629f39d466dfc

Observation 08b27bca-02f0-49a5-9cbb-97cca54c4648 · outbound

This paper cites Ecapa2: A hybrid neural network archi- tecture and training strategy for robust speaker embeddings,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Ecapa2: A hybrid neural network archi- tecture and training strategy for robust speaker embeddings,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:44.343895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:40.303176Z digest=sha256:2a39db40071e41453dee22cc4cfc7d6747860052959903499dc277e1c7ad24bb

Observation 4cb35800-a4d5-4da0-8aa0-d612e50c931e · outbound

This paper cites Multi-Level Speaker Representation for Target Speaker Extraction.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Multi-Level Speaker Representation for Target Speaker Extraction

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:58:42.148745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:40.451824Z digest=sha256:c897f2c0d530da9951feda93cf27528b9efa1f842b1b2b3bf14b4213617d1ddb

Observation 053c27d9-27b9-4d9a-a5eb-f6d36f47f9b4 · outbound

This paper cites Dpccn: Densely-connected pyramid complex convolutional network for robust speech separation and extraction,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Dpccn: Densely-connected pyramid complex convolutional network for robust speech separation and extraction,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:44.054911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:40.578466Z digest=sha256:d5ef78afc28f53564ac3a0f2fbb2dbc22bd6848b32108663d047fa6584bdd50e

Observation 9850ff13-8b76-4491-a351-cabbcf7af160 · outbound

This paper cites Mc-spex: Towards effective speaker extraction with multi-scale interfu- sion and conditional speaker modulation,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Mc-spex: Towards effective speaker extraction with multi-scale interfu- sion and conditional speaker modulation,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:43.814757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:40.670334Z digest=sha256:51250da75f14cbb5fc61e1dfc1340aa05d1759da8a4905f93421525b4fa7d870

Observation de256083-5227-41d8-b2ef-3a28a4986f1a · outbound

This paper cites Tar- get speech extraction with pre-trained self-supervised learning models,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Tar- get speech extraction with pre-trained self-supervised learning models,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:47.697564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:40.797574Z digest=sha256:eb694a7bc9608ba246f803f84a735ab50d0d5330a2d7f1aa4ae49aa705edca10

Observation 531299cb-327f-401b-804b-7061af17d7fb · outbound

This paper cites Tf- gridnet: Making time-frequency domain models great again for monaural speaker separation,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Tf- gridnet: Making time-frequency domain models great again for monaural speaker separation,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:43.595248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:40.942430Z digest=sha256:08a21f358908d47e82e33587d3ef011429acd5a0946c2bc9f1862712b622d96c

Observation 8e319a68-b513-4e42-8018-03d56c53fe1a · outbound

This paper cites V oxceleb: A large-scale speaker identification dataset,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling V oxceleb: A large-scale speaker identification dataset,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:43.348936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:41.063893Z digest=sha256:2c8a60165eebe5b24239b41d33de6eb384ac877c46ad19555f0144876b9f0f70

Observation 4202cf03-9f11-4a00-80b4-ede8b5b33f86 · outbound

This paper cites an unresolved cited work.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Unresolved cited work

Reference 128

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:58:52.159552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:34.919879Z digest=sha256:8a88071d801033f54092d6549fa6183aaabaf2705d518ba196b39a599b9ebd12

Pith citing papers

Observation 460796ea-b9b2-49bc-a537-08fdb67af52d · inbound

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling cites this paper.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:58:42.987162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:58:34.654304Z digest=sha256:855612bc51a3d1c996344552c26330adc83faede16a8951f4795cc47a0f12c03