Pith. sign in

Paper Citation Record · LEDGER

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions

As of 11 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2502.08191.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.08191 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T10:06:36.061223Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact3
  • verified fuzzy25
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3747b99a-2b8c-4218-974f-f3bf502692f7 · outbound

This paper cites Some experiments on the recognition of speech, with one and with two ears,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Some experiments on the recognition of speech, with one and with two ears,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:35.923599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:35.923599Z digest=sha256:05af6c9bcb5215279793da9308bf7669a4eb114d33e3a55a265315c4adc61712

Observation 089cb7e3-f9e3-4167-9c3a-f6ab29ef4689 · outbound

This paper cites The cocktail party phenomenon revisited: The importance of working memory capacity,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions The cocktail party phenomenon revisited: The importance of working memory capacity,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.571334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:35.928089Z digest=sha256:dea4b8821173a55559ac5ce1dbfb65640fab2d64a34321a7baf314c999d3108c

Observation 2bea3240-35d6-4ad0-8511-c50204933188 · outbound

This paper cites An event-related potential study of selective auditory attention in children and adults,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions An event-related potential study of selective auditory attention in children and adults,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.556553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:35.932253Z digest=sha256:5558213d39bcb00ceae326e9cdaf918bfa04b561d44c0ef6e31be2b9292e6c6a

Observation 5857d84c-5ee4-4e63-abce-94cf6a016643 · outbound

This paper cites Selective cortical representation of attended speaker in multi-talker speech perception,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Selective cortical representation of attended speaker in multi-talker speech perception,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.543202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:35.937068Z digest=sha256:5e1a9bd4b35f0d94884016e90b1e758741549f31a413bef2f5bcab31f2f51d90

Observation 2e61c42f-8f5d-4c2d-a605-fcd7010b3ed8 · outbound

This paper cites Supervised speech separation based on deep learning: An overview,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Supervised speech separation based on deep learning: An overview,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:35.941317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:35.941317Z digest=sha256:97ea588838772e3568882462b857077a3124b2bbf062ab3483d34b7bdb200416

Observation eef047c9-031f-447f-88bf-8363d0e36557 · outbound

This paper cites Dual-path rnn: efficient long sequence modeling for time-domain single- channel speech separation,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Dual-path rnn: efficient long sequence modeling for time-domain single- channel speech separation,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.520393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:35.946093Z digest=sha256:c01b42bdcc467b9776814f12ec8676550c29a4b3b5645996fd89ae674b3bffb9

Observation c6eeedb6-77db-4bcb-9040-bc30aa612f7f · outbound

This paper cites Dual-Path Transformer Network: Direct Context-Aware Modeling for End-to-End Monaural Speech Separation.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Dual-Path Transformer Network: Direct Context-Aware Modeling for End-to-End Monaural Speech Separation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:35.950643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:35.950643Z digest=sha256:bb30211b4aa09864c7aecae1e176426083ea596546e1a823211d60d383037d01

Observation 92dc1bb7-45be-4719-9d83-723f08d71608 · outbound

This paper cites Tf-gridnet: Making time-frequency domain models great again for monaural speaker separation,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Tf-gridnet: Making time-frequency domain models great again for monaural speaker separation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.507409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:35.955445Z digest=sha256:5ec47e698e1b31b606c830628d3cbe94b349428efe6e68d9fa37a350637a5318

Observation 221bdceb-83a8-44c7-a5f5-b735fd482ccb · outbound

This paper cites CrossNet: Leveraging Global, Cross-Band, Narrow-Band, and Positional Encoding for Single- and Multi-Channel Speaker Separation.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions CrossNet: Leveraging Global, Cross-Band, Narrow-Band, and Positional Encoding for Single- and Multi-Channel Speaker Separation

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-08T10:06:36.189043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:35.959087Z digest=sha256:964d902826e92bb8ffa70ef00b436fe2c83c6bd1a7f9b56dda3fbfa487e0f903

Observation b94022f0-c4fe-48e1-8e4b-33f3e08fa950 · outbound

This paper cites Neural target speech extraction: An overview,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Neural target speech extraction: An overview,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.493809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:35.963133Z digest=sha256:25322192707ae19e3f399a425a2bbd36f3d2547524745f343baafd02616bacca

Observation 2210c1a4-6f72-4a10-a827-355990ffd65a · outbound

This paper cites VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:35.967023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:35.967023Z digest=sha256:154bdc2348bb8ba26ea4cced76086ff019406262b8679e1e4e0b538632afe7bb

Observation 48c25eb6-4f8b-47b3-b2fb-195fdaa8e166 · outbound

This paper cites Time-domain speaker extraction network,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Time-domain speaker extraction network,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.480225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:35.971084Z digest=sha256:9e21d341637197235ff8916a95365e2ce43218bfac1d8c5ee2b513064588218e

Observation 6651f7ab-cec5-42a8-b726-e770ba0fb542 · outbound

This paper cites Spex: Multi-scale time domain speaker extraction network,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Spex: Multi-scale time domain speaker extraction network,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.467488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:35.975088Z digest=sha256:f748f9ba5157a9f149cd663713ab08b2b6369a666e829c3fddeb80ff2eb298a2

Observation 6b72a265-deb3-4807-b0da-30af340934d6 · outbound

This paper cites SpEx+: A Complete Time Domain Speaker Extraction Network.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions SpEx+: A Complete Time Domain Speaker Extraction Network

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:35.978757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:35.978757Z digest=sha256:e65b25cff8de13f673aca9c73e91ebf2207e7ab0ca01a5059cf6b94cb6eb9955

Observation c9959fb2-8b07-40a2-adbd-4e8b258d9cac · outbound

This paper cites Adaptive-spex: Local and global perceptual modeling with speaker adaptation for target speaker extraction,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Adaptive-spex: Local and global perceptual modeling with speaker adaptation for target speaker extraction,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.454231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:35.982669Z digest=sha256:b7829ace73fda0b5f99c91bb30e4d961d2fd2c507d24ec65f5d1a7e459a260a8

Observation 4682cbad-a6bd-4abd-a463-b0d7ff93328a · outbound

This paper cites Multi-stage speaker extraction with utterance and frame-level reference signals,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Multi-stage speaker extraction with utterance and frame-level reference signals,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.441070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:35.986358Z digest=sha256:db3e9515674366d7aa7b62a26c5b69a564a5e9ba4275f86d644971b7290c0486

Observation 57a39665-19a3-422c-8372-87685bc5b386 · outbound

This paper cites Neural speaker extraction with speaker-speech cross-attention network.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Neural speaker extraction with speaker-speech cross-attention network

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.427846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:35.990204Z digest=sha256:31d0abdc25f9a0325a71ab5a5fc7ac7772c5beffdcec0b9c529217bd1992d37b

Observation 1e563c58-4a7d-4181-b7a9-78217b2db8e5 · outbound

This paper cites Robust Speaker Extraction Network Based on Iterative Refined Adaptation.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Robust Speaker Extraction Network Based on Iterative Refined Adaptation

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-08T10:06:36.140824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:35.993768Z digest=sha256:ea7a9f79891f390888607480712c57132facc565f398ce3b53ab57af46a1df23

Observation 0ace271d-d7ea-45ab-8126-78bae64fccb3 · outbound

This paper cites Target speaker extraction with ultra-short reference speech by ve-ve framework,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Target speaker extraction with ultra-short reference speech by ve-ve framework,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.414639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:35.997632Z digest=sha256:153bc1cf33313034e240f8429c4e6acbf96c3910a256ec1f432bf78a96c62e51

Observation 2cb27952-e0e2-48a1-b9ac-741ab2cd8a21 · outbound

This paper cites X-sepformer: End-to-end speaker extraction network with explicit optimization on speaker confusion,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions X-sepformer: End-to-end speaker extraction network with explicit optimization on speaker confusion,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.400779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:36.001194Z digest=sha256:23982f717d68e2d1c72aed422b258af5b53cc08387e98a09e2fbff06274eaac2

Observation 1bc63093-491e-4ee0-90c1-082d190792a7 · outbound

This paper cites Target speech extraction with pre-trained self-supervised learning models,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Target speech extraction with pre-trained self-supervised learning models,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.387373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:36.005004Z digest=sha256:30bd4dbc0aee8b1361c5492707e064d0aa46a0285e7a0994c5b79109439f53e0

Observation 49c8b380-dd29-4ecb-b1eb-3cadf93c1fc4 · outbound

This paper cites Speakerbeam: Speaker aware neural network for target speaker extraction in speech mixtures,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Speakerbeam: Speaker aware neural network for target speaker extraction in speech mixtures,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.372641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:36.008561Z digest=sha256:ba57d453845000cd3b57f74c7204911ec5f7cb255cf2e87bdbc5d77a8738f5fb

Observation 437363be-9fa3-4d84-9def-bbffade4d16b · outbound

This paper cites X-tf-gridnet: A time–frequency domain target speaker extraction network with adaptive speaker embedding fusion,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions X-tf-gridnet: A time–frequency domain target speaker extraction network with adaptive speaker embedding fusion,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.357442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:36.012214Z digest=sha256:a6b6e5d735e500125893f12b6cfb156c820755389e5a45593d802682feb1c356

Observation 008e86cc-b5c9-4fb0-a838-e68530ce945e · outbound

This paper cites Single-channel speech extraction using speaker inventory and attention network,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Single-channel speech extraction using speaker inventory and attention network,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.342905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:36.015841Z digest=sha256:2dc75acfe298472b2c687d070d99c4cfff59fb06bebd78d90e096b3b9c03d5c0

Observation 0348d1ef-8c9e-472a-8489-c33df89c7751 · outbound

This paper cites Sef-net: Speaker embedding free target speaker extraction network,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Sef-net: Speaker embedding free target speaker extraction network,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.329135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:36.019330Z digest=sha256:73972f0981492bb18f970fab2e98586dcb39acaedf2417390f8fbea376157747

Observation cf045db9-2cf4-43e4-8829-9899f3e07ab9 · outbound

This paper cites Target speaker extraction by directly exploiting contextual information in the time-frequency domain,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Target speaker extraction by directly exploiting contextual information in the time-frequency domain,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.316022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:36.023295Z digest=sha256:37ce4a6381425feef9c342ad498e03a41c315e1103e427424f926264b7e692e1

Observation 8450e850-d73e-4009-bd77-31f799426d3b · outbound

This paper cites On the importance of power compression and phase estimation in monaural speech dereverberation,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions On the importance of power compression and phase estimation in monaural speech dereverberation,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.302169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:36.026941Z digest=sha256:75885e54cb0e6dee1325f3fe783ca96d377ec3ad45c0332fa11ef18df90e8564

Observation 3d76b5a3-25e9-42c8-b7a8-6ea956d8929b · outbound

This paper cites Root mean square layer normalization,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Root mean square layer normalization,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:36.030834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:36.030834Z digest=sha256:66e2477f59be4e0e016a5dd97f49e9c4eef27eccc6bf9bdfe0844f62db836581

Observation f21e3ffd-f1ea-4a4b-a201-08815658a7e1 · outbound

This paper cites Single image reflection separation via component synergy,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Single image reflection separation via component synergy,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.279476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:36.034676Z digest=sha256:9dc5320c83ef68bfb13e01a74d768c1309858588ccf35826b7d4d610891a7933

Observation 833b93e9-6b0d-45eb-b5b9-fdddd2923c9d · outbound

This paper cites Squeeze-and-excitation networks,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Squeeze-and-excitation networks,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:36.038354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:36.038354Z digest=sha256:442df7ab09484b7985b26be5f1ef06787c3ccd274ba80ca6989ead8dad89d977

Observation 988293ba-2ded-4372-8ebd-54de16006b57 · outbound

This paper cites Sdr–half-baked or well done?.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Sdr–half-baked or well done?

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.257046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:36.041955Z digest=sha256:9429b5525822864bd32373b654750041936abc757977a670557073a5e624583d

Observation ecfaf012-9d1a-4967-b870-84dc9e0d61fb · outbound

This paper cites Csr-i (wsj0) complete ldc93s6a,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Csr-i (wsj0) complete ldc93s6a,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.243969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:36.045599Z digest=sha256:0c67cc463bb17a495bd1c760ddc80deedde8fcc543cf6269491372799bdd9b26

Observation fe23e7e1-d0a2-492b-a0ef-d66c5a0e5c28 · outbound

This paper cites WHAM!: Extending Speech Separation to Noisy Environments.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions WHAM!: Extending Speech Separation to Noisy Environments

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:36.049963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:36.049963Z digest=sha256:b7c65e9126e8ed640627e4820907e1d23f1cc47db9deee34012f05b238417998

Observation a15ad033-fe13-4d55-851e-d81c5112c7ad · outbound

This paper cites Whamr!: Noisy and reverberant single-channel speech separation,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Whamr!: Noisy and reverberant single-channel speech separation,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.230796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:36.053896Z digest=sha256:789f098da8dbeb9ab3d4fe711b5459727181d80c0cce5f6d92e823ae890776d6

Observation 0a2cad35-1e31-4976-8d1a-f57a28e0ed81 · outbound

This paper cites Target Confusion in End-to-end Speaker Extraction: Analysis and Approaches.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Target Confusion in End-to-end Speaker Extraction: Analysis and Approaches

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-08T10:06:36.105012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:36.057455Z digest=sha256:ce25d76a593b31a5d2d4ca6862711c8cf7252d9cf6ccff576a255e37edb4a239

Observation 4d426059-eea5-4d34-b37b-064985ae43e2 · outbound

This paper cites Attention is all you need,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Attention is all you need,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.217503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T10:06:36.061223Z digest=sha256:5abea75bb1f1dfeb5f37fe59335c5c1dac3d1f8e1a67bc18c8c33bff9c93e806

Pith citing papers

No inbound Pith citation observations are available.