Pith. sign in

Paper Citation Record · LEDGER

VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:1810.04826.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1810.04826 v6

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T13:38:44.171139Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T12:57:07.589361Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 45f9cf6b-e8a3-4aa0-9f84-112abc83ea01 · inbound

End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning cites this paper.

End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-14T13:38:44.171139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T13:38:44.171139Z digest=sha256:3fcdab7fec9f91ef751be3d607c7eff1aec0a25aba364d766a7e5ed4f0a9fbc0

Observation 8841318e-8b8e-458b-876c-4e54099e768f · inbound

X-CrossNet: A complex spectral mapping approach to target speaker extraction with cross attention speaker embedding fusion cites this paper.

X-CrossNet: A complex spectral mapping approach to target speaker extraction with cross attention speaker embedding fusion VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T15:53:11.061029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:53:11.061029Z digest=sha256:095f0b03ef843b5a18660f4ecbbd50e12b6faf57f59fde79b3f0e88c1f731e88

Observation faba0c50-5095-49cc-a2c3-325524baf5ec · inbound

Comprehensive Audio Query Handling System with Integrated Expert Models and Contextual Understanding cites this paper.

Comprehensive Audio Query Handling System with Integrated Expert Models and Contextual Understanding VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T21:56:42.003212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:56:42.003212Z digest=sha256:6ff082ba1e0b649ee45cda4d65a0150aa8a0a708bf8809e382de889775a67931

Observation 387d996f-4d08-4dcf-861a-6178b55bc93d · inbound

Metis: A Foundation Speech Generation Model with Masked Generative Pre-training cites this paper.

Metis: A Foundation Speech Generation Model with Masked Generative Pre-training VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-09T05:54:18.770915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:54:18.770915Z digest=sha256:4a5d8eb8673e61a29d638959d58a7f52ae0adfc032721dd3d99e29efa9d7fc6b

Observation 2210c1a4-6f72-4a10-a827-355990ffd65a · inbound

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions cites this paper.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:35.967023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:35.967023Z digest=sha256:4fa055e4bd84a94d412c39d0f5d6b62fe6b7abfc088397bd28be52e83e2a16d4

Observation 25102226-6c02-4c11-ad2a-5d0b34580ed6 · inbound

FlowTSE: Target Speaker Extraction with Flow Matching cites this paper.

FlowTSE: Target Speaker Extraction with Flow Matching VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:37:09.365680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:37:09.365680Z digest=sha256:40114642bf64b1f3e8441a443b0e807f03db3d64e42840356bb9d2c7df61483a

Observation 874644e3-00e2-47b6-843e-b763ee0c6c55 · inbound

UniFlow: Unifying Speech Front-End Tasks via Continuous Generative Modeling cites this paper.

UniFlow: Unifying Speech Front-End Tasks via Continuous Generative Modeling VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-05T22:07:46.317445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:07:46.317445Z digest=sha256:4895ed00f992b9400a796d224a374b10c40ab059d0e45dfd10c5afe6039ae086

Observation 277a7ae1-61ff-4726-84bc-c37cbb9ff803 · inbound

Fundamentals of Data-Driven Approaches to Acoustic Signal Detection, Filtering, and Transformation cites this paper.

Fundamentals of Data-Driven Approaches to Acoustic Signal Detection, Filtering, and Transformation VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-05T14:24:34.109330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:24:34.109330Z digest=sha256:af64513a9ae248f3df105182f9529e9648ed5a1e796efaecde9375e40641f78c

Observation d2f01861-67ed-4bcc-9e2f-22a0fc83e479 · inbound

PS4: Proxy-Supervised Joint Training for Real Target Speaker Extraction cites this paper.

PS4: Proxy-Supervised Joint Training for Real Target Speaker Extraction VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:57:07.591118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-10T12:54:17.490105Z digest=sha256:18a7c4a573414c5032f1febd08c1297e19400cdb125eaae735e42dd3cf43b841