Pith. sign in

Paper Citation Record · LEDGER

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification

As of 18 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 1 inbound Pith citation observation for arXiv:2507.21642.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21642 v2

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:36:19.265120Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:36:14.939053Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T12:36:19.393272Z

Reference resolution

46 of 46 outbound references displayed

  • verified exact0
  • verified fuzzy40
  • unresolved5
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0d7fdc80-86a8-45d4-8f83-7ff408980450 · outbound

This paper cites In-the-wild (ITW) refers to data that is not recorded or collected in a highly controlled environ- ment [1].

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification In-the-wild (ITW) refers to data that is not recorded or collected in a highly controlled environ- ment [1]

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.347827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:14.694066Z digest=sha256:522fe5691c8b0c54b32152a29935991eca2a828cc578a241b016b404b6e71e39

Observation d39270fe-4d4d-49d2-aa0c-7197b77f5b29 · outbound

This paper cites Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification

Reference 2

Resolution
malformed identifier
local_arxiv, observed 2026-08-06T12:36:19.486175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:14.939053Z digest=sha256:0ec4ae44f07cac678e9699c11d76fcb95f9c4671af7dafeff0d3e03880ef63c0

Observation 3d8a8247-85f7-42fe-9e62-1715c538b78f · outbound

This paper cites For the first training stage, simple filename and label pairs are derived from non-ITW datasets.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification For the first training stage, simple filename and label pairs are derived from non-ITW datasets

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.320435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:15.203573Z digest=sha256:6e1d876e93d92f602f291d9119fac29676e516bd563a4a086510a5c39f394781

Observation deb1ed1e-34e7-4258-87eb-c7cb1c86c98d · outbound

This paper cites an unresolved cited work.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:36:20.311411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:15.367337Z digest=sha256:4ddbeb09076fb7bcd22b25a9aacefff24d67af4f2fe736eb8d3605674d6d7dc8

Observation 18f3e325-9386-4531-9dd6-f7eb3020ba32 · outbound

This paper cites an unresolved cited work.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:36:20.302330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:15.474633Z digest=sha256:59b693021d9be91361c16fc6d727dca3c458a5285ce970e6949fc4898720c391

Observation b202d257-f806-44c3-a2b5-8e725b0e3ef8 · outbound

This paper cites noisiness.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification noisiness

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.293300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:15.633256Z digest=sha256:db9405c9a71f4111c2bfba806c91a5f3c65e4568a02b69bc4555a4660dbd8512

Observation 0eab6f77-b80c-489d-b6d9-13c2aa1b0544 · outbound

This paper cites It was shown that the AITW dataset can be used to train classifiers on multispeaker, foreign language, background music, noise and synthetic speech labels.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification It was shown that the AITW dataset can be used to train classifiers on multispeaker, foreign language, background music, noise and synthetic speech labels

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.284045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:15.759015Z digest=sha256:7315127685c05c606a9bf026737971cfa9c40ef28945e6c5efcf85d343237519

Observation 662c5dbe-1be9-445e-801f-2669df316465 · outbound

This paper cites An open-source speaker gender detection framework for moni- toring gender equality,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification An open-source speaker gender detection framework for moni- toring gender equality,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.213158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:17.012882Z digest=sha256:6d91b948edaa5d51cf776f7dde2a771a99f07077293afcb8eff09accf953fb05

Observation f66a5262-6e0e-4108-a8df-557f08781bcd · outbound

This paper cites Label Studio: Data labeling software,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Label Studio: Data labeling software,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.203845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:17.188737Z digest=sha256:83ed0ac65095dd97ca1fa0f1d2315c31753a2e337c6b7c081c97ce4a2450f3b3

Observation 80080ab1-ade4-4845-89ee-65c9459317ce · outbound

This paper cites V oxceleb2: Deep speaker recognition,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification V oxceleb2: Deep speaker recognition,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.274372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:15.952302Z digest=sha256:7562a2b84d943fac8e090d405290e7b5e584dbfa756a6264be1ca5dccf700a99

Observation 7f0c58ca-f948-4e55-ad21-a87f4cf104a9 · outbound

This paper cites YODAS: Youtube-Oriented Dataset for Audio and Speech,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification YODAS: Youtube-Oriented Dataset for Audio and Speech,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.263840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:16.062098Z digest=sha256:45c39c2546b645f2a7c175cbbdbef4662a6ccc60922818092dae6f5a607dc269

Observation d4a39cdd-a493-4d96-90f3-f2082407f8dd · outbound

This paper cites Emilia: An Extensive, Multilingual, and Diverse Speech Dataset for Large-Scale Speech Generation.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Emilia: An Extensive, Multilingual, and Diverse Speech Dataset for Large-Scale Speech Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T12:36:16.250981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:36:16.250981Z digest=sha256:88ae5f89b8db7332f88129ce22de6f598a9d695cd85116388c877975e7c99e17

Observation d42dd31b-d7a2-4b0c-ac1c-20430aa0225f · outbound

This paper cites Speak, read and prompt: High-fidelity text-to-speech with min- imal supervision,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Speak, read and prompt: High-fidelity text-to-speech with min- imal supervision,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.253100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:16.377864Z digest=sha256:95e354048112b03903111887cf6d6f8822699a9d1e1101633a8a5cf13cd030c8

Observation 4e3daa7b-011d-4649-8449-6688a6aeeeb6 · outbound

This paper cites synthetic speech and non-target languages in many cases).

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification synthetic speech and non-target languages in many cases)

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.339065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:14.761813Z digest=sha256:a67de4d5fb38fc470aae8c3e9a3b72881f21fc70b42bf5e0c88b4fe5310f3e00

Observation 788c4853-3cc5-42c1-b3fe-f9e889214510 · outbound

This paper cites mhubert-147: A compact multilingual hubert model,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification mhubert-147: A compact multilingual hubert model,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.242353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:16.582222Z digest=sha256:95c7995915bbc7a15fe361a164eba9587927e3e90553e57f77937feab4f447e4

Observation 3a99ca04-b170-4b91-9a1c-1a49d79b0ac9 · outbound

This paper cites Sentence level intelligibility eval- uation for mandarin text-to-speech systems using semantically unpredictable sentences,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Sentence level intelligibility eval- uation for mandarin text-to-speech systems using semantically unpredictable sentences,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.232020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:16.739548Z digest=sha256:1413c89a349f095547e80af758576e646792398126cd7e163c309abdb3bfbe02

Observation 8095009c-91eb-4157-b76f-2d882ec09be2 · outbound

This paper cites Data-filtering methods for self-training of auto- matic speech recognition systems,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Data-filtering methods for self-training of auto- matic speech recognition systems,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.222623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:16.856553Z digest=sha256:b19c74ed027fceb359cf5d1a470975b89f03205998066ce58a4ab8bc1a838b38

Observation dfd21f4b-5d2b-4ef0-ac35-e1014274d4fb · outbound

This paper cites Attention is All you Need,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Attention is All you Need,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.026366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:17.993141Z digest=sha256:c13e5036a11dbbc09f53070344b2d804b7ba9f08afac5c5e7892bbb6ba522465

Observation b9c5a3b8-08c8-40f9-9502-6c92053cb87b · outbound

This paper cites pyannote.audio 2.1 speaker diarization pipeline: prin- ciple, benchmark, and recipe,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification pyannote.audio 2.1 speaker diarization pipeline: prin- ciple, benchmark, and recipe,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.194492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:17.257923Z digest=sha256:0bb6dd9296633bcfcb5a3a9c7b95144e3d9027091f30acff1bf161b52130d830

Observation 6c8d09a5-d94e-415b-ae78-4a814d3773de · outbound

This paper cites FoR: A dataset for synthetic speech detection,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification FoR: A dataset for synthetic speech detection,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.092229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:17.347739Z digest=sha256:463b8ddb5c43aecc04d6c5f2cc5dbbeb6112866856089fd88cc618a293279e7c

Observation 9a699321-9cf0-4061-93d0-a07c96283c51 · outbound

This paper cites AASIST: Audio Anti-Spoofing Us- ing Integrated Spectro-Temporal Graph Attention Networks,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification AASIST: Audio Anti-Spoofing Us- ing Integrated Spectro-Temporal Graph Attention Networks,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.083301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:17.436804Z digest=sha256:e511c906f3fb1440fa22768ec531a6789526093a8eeef248c20ecc7026a1c7f4

Observation d2e535ba-d3a9-4939-8abf-99a1931349bd · outbound

This paper cites This as- sumption is validated later in our results, cf.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification This as- sumption is validated later in our results, cf

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.329846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:15.048489Z digest=sha256:53d20d0abd95c3be6bf45a17a1b9e3d568fe4128e617e25d853f9af191e3d32a

Observation 431ed380-c6a3-4748-9ba2-f678217b81d3 · outbound

This paper cites Robust speech recognition via large-scale weak su- pervision,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Robust speech recognition via large-scale weak su- pervision,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.073766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:17.518641Z digest=sha256:12fb435c42630d8ccb28a22ac5699b9598a754c5a9791574214de5aebfb5d7f4

Observation 7a73256a-16f2-4004-b601-e7a39733765d · outbound

This paper cites Perceive and predict: Self-supervised speech representation based loss func- tions for speech enhancement,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Perceive and predict: Self-supervised speech representation based loss func- tions for speech enhancement,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.064610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:17.608036Z digest=sha256:e4ca6d59188fc843c727bcdbed861543c7eaeac5db4c53e0efa005067254bc53

Observation 1d4fe041-0f67-4bbe-9392-c30dc86605ce · outbound

This paper cites Transcription-free fine-tuning of speech separation models for noisy and reverberant multi-speaker automatic speech recognition,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Transcription-free fine-tuning of speech separation models for noisy and reverberant multi-speaker automatic speech recognition,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.055009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:17.687623Z digest=sha256:b114ba1642f0f6036a7aea7a0533d39d142de4b3d7d6865c6080a6fdf7f67add

Observation 4bd74790-bffe-434a-8f27-88d484b90167 · outbound

This paper cites Non-intrusive speech intelligibility pre- diction for hearing-impaired users using intermediate asr features and human memory models,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Non-intrusive speech intelligibility pre- diction for hearing-impaired users using intermediate asr features and human memory models,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.045407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:17.818034Z digest=sha256:517dc1fa277d607c3bc991ef88f43c2854059e2ddc4a5e6e9d8bf08dd790bae9

Observation 9a794ed0-4674-4f02-afdc-075e8d938a42 · outbound

This paper cites Heterogeneity over homogeneity: Investigating multilingual speech pre-trained models for detecting audio deepfake,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Heterogeneity over homogeneity: Investigating multilingual speech pre-trained models for detecting audio deepfake,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.036295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:17.904511Z digest=sha256:498cb8fdde6a712fbaddd3e811b0fbb832ebc3789a069750d73b9684201e6fa5

Observation f78d6fe7-9e8c-4e81-9b18-d2446077e1d7 · outbound

This paper cites Nisqa: A deep cnn-self-attention model for multidimensional speech quality pre- diction with crowdsourced datasets,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Nisqa: A deep cnn-self-attention model for multidimensional speech quality pre- diction with crowdsourced datasets,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.016269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:18.073975Z digest=sha256:1ad0fb0d79a0de6b687c6a6308341eec412d0a9929d2d47814b89ffc92936975

Observation 19dd9b2a-49d7-4688-a944-9e8599a75257 · outbound

This paper cites Hallucination in perceptual metric-driven speech enhancement networks,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Hallucination in perceptual metric-driven speech enhancement networks,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.005475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:18.121425Z digest=sha256:0317b6bfa3f615d30ade1270b3777770c3f9d73b96967079e6828c8f1681d0e0

Observation e2518e7a-a00f-478f-abe5-525e202aa34b · outbound

This paper cites An overview of multi-task learning in deep neural networks,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification An overview of multi-task learning in deep neural networks,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.994398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:18.183181Z digest=sha256:d0eca9058c6cd3fe992bfdf1bd559cac1197c4ad5fbc95218cc94ba4b54985eb

Observation 05818646-6738-4a57-9ddd-199d562323d4 · outbound

This paper cites BEATs: Audio pre-training with acoustic to- kenizers,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification BEATs: Audio pre-training with acoustic to- kenizers,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.985310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:18.246729Z digest=sha256:89b307f751277a8d4b476944a9febbf03f39c8d841f040609a3ed352d5cf5a28

Observation aebc4646-88f0-4a16-b3e8-57bcbf75fec7 · outbound

This paper cites Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.976010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:18.331090Z digest=sha256:44ad6cfaf3b5a328a9ff02c39b267ad18497e4b02074440ba152642d9f62644c

Observation 014a8bb2-4168-48c3-a442-168f1bee4121 · outbound

This paper cites Wavesplit: End-to-end speech separation by speaker clustering,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Wavesplit: End-to-end speech separation by speaker clustering,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.966245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:18.453605Z digest=sha256:c8a308151e83fa188ea306c0acf51a42c6cbeb4acd12e666b79eb590f88d18b3

Observation fb1510eb-4c64-47b5-aada-3fdee4d7de85 · outbound

This paper cites The AMI Meeting Corpus: A Pre- announcement,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification The AMI Meeting Corpus: A Pre- announcement,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T12:36:18.476532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:36:18.476532Z digest=sha256:5f8f4bfbebe14e3533ea56f2eae255c6922a73aa8f615bb65170f4eb074df1df

Observation e2ccdfa0-b169-4d5b-b4b3-0a34d3d08f00 · outbound

This paper cites M2Met: The ICASSP 2022 Multi-Channel Multi-Party Meeting Transcription Challenge,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification M2Met: The ICASSP 2022 Multi-Channel Multi-Party Meeting Transcription Challenge,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.950665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:18.591304Z digest=sha256:c9137e0120cac6ea22d83f30b0d54e3bd5de750a140da3076e16c4d53add3410

Observation 5a0b28c5-50c1-44f0-bc0d-ddb1dd0f03e8 · outbound

This paper cites MLS: A Large-Scale Multilingual Dataset for Speech Research,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification MLS: A Large-Scale Multilingual Dataset for Speech Research,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.940318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:18.612570Z digest=sha256:11be64d92dda1c75adb5d61859c66806201b33c8d19cb9ea0cba96f098605388

Observation 79768340-eee1-4774-b730-8ef1af09d142 · outbound

This paper cites MUSAN: A Music, Speech, and Noise Corpus.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification MUSAN: A Music, Speech, and Noise Corpus

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T12:36:18.698419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:36:18.698419Z digest=sha256:7d0b5a50e1093c969c17e2d286a0fb93217044a5daf79c1cb88bce01dc0460a1

Observation 153e9b18-4ea5-417b-b93c-d74ff886df1b · outbound

This paper cites An open dataset of synthetic speech,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification An open dataset of synthetic speech,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.930422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:18.763442Z digest=sha256:4007a52367507f25a3d284723513c0a84d2be4f5de601c9de71ffa528ae39348

Observation 3d43015b-4aec-426c-a567-8d67381e9d58 · outbound

This paper cites OpenMIC-2018: An Open Dataset for Multiple Instrument Recognition,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification OpenMIC-2018: An Open Dataset for Multiple Instrument Recognition,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.920880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:18.817911Z digest=sha256:b11d3660ef325c6d798f2c85d3e52351a7a81a61e9ccaef05f44d1fdf60dd143

Observation 94bc4d53-bd70-44fb-8698-49640fe024cb · outbound

This paper cites Learning sound event classifiers from web audio with noisy labels,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Learning sound event classifiers from web audio with noisy labels,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.910904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:18.927422Z digest=sha256:f82ef24366b06184cf17560818cdd0afe557da859c3fec060ac59b716b68de5c

Observation 071634a6-5ab4-47b2-82e3-4d333f345586 · outbound

This paper cites The Diverse Environments Multi-channel Acoustic Noise Database (DEMAND): A database of multichannel environmental noise recordings,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification The Diverse Environments Multi-channel Acoustic Noise Database (DEMAND): A database of multichannel environmental noise recordings,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.900996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:18.964872Z digest=sha256:c6689331aa219dcee6c2840a134c0c36c56ae1e167c7e401eba419c453eced51

Observation 1bd364b6-7016-4a1a-b27a-88cae80a4116 · outbound

This paper cites Adam: A method for stochastic opti- mization,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Adam: A method for stochastic opti- mization,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.890136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:19.040456Z digest=sha256:0fe64493dcc5155b048948ff07b7bdf0d999abaa577f4465e1cd0904337eeafa

Observation a5aa999c-6a45-46ea-9f94-54606eccca78 · outbound

This paper cites Data augmen- tation for speech separation,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Data augmen- tation for speech separation,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.879696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:19.120411Z digest=sha256:122acfec7bcadded5aa1aa6617c68dceece2e658a2b4903e136477eea2405ade

Observation e52c97ed-3e81-4024-aec6-8138acbdf9b0 · outbound

This paper cites Dnsmos: A non- intrusive perceptual objective speech quality metric to evaluate noise suppressors,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Dnsmos: A non- intrusive perceptual objective speech quality metric to evaluate noise suppressors,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.870112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:19.139425Z digest=sha256:434a953fd496baa7530295f12646e9c954f97d3e98eb239b156f31c6a1aa67ce

Observation 9b35680e-7f3d-4ca2-95a1-ec19dc7d79f5 · outbound

This paper cites Convolutional recurrent neural networks for poly- phonic sound event detection,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Convolutional recurrent neural networks for poly- phonic sound event detection,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.834641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:19.156870Z digest=sha256:872120b73e25c07aa074031ecb9d51ed62963edcf4e4ae264191e9e452ab176c

Observation 2a14e91f-c44d-491f-ab8d-9d88d3921049 · outbound

This paper cites Mlaad: The multi- language audio anti-spoofing dataset,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Mlaad: The multi- language audio anti-spoofing dataset,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.652398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:19.265120Z digest=sha256:4eef03d3ef1c7206c71b664f105768c2e160462b0540fc4c9f1997865bfb445c

Pith citing papers

Observation d39270fe-4d4d-49d2-aa0c-7197b77f5b29 · inbound

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification cites this paper.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification

Reference 2

Resolution
malformed identifier
local_arxiv, observed 2026-08-06T12:36:19.486175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:36:14.939053Z digest=sha256:0ec4ae44f07cac678e9699c11d76fcb95f9c4671af7dafeff0d3e03880ef63c0