Pith. sign in

Paper Citation Record · LEDGER

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection

As of 13 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 1 inbound Pith citation observation for arXiv:2509.04161.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.04161 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:24:21.774639Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-09T19:27:59.124425Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T15:41:32.859362Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy39
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c70deeb6-e849-41a6-a9ad-f82a07e070c5 · outbound

This paper cites Asvspoof 2019: Future horizons in spoofed and fake audio detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Asvspoof 2019: Future horizons in spoofed and fake audio detection,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.301745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.408376Z digest=sha256:b6fc5ae16b3871a405c7d0a1317affddfd9fbd07566efb3040991d6ecabb33d3

Observation 24e11cd6-bc45-4e20-af29-f71bea87989e · outbound

This paper cites Asvspoof 2021: accelerating progress in spoofed and deepfake speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Asvspoof 2021: accelerating progress in spoofed and deepfake speech detection,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.276822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.414020Z digest=sha256:e341875db80dfcc8dcd467c8c5f6e048f535d30c134dcb56123317a41689acac

Observation 09a51a00-6777-465e-a99c-eced5c22e306 · outbound

This paper cites ASVspoof 5: Crowdsourced Speech Data, Deepfakes, and Adversarial Attacks at Scale.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection ASVspoof 5: Crowdsourced Speech Data, Deepfakes, and Adversarial Attacks at Scale

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T10:24:21.420472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:24:21.420472Z digest=sha256:c705d1ec7ec0a9499f383b9536c3ee05c476b1a0179c1a24759b96aa80194ddf

Observation 7610222b-8f79-4dc3-98e2-65bb3738040f · outbound

This paper cites Robust audio anti-spoofing with fusion-reconstruction learning on multi-order spectrograms,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Robust audio anti-spoofing with fusion-reconstruction learning on multi-order spectrograms,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.253666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.428804Z digest=sha256:fe263a7217d6d9eccc73f08f923a67bb2037e9d148583279c8fc86717294754a

Observation 313ac6e4-e15b-4fd3-9d91-491e9935e3c7 · outbound

This paper cites A comparative study on recent neural spoofing countermeasures for synthetic speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection A comparative study on recent neural spoofing countermeasures for synthetic speech detection,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.234638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.441067Z digest=sha256:536a6cea0f4dd3cfec9e7e4e5528154c70eba15d7fa41f39c756a0760be57f61

Observation b9c69ebe-abab-47b5-8bf1-9461210140fe · outbound

This paper cites Channel-wise gated res2net: Towards robust detection of synthetic speech attacks,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Channel-wise gated res2net: Towards robust detection of synthetic speech attacks,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.212866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.448595Z digest=sha256:e843e9be02f9976db694b703b7da10c7236b790ac39bb193597c38cae2bede5e

Observation f779088e-a02c-4458-8928-d3d9cdd052c3 · outbound

This paper cites Fastaudio: A learnable audio front-end for spoof speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Fastaudio: A learnable audio front-end for spoof speech detection,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.185669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.460691Z digest=sha256:c6f8878c9ae939137466800d07ac9c2947dbca85c47bdc16919337c09a1eeb31

Observation 53c12967-7231-4a49-92bb-0a3ef463c365 · outbound

This paper cites The effect of silence and dual-band fusion in anti-spoofing system,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection The effect of silence and dual-band fusion in anti-spoofing system,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.152892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.468943Z digest=sha256:adcea75c97a6739c96fa26e20d2c8a4fd30193973800427794df75e7b0f61ef7

Observation 09fc9929-97bd-4a22-857c-27ee94efc48d · outbound

This paper cites Towards end-to-end synthetic speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Towards end-to-end synthetic speech detection,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.131004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.474911Z digest=sha256:d3a0e1ac8ccd2adb7a5a9dd9a1034821990421746322c69dffbc42b831d91104

Observation 7e98b17d-5cbc-4802-8b61-c0c936ae7b81 · outbound

This paper cites Aasist: Audio anti-spoofing using integrated spectro-temporal graph attention networks,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Aasist: Audio anti-spoofing using integrated spectro-temporal graph attention networks,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.104623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.485405Z digest=sha256:454ccab781047ace61d4169b6d6f9a0e55f945957e96d2f6389d8ac242a6331a

Observation 1d8b3745-16f2-4547-a2c7-439829104472 · outbound

This paper cites Discriminative frequency information learning for end-to-end speech anti-spoofing,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Discriminative frequency information learning for end-to-end speech anti-spoofing,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.074341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.490772Z digest=sha256:3515ee8768709796a79dad3f9c702f58a048fdf2aaa4edc2fedc84d5dc1a47b1

Observation 8a3306aa-c5fb-4575-a448-da837e11deee · outbound

This paper cites Robust data2vec: Noise-robust speech representation learning for asr by com- bining regression and improved contrastive learning,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Robust data2vec: Noise-robust speech representation learning for asr by com- bining regression and improved contrastive learning,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.050984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.497043Z digest=sha256:c9f947e18d9ca93b4d7d9cb27efdc4d524886d11c742f8b4cd4621e6c2af693a

Observation 233bcfef-9b05-45e5-a7eb-3d707debecbd · outbound

This paper cites Self-supervised learning with cluster-aware-dino for high-performance robust speaker verification,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Self-supervised learning with cluster-aware-dino for high-performance robust speaker verification,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.023588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.503153Z digest=sha256:ec68562022e4fc85c8464e442b7255786027a71d7472a2a9b7036ff0cdaf0a04

Observation 5d8db485-0f8b-461d-8f8b-6e9f14e405af · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection wav2vec 2.0: A framework for self-supervised learning of speech representations,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.000993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.511661Z digest=sha256:42ef3ae24044efedb726255f5ca58084f66ac41970151a79d01e9880d138042b

Observation b329de89-acad-41eb-a80b-bf587c782dd6 · outbound

This paper cites Hubert: Self-supervised speech representation learning by masked prediction of hidden units,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Hubert: Self-supervised speech representation learning by masked prediction of hidden units,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T10:24:21.519675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:24:21.519675Z digest=sha256:ff99312b9e708192941158849d50b5e4f304144bb856470629a53b2f16e4c443

Observation c00d1e3e-d147-4ed4-b3b5-6f6c52e20dfe · outbound

This paper cites Wavlm: Large-scale self-supervised pre- training for full stack speech processing,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Wavlm: Large-scale self-supervised pre- training for full stack speech processing,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.961333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.527515Z digest=sha256:d3ca8b4d82f92865a5f54872d6b8652dcc24440abac61d037138d1bf3dd301d9

Observation 7bcb0426-331c-4d2f-b582-39ffd30ca0fe · outbound

This paper cites Robust spoof speech detection based on multi-scale feature aggregation and dynamic convolution,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Robust spoof speech detection based on multi-scale feature aggregation and dynamic convolution,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.929032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.536857Z digest=sha256:d70b9c67a29e1747cbe6e316eff46886cf6cbe43726a590bea8555b488e12eef

Observation 65e212ef-e50f-488d-a048-d3225b0b0feb · outbound

This paper cites Automatic speaker verification spoofing and deepfake detection using wav2vec 2.0 and data augmenta- tion,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Automatic speaker verification spoofing and deepfake detection using wav2vec 2.0 and data augmenta- tion,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.906918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.552238Z digest=sha256:a82048e2409552c5a2e954b3929d00c224c6c9a22cb62f3a666635f1ee2f6d15

Observation 1d9f4958-6701-45e7-a0f6-f861346fc972 · outbound

This paper cites Investigating self-supervised front ends for speech spoofing countermeasures,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Investigating self-supervised front ends for speech spoofing countermeasures,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.866635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.561268Z digest=sha256:2a65ca06ef70e579123ab7f5af17121aacd2f844e8e868421717f643a5836d3b

Observation cd1d6e70-f61b-4430-8906-bc50e621ed06 · outbound

This paper cites Lora: Low-rank adaptation of large language models,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Lora: Low-rank adaptation of large language models,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.842890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.567817Z digest=sha256:18dd31af84fd8171e3c00df37ffad70d5926b8cfcaf9c714f48fae5768e50078

Observation b25906c1-fa8d-4b93-8685-cad55e7f03e8 · outbound

This paper cites Parameter-efficient transfer learning for nlp,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Parameter-efficient transfer learning for nlp,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.817025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.575475Z digest=sha256:0596edaa90bd4d51a6f4f961dcb9c615e9e5d2e1c8fc3e2fc721376b97d0c766

Observation 6a8a19f9-1d50-4fc6-99b6-7edd2978f635 · outbound

This paper cites Audio deepfake detection with self- supervised xls-r and sls classifier,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Audio deepfake detection with self- supervised xls-r and sls classifier,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.784447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.582486Z digest=sha256:3085ae38ac7f9aea4c77c64a0b30e2cb7348ad50d658064930701d134a717e6d

Observation 0c23d29d-6f25-480b-b1b9-da2f266669d3 · outbound

This paper cites Attentive merging of hidden embeddings from pre-trained speech model for anti-spoofing detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Attentive merging of hidden embeddings from pre-trained speech model for anti-spoofing detection,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.754575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.589077Z digest=sha256:379358590cde859241fa2387ebd6a25263fb3d406c6da019f5ed3d948232d4ca

Observation a3f4acad-1ad8-4a97-8bb6-34d1718867cf · outbound

This paper cites Outrageously large neural networks: The sparsely-gated mixture-of-experts layer,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Outrageously large neural networks: The sparsely-gated mixture-of-experts layer,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.727153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.597639Z digest=sha256:4195525d3e4c1b53d4ee87b0bc77792859a1544df008038d6403eac544a698a3

Observation 6d860ee5-5eb3-43d8-96c4-d83eddc48c9f · outbound

This paper cites Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.697394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.610876Z digest=sha256:ad34f414b812b6722f0c6bafa346e44418cdbeaa68c02b412d5a585cd620b29b

Observation d961a671-a693-4f34-82dd-a6941668f64c · outbound

This paper cites Hifi-gan: Generative adversarial networks for efficient and high fidelity speech synthesis,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Hifi-gan: Generative adversarial networks for efficient and high fidelity speech synthesis,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.661369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.618851Z digest=sha256:d71b6a4521953d764f062978a82c194f16eb2a7b1cbd25f2c325f4aea1189e55

Observation 7fab0d9b-c156-4d25-989b-a3fea427a85b · outbound

This paper cites Fastdiff: A fast conditional diffusion model for high-quality speech synthesis,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Fastdiff: A fast conditional diffusion model for high-quality speech synthesis,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.633203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.624994Z digest=sha256:502c0f6f2571dea398cc5403c425669f2f3a3d8246ef323bef1107a526fc2f48

Observation fa174add-e3d3-4ae7-bf5f-af9339b3057d · outbound

This paper cites Freevc: Towards high-quality text-free voice conversion,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Freevc: Towards high-quality text-free voice conversion,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.600627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.632325Z digest=sha256:687d16b823b13142b1d44b6d54e7142e550bfd1cba989875d8daa9c610af0a9a

Observation aedd6042-2cc0-4a3f-a886-dc06bcd8e49d · outbound

This paper cites Asvspoof 2019: A large- scale public database of synthesized, converted and replayed speech,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Asvspoof 2019: A large- scale public database of synthesized, converted and replayed speech,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.546545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.638598Z digest=sha256:aaa2b784ddc8ed28424ccdabcd301aba7a26e55b4721925486bc32d1b98b4d0d

Observation 9cdcd84e-a8fa-4c44-ba8d-ff4528ec75d9 · outbound

This paper cites Asvspoof 2021: Towards spoofed and deepfake speech detection in the wild,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Asvspoof 2021: Towards spoofed and deepfake speech detection in the wild,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.524777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.645949Z digest=sha256:1d3678ae38648ec57eaa80058afaa29825f8b7a6f996db82db53ae1de46a3512

Observation f0870546-ba7c-40c5-979a-c21c66db1b5a · outbound

This paper cites Does audio deepfake detection generalize?,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Does audio deepfake detection generalize?,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T10:24:21.653089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:24:21.653089Z digest=sha256:ac032d3c4af6099ced41cb96ce75b7f31c948cb8f19987790ff73a61798ea050

Observation 33e3a0e4-df49-433f-bf4a-6bb631854fcf · outbound

This paper cites t-dcf: a detection cost function for the tandem assessment of spoofing countermeasures and automatic speaker verification,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection t-dcf: a detection cost function for the tandem assessment of spoofing countermeasures and automatic speaker verification,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.486726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.670151Z digest=sha256:291be761ea3be9a4fe26bbb105a2d08fd55edec4d31c7ef6dd4b4ea9f7c8f51c

Observation 01b0cda1-daa3-4201-b06f-1a95455a02e3 · outbound

This paper cites Rawboost: A raw data boosting and augmentation method applied to automatic speaker verification anti- spoofing,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Rawboost: A raw data boosting and augmentation method applied to automatic speaker verification anti- spoofing,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.436447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.677332Z digest=sha256:9a7cdeb469d3648edab9a8c05e7cb4c9037c1d0103d23749ca61c29c9e1454e4

Observation ab76689c-a2af-4eaf-a937-958698e4afd7 · outbound

This paper cites Improving short utterance anti-spoofing with aasist2,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Improving short utterance anti-spoofing with aasist2,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.401284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.692373Z digest=sha256:6d23544653ecaa35c9d7cd7437afd7dc2051ea13bba98febe8bef2cc66fa8209

Observation f77714c3-81f4-4924-9b58-4c158c99b7fc · outbound

This paper cites A conformer-based classifier for variable-length utterance processing in anti-spoofing,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection A conformer-based classifier for variable-length utterance processing in anti-spoofing,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.378433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.711964Z digest=sha256:37cad47614bdcd5e8d300be558d772e60dcf4da159f9fde210374a579b210cf0

Observation ca664df3-d1ec-4a9f-9947-994fd6e85791 · outbound

This paper cites Audio deepfake detection with self-supervised wavlm and multi-fusion attentive classi- fier,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Audio deepfake detection with self-supervised wavlm and multi-fusion attentive classi- fier,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.332612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.723573Z digest=sha256:4d6d6d2a5b7e299af64db084bc866a595093e7557df7348857f94876783ca3a0

Observation 8bb61193-c48a-465c-94a6-15c24fa27804 · outbound

This paper cites One class learning with adaptive centroid shift for audio deepfake detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection One class learning with adaptive centroid shift for audio deepfake detection,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.300309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.735335Z digest=sha256:104c29a502629fe77e181b8a75a28e66a2cc9ade983cf4717e85989eeebbb59a

Observation 5be16f9c-d688-4e58-af75-6d045fd582b8 · outbound

This paper cites Temporal-channel modeling in multi-head self-attention for synthetic speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Temporal-channel modeling in multi-head self-attention for synthetic speech detection,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.273266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.740808Z digest=sha256:e7c609e5b4955901f3cb7e858156e6e598da4236023ad390e34aca36481dccbc

Observation 16b44864-4e4b-4a03-bc9d-291c03435c4d · outbound

This paper cites A robust audio deepfake detection system via multi-view feature,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection A robust audio deepfake detection system via multi-view feature,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.221833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.748635Z digest=sha256:639a85df77396e1381b1e0b80e85812194eff9509bc93d2b418e1c787d5e7de0

Observation 982ae0e7-4a3d-4158-a9a4-e1788409b501 · outbound

This paper cites Spoofed training data for speech spoofing countermeasure can be efficiently created using neural vocoders,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Spoofed training data for speech spoofing countermeasure can be efficiently created using neural vocoders,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.172178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.757104Z digest=sha256:84ea2febf62c22c8944f43c0ac23bc3831f44366f02eb83416072c80f9dc3d21

Observation e52fcafd-748c-402c-851c-12c99b215171 · outbound

This paper cites Can large-scale vocoded spoofed data improve speech spoofing countermeasure with a self-supervised front end?,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Can large-scale vocoded spoofed data improve speech spoofing countermeasure with a self-supervised front end?,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.131597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.767444Z digest=sha256:c41e26043bbe306a9782ef3a914b2bb6920dbd2e9a1c267de82b82df7a0b0b07

Observation 4a7180f7-0ce7-4e35-84fd-32e0a20b35d3 · outbound

This paper cites One-class knowl- edge distillation for spoofing speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection One-class knowl- edge distillation for spoofing speech detection,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.105531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T10:24:21.774639Z digest=sha256:213d72daf9638ac58746d23ee67c64d00007ec31a4db5047a0bb57e88802e97b

Pith citing papers

Observation e60da123-abcf-460e-a92e-e8dbb05497e8 · inbound

Alethia: A Foundational Encoder for Voice Deepfakes cites this paper.

Alethia: A Foundational Encoder for Voice Deepfakes Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:32.948686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:2164a9073a13426db182b324c5eac45b0a2c4e3cf0b1ca123d940b742b67bcc6