Pith. sign in

Paper Citation Record · LEDGER

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection

As of 10 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 1 inbound Pith citation observation for arXiv:2509.04161.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.04161 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:24:21.774639Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-09T19:27:59.124425Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T15:41:32.859362Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy39
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c70deeb6-e849-41a6-a9ad-f82a07e070c5 · outbound

This paper cites Asvspoof 2019: Future horizons in spoofed and fake audio detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Asvspoof 2019: Future horizons in spoofed and fake audio detection,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.301745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.408376Z digest=sha256:968d1754c1badb03ab0de8331050e23ba5060a4e22c0965382da1c19d69418df

Observation 24e11cd6-bc45-4e20-af29-f71bea87989e · outbound

This paper cites Asvspoof 2021: accelerating progress in spoofed and deepfake speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Asvspoof 2021: accelerating progress in spoofed and deepfake speech detection,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.276822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.414020Z digest=sha256:7c7ddc5f387c87278ab415888d311e320fef6d043c8ee3529fb8200d87c9a92b

Observation 09a51a00-6777-465e-a99c-eced5c22e306 · outbound

This paper cites ASVspoof 5: Crowdsourced Speech Data, Deepfakes, and Adversarial Attacks at Scale.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection ASVspoof 5: Crowdsourced Speech Data, Deepfakes, and Adversarial Attacks at Scale

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T10:24:21.420472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:24:21.420472Z digest=sha256:0279f3abb4e00e176de137af4779081d68213002444bd7db69906657b443a129

Observation 7610222b-8f79-4dc3-98e2-65bb3738040f · outbound

This paper cites Robust audio anti-spoofing with fusion-reconstruction learning on multi-order spectrograms,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Robust audio anti-spoofing with fusion-reconstruction learning on multi-order spectrograms,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.253666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.428804Z digest=sha256:f2600cb785dc6ce1c739b748b68e8dc210a8b8711be9cedc3506c7d58a1e1738

Observation 313ac6e4-e15b-4fd3-9d91-491e9935e3c7 · outbound

This paper cites A comparative study on recent neural spoofing countermeasures for synthetic speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection A comparative study on recent neural spoofing countermeasures for synthetic speech detection,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.234638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.441067Z digest=sha256:966861c054dd6969c1e7991733f71748a47a03f2beef1a06d504d6851f58b8e2

Observation b9c69ebe-abab-47b5-8bf1-9461210140fe · outbound

This paper cites Channel-wise gated res2net: Towards robust detection of synthetic speech attacks,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Channel-wise gated res2net: Towards robust detection of synthetic speech attacks,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.212866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.448595Z digest=sha256:d17815992656ecd12007e48bd4833dfbdea9357f2791078d9f27047ee5068fa7

Observation f779088e-a02c-4458-8928-d3d9cdd052c3 · outbound

This paper cites Fastaudio: A learnable audio front-end for spoof speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Fastaudio: A learnable audio front-end for spoof speech detection,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.185669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.460691Z digest=sha256:d859041a47a11020706cf1f30f5ed48b6ac838613c9eb1c033dc910ef7772974

Observation 53c12967-7231-4a49-92bb-0a3ef463c365 · outbound

This paper cites The effect of silence and dual-band fusion in anti-spoofing system,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection The effect of silence and dual-band fusion in anti-spoofing system,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.152892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.468943Z digest=sha256:330cf6d280f57ec5c92812818daa84eddbba902852481250e469c3c975d7f16b

Observation 09fc9929-97bd-4a22-857c-27ee94efc48d · outbound

This paper cites Towards end-to-end synthetic speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Towards end-to-end synthetic speech detection,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.131004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.474911Z digest=sha256:ea2328b52270cb0405fd527c023f83a62ae384ec11422d51dde8bc3cdf40bba4

Observation 7e98b17d-5cbc-4802-8b61-c0c936ae7b81 · outbound

This paper cites Aasist: Audio anti-spoofing using integrated spectro-temporal graph attention networks,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Aasist: Audio anti-spoofing using integrated spectro-temporal graph attention networks,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.104623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.485405Z digest=sha256:2a4784e64d168c30af1e16cb81dc8ec1411ac8f8743e9f6f3722c276be2ae066

Observation 1d8b3745-16f2-4547-a2c7-439829104472 · outbound

This paper cites Discriminative frequency information learning for end-to-end speech anti-spoofing,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Discriminative frequency information learning for end-to-end speech anti-spoofing,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.074341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.490772Z digest=sha256:f7433fbe9e2cd5c231af22d339453d3f6b8d4ed2cf5313c0c72157531a742dcb

Observation 8a3306aa-c5fb-4575-a448-da837e11deee · outbound

This paper cites Robust data2vec: Noise-robust speech representation learning for asr by com- bining regression and improved contrastive learning,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Robust data2vec: Noise-robust speech representation learning for asr by com- bining regression and improved contrastive learning,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.050984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.497043Z digest=sha256:3d14f0d73ad144cc39512e0643b5306335374ca9ca27e5352fd3a2ae0111bfd3

Observation 233bcfef-9b05-45e5-a7eb-3d707debecbd · outbound

This paper cites Self-supervised learning with cluster-aware-dino for high-performance robust speaker verification,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Self-supervised learning with cluster-aware-dino for high-performance robust speaker verification,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.023588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.503153Z digest=sha256:63476dc8fa2c56ce15ff364e889f6da3d25a053aa515735d8c82042ec2552973

Observation 5d8db485-0f8b-461d-8f8b-6e9f14e405af · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection wav2vec 2.0: A framework for self-supervised learning of speech representations,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.000993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.511661Z digest=sha256:8c91c5d346b25f7edee6999408021840221a9c0d31789e421941ca55eb833909

Observation b329de89-acad-41eb-a80b-bf587c782dd6 · outbound

This paper cites Hubert: Self-supervised speech representation learning by masked prediction of hidden units,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Hubert: Self-supervised speech representation learning by masked prediction of hidden units,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T10:24:21.519675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:24:21.519675Z digest=sha256:ff99312b9e708192941158849d50b5e4f304144bb856470629a53b2f16e4c443

Observation c00d1e3e-d147-4ed4-b3b5-6f6c52e20dfe · outbound

This paper cites Wavlm: Large-scale self-supervised pre- training for full stack speech processing,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Wavlm: Large-scale self-supervised pre- training for full stack speech processing,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.961333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.527515Z digest=sha256:dcc2129ffbcbe68a3537eb3dd89d6c16d710d0f4295a1719944a2a5dff381cc0

Observation 7bcb0426-331c-4d2f-b582-39ffd30ca0fe · outbound

This paper cites Robust spoof speech detection based on multi-scale feature aggregation and dynamic convolution,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Robust spoof speech detection based on multi-scale feature aggregation and dynamic convolution,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.929032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.536857Z digest=sha256:749ea45bd5bf08c868e21e453f4087b672abe88875963e4ccac7ee48abde7ea2

Observation 65e212ef-e50f-488d-a048-d3225b0b0feb · outbound

This paper cites Automatic speaker verification spoofing and deepfake detection using wav2vec 2.0 and data augmenta- tion,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Automatic speaker verification spoofing and deepfake detection using wav2vec 2.0 and data augmenta- tion,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.906918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.552238Z digest=sha256:b79154253160ac815369920f0e3094cfc372cc3b6440e116a041473791746296

Observation 1d9f4958-6701-45e7-a0f6-f861346fc972 · outbound

This paper cites Investigating self-supervised front ends for speech spoofing countermeasures,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Investigating self-supervised front ends for speech spoofing countermeasures,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.866635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.561268Z digest=sha256:1259526d1738ade34e699fb9c490f70abc042c49861bcf3a275bf87be159ba59

Observation cd1d6e70-f61b-4430-8906-bc50e621ed06 · outbound

This paper cites Lora: Low-rank adaptation of large language models,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Lora: Low-rank adaptation of large language models,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.842890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.567817Z digest=sha256:1f75f9b4733234737a2d4dc7b643a209309b4e29c7023a82c8bedc07b5cf3cf4

Observation b25906c1-fa8d-4b93-8685-cad55e7f03e8 · outbound

This paper cites Parameter-efficient transfer learning for nlp,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Parameter-efficient transfer learning for nlp,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.817025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.575475Z digest=sha256:f83f68e674b4d1fad33687b63d6f477541d8d20b96388f82c5dd49d0ba4dcce6

Observation 6a8a19f9-1d50-4fc6-99b6-7edd2978f635 · outbound

This paper cites Audio deepfake detection with self- supervised xls-r and sls classifier,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Audio deepfake detection with self- supervised xls-r and sls classifier,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.784447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.582486Z digest=sha256:982d7c9c9e1f865827fc4f2f0a5a5046376e79dded990c0166558a5f85f61f7e

Observation 0c23d29d-6f25-480b-b1b9-da2f266669d3 · outbound

This paper cites Attentive merging of hidden embeddings from pre-trained speech model for anti-spoofing detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Attentive merging of hidden embeddings from pre-trained speech model for anti-spoofing detection,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.754575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.589077Z digest=sha256:1a5c14a96a4c59e2fe92a4d9acf3c421f69106bbca898694fea5f457a3147daf

Observation a3f4acad-1ad8-4a97-8bb6-34d1718867cf · outbound

This paper cites Outrageously large neural networks: The sparsely-gated mixture-of-experts layer,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Outrageously large neural networks: The sparsely-gated mixture-of-experts layer,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.727153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.597639Z digest=sha256:e824a6c4ba5f34d1a204e7e9c77bf516a8b8d2bf5f689e3bec789a8f4b60853f

Observation 6d860ee5-5eb3-43d8-96c4-d83eddc48c9f · outbound

This paper cites Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.697394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.610876Z digest=sha256:92e37f81e2ef061c4d7e021fcb893b9c8f71cc3efea6da40e47a846df3e2875e

Observation d961a671-a693-4f34-82dd-a6941668f64c · outbound

This paper cites Hifi-gan: Generative adversarial networks for efficient and high fidelity speech synthesis,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Hifi-gan: Generative adversarial networks for efficient and high fidelity speech synthesis,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.661369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.618851Z digest=sha256:1f4d9ad0b507c63708be5eaae0ada9e00622103e077991f180bf9e14ec4912e1

Observation 7fab0d9b-c156-4d25-989b-a3fea427a85b · outbound

This paper cites Fastdiff: A fast conditional diffusion model for high-quality speech synthesis,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Fastdiff: A fast conditional diffusion model for high-quality speech synthesis,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.633203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.624994Z digest=sha256:0bbfe7322a092bb2a8211147dcf8403d467106e132e936688df6a162c840d79b

Observation fa174add-e3d3-4ae7-bf5f-af9339b3057d · outbound

This paper cites Freevc: Towards high-quality text-free voice conversion,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Freevc: Towards high-quality text-free voice conversion,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.600627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.632325Z digest=sha256:bf6cbad381c5c1f514e0364009e199997e62d86996cc83c07946302b0331c7ca

Observation aedd6042-2cc0-4a3f-a886-dc06bcd8e49d · outbound

This paper cites Asvspoof 2019: A large- scale public database of synthesized, converted and replayed speech,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Asvspoof 2019: A large- scale public database of synthesized, converted and replayed speech,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.546545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.638598Z digest=sha256:884bedf377ac1b1b3d8924c0c2c07d8df71dc2ad31bcde225557783fbaf9e227

Observation 9cdcd84e-a8fa-4c44-ba8d-ff4528ec75d9 · outbound

This paper cites Asvspoof 2021: Towards spoofed and deepfake speech detection in the wild,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Asvspoof 2021: Towards spoofed and deepfake speech detection in the wild,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.524777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.645949Z digest=sha256:022ad290b2c7c18f6be15557850d887fc583f94d4efeb16bb429261ac86100bc

Observation f0870546-ba7c-40c5-979a-c21c66db1b5a · outbound

This paper cites Does audio deepfake detection generalize?,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Does audio deepfake detection generalize?,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T10:24:21.653089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:24:21.653089Z digest=sha256:ac032d3c4af6099ced41cb96ce75b7f31c948cb8f19987790ff73a61798ea050

Observation 33e3a0e4-df49-433f-bf4a-6bb631854fcf · outbound

This paper cites t-dcf: a detection cost function for the tandem assessment of spoofing countermeasures and automatic speaker verification,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection t-dcf: a detection cost function for the tandem assessment of spoofing countermeasures and automatic speaker verification,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.486726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.670151Z digest=sha256:23b4e35ddba43ea9e2b1031717ac056858f1054340be790398a9eb700635e0de

Observation 01b0cda1-daa3-4201-b06f-1a95455a02e3 · outbound

This paper cites Rawboost: A raw data boosting and augmentation method applied to automatic speaker verification anti- spoofing,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Rawboost: A raw data boosting and augmentation method applied to automatic speaker verification anti- spoofing,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.436447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.677332Z digest=sha256:a24aeac4f6bb7497eb50e5a58623ddbd30366138b259546e4394d49688908888

Observation ab76689c-a2af-4eaf-a937-958698e4afd7 · outbound

This paper cites Improving short utterance anti-spoofing with aasist2,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Improving short utterance anti-spoofing with aasist2,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.401284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.692373Z digest=sha256:bd7add66ce9bbd1d326845697535222e170dce5594f2bb921c9c75f524a7deb5

Observation f77714c3-81f4-4924-9b58-4c158c99b7fc · outbound

This paper cites A conformer-based classifier for variable-length utterance processing in anti-spoofing,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection A conformer-based classifier for variable-length utterance processing in anti-spoofing,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.378433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.711964Z digest=sha256:006420dbdba3225c90d6e454e164ada38ddcb278833855b197107d160649eea8

Observation ca664df3-d1ec-4a9f-9947-994fd6e85791 · outbound

This paper cites Audio deepfake detection with self-supervised wavlm and multi-fusion attentive classi- fier,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Audio deepfake detection with self-supervised wavlm and multi-fusion attentive classi- fier,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.332612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.723573Z digest=sha256:240a0f1b19e7d9eefe8d01a66ca9274526a5a7be51c5d51ee326c900a90ae7e1

Observation 8bb61193-c48a-465c-94a6-15c24fa27804 · outbound

This paper cites One class learning with adaptive centroid shift for audio deepfake detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection One class learning with adaptive centroid shift for audio deepfake detection,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.300309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.735335Z digest=sha256:d259b05c19d70d8d2e12b10d01a71a3b63ef34e5293f67eb25a6992f7de73cb6

Observation 5be16f9c-d688-4e58-af75-6d045fd582b8 · outbound

This paper cites Temporal-channel modeling in multi-head self-attention for synthetic speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Temporal-channel modeling in multi-head self-attention for synthetic speech detection,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.273266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.740808Z digest=sha256:e250ba073568956ee566df8b7754ce61948416b8a00cc1c1e19530c2084be726

Observation 16b44864-4e4b-4a03-bc9d-291c03435c4d · outbound

This paper cites A robust audio deepfake detection system via multi-view feature,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection A robust audio deepfake detection system via multi-view feature,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.221833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.748635Z digest=sha256:f73f0518706d65fbd16fb8fb7fc9efbc76713c562ab6f1b55edbe187731aed56

Observation 982ae0e7-4a3d-4158-a9a4-e1788409b501 · outbound

This paper cites Spoofed training data for speech spoofing countermeasure can be efficiently created using neural vocoders,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Spoofed training data for speech spoofing countermeasure can be efficiently created using neural vocoders,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.172178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.757104Z digest=sha256:ff6692663ea3d20929487567b614c0ee075605ecd079568f7c6dfd14bc559372

Observation e52fcafd-748c-402c-851c-12c99b215171 · outbound

This paper cites Can large-scale vocoded spoofed data improve speech spoofing countermeasure with a self-supervised front end?,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Can large-scale vocoded spoofed data improve speech spoofing countermeasure with a self-supervised front end?,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.131597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.767444Z digest=sha256:285a02358546621c915e506381c3990e41573918a7bac8ed4354a89fbf2b5343

Observation 4a7180f7-0ce7-4e35-84fd-32e0a20b35d3 · outbound

This paper cites One-class knowl- edge distillation for spoofing speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection One-class knowl- edge distillation for spoofing speech detection,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.105531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:24:21.774639Z digest=sha256:65a8a87481cd0face2ace6f427c5c0ca51a93f160b1b659601ddd2430fab29d5

Pith citing papers

Observation e60da123-abcf-460e-a92e-e8dbb05497e8 · inbound

Alethia: A Foundational Encoder for Voice Deepfakes cites this paper.

Alethia: A Foundational Encoder for Voice Deepfakes Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:32.948686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:5a448efe96e6f666ac785912eca036f73d8b0c2aee02b96ef9c101d29ee43555