Pith. sign in

Paper Citation Record · LEDGER

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning

As of 15 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 1 inbound Pith citation observation for arXiv:2412.00175.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.00175 v3

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T05:44:30.140670Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T14:34:11.253445Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T14:34:12.126271Z

Reference resolution

60 of 60 outbound references displayed

  • verified exact2
  • verified fuzzy42
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0f1a3568-504f-403d-9794-aed617ec33e9 · outbound

This paper cites MesoNet: A compact facial video forgery detection network.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning MesoNet: A compact facial video forgery detection network

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.924180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.902867Z digest=sha256:623775134617d844e20f1d7d987902f2142cab3ea7dc1f2230d39cc0d7b422cd

Observation 37e61685-63f3-40c6-a3e0-c1df0226a21f · outbound

This paper cites Lost in translation: Lip- sync deepfake detection from audio-video mismatch.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Lost in translation: Lip- sync deepfake detection from audio-video mismatch

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.912186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.907182Z digest=sha256:7813cdfcf311904d01de7bf1b55f828b2b57930a79d2a2ce7cface92659b621d

Observation a5321d64-0dcc-4c1d-ab08-171b41a7d054 · outbound

This paper cites Is synthetic voice detection research going into the right direction? InCVPRW, pages 71–80, 2022.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Is synthetic voice detection research going into the right direction? InCVPRW, pages 71–80, 2022

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.901135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.910553Z digest=sha256:1ccffc8a236d9ce137ca20b0f29b89308d8ac3c8722436feaa91b5314e958799

Observation ca19337e-d4c9-4cd0-a22b-c7f0ef922525 · outbound

This paper cites Glitch in the ma- trix: A large scale benchmark for content driven audio-visual forgery detection and localization.Comput.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Glitch in the ma- trix: A large scale benchmark for content driven audio-visual forgery detection and localization.Comput

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.889636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.914197Z digest=sha256:7c9a93d176797216729efa8c53de98f7a7a5aae4c2a1763ae41ee1af9a8be2c4

Observation 41e03800-cb17-4061-a8fe-5123ee75ae57 · outbound

This paper cites MARLIN: Masked autoencoder for facial video rep- resentation learning, 2023.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning MARLIN: Masked autoencoder for facial video rep- resentation learning, 2023

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.877434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.918272Z digest=sha256:11ee2724aff0a13faaf4e5d2ce3f07fec05432458e18e4a6c9f358b484318464

Observation 204c936b-d5a6-4257-aab1-e31ce11da211 · outbound

This paper cites A V-Deepfake1M: A large-scale LLM-driven audio-visual deepfake dataset, 2024.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning A V-Deepfake1M: A large-scale LLM-driven audio-visual deepfake dataset, 2024

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.866718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.921956Z digest=sha256:264d7b78aae71d4d2bb7f887634c5eb4468ed899277a9fd839f45911984b5051

Observation e0d3c3d5-aa07-4aaa-b7a0-db4cddaee990 · outbound

This paper cites an unresolved cited work.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-12T05:44:30.854772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.925819Z digest=sha256:be1239cc11d1248f7ab7ef0cef2a115f45a31a6b34c389c2c4d0e184b390267d

Observation 2f537494-10f5-446d-96c9-ef6ace1da91e · outbound

This paper cites What makes fake images detectable? understanding prop- erties that generalize.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning What makes fake images detectable? understanding prop- erties that generalize

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.844018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.929916Z digest=sha256:3e7ea02e236823be577e600432cd26daa3201e73d34751baac0c6e5595e44a72

Observation ddcda2a4-a8a4-4a14-b71d-20ec75219149 · outbound

This paper cites Xception: Deep learning with depthwise separable convolutions.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Xception: Deep learning with depthwise separable convolutions

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.832907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.933869Z digest=sha256:136deeb5028caf6d86af87eb08fc67dc7d55168adb50bf40e28f9323ce8062a7

Observation 601c765b-d82a-4590-b1fb-204f6be653ca · outbound

This paper cites Not made for each other-audio- visual dissonance-based deepfake detection and localization.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Not made for each other-audio- visual dissonance-based deepfake detection and localization

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.820744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.937924Z digest=sha256:6e99e981c9126d0a2b4ac925c4e871872f2141b5f75b865772ef285019cca54b

Observation 02b1325f-7e4f-4c1d-9e0e-c2c915a5755e · outbound

This paper cites an unresolved cited work.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-12T05:44:30.807771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.941961Z digest=sha256:e2e95200a74d334571db554d5104fa49d4bfb6806acb880095803e181200e496

Observation 7522d123-ddf9-45af-8cad-ccd1a55ba0c7 · outbound

This paper cites Combining EfficientNet and vision transformers for video deepfake detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Combining EfficientNet and vision transformers for video deepfake detection

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.794869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.946772Z digest=sha256:94449d17a307aa54334afd225b9ede0531bb16b83df018acea17d98498871448

Observation 0e91a96a-2aea-46f4-a7b1-e1588ed3383b · outbound

This paper cites Raising the bar of AI-generated image detection with CLIP.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Raising the bar of AI-generated image detection with CLIP

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.778987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.951523Z digest=sha256:4453408adf891768b39e20a0fde96c7cfe404a9bc994c561037c3d1b210dc193

Observation f59e3463-fe24-48a2-ac7d-ac2400aaa6f9 · outbound

This paper cites Zero-shot detection of AI-generated im- ages.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Zero-shot detection of AI-generated im- ages

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.766713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.955266Z digest=sha256:cb2ddab0d7df3c42af56974d974b594fda758fd48ec03fbb307bba5f56f733c1

Observation 4ff1bd2d-e13c-4e45-8d4b-bfbcf5838d64 · outbound

This paper cites Real time speech enhancement in the waveform domain.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Real time speech enhancement in the waveform domain

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.754457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.958972Z digest=sha256:d17b1dea6d0ebf07944eecdb75d1f5bf790df0926e80454b65283ef21b7cff41

Observation 9b9bf00c-b21a-4975-b554-beb863f04c45 · outbound

This paper cites The DeepFake Detection Challenge (DFDC) Dataset.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning The DeepFake Detection Challenge (DFDC) Dataset

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:29.962790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:29.962790Z digest=sha256:d3e598b50cc7918f33eb7ad88285343ffd6830d85d3f8aeefe1994f35949ed5c

Observation b9dc861f-40ff-4064-b2db-040a5d505a29 · outbound

This paper cites Self- supervised video forensics by audio-visual anomaly detec- tion.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Self- supervised video forensics by audio-visual anomaly detec- tion

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.742861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.967173Z digest=sha256:e9d34cc447b8b1b9383f09fb49f7e0e7971a3c928c162fc8d2fc4a724feef651

Observation 94eb28c8-6cf5-4d2f-b66a-9e4fe7874ada · outbound

This paper cites Lips don’t lie: A generalisable and robust approach to face forgery detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Lips don’t lie: A generalisable and robust approach to face forgery detection

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.730819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.970864Z digest=sha256:6878234d58c91c19280e9c16a21612ca588f4ee7c13d5b38a1e99fce6c3af825

Observation 61dc8e57-70cd-4bca-b411-02c32fc4731d · outbound

This paper cites Leveraging real talking faces via self- supervision for robust forgery detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Leveraging real talking faces via self- supervision for robust forgery detection

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.719091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.975148Z digest=sha256:c2074004e403a8d04f22682257218802c656ac252121c957852552632d8527b7

Observation 85c606d9-92f6-44ac-964a-1ca897a6c2f6 · outbound

This paper cites AVTENet: A Human-Cognition-Inspired Audio-Visual Transformer-Based Ensemble Network for Video Deepfake Detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning AVTENet: A Human-Cognition-Inspired Audio-Visual Transformer-Based Ensemble Network for Video Deepfake Detection

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-12T05:44:30.367926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.979061Z digest=sha256:d6fc7a8ea55200168edee313d40f24db393463bcc3d4be614a0c37d6f8212338

Observation 193d8633-31df-4c3c-a4a6-51e11fdb8b32 · outbound

This paper cites Implicit identity driven deepfake face swapping detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Implicit identity driven deepfake face swapping detection

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.708387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.983103Z digest=sha256:915dbe153da3fa589b10ab7e1b8ca4d886ebbb708c25afd25c3e39c76d8d9573

Observation ac5c7317-76a3-4300-8343-73765124b811 · outbound

This paper cites Weiss, Quan Wang, Jonathan Shen, Fei Ren, Zhifeng Chen, Patrick Nguyen, Ruoming Pang, Ig- nacio L ´opez-Moreno, and Yonghui Wu.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Weiss, Quan Wang, Jonathan Shen, Fei Ren, Zhifeng Chen, Patrick Nguyen, Ruoming Pang, Ig- nacio L ´opez-Moreno, and Yonghui Wu

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.697002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.986480Z digest=sha256:b8823bc0dec062fe6dc252235836ca679a2997f0c66e90326e4d00ef85c9545e

Observation 3b7b43e2-c716-49d5-b0ac-c0690dd3e926 · outbound

This paper cites an unresolved cited work.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-12T05:44:30.685497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.989809Z digest=sha256:1b177070a8401969c9f232b0c6132285551f8363f003155ab824ec096d7fa2fc

Observation 9ee132c4-b704-4c2e-95d3-8b8d019f0cdc · outbound

This paper cites Conditional variational autoencoder with adversarial learning for end-to- end text-to-speech.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Conditional variational autoencoder with adversarial learning for end-to- end text-to-speech

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.674615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:29.993621Z digest=sha256:3a71d39c78f4d3701f550790cc0e90c3d2294fac4f9de9b73888b448f4f4b278

Observation 6f07a43b-8e35-4e9d-b9c3-d0c55d9f97cd · outbound

This paper cites DeepFakes: a New Threat to Face Recognition? Assessment and Detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning DeepFakes: a New Threat to Face Recognition? Assessment and Detection

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:29.997426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:29.997426Z digest=sha256:756c27bdf17002e49f59d883678db07fad30309033a6f1f18fe29d4aa736b479

Observation 07a55832-9392-4694-8f79-7ab5cd4afac2 · outbound

This paper cites Fast face-swap using convolutional neural networks.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Fast face-swap using convolutional neural networks

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.662715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.001658Z digest=sha256:b2f200a91abfe34d437b9ac3d40d1dd50acabbce97fec7c4970cd6578e349339

Observation 2f3969ed-538d-4505-85f6-cf0859757533 · outbound

This paper cites DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.005614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.005614Z digest=sha256:73d5c85b7b480f49004d4375bfed68e0d2b40704e5f4efb523d67a6a7e88f668

Observation d83fdd33-7043-40ff-bab6-37a75629e3b0 · outbound

This paper cites KoDF: A large-scale Korean deepfake detection dataset.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning KoDF: A large-scale Korean deepfake detection dataset

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.650538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.010154Z digest=sha256:c6c632a7646cc705f21abe04ae770b60de5fdd7c580e84b11e65e9698f2f8b37

Observation 6f7b9195-4c79-4c78-a07e-383e73b100e3 · outbound

This paper cites Zero-Shot Fake Video Detection by Audio-Visual Consistency.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Zero-Shot Fake Video Detection by Audio-Visual Consistency

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.015336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.015336Z digest=sha256:cd1b6e971e247f61193c8d292bbe6f527d96d35fc89d748b6d9d720b6f0e9d9a

Observation d103036f-1b70-4766-8d34-16be070831ac · outbound

This paper cites SpeechForensics: Audio-visual speech representation learn- ing for face forgery detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning SpeechForensics: Audio-visual speech representation learn- ing for face forgery detection

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.638655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.019867Z digest=sha256:17cbbc538dfbe3d674b49e9468198c2c89c1fb8ee0f85b70f15f9b35e0203e50

Observation 864ae0bf-3730-4f9a-a9dc-ef63542c0bf0 · outbound

This paper cites Lips are lying: Spotting the temporal inconsistency between audio and visual in lip- syncing deepfakes.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Lips are lying: Spotting the temporal inconsistency between audio and visual in lip- syncing deepfakes

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.626444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.024385Z digest=sha256:40fcea291b50d538ec2467fe3445be61b6fa446b9cb32a96fb7920b86d594fb4

Observation 68f3aa07-355d-4ca7-b494-0495de107945 · outbound

This paper cites When Synthetic Traces Hide Real Content: Analysis of Stable Diffusion Image Laundering.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning When Synthetic Traces Hide Real Content: Analysis of Stable Diffusion Image Laundering

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-12T05:44:30.318622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.028583Z digest=sha256:d337b4590f1db4532240df3a22eb15a3685899da091d17c6a9fae304d2fdd5d4

Observation 2953a0a3-8d77-4e60-acc9-6b1b3fe83f5c · outbound

This paper cites TGIF: Text-Guided Inpainting Forgery Dataset.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning TGIF: Text-Guided Inpainting Forgery Dataset

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.032962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.032962Z digest=sha256:faca209b487705cd2966ce08b9258bf296712bf5688e4dd571552ac35af5e6ed

Observation 1fa776b0-864a-4663-a75b-a74cbc5bda9e · outbound

This paper cites Do GANs leave artificial fingerprints? In MIPR, pages 506–511, 2019.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Do GANs leave artificial fingerprints? In MIPR, pages 506–511, 2019

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.614526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.037292Z digest=sha256:0a484f79a488100930d3984b9befab1120f0b85725d4e97eb1970d9b72016b4d

Observation 5ff8c039-2b88-415d-81d0-c3259f4fd4a5 · outbound

This paper cites Speech is Silver, Silence is Golden: What do ASVspoof-trained Models Really Learn?.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Speech is Silver, Silence is Golden: What do ASVspoof-trained Models Really Learn?

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.041044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.041044Z digest=sha256:443a30e34b9884d8a1e24bc15777c181499343a5083729a2d6daafa07eabc12a

Observation aeb44556-7661-46de-a740-a292c5702c81 · outbound

This paper cites M ¨uller, Piotr Kawa, Wei Herng Choong, Edres- son Casanova, Eren G ¨olge, Thorsten M ¨uller, Piotr Syga, Philip Sperl, and Konstantin B¨ottinger.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning M ¨uller, Piotr Kawa, Wei Herng Choong, Edres- son Casanova, Eren G ¨olge, Thorsten M ¨uller, Piotr Syga, Philip Sperl, and Konstantin B¨ottinger

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.602844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.045120Z digest=sha256:26ca425d74172fd0aa46956401370f0ab15163756edfc336dedc539291f19753

Observation 969af691-c1d9-44d7-a802-6dfc7f92a78c · outbound

This paper cites FSGAN: Subject agnostic face swapping and reenactment.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning FSGAN: Subject agnostic face swapping and reenactment

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.591140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.048924Z digest=sha256:1497e0ea222c6442b25d1ba9c7826d48cc2f640f4a7935bfaf460aae98164470

Observation ea3a47d5-9301-44c7-a21c-da8c69fd17b4 · outbound

This paper cites Towards uni- versal fake image detectors that generalize across generative models.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Towards uni- versal fake image detectors that generalize across generative models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.578222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.052119Z digest=sha256:e54ffaee9958bc0a68299cf31300b3252aabe4240b3ffeba8b0314c6cc3ac650

Observation 5ae1e3f7-bfde-4e9b-90d6-bd3990cf174d · outbound

This paper cites A VFF: Audio-visual feature fusion for video deepfake detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning A VFF: Audio-visual feature fusion for video deepfake detection

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.566161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.055580Z digest=sha256:a51fd16c79aea7d2d5c2330ad4af78b3e8ec6a9ec7dd341d5028df557a6aa6d5

Observation 9b46e0a9-8ec3-41f8-83bc-25fd6f4fbfd0 · outbound

This paper cites Towards generalisable and cali- brated audio deepfake detection with self-supervised repre- sentations.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Towards generalisable and cali- brated audio deepfake detection with self-supervised repre- sentations

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.556304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.059228Z digest=sha256:19e02866205f28425990074a536a1f6df258a4c8c0b6765d89044a2932a27511

Observation e781906e-45f3-4e1c-a5e2-6ad8ba2870cb · outbound

This paper cites Training-free deepfake voice recognition by leveraging large-scale pre-trained models.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Training-free deepfake voice recognition by leveraging large-scale pre-trained models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.546248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.062350Z digest=sha256:85364257dca18dbae85011fe336124cc72353d19572c882f243e6755528b7521

Observation b548f3d9-54c6-406e-862c-fca0b7e40564 · outbound

This paper cites Nambood- iri, and C.V.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Nambood- iri, and C.V

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.535050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.065597Z digest=sha256:ff8c7afd170e89dd701d2dd6585ec08e340097e26ac3512b58d3a180618c6b0c

Observation a5b3b9be-cb06-4225-af1c-6ac3e4675cc4 · outbound

This paper cites Aligned Datasets Improve Detection of Latent Diffusion-Generated Images.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Aligned Datasets Improve Detection of Latent Diffusion-Generated Images

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.069372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.069372Z digest=sha256:4770fbea514e6026f652d4d0331302a155039c1bf5d73dea52f6d20b86b79658

Observation 80a80d7d-94d0-4ed7-a609-0174cd2ae0c4 · outbound

This paper cites Detecting Deepfakes Without Seeing Any.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Detecting Deepfakes Without Seeing Any

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.073469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.073469Z digest=sha256:daec3b53eb2af9800c25a92ba5d9a27a83ca7ed7f454a2d7358145c75d76a7bc

Observation ae187751-f545-462d-ba06-4a0eaca1bf26 · outbound

This paper cites AEROB- LADE: Training-free detection of latent diffusion images us- ing autoencoder reconstruction error.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning AEROB- LADE: Training-free detection of latent diffusion images us- ing autoencoder reconstruction error

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.523986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.077820Z digest=sha256:ea312dca6d70da5786e0846f07e3329e01ef78868611b5f8476941bc55c1f992

Observation d02d19e7-b2af-40a2-8a8e-11703e729397 · outbound

This paper cites FaceForen- sics++: Learning to detect manipulated facial images.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning FaceForen- sics++: Learning to detect manipulated facial images

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.512076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.081672Z digest=sha256:62646122e9854a354c9eb3a6b923ba059c0a5d53a7d61d20f698d4dcd0aec40e

Observation cf3cfcc4-06bd-4920-9546-450aed995371 · outbound

This paper cites Hosler, Paolo Bestagini, Matthew C.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Hosler, Paolo Bestagini, Matthew C

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.500830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.085554Z digest=sha256:e20752f5300f70014ed7af944f19536a0d008c7471cd0ce7b47bce2caa684d67

Observation 15fcc1e4-fd8d-4958-a636-fb21c09db0b8 · outbound

This paper cites A V-Lip-Sync+: Lever- aging A V-HuBERT to exploit multimodal inconsistency for video deepfake detection.CoRR, abs/2311.02733, 2023.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning A V-Lip-Sync+: Lever- aging A V-HuBERT to exploit multimodal inconsistency for video deepfake detection.CoRR, abs/2311.02733, 2023

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.089207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.089207Z digest=sha256:ef38325253a39249bfdd7e1d444067bf8f698aa1fe43b2e75bcb2e1e9a072ce9

Observation 6134641f-d05b-46e0-a138-e99f04eec00d · outbound

This paper cites Learning audio-visual speech representation by masked multimodal cluster prediction.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Learning audio-visual speech representation by masked multimodal cluster prediction

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.489852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.093019Z digest=sha256:6bd50a3637242ed402ceee2b8ff824243cbb877716ba2fff4e32bda9a74a072e

Observation 1c02577e-a80d-48f7-8c25-d8e8650742b4 · outbound

This paper cites Detecting deep- fakes with self-blended images.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Detecting deep- fakes with self-blended images

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.477035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.096651Z digest=sha256:780d57fd27478bbb6e4b20313bc5b1e91f686bec05b65d69007109569bbb917d

Observation 1b0066b4-d6ec-4b9f-a7d7-67e4c79f2bb8 · outbound

This paper cites DeCLIP: Decoding CLIP representations for deepfake localization.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning DeCLIP: Decoding CLIP representations for deepfake localization

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.100607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.100607Z digest=sha256:1ecb5d9435adc97655238dcd942182d500dd9faa3a0a4108fb3345603a40344a

Observation 92b24c84-6659-4cbd-88c0-6b6ce0746084 · outbound

This paper cites Lip reading sentences in the wild.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Lip reading sentences in the wild

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.465517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.105280Z digest=sha256:3ace3fccdf3e6cf9d505e1c9c06761023c4bdb50f86fe20766af8b8af4433ce1

Observation 870152d8-6ac1-4255-b6bf-b0ff5c6fe1bc · outbound

This paper cites End-to-end anti-spoofing with RawNet2.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning End-to-end anti-spoofing with RawNet2

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.454182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.109954Z digest=sha256:05e5a9100bc4fb37afff1830b8e7d7a591c7f510ec90c5423223d4578809e439

Observation d415a12d-aedf-435d-92bd-0b82b23a53da · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Representation Learning with Contrastive Predictive Coding

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.114598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.114598Z digest=sha256:288eeca85e246db42a65ebe3ba358402dbb9ab55ecc57c14131e4b57d1e8a494

Observation 54636c10-493f-46fe-8399-0cff0a423dc3 · outbound

This paper cites Tan, and Haizhou Li.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Tan, and Haizhou Li

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.442071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.119711Z digest=sha256:7145268fd00cfb449ea5d9d76819fa060146084aea4ab5113cd193a0b898d829

Observation 1747dde7-7503-4e95-b232-e51c5471ec7d · outbound

This paper cites an unresolved cited work.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-12T05:44:30.429625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.124304Z digest=sha256:dd4d3b86761a38ce1da2996e2be7a752b926e1454bffa086cd3300726375d4d6

Observation b02d6ff3-e91c-46b0-bbe0-df84d750d801 · outbound

This paper cites DF40: Toward Next-Generation Deepfake Detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning DF40: Toward Next-Generation Deepfake Detection

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.128411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.128411Z digest=sha256:ed719beb2fc79e81809187cd8c2f2fe2f243c7af9f251dfdee192d005d1d11c9

Observation 3fd4eb5b-b339-40f2-9821-dadb6a384bd9 · outbound

This paper cites 10 A V oiD-DF: Audio-visual joint learning for detecting deep- fake.IEEE Trans.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning 10 A V oiD-DF: Audio-visual joint learning for detecting deep- fake.IEEE Trans

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.417412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.132785Z digest=sha256:2ee91d67cc0e0f3d8e29264b8f6d9c389c6414386ec983d0bfa1f6ab8df94121

Observation 01786df4-5813-4367-b56e-a9c5706684c4 · outbound

This paper cites Attributing fake images to GANs: Learning and analyzing GAN fingerprints.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Attributing fake images to GANs: Learning and analyzing GAN fingerprints

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.405087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.136494Z digest=sha256:1160c58e756bea737911848a3852803dcebaa2360db509d510eb1b1c5a6b6207

Observation 56102646-6931-4131-878f-42212b6c5254 · outbound

This paper cites Exploring temporal coherence for more general video face forgery detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Exploring temporal coherence for more general video face forgery detection

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.392580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:44:30.140670Z digest=sha256:0d2734c9ebcc788a1b81d8e431b0757d8216becf138269c0c7b4800cacb2d202

Pith citing papers

Observation cf789b32-7ebb-4b20-a463-9edcb03fb8f5 · inbound

Unmasking Synthetic Realities in Generative AI: A Comprehensive Review of Adversarially Robust Deepfake Detection Systems cites this paper.

Unmasking Synthetic Realities in Generative AI: A Comprehensive Review of Adversarially Robust Deepfake Detection Systems Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning

Reference 200

Resolution
verified exact
local_arxiv, observed 2026-08-06T14:34:12.131864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T14:34:11.253445Z digest=sha256:e42e5465e09ea668faeca689d3bfbcb20467af390ce169a937aa5cac76ee42c0