Pith. sign in

Paper Citation Record · LEDGER

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning

As of 14 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 1 inbound Pith citation observation for arXiv:2412.00175.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.00175 v3

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T05:44:30.140670Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T14:34:11.253445Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T14:34:12.126271Z

Reference resolution

60 of 60 outbound references displayed

  • verified exact2
  • verified fuzzy42
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0f1a3568-504f-403d-9794-aed617ec33e9 · outbound

This paper cites MesoNet: A compact facial video forgery detection network.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning MesoNet: A compact facial video forgery detection network

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.924180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.902867Z digest=sha256:8e5c1846154acdbda2f5b980093f70fafb5f3c855ab64ce0ba541e956f5d7aaf

Observation 37e61685-63f3-40c6-a3e0-c1df0226a21f · outbound

This paper cites Lost in translation: Lip- sync deepfake detection from audio-video mismatch.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Lost in translation: Lip- sync deepfake detection from audio-video mismatch

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.912186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.907182Z digest=sha256:72a6145fea978bc47780b0c615815f4bf06e5403a7a09d75f4184750d30e40f1

Observation a5321d64-0dcc-4c1d-ab08-171b41a7d054 · outbound

This paper cites Is synthetic voice detection research going into the right direction? InCVPRW, pages 71–80, 2022.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Is synthetic voice detection research going into the right direction? InCVPRW, pages 71–80, 2022

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.901135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.910553Z digest=sha256:db54afb5ae77d4287458c401e04651d65a6285b21d34e4ced64b1dc0bc72dc42

Observation ca19337e-d4c9-4cd0-a22b-c7f0ef922525 · outbound

This paper cites Glitch in the ma- trix: A large scale benchmark for content driven audio-visual forgery detection and localization.Comput.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Glitch in the ma- trix: A large scale benchmark for content driven audio-visual forgery detection and localization.Comput

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.889636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.914197Z digest=sha256:2a62ad8ca1c85eb393636d393d767877cfa4aefaea293cbbc71711735983f503

Observation 41e03800-cb17-4061-a8fe-5123ee75ae57 · outbound

This paper cites MARLIN: Masked autoencoder for facial video rep- resentation learning, 2023.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning MARLIN: Masked autoencoder for facial video rep- resentation learning, 2023

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.877434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.918272Z digest=sha256:b47c8a9cfb987d917a747b324063a085f69f5fabd6ad847d41d1b80e643392b2

Observation 204c936b-d5a6-4257-aab1-e31ce11da211 · outbound

This paper cites A V-Deepfake1M: A large-scale LLM-driven audio-visual deepfake dataset, 2024.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning A V-Deepfake1M: A large-scale LLM-driven audio-visual deepfake dataset, 2024

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.866718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.921956Z digest=sha256:34de110ddfcd0a830580b7bef61048a1a3659074d72651a45bf9c35bf94e741c

Observation e0d3c3d5-aa07-4aaa-b7a0-db4cddaee990 · outbound

This paper cites an unresolved cited work.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-12T05:44:30.854772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.925819Z digest=sha256:88aba824a20f9d0d727828998bd748b1ae51325bd7fadb9686d9fbc41c3a48f4

Observation 2f537494-10f5-446d-96c9-ef6ace1da91e · outbound

This paper cites What makes fake images detectable? understanding prop- erties that generalize.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning What makes fake images detectable? understanding prop- erties that generalize

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.844018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.929916Z digest=sha256:10f2e2947659e29c89430cc5163b82eb6e54a69be00c21a2e72912b177b6a758

Observation ddcda2a4-a8a4-4a14-b71d-20ec75219149 · outbound

This paper cites Xception: Deep learning with depthwise separable convolutions.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Xception: Deep learning with depthwise separable convolutions

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.832907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.933869Z digest=sha256:86829f5327880f3c5778c99f6beb5dc8c3c75496a4b5327aa46a2300ff6e7fda

Observation 601c765b-d82a-4590-b1fb-204f6be653ca · outbound

This paper cites Not made for each other-audio- visual dissonance-based deepfake detection and localization.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Not made for each other-audio- visual dissonance-based deepfake detection and localization

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.820744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.937924Z digest=sha256:9f6793306bcb55bdf9ee6771b61601f05f7f8a4fff19291aa02534bb19147814

Observation 02b1325f-7e4f-4c1d-9e0e-c2c915a5755e · outbound

This paper cites an unresolved cited work.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-12T05:44:30.807771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.941961Z digest=sha256:5baaabca0c49b1c4a4bfb7e03d2dac7f3698a7142f38f0016b1c15d9162a468e

Observation 7522d123-ddf9-45af-8cad-ccd1a55ba0c7 · outbound

This paper cites Combining EfficientNet and vision transformers for video deepfake detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Combining EfficientNet and vision transformers for video deepfake detection

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.794869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.946772Z digest=sha256:64f56e679de0bdbecdcfd36c758f22e273707b82aab5d68b3c08b7f0cb49a74e

Observation 0e91a96a-2aea-46f4-a7b1-e1588ed3383b · outbound

This paper cites Raising the bar of AI-generated image detection with CLIP.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Raising the bar of AI-generated image detection with CLIP

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.778987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.951523Z digest=sha256:f29a5ed5278f56dc634f99757d0d881010930d5c78c26352883e1f2c77e14ef2

Observation f59e3463-fe24-48a2-ac7d-ac2400aaa6f9 · outbound

This paper cites Zero-shot detection of AI-generated im- ages.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Zero-shot detection of AI-generated im- ages

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.766713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.955266Z digest=sha256:beacf884257e5640b003908bd0daf1125308b05985b545f272d04d4cdde1ac7d

Observation 4ff1bd2d-e13c-4e45-8d4b-bfbcf5838d64 · outbound

This paper cites Real time speech enhancement in the waveform domain.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Real time speech enhancement in the waveform domain

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.754457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.958972Z digest=sha256:e34b1be19a3179e930b6792442017d561d45b30bddce141fc7df7281ecaf863c

Observation 9b9bf00c-b21a-4975-b554-beb863f04c45 · outbound

This paper cites The DeepFake Detection Challenge (DFDC) Dataset.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning The DeepFake Detection Challenge (DFDC) Dataset

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:29.962790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:29.962790Z digest=sha256:f992320cd4a95a40ebec9f4b424bbada69466aab0f2209dc63a6bf99c654b7cf

Observation b9dc861f-40ff-4064-b2db-040a5d505a29 · outbound

This paper cites Self- supervised video forensics by audio-visual anomaly detec- tion.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Self- supervised video forensics by audio-visual anomaly detec- tion

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.742861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.967173Z digest=sha256:63594f667e5af2190c468ad7a5aee86cca0badc49f34796f21573e368c9e6026

Observation 94eb28c8-6cf5-4d2f-b66a-9e4fe7874ada · outbound

This paper cites Lips don’t lie: A generalisable and robust approach to face forgery detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Lips don’t lie: A generalisable and robust approach to face forgery detection

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.730819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.970864Z digest=sha256:9008631d14cba626bcf3f91a10b34b997876b6861cd336b1207650f1359bae14

Observation 61dc8e57-70cd-4bca-b411-02c32fc4731d · outbound

This paper cites Leveraging real talking faces via self- supervision for robust forgery detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Leveraging real talking faces via self- supervision for robust forgery detection

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.719091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.975148Z digest=sha256:cbd525f26f06e96dc53ba35a1c17c30c983d8174df83926956249538fdd0e1dd

Observation 85c606d9-92f6-44ac-964a-1ca897a6c2f6 · outbound

This paper cites AVTENet: A Human-Cognition-Inspired Audio-Visual Transformer-Based Ensemble Network for Video Deepfake Detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning AVTENet: A Human-Cognition-Inspired Audio-Visual Transformer-Based Ensemble Network for Video Deepfake Detection

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-12T05:44:30.367926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.979061Z digest=sha256:3964d573d1271c116657ac439212cece262a839c127ed2abf224492cd0c5de93

Observation 193d8633-31df-4c3c-a4a6-51e11fdb8b32 · outbound

This paper cites Implicit identity driven deepfake face swapping detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Implicit identity driven deepfake face swapping detection

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.708387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.983103Z digest=sha256:674c72f8ee568f14052d379240b53eae41069aa5c19aa704b1726d4f35bf41f8

Observation ac5c7317-76a3-4300-8343-73765124b811 · outbound

This paper cites Weiss, Quan Wang, Jonathan Shen, Fei Ren, Zhifeng Chen, Patrick Nguyen, Ruoming Pang, Ig- nacio L ´opez-Moreno, and Yonghui Wu.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Weiss, Quan Wang, Jonathan Shen, Fei Ren, Zhifeng Chen, Patrick Nguyen, Ruoming Pang, Ig- nacio L ´opez-Moreno, and Yonghui Wu

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.697002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.986480Z digest=sha256:2dedcc4e7af958c8308cbd538baa5c95ade0b32e87f6916a4bd0ff8756cbf2a7

Observation 3b7b43e2-c716-49d5-b0ac-c0690dd3e926 · outbound

This paper cites an unresolved cited work.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-12T05:44:30.685497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.989809Z digest=sha256:4eae5a65e0f7acaf3c7c790599aaa767b238b97e11ddfb7ed6a30ffee66fa6cf

Observation 9ee132c4-b704-4c2e-95d3-8b8d019f0cdc · outbound

This paper cites Conditional variational autoencoder with adversarial learning for end-to- end text-to-speech.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Conditional variational autoencoder with adversarial learning for end-to- end text-to-speech

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.674615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:29.993621Z digest=sha256:ab2d26cf6fc261f639246dee02574948b72d7569ee0f22524ec4a16adfaa05ff

Observation 6f07a43b-8e35-4e9d-b9c3-d0c55d9f97cd · outbound

This paper cites DeepFakes: a New Threat to Face Recognition? Assessment and Detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning DeepFakes: a New Threat to Face Recognition? Assessment and Detection

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:29.997426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:29.997426Z digest=sha256:956a6dc5ec9091439de861dec90126275851072e509d6eb4810cdeb4565e68c5

Observation 07a55832-9392-4694-8f79-7ab5cd4afac2 · outbound

This paper cites Fast face-swap using convolutional neural networks.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Fast face-swap using convolutional neural networks

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.662715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.001658Z digest=sha256:aceab9ca7f5adedd8ac73ffd017ea613d61dcd2d6b585430ca919d90fadcb8e6

Observation 2f3969ed-538d-4505-85f6-cf0859757533 · outbound

This paper cites DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.005614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.005614Z digest=sha256:d80f21479433f3cb5dda44dfec17553d320aec689651ddadb78b06abc2ce7723

Observation d83fdd33-7043-40ff-bab6-37a75629e3b0 · outbound

This paper cites KoDF: A large-scale Korean deepfake detection dataset.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning KoDF: A large-scale Korean deepfake detection dataset

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.650538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.010154Z digest=sha256:f3e609916c87eded6976c394399edf6ee688fe5c9a2c2ae8c98f13e1a4937534

Observation 6f7b9195-4c79-4c78-a07e-383e73b100e3 · outbound

This paper cites Zero-Shot Fake Video Detection by Audio-Visual Consistency.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Zero-Shot Fake Video Detection by Audio-Visual Consistency

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.015336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.015336Z digest=sha256:40fca8348cc9eab4c2dbe9cf3434be212473231401a5b618f97818bf94ffa571

Observation d103036f-1b70-4766-8d34-16be070831ac · outbound

This paper cites SpeechForensics: Audio-visual speech representation learn- ing for face forgery detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning SpeechForensics: Audio-visual speech representation learn- ing for face forgery detection

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.638655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.019867Z digest=sha256:463703debe2a0caa02465dbd67e9e9287a11002ed11ddbbef7972d1b5f337cac

Observation 864ae0bf-3730-4f9a-a9dc-ef63542c0bf0 · outbound

This paper cites Lips are lying: Spotting the temporal inconsistency between audio and visual in lip- syncing deepfakes.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Lips are lying: Spotting the temporal inconsistency between audio and visual in lip- syncing deepfakes

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.626444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.024385Z digest=sha256:49f4020d27c4c7a51590c4f868f38741a3c251bcc51ebdd1dd4f27541168f32d

Observation 68f3aa07-355d-4ca7-b494-0495de107945 · outbound

This paper cites When Synthetic Traces Hide Real Content: Analysis of Stable Diffusion Image Laundering.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning When Synthetic Traces Hide Real Content: Analysis of Stable Diffusion Image Laundering

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-12T05:44:30.318622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.028583Z digest=sha256:01fa3cd523f10134b7873dbff3f35ffccc9c92855865da1d19556ef9ed839504

Observation 2953a0a3-8d77-4e60-acc9-6b1b3fe83f5c · outbound

This paper cites TGIF: Text-Guided Inpainting Forgery Dataset.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning TGIF: Text-Guided Inpainting Forgery Dataset

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.032962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.032962Z digest=sha256:a0a541180c910f4dd6c4c310d9a11306c44f25c7a296b4b8fbeddded1fe6475a

Observation 1fa776b0-864a-4663-a75b-a74cbc5bda9e · outbound

This paper cites Do GANs leave artificial fingerprints? In MIPR, pages 506–511, 2019.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Do GANs leave artificial fingerprints? In MIPR, pages 506–511, 2019

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.614526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.037292Z digest=sha256:e2f4c8758eb23181e9793cd0d18940259aea216d8f435a7ce78d5208eeb0316b

Observation 5ff8c039-2b88-415d-81d0-c3259f4fd4a5 · outbound

This paper cites Speech is Silver, Silence is Golden: What do ASVspoof-trained Models Really Learn?.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Speech is Silver, Silence is Golden: What do ASVspoof-trained Models Really Learn?

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.041044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.041044Z digest=sha256:9789aa9f9f0775949af94e058a09b546abe55d953030f71cc08fea8036102ea4

Observation aeb44556-7661-46de-a740-a292c5702c81 · outbound

This paper cites M ¨uller, Piotr Kawa, Wei Herng Choong, Edres- son Casanova, Eren G ¨olge, Thorsten M ¨uller, Piotr Syga, Philip Sperl, and Konstantin B¨ottinger.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning M ¨uller, Piotr Kawa, Wei Herng Choong, Edres- son Casanova, Eren G ¨olge, Thorsten M ¨uller, Piotr Syga, Philip Sperl, and Konstantin B¨ottinger

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.602844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.045120Z digest=sha256:90ff7a2ba82b9b7d972552451ab52b4292e52efbef25b1493899f13894591020

Observation 969af691-c1d9-44d7-a802-6dfc7f92a78c · outbound

This paper cites FSGAN: Subject agnostic face swapping and reenactment.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning FSGAN: Subject agnostic face swapping and reenactment

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.591140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.048924Z digest=sha256:588cb04092b9a80cf714ee5ce46e9262e15c6ed2979e329dba09b9821d6535b7

Observation ea3a47d5-9301-44c7-a21c-da8c69fd17b4 · outbound

This paper cites Towards uni- versal fake image detectors that generalize across generative models.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Towards uni- versal fake image detectors that generalize across generative models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.578222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.052119Z digest=sha256:9e8407afab86fd0c7d239055a0b74530cc0cf283869bb5f82453d0206de5bbc0

Observation 5ae1e3f7-bfde-4e9b-90d6-bd3990cf174d · outbound

This paper cites A VFF: Audio-visual feature fusion for video deepfake detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning A VFF: Audio-visual feature fusion for video deepfake detection

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.566161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.055580Z digest=sha256:8034c3e1d60b8f5eec74d5cfde44a95fd29f387e923108439ead249c819e8cd8

Observation 9b46e0a9-8ec3-41f8-83bc-25fd6f4fbfd0 · outbound

This paper cites Towards generalisable and cali- brated audio deepfake detection with self-supervised repre- sentations.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Towards generalisable and cali- brated audio deepfake detection with self-supervised repre- sentations

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.556304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.059228Z digest=sha256:919e916e205e1129484eed9be4f0b2ade750928594330a9a8202cd3b25b36d05

Observation e781906e-45f3-4e1c-a5e2-6ad8ba2870cb · outbound

This paper cites Training-free deepfake voice recognition by leveraging large-scale pre-trained models.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Training-free deepfake voice recognition by leveraging large-scale pre-trained models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.546248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.062350Z digest=sha256:f355d489a8f34c2f85498aa780df07d0c9f82b5fea79e1897920384915383510

Observation b548f3d9-54c6-406e-862c-fca0b7e40564 · outbound

This paper cites Nambood- iri, and C.V.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Nambood- iri, and C.V

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.535050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.065597Z digest=sha256:95ba99214c492ebc918c14cfdfc5606c64588e8ad14ae6adc6e017a18bf6b0ef

Observation a5b3b9be-cb06-4225-af1c-6ac3e4675cc4 · outbound

This paper cites Aligned Datasets Improve Detection of Latent Diffusion-Generated Images.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Aligned Datasets Improve Detection of Latent Diffusion-Generated Images

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.069372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.069372Z digest=sha256:8b893ad19df3a9d5f95487c4afb4fa3606133e74ca8434dc1c6b59618711fa7a

Observation 80a80d7d-94d0-4ed7-a609-0174cd2ae0c4 · outbound

This paper cites Detecting Deepfakes Without Seeing Any.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Detecting Deepfakes Without Seeing Any

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.073469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.073469Z digest=sha256:cec0f196a0f3236afe7ac9921b1269223bcd8464c4990496d7a172d630355e56

Observation ae187751-f545-462d-ba06-4a0eaca1bf26 · outbound

This paper cites AEROB- LADE: Training-free detection of latent diffusion images us- ing autoencoder reconstruction error.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning AEROB- LADE: Training-free detection of latent diffusion images us- ing autoencoder reconstruction error

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.523986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.077820Z digest=sha256:acba854f28d2892bc99701490f74dc979a5a04fe8a8fd3ae99c821e6a5fb1073

Observation d02d19e7-b2af-40a2-8a8e-11703e729397 · outbound

This paper cites FaceForen- sics++: Learning to detect manipulated facial images.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning FaceForen- sics++: Learning to detect manipulated facial images

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.512076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.081672Z digest=sha256:f16d3855f35a22334a35e9714fc5349f91b84a9250d3313e0c1046f07038c3a9

Observation cf3cfcc4-06bd-4920-9546-450aed995371 · outbound

This paper cites Hosler, Paolo Bestagini, Matthew C.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Hosler, Paolo Bestagini, Matthew C

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.500830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.085554Z digest=sha256:58e1118dfe73dce700129076e6d635a29a2fe1dec92f73261a2ec81ee8c7877f

Observation 15fcc1e4-fd8d-4958-a636-fb21c09db0b8 · outbound

This paper cites A V-Lip-Sync+: Lever- aging A V-HuBERT to exploit multimodal inconsistency for video deepfake detection.CoRR, abs/2311.02733, 2023.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning A V-Lip-Sync+: Lever- aging A V-HuBERT to exploit multimodal inconsistency for video deepfake detection.CoRR, abs/2311.02733, 2023

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.089207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.089207Z digest=sha256:3a6a75fa4941dad16ebe2e11a3677fd94857dab96f12cb9d84b96ebb91a59a0a

Observation 6134641f-d05b-46e0-a138-e99f04eec00d · outbound

This paper cites Learning audio-visual speech representation by masked multimodal cluster prediction.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Learning audio-visual speech representation by masked multimodal cluster prediction

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.489852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.093019Z digest=sha256:cd4e176a41a109a7064b0417b4fde0bc4c6bea3dbfc69c79ecc806bb4d464c5d

Observation 1c02577e-a80d-48f7-8c25-d8e8650742b4 · outbound

This paper cites Detecting deep- fakes with self-blended images.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Detecting deep- fakes with self-blended images

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.477035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.096651Z digest=sha256:1c12d5908412a4a1c280abaaa983f693d802be0594f2ae0f9b4377482ad74053

Observation 1b0066b4-d6ec-4b9f-a7d7-67e4c79f2bb8 · outbound

This paper cites DeCLIP: Decoding CLIP representations for deepfake localization.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning DeCLIP: Decoding CLIP representations for deepfake localization

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.100607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.100607Z digest=sha256:064872e7f00f37bf2e7880febae292f8cba4e17a80ceac4c038a79e978d9436b

Observation 92b24c84-6659-4cbd-88c0-6b6ce0746084 · outbound

This paper cites Lip reading sentences in the wild.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Lip reading sentences in the wild

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.465517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.105280Z digest=sha256:3b85924d974c22f6bf1273e3024ffc4d8a219afaf2779d7950023f47df8c5c2d

Observation 870152d8-6ac1-4255-b6bf-b0ff5c6fe1bc · outbound

This paper cites End-to-end anti-spoofing with RawNet2.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning End-to-end anti-spoofing with RawNet2

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.454182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.109954Z digest=sha256:5404962f7d09fe83756a966433e6258ede39ce23fcb50f742823856a65123bd5

Observation d415a12d-aedf-435d-92bd-0b82b23a53da · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Representation Learning with Contrastive Predictive Coding

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.114598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.114598Z digest=sha256:c445d6e4d980b2b4bc4e5254e88396d796a029f4e1daf5699288cd718c595014

Observation 54636c10-493f-46fe-8399-0cff0a423dc3 · outbound

This paper cites Tan, and Haizhou Li.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Tan, and Haizhou Li

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.442071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.119711Z digest=sha256:ab49430a67cf5daaa6def629be77892ca0aace3377ad3f9e7ae8e76e817cced6

Observation 1747dde7-7503-4e95-b232-e51c5471ec7d · outbound

This paper cites an unresolved cited work.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-12T05:44:30.429625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.124304Z digest=sha256:8e2f4ee12686f886f0350da3a374dc8ca5e39e4246111eb18ab331325dc470f9

Observation b02d6ff3-e91c-46b0-bbe0-df84d750d801 · outbound

This paper cites DF40: Toward Next-Generation Deepfake Detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning DF40: Toward Next-Generation Deepfake Detection

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-12T05:44:30.128411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:44:30.128411Z digest=sha256:dd9a4daa0799d7d54da7f6ccfc536190b9b01cc0e7bb7d439478cc23ef7697aa

Observation 3fd4eb5b-b339-40f2-9821-dadb6a384bd9 · outbound

This paper cites 10 A V oiD-DF: Audio-visual joint learning for detecting deep- fake.IEEE Trans.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning 10 A V oiD-DF: Audio-visual joint learning for detecting deep- fake.IEEE Trans

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.417412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.132785Z digest=sha256:a12dfd1c51ee33318541ac4b40a3e79035eeec225d3672bfb6f4262114c6b0be

Observation 01786df4-5813-4367-b56e-a9c5706684c4 · outbound

This paper cites Attributing fake images to GANs: Learning and analyzing GAN fingerprints.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Attributing fake images to GANs: Learning and analyzing GAN fingerprints

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.405087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.136494Z digest=sha256:87e64c6eba7a342d4cd3e1ce6e56a683876c54131ff1a55205c66e5d8add5be1

Observation 56102646-6931-4131-878f-42212b6c5254 · outbound

This paper cites Exploring temporal coherence for more general video face forgery detection.

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning Exploring temporal coherence for more general video face forgery detection

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:44:30.392580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:44:30.140670Z digest=sha256:7870fc1c21ca2cf69d96d0970a8db4b86a65fbcd4e4cff50942d91ebea6bbf77

Pith citing papers

Observation cf789b32-7ebb-4b20-a463-9edcb03fb8f5 · inbound

Unmasking Synthetic Realities in Generative AI: A Comprehensive Review of Adversarially Robust Deepfake Detection Systems cites this paper.

Unmasking Synthetic Realities in Generative AI: A Comprehensive Review of Adversarially Robust Deepfake Detection Systems Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning

Reference 200

Resolution
verified exact
local_arxiv, observed 2026-08-06T14:34:12.131864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T14:34:11.253445Z digest=sha256:349155ffd7ff7f42c3b54cfbf7d9a54c7069efd65071588babac8672de484498