Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T19:57:43.285908Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 93 of 93 outbound references and 11 inbound Pith citation observations for arXiv:2411.10193.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T19:57:43.285908Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T05:44:30.005614Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T14:07:02.288478Z
93 of 93 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ba01d280-6c6d-4288-8824-d622934a0dc5 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Mesonet: a compact facial video forgery detection network
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1151922-57ea-4923-b5e7-fa222dbbfae5 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Deep audio-visual speech recognition
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 828b5171-0cd1-4e72-aec0-54453fecb232 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization LRS3-TED: a large-scale dataset for visual speech recognition
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c509f8c8-f2bd-4b6f-85c7-63c96cfbf0fd · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Transitions in neural oscillations reflect pre- diction errors generated in audiovisual speech.Na- ture neuroscience, 14(6):797–801, 2011
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec047c76-d001-42ee-b4db-eb30974fa3e7 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Hear Me Out: Fusional Approaches for Audio Augmented Temporal Action Localization
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccd8201c-efa2-444d-bfad-649975a7270d · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Phoneme-to- viseme mappings: the good, the bad, and the ugly
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1477da99-97c7-4dfb-9fb4-a43a6fd92262 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Lost in trans- lation: Lip-sync deepfake detection from audio- video mismatch
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1221b4f-f42b-4d64-b23b-02301d3d5f55 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Av-deepfake1m: A large-scale llm-driven audio-visual deepfake dataset
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e696f7b7-445d-47e5-9cb6-678e484b7745 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Glitch in the matrix: A large scale benchmark for content driven audio–visual forgery detection and localization
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 821fc75e-e240-4207-962b-3a2555be233c · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Do you really mean that? con- tent driven audio-visual deepfake dataset and mul- timodal method for temporal forgery localiza- tion
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d4cbacb-e155-40f2-ad9e-501ea836ce7b · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization End- to-end reconstruction-classification learning for face forgery detection
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dd11b83-9da6-41c9-b8ea-f61fc5d1582d · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Chandrasekaran and A
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea3579cf-100f-4f19-b476-b0e23e09caba · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization What you see depends on what you hear: Temporal averaging and crossmodal in- tegration
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fb2ee97-c599-4f31-bf5d-81442ed91bad · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Voice-face ho- mogeneity tells deepfake
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c18a556a-c7e7-4820-a2b9-2aa1177a9adb · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Not made for each other-audio-visual dissonance-based deepfake de- tection and localization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f699a059-1b0d-4958-929d-2cb7bd7f8bdc · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Vox- Celeb2: Deep speaker recognition
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a1af069b-7321-490b-95e3-84634ff16a2c · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Lip reading in the wild
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6c21c2bd-d294-4363-8173-59661ba0919c · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Combin- ing efficientnet and vision transformers for video deepfake detection
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 69f40b8d-c460-40c3-ada1-7031d3d862b7 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Audio-visual person-of-interest deepfake detection
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0f0cdda5-73a8-4213-aa38-9ee2ff227120 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Imagenet: A large-scale hi- erarchical image database
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 00b02011-e2ab-4a8f-9e73-b72be070ad96 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization The DeepFake Detection Challenge (DFDC) Dataset
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22db688a-6f2d-411c-97be-e116aad56a93 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Taming transformers for high-resolution im- age synthesis
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation baed580b-9dc3-48d1-a5c5-39454dc8f02a · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Self-supervised video forensics by audio-visual anomaly detection
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9c06d2ac-5a8e-43b1-9045-1e8eb5568f08 · outbound
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation dab7ee5a-300a-458e-88a4-3dee2f931577 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Rational decisions
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1d1e77ac-7331-4cb8-96bf-a9322fe940b6 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Bootstrap your own latent-a new approach to self-supervised learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e0f1f88b-835e-48b7-92fa-4c93a9e0f5d9 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Deepfake video detection using audio- visual consistency
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5ec760ee-13e0-4616-b656-f919a33436ee · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Delving into the local: Dynamic inconsistency learning for deep- fake video detection
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 07767f8c-4def-4733-b97b-b8f0562a3030 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Leveraging real talking faces via self-supervision for robust forgery detec- tion
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3caaede2-915e-46f2-a7fb-292e27946345 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Lips don’t lie: A generalisable and robust approach to face forgery detection
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 23ead194-4691-43ea-8cea-a87b0e024c27 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization De- tection of fake images via the ensemble of deep representations from multi color spaces
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 123db3e9-4e4d-48a2-a0aa-32c47a58ebb7 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Avfakenet: A unified end-to-end dense swin transformer deep learning model for audio- visual deepfakes detection
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 55b1aaac-a140-4697-b961-17b7222667d7 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Deeperforensics-1.0: A large-scale dataset for real-world face forgery de- tection
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 72e2ad54-abcd-42d0-84d3-daf329e5ea0f · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Con- textual cross-modal attention for audio-visual deepfake detection and localization
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 34f77d1b-062f-4df0-a60f-8c76ce62a0c8 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization The Kinetics Human Action Video Dataset
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33a3881c-034c-4b63-a1a7-85ad53ce3fbd · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Fakeavceleb: A novel audio-video multimodal deepfake dataset
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 128aad35-bf40-4b2f-8a22-d537538d6ff6 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Deep- fakes: a new threat to face recognition? assess- ment and detection
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9446f9ec-fdb8-4fac-bc67-01faaf4750f7 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Kodf: A large-scale korean deepfake detection dataset
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9100865c-9700-49b6-bad7-8fe3530a42c8 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Layer normalization
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 30998388-2085-4b43-bf51-7e0ca0e0bb0a · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Face x-ray for more general face forgery detection
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2795419c-c1ad-4f57-9679-787517bac0a5 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Spatio-temporal catcher: A self-supervised transformer for deepfake video detection
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation bc8a5313-d2de-4a7f-98a3-7455c11e82ce · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Zero-shot fake video de- tection by audio-visual consistency
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 76094af0-f5b3-41ae-b72e-c166b7fd1d7b · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Celeb-df: A large-scale challenging dataset for deepfake forensics
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d51e9110-c6bd-495a-b85b-027388a9ad0b · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Bmn: Boundary-matching network for temporal action proposal generation
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a17a86e6-b252-433a-9ef0-df14ea693e57 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Focal loss for dense object detection
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7a519223-7193-4e09-97c5-753c7f2c48fe · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Visual speech recognition for multiple languages in the wild
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b974cab0-f249-4722-9cf9-bcbcf2d2ac50 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Deepfakes generation and detection: State- of-the-art, open challenges, countermeasures, and way forward
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a9bb3df7-8242-421b-afab-390bb0433827 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Emotions don’t lie: An audio-visual deepfake de- tection method using affective cues
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e1baa66b-138a-4536-b6a5-ea3083ca0c7a · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Df-platter: Multi-face heterogeneous deepfake dataset
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation dfc75e8e-dbc3-4116-8f63-43df994f805f · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization A neural basis for interindividual differences in the mcgurk effect, a multisensory speech illusion.Neu- roimage, 59(1):781–787, 2012
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b1696474-fa3f-42fe-877b-1097f2b4397b · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Activity Graph Transformer for Temporal Action Localization
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf3fd579-3604-42dc-9051-37c14cf3a53a · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Frade: Forgery-aware audio- distilled multimodal learning for deepfake detec- tion
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 15b9793b-a54f-47f6-857b-5463f88c8646 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Seeing what you hear: Cross- modal illusions and perception
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6f2caed7-7260-4e84-9fd9-dbe25f435ce2 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Avff: Audio-visual feature fusion for video deepfake de- tection
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f3ecac87-f8dd-4142-ba71-d8a010b2c416 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Deep- fake generation and detection: A benchmark and survey
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba1211c5-7536-46e7-a7e9-9dea4da2337e · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Powerset multi-class cross entropy loss for neural speaker diarization
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80cb1353-d837-4424-95ee-54c02dfeea93 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Thinking in frequency: Face forgery detection by mining frequency-aware clues
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f939d77a-9b46-493b-a82a-9045a855e781 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Multimodaltrace: Deepfake detection us- ing audiovisual representation learning
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 329e15fd-4d41-4a0e-86fe-881e18cf9f75 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Detecting Deepfakes Without Seeing Any
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91ddaaa7-7b49-44a8-8f89-047904d854d2 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Faceforensics++: Learning to detect manipulated facial images
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f6185d66-bcc0-4bbc-88dd-8a96f22f67ab · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Lip sync matters: A novel multimodal forgery detector
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9d1cf413-d3d9-4cb3-b389-8ca124c88ef2 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Visual illusion induced by sound
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 303df92f-315a-472c-8734-3c2767c40494 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Learning audio-visual speech representation by masked multimodal clus- ter prediction
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e037fbe6-c3dc-4e02-b3a3-0e3912b881d6 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Tridet: Temporal ac- tion detection with relative boundary modeling
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e8caeb53-edcf-4e59-8dee-008f1461e8c7 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization De- tecting deepfakes with self-blended images
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 75c2d7f8-85ad-4b4a-920c-e4dde18c964a · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Locate and ver- ify: A two-stream network for improved deepfake detection
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 874ad5b2-a553-4c21-85a1-bea3412b0123 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization The contribution of visual information to the perception of speech in noise with and without informative temporal fine structure
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5fe41125-7665-4625-939e-331c12026a91 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Visual contribution to speech intelligibility in noise
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 143b8ad7-dc2a-42f9-8d1e-2beda29d4862 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Learning on gradients: Generalized artifacts representation for gan-generated images detection
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0ed2dd83-ab49-4660-aec6-7281ee692def · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Attention is all you need
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86cfc1e0-69a2-4970-803c-3a463ef1ed7b · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Videomae v2: Scaling video masked autoencoders with dual masking
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 81309a39-bd8b-44ff-b3e4-0ae69a8adf35 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Building robust video-level deepfake detection via audio- visual local-global interactions
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6cb7def3-e3dd-465a-9962-4b2a978190b2 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Audio-visual deep- fake detection using articulatory representation learning
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c3b3b261-43bb-46e4-8398-ff632abee857 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization What you see is what you hear: sounds alter the contents of visual per- ception
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation acd175f2-4e6a-4f3a-bb1a-e9e2b52ac72a · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Avoid-df: Audio-visual joint learning for detecting deepfake
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6097e91f-0892-40c8-bf82-6e61e843dd01 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Expos- ing deep fakes using inconsistent head poses
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 91778712-caf1-48d4-9c23-be4d264ebf1b · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Pvass-mdd: predictive visual-audio alignment self-supervision for multimodal deepfake detection
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4c63d14e-9419-4b77-bf3b-2e33525c25e8 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Ac- tionformer: Localizing moments of actions with transformers
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation dca4b936-a2e0-4a10-b38c-09c9f991b6c0 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Video- llama: An instruction-tuned audio-visual lan- guage model for video understanding
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 140244de-1acb-40f0-af06-5e4ebbacff0b · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Um- maformer: A universal multimodal-adaptive transformer framework for temporal forgery local- ization
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1b9a6618-3e17-480e-940e-b685275af55b · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Joint audio-visual attention with contrastive learning for more general deepfake detection
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 556ae367-c25b-4bba-a206-dd7c2c67cf69 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Exploring temporal co- herence for more general video face forgery de- tection
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 186c60e4-8edd-49de-a06f-0e8b077fa5e3 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Distance-iou loss: Faster and better learning for bounding box regression
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2ccd7527-90c6-4436-8597-75225c08094f · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Joint audio- visual deepfake detection
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 970e70eb-b674-44f6-ab45-932773b56d88 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Wilddeepfake: A chal- lenging real-world dataset for deepfake detection
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1d2e6818-3e17-4b91-945f-d50179b9694f · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Cross- modality and within-modality regularization for audio-visual deepfake detection
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 40f4ca2d-d003-428c-8292-3bb3bfa32533 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Unresolved cited work
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a51072a7-9b24-4a70-b4d8-53248fabc08a · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Unresolved cited work
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 310d7c0c-8b90-4791-a1d4-b8dfcf45e59c · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Since we apply the manual screening process on synthesized videos, the final video count is more than 20,000
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c71b1ba0-f73b-4698-b244-9d98710a070e · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Unresolved cited work
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 222fe795-9d65-45a4-85ec-9a3f8a14d386 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization In addition, DiMoDif outperforms A VFF un- der all perturbation scenarios and at all intensity levels
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9a41d238-2c84-4614-b317-998cf3c071cf · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Table 12 presents the corresponding results
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation bcb532c3-57a7-4c9a-a5e0-490816961776 · outbound
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization Unresolved cited work
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2f3969ed-538d-4505-85f6-cf0859757533 · inbound
Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2840c89-2355-444f-bdac-136f94fd0a5d · inbound
Survey on AI-Generated Media Detection: From Non-MLLM to MLLM DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization
Reference 163
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96394362-d0dc-478a-bf7d-d7b9b080d66c · inbound
DeepFake Doctor: Diagnosing and Treating Audio-Video Fake Detection DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a89fe335-7651-4d5e-a693-ba01739559a8 · inbound
Context-aware TFL: A Universal Context-aware Contrastive Learning Framework for Temporal Forgery Localization DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb3528dc-8b71-41c0-b153-3295687cd0cb · inbound
Unmasking Synthetic Realities in Generative AI: A Comprehensive Review of Adversarially Robust Deepfake Detection Systems DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization
Reference 180
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c29502d-5ae7-445d-8642-809ed8fb6505 · inbound
Generalizing Video DeepFake Detection by Self-generated Audio-Visual Pseudo-Fakes DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b034ae34-662b-41a5-ac29-7e8303e9d928 · inbound
Inconsistency-aware Multimodal Schr\"odinger Bridge for Deepfake Localization DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 739272c8-217b-465e-8123-fb9e5cf0c772 · inbound
MG-RWKV: Multi-Grained Context-Aware RWKV for Temporal Forgery Localization DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 37442a88-3105-480f-9aa7-a64e3f0ae0ed · inbound
EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed579bdb-c456-4701-acfd-a3a3eded52d6 · inbound
UniSkip-Mamba: A Frequency-Aware State Space Model for Audio-Visual Temporal Forgery Localization DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd5aa6ed-9377-4bee-8c26-4f1972756867 · inbound
UniSkip-Mamba: A Frequency-Aware State Space Model for Audio-Visual Temporal Forgery Localization DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.