Pith. sign in

Paper Citation Record · LEDGER

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation

As of 18 August 2026, this Paper Citation Record lists 74 of 74 outbound references and 1 inbound Pith citation observation for arXiv:2505.15233.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15233 v1

Coverage vector

measured 74 of 74 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:27:00.996471Z

measured 75 of 75 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-20T15:08:25.309094Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T15:13:25.040118Z

Reference resolution

74 of 74 outbound references displayed

  • verified exact2
  • verified fuzzy53
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bb4cd9f1-cff9-435c-89b3-9c71e68d3766 · outbound

This paper cites Mesonet: a compact facial video forgery detection network.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Mesonet: a compact facial video forgery detection network

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:04.335259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:54.074983Z digest=sha256:0ba75d634af0969348f88aff4bdae8bad725f19446bd9d3b8b9a6968cc0f4088

Observation c58e74d7-cdc7-43a4-8e66-492732a5fe8c · outbound

This paper cites A review of modern audio deepfake detection methods: challenges and future directions.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation A review of modern audio deepfake detection methods: challenges and future directions

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:04.316636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:54.150530Z digest=sha256:ac5f1c450a9b9fa1598f64fe28d97a08370ffcee5ea9e2388f7a08a5a20698a2

Observation 1f187ee3-29bd-4605-9309-632285bf018e · outbound

This paper cites Do you really mean that? content driven audio-visual deepfake dataset and multimodal method for temporal forgery localization.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Do you really mean that? content driven audio-visual deepfake dataset and multimodal method for temporal forgery localization

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:04.292173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:54.322384Z digest=sha256:4435682abc569a0703fd74998e80d7a747fd040ac26a0a61dbc0ecc2b5c73edc

Observation fe38a54f-d59a-4ce4-8404-a8aa568b0c69 · outbound

This paper cites Quo vadis, action recognition? a new model and the kinetics dataset.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Quo vadis, action recognition? a new model and the kinetics dataset

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:54.433137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:54.433137Z digest=sha256:59e551a14250c74950eee3a9fadb61de323eca6aa54294dd7dc19b22260e04ef

Observation d6ce7981-1a59-44ce-bea4-20bcf51e17cb · outbound

This paper cites an unresolved cited work.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:27:04.263567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:54.539628Z digest=sha256:43eabb49543f80c4a11a2ed080c143e4f3df1aee0d195619d155fb7ca10f59ab

Observation 933e19fd-3ba1-4275-9383-5013726f029d · outbound

This paper cites Self-supervised learning of adversarial example: Towards good generalizations for deepfake detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Self-supervised learning of adversarial example: Towards good generalizations for deepfake detection

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:04.246938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:54.649874Z digest=sha256:0ac497b90e387461f9fdee871292b9b588b2b92beeb220f773b3aee6fc27d5b9

Observation 32a37ffa-7e7c-4397-b206-ab43e5cbd069 · outbound

This paper cites A Simple Framework for Contrastive Learning of Visual Representations.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation A Simple Framework for Contrastive Learning of Visual Representations

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:54.756360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:54.756360Z digest=sha256:b94f8cf1b04dfedb513a36eea2c1ce8ab83c307403e355edadfa14cfa50f0af2

Observation 50ac14fd-fba5-4306-a83f-f50b062b16be · outbound

This paper cites Exploring simple siamese representation learning.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Exploring simple siamese representation learning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:04.230249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:54.856339Z digest=sha256:bc5bb4ff9a476d6f07f9cc33346cae2f3b4acc9bba8e1b918a2028e43e0784f8

Observation a8f0832a-cbc0-436c-8860-320875445262 · outbound

This paper cites Sophia Koepke, Ying Shan, and Zeynep Akata.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Sophia Koepke, Ying Shan, and Zeynep Akata

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:04.213684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:54.977105Z digest=sha256:888b9492f96ba29e0705ef7f9f9e31aec2ee0b64c6a82aa98ddd8fc8ed8c7f79

Observation 64487f5f-8e60-4cc6-bc87-df0abcfee9a4 · outbound

This paper cites V oice-face homogeneity tells deepfake.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation V oice-face homogeneity tells deepfake

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:04.197708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:55.048303Z digest=sha256:adadcac00bcdd757d89f19180e66d6c9addc2a77818e1bc9ab4f4dcb5e5eff5c

Observation 5769a50b-9188-4a15-8163-a547698489f2 · outbound

This paper cites Can We Leave Deepfake Data Behind in Training Deepfake Detector?.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Can We Leave Deepfake Data Behind in Training Deepfake Detector?

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:55.184286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:55.184286Z digest=sha256:ffecd1eeae44cc1ced0d2a73dbe4fae48a991e6fc58ffdcff428e1e8652afd33

Observation 129e1880-4817-4737-ae70-a38b7652923e · outbound

This paper cites Xception: Deep learning with depthwise separable convolutions.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Xception: Deep learning with depthwise separable convolutions

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:04.180876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:55.302595Z digest=sha256:1c7487af92017a8b764e481b93954856412bc95ece437492410dce8dfd92dc9e

Observation 8284eb72-ad9e-472d-a7aa-63f5969c98d0 · outbound

This paper cites Not made for each other- audio-visual dissonance-based deepfake detection and localization.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Not made for each other- audio-visual dissonance-based deepfake detection and localization

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:04.155779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:55.405611Z digest=sha256:48ad992806a9430286b64e621cba6b0aab172b18f3ec0b39c4a68e0b877a6a95

Observation 56d9c05b-5efd-4049-9806-edda9d0c0262 · outbound

This paper cites Audio-visual person-of-interest deepfake detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Audio-visual person-of-interest deepfake detection

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:04.138476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:55.485300Z digest=sha256:eccaf7a3ddb07cd44eeded51976fa570a4877a3d64adc90fc06174851e5e0cc1

Observation 5eb7d261-3663-4989-872a-9a2da6120e54 · outbound

This paper cites Dong, Jin Wang, Renhe Ji, Jiajun Liang, Haoqiang Fan, and Zheng Ge.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Dong, Jin Wang, Renhe Ji, Jiajun Liang, Haoqiang Fan, and Zheng Ge

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:04.119110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:55.597172Z digest=sha256:c3be8240caa62a214c8fde1e584af2eeb70e52f4dc4567eb0ca65874689ef9ae

Observation 5b2b32de-9a52-488c-848a-68dffb596ebb · outbound

This paper cites www.github.com/MarekKowalski/FaceSwap Accessed 2021-04-24.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation www.github.com/MarekKowalski/FaceSwap Accessed 2021-04-24

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:04.099701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:55.697949Z digest=sha256:e0bd6b95fe40eb4f9038237674ac3ff402778501f5911dff3d13c94be8af9c6e

Observation 85c18813-1f05-4efb-ba7f-1188277e5ec9 · outbound

This paper cites Self-supervised video forensics by audio-visual anomaly detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Self-supervised video forensics by audio-visual anomaly detection

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:04.079789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:55.778876Z digest=sha256:f5c45a94ee3cb631214ac1ebbab55497fbfcb994e1c6b3906fd9f31f0e28497a

Observation 2a551dd0-faf3-411d-88b2-5fbe681f7e98 · outbound

This paper cites Joint 3D Face Reconstruction and Dense Alignment with Position Map Regression Network.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Joint 3D Face Reconstruction and Dense Alignment with Position Map Regression Network

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:27:01.797723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:55.856641Z digest=sha256:0b98a80eec4d06574bc94556f2d6f079e74c6199157b1b3dbc35543fa9ea7c2e

Observation 910b8953-1a46-46c7-80d4-6a8afc3b5515 · outbound

This paper cites Imagebind one embedding space to bind them all.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Imagebind one embedding space to bind them all

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:04.058641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:55.927230Z digest=sha256:dd31e7d89ea29e0dd2aeb926df7c6f3d1f5e78d3441b6fd81d1ddc8645fd465a

Observation 1d42e731-48c8-4c7c-adb2-ceddac4857b8 · outbound

This paper cites Leveraging real talking faces via self-supervision for robust forgery detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Leveraging real talking faces via self-supervision for robust forgery detection

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:04.039846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:56.007650Z digest=sha256:68d5378ad57643e5c5e9cdd59baba4702c398deeef6b75ba76b874dfd01ab135

Observation 242ef530-74f6-4b75-955b-c98b0fde5954 · outbound

This paper cites Lips don’t lie: A generalisable and robust approach to face forgery detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Lips don’t lie: A generalisable and robust approach to face forgery detection

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:04.016350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:56.106380Z digest=sha256:2996356110ee8abaa6608498434e2007dee2c2cb39086e4582a26812c8fc6b0d

Observation 459ebd5f-63b3-4792-9178-d88286def08e · outbound

This paper cites Masked autoencoders are scalable vision learners.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Masked autoencoders are scalable vision learners

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.992759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:56.188215Z digest=sha256:125b047e008898fdff9e1012a980f3d2846a85bf8f1eb73792dd9daf84a81f29

Observation c7900e5d-7aea-49b9-9312-8e74246eb0b9 · outbound

This paper cites Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.972701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:56.329894Z digest=sha256:505028eaf53d3dc12145b757e20d2e5d7ec94e394f723b742eb2416f326c10ae

Observation 105c5295-9d6b-4502-bb53-817ee1220e08 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Lora: Low-rank adaptation of large language models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:56.400893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:56.400893Z digest=sha256:f3f63c4783668a7131d4b5e4965547dd2b498c2e24b44455c02290a962a5c69b

Observation 02a2a4a0-ad58-4f6e-846c-7aec0d9ea89c · outbound

This paper cites Avfakenet: A unified end-to-end dense swin transformer deep learning model for audio-visual deepfakes detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Avfakenet: A unified end-to-end dense swin transformer deep learning model for audio-visual deepfakes detection

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.944903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:56.513786Z digest=sha256:aa9f2a6320b25c82c5b040d83bcd77f8ce6babe0aea7f7b50096d80f3dc74547

Observation 4c94acf6-c13d-49e4-b28c-c07b57924b0c · outbound

This paper cites Information theory and statistical mechanics.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Information theory and statistical mechanics

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:56.596373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:56.596373Z digest=sha256:f2fa1b47f283308bddfed97a3d4d88392d50e5a0266921c9b232f1c430ed9749

Observation ceb33f51-45cd-4bcf-8c1c-5be15e06db80 · outbound

This paper cites an unresolved cited work.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:27:03.913566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:56.679484Z digest=sha256:32dcb9709e27b84b9dea833ad05a96785a47439787885455da12ab550d938a1c

Observation 6a14a426-5563-4cd0-8978-90fd0c27ee70 · outbound

This paper cites FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:56.756989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:56.756989Z digest=sha256:0ae741c9c831ecb0dc5d7cbe06303a58ad30f520ecb0451d7ecc667e876e5d1e

Observation aed45079-acd6-41b8-a7b5-9021e513db61 · outbound

This paper cites Fast face-swap using convolu- tional neural networks.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Fast face-swap using convolu- tional neural networks

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.893130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:56.853865Z digest=sha256:79ed6581e153c1e67b1f8551d69b9e055d15294793209b8a74734df0d045acc1

Observation 9a0605a9-493f-4e9d-8144-64663846db63 · outbound

This paper cites Face x-ray for more general face forgery detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Face x-ray for more general face forgery detection

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.874270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:56.962771Z digest=sha256:f75f3a00d016cc345f06007dbf3ebd202967e2ea28188e923c2ac583b2485c24

Observation 15b3d7be-6001-44bf-b67b-a03298ab009d · outbound

This paper cites A Survey on Speech Deepfake Detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation A Survey on Speech Deepfake Detection

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:57.093958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:57.093958Z digest=sha256:f955ebc8fbfcd4058f17061dc37f07302930a58aa47d85f25e23bf55e9596841

Observation a63c7869-feb6-493e-b625-b0cb81abfd36 · outbound

This paper cites Celeb-df: A large-scale challenging dataset for deepfake forensics.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Celeb-df: A large-scale challenging dataset for deepfake forensics

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.853254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:57.170497Z digest=sha256:eac63e713401c595d09313faa2cc22d332e241ac56a6e66f993fef7f514af3a6

Observation 22291b3e-f02a-49b8-9700-1ad0cdd5a626 · outbound

This paper cites Lips are lying: Spotting the temporal inconsistency between audio and visual in lip-syncing deepfakes.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Lips are lying: Spotting the temporal inconsistency between audio and visual in lip-syncing deepfakes

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.836537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:57.249489Z digest=sha256:1f1d3ae83d5589fe8360411cb5b9fe8a7f40600bd87253217ddb134ff67e0069

Observation b27d098b-b827-4513-b161-e801418e734b · outbound

This paper cites Exploiting visual artifacts to expose deepfakes and face manipulations.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Exploiting visual artifacts to expose deepfakes and face manipulations

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.817396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:57.377989Z digest=sha256:cb5cb5f0955f17859441c2eef42013cee2bdf1cec6316f3951e55a0e64eaac17

Observation ef0db15d-c892-49c4-9f39-138ae7df8e3c · outbound

This paper cites Emo- tions don’t lie: An audio-visual deepfake detection method using affective cues.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Emo- tions don’t lie: An audio-visual deepfake detection method using affective cues

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.798323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:57.468124Z digest=sha256:3827db506deb924490312ad62a3b8f375926bffacfda3e0bf0af3e6790236448

Observation 9bc000c9-199a-4dc6-81f0-9058d7085044 · outbound

This paper cites Does audio deepfake detection generalize? arXiv preprint arXiv:2203.16263, 2022.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Does audio deepfake detection generalize? arXiv preprint arXiv:2203.16263, 2022

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:57.564874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:57.564874Z digest=sha256:7d91138b079b549eb528ba6bf4796c175e0b2d2c53aac89a72c8348170876b3e

Observation 1761e696-aaa8-4fd5-9b5c-63c1bf77c4c1 · outbound

This paper cites Nguyen, Junichi Yamagishi, and Isao Echizen.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Nguyen, Junichi Yamagishi, and Isao Echizen

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.780765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:57.664207Z digest=sha256:b02fd03942ee56b1953aadddc73bd2738e83e09e409502fcfb5b0d3966c86ce2

Observation f62a7ef7-de00-4d25-b642-f086ef6c0a94 · outbound

This paper cites Towards universal fake image detectors that generalize across generative models.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Towards universal fake image detectors that generalize across generative models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.759973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:57.758259Z digest=sha256:52361809bd98ba11db873c04b0798d9f5b92747e27d50cffcce54aaecd054541

Observation b33bf91a-4459-486a-a57f-7206d11da1d0 · outbound

This paper cites Avff: Audio-visual feature fusion for video deepfake detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Avff: Audio-visual feature fusion for video deepfake detection

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.738263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:57.831456Z digest=sha256:6301f8dce1634374cb30ff6b6542cb1d2891f5639f3b898428852c66b770d409

Observation 2e33d572-82a6-49c0-a661-9a5ed21294cf · outbound

This paper cites Gpt-4: A large multimodal model.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Gpt-4: A large multimodal model

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.709277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:57.907416Z digest=sha256:0abd3ca9fe889857e249deee3ced705d8b6bcd85217196b6bab7a23a64b58031

Observation a961ec6c-c6cd-4f8e-8810-3f959b1956e7 · outbound

This paper cites Deepfake generation and detection: A benchmark and survey.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Deepfake generation and detection: A benchmark and survey

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:57.992374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:57.992374Z digest=sha256:10ce42e0e4162195f336d6845621131cbb367e2dfedaee2f6923c8c5669186b5

Observation 4939c70d-d9cf-4849-abff-6806947f0f82 · outbound

This paper cites Namboodiri, and C.V.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Namboodiri, and C.V

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.690146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:58.114333Z digest=sha256:3b9cd2978497b58efad0700c1ffa5dcbc85a0c52c4d3dc9e74920bce1a0ac042

Observation ce7db574-1fc5-4d78-8ee6-13722ead2071 · outbound

This paper cites Audio-visual deep neural network for robust person verification.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Audio-visual deep neural network for robust person verification

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.667273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:58.192243Z digest=sha256:fdf7d41940e21b7ee69c0287085ec442e3623b9d8fd02a49c9802bfe23ebf443

Observation 998addff-3393-417d-b781-5e3d60ff255a · outbound

This paper cites Learning transferable visual models from natural language supervision.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Learning transferable visual models from natural language supervision

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.644870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:58.255476Z digest=sha256:b8a386e84706cfbaa07f9276495854367f79e96304b6c01e7e23b6ff1c992645

Observation ed3b185b-2c25-44f5-8a5d-80711d7b9155 · outbound

This paper cites Robust speech recognition via large-scale weak supervision.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Robust speech recognition via large-scale weak supervision

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:58.327095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:58.327095Z digest=sha256:d3799f8f3715ff78c2b51bc40226ecd5c23f4dd06d461934375597f27d0aea8f

Observation 27b2ac9b-a357-416b-8236-577cc5b91d59 · outbound

This paper cites Faceforensics++: Learning to detect manipulated facial images.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Faceforensics++: Learning to detect manipulated facial images

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.606717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:58.398378Z digest=sha256:f14f955ea1251a50ac7fe7e43aff17aafa93d59da318a3dbe99ccb3518a5dfe0

Observation 4221d3cd-8174-470c-804c-b320f7db35c7 · outbound

This paper cites A comprehensive overview of deepfake: Generation, detection, datasets, and opportunities.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation A comprehensive overview of deepfake: Generation, detection, datasets, and opportunities

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.586152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:58.461992Z digest=sha256:4c22f0044e9c0a64f1eaf8eb308cb559c5f07f4faf8f5f553d74ce56e0326122

Observation 5bdf06cd-1658-4f2d-a4ef-2585d7efe552 · outbound

This paper cites Detecting deepfakes with self-blended images.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Detecting deepfakes with self-blended images

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:58.518035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:58.518035Z digest=sha256:dd7a27c4df42bfa6eb295a5e875b4240750c9588fd3ec9931810bdb09e42ce4d

Observation 1a3c8a52-071e-4a1b-9382-492116dc065e · outbound

This paper cites Representative forgery mining for fake face detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Representative forgery mining for fake face detection

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.548967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:58.570297Z digest=sha256:b22bf79bac5bb5bc2b951cd5b061129874ed654f3a9ae0ab193f84f5a345411f

Observation 97f193b6-0890-403e-8b58-e6db7fecf709 · outbound

This paper cites Exploring Depth Information for Detecting Manipulated Face Videos.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Exploring Depth Information for Detecting Manipulated Face Videos

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:27:01.335494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:58.647254Z digest=sha256:053c0801000741bc3d13bf09aab92dc3636bfbd5d25674c72b2fcfa4c2f8c475

Observation ae051bda-29c8-4dc1-a08a-1e6535fcd5a8 · outbound

This paper cites Tan, and Haizhou Li.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Tan, and Haizhou Li

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.522275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:58.706727Z digest=sha256:6902ffe290fa88d802e56dab77eb854119570332c492cef2f5214b0f0b69c67e

Observation 0ad396e1-9008-44d0-909e-bce6a7cfc442 · outbound

This paper cites Deep spatial gradient and temporal depth learning for face anti-spoofing.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Deep spatial gradient and temporal depth learning for face anti-spoofing

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.497856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:58.768804Z digest=sha256:defd8fa3ae852f6b16da9ca8516aa4ae7a78d9ed2ec7a6edbf18e0a5da8be6ce

Observation 8654a873-3fb6-43db-b586-f4a3d6f549a9 · outbound

This paper cites Altfreezing for more general video face forgery detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Altfreezing for more general video face forgery detection

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.470186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:58.844117Z digest=sha256:d90256508e3a3c3c9182b661044670f3f651f417d26a2aee923316d0666f3dae

Observation febc692d-c351-4c3a-85a6-300c68ce4ca3 · outbound

This paper cites Deepfake Video Detection Using Convolutional Vision Transformer.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Deepfake Video Detection Using Convolutional Vision Transformer

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:58.916792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:58.916792Z digest=sha256:268005ba4e2bf70fec8620b3c75d46c9893a9fa24fcdbc2cd88dd04d6160f46f

Observation ddcbe21a-8038-433e-9d37-8d3b1f313f5f · outbound

This paper cites Binaural audio-visual localization.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Binaural audio-visual localization

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.445286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:58.988970Z digest=sha256:87182e34ff52cc353759629048193045ac4f5fcccea6e6fa19420a940ec7d4b9

Observation b4697dcc-b1a5-47a8-ad45-b2d8fffe439d · outbound

This paper cites Identity- driven multimedia forgery detection via reference assistance.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Identity- driven multimedia forgery detection via reference assistance

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.421904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:59.078297Z digest=sha256:9a59a899f05affbdde514304d225494eecaa631931e5b2ff2156f024fdd923a4

Observation 15ce0a97-bc90-4e7e-ac37-97abf628d8bb · outbound

This paper cites Tall: Thumbnail layout for deepfake video detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Tall: Thumbnail layout for deepfake video detection

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.390976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:59.161502Z digest=sha256:3c7d5524213c5a81bfdbc74251655878261173aed9a46a2e491627310df183e2

Observation c6811e39-49e0-4c8c-af3b-f003898121cd · outbound

This paper cites Transcending forgery specificity with latent space augmentation for generalizable deepfake detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Transcending forgery specificity with latent space augmentation for generalizable deepfake detection

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.362476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:59.229390Z digest=sha256:5e8371975c6c1811a8d2d64d1fcb8438b806ee89e2188ed88be3c2a5a48c7408

Observation 971ad1d4-6480-44f9-a694-e2f7f915d7e2 · outbound

This paper cites DF40: Toward Next-Generation Deepfake Detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation DF40: Toward Next-Generation Deepfake Detection

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:59.305225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:59.305225Z digest=sha256:53f34f61bebab7c53b2fac4520b872a4ff9c08086319ce4b8b9e020ae15b9dcf

Observation 1d877f55-ac55-496c-9d20-c55c0911f808 · outbound

This paper cites GPT-ImgEval: A Comprehensive Benchmark for Diagnosing GPT4o in Image Generation.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation GPT-ImgEval: A Comprehensive Benchmark for Diagnosing GPT4o in Image Generation

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:59.364495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:59.364495Z digest=sha256:33914dc72ab3557049d71c35a79eec618181596234bd4a38f32fe9779f87d729

Observation d25b769b-48ca-4156-bb5f-d47bcb261003 · outbound

This paper cites Ucf: Uncovering common features for generalizable deepfake detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Ucf: Uncovering common features for generalizable deepfake detection

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.337860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:59.442188Z digest=sha256:b0073886d77ff367f76b915468fc00fd8733185c6703f58175e5d6bc4d2080d3

Observation 9c056045-81dc-401a-a5b1-e90a2e2f039f · outbound

This paper cites Generalizing Deepfake Video Detection with Plug-and-Play: Video-Level Blending and Spatiotemporal Adapter Tuning.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Generalizing Deepfake Video Detection with Plug-and-Play: Video-Level Blending and Spatiotemporal Adapter Tuning

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:59.498474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:59.498474Z digest=sha256:b67d097fa60077eb33fbd16b033eba19f5d211c30558d1a783911bab2ace25fa

Observation ad1cfbbd-6071-4c67-919d-5c751b1c3a31 · outbound

This paper cites Avoid-df: Audio-visual joint learning for detecting deepfake.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Avoid-df: Audio-visual joint learning for detecting deepfake

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:03.032256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:59.566845Z digest=sha256:e1a865b3eb4586324cf56870934b652374314640ba017b9c778fb173a894b4e5

Observation be58bad2-38dc-4944-941b-4c71e3c5c80e · outbound

This paper cites Exposing deep fakes using inconsistent head poses.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Exposing deep fakes using inconsistent head poses

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:02.912242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:59.676740Z digest=sha256:cb21c3d32c72888a7e83d24bb9bb0ee5f1c25254a0fd7c1f6c8ea7972afc4e9a

Observation 4939e96a-4e65-4f9e-b493-fec38e132990 · outbound

This paper cites Audio Deepfake Detection: A Survey.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Audio Deepfake Detection: A Survey

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:59.790961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:59.790961Z digest=sha256:31d9eabaec171bc7a1c56d0b90de14f12850a08de7ba4c34fc09cfc86bdd8ef4

Observation f57e4aa7-b32c-4034-9cac-eafa94f7c76a · outbound

This paper cites Learning natural consistency representation for face forgery video detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Learning natural consistency representation for face forgery video detection

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:02.835435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:26:59.977050Z digest=sha256:f0fac9744e4d8a7b6b55ba7ac39b3e57f0b88d7e0ec069e9dd058d7260b21bb3

Observation 0a1ded16-7871-485e-a9f1-31875ef21b89 · outbound

This paper cites Inclusion 2024 Global Multimedia Deepfake Detection Challenge: Towards Multi-dimensional Face Forgery Detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Inclusion 2024 Global Multimedia Deepfake Detection Challenge: Towards Multi-dimensional Face Forgery Detection

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T15:27:00.152871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:27:00.152871Z digest=sha256:4f77f113538f199e39051b44b0c07b7ab8496016a33feb1ef61028c2feb7e7a3

Observation 7b594adb-5bb3-4472-b87f-5e8c0bae3f49 · outbound

This paper cites Multi-attentional deepfake detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Multi-attentional deepfake detection

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:02.717608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:27:00.352741Z digest=sha256:5eb6b8efcab29f80e3417003bdeec3a4e42bd5ba97170c5e36eaeb13ce143cde

Observation e5f35ecf-5e29-4f4a-aa9d-05a896ab8679 · outbound

This paper cites Attention-based spatial-temporal multi-scale network for face anti-spoofing.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Attention-based spatial-temporal multi-scale network for face anti-spoofing

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:02.577494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:27:00.559647Z digest=sha256:a3ee7e63954aaa825c822c3623bfc252653c1e096019e386c465d4f7b6327850

Observation 98e6c56d-f524-4cca-8c56-5edb2bfa5436 · outbound

This paper cites Exploring temporal coherence for more general video face forgery detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Exploring temporal coherence for more general video face forgery detection

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:02.445239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:27:00.736070Z digest=sha256:79792d3cda30ade3b36218fb418bd461225748c1f7c476ab756210404605c7c5

Observation ddf7a907-6417-4d79-bc02-88e3e060afbe · outbound

This paper cites Learning deep features for discriminative localization.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Learning deep features for discriminative localization

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:02.349469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:27:00.875898Z digest=sha256:fb46ef419ff9862c793c93915754080c0c400c2f80bd9dd4827bcc0e180400f4

Observation 80b7e914-261f-4337-9107-8595853eb3e8 · outbound

This paper cites Makelttalk: speaker-aware talking-head animation.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Makelttalk: speaker-aware talking-head animation

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:02.294628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:27:00.979116Z digest=sha256:0051355932eb2b5aad20eefbdc5044fa9251e6b1ab6b1934b7019c7de9e2313b

Observation 80b73bbd-895d-4d40-88b8-38d16dff6ea2 · outbound

This paper cites Joint audio-visual deepfake detection.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Joint audio-visual deepfake detection

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:02.134020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:27:00.986411Z digest=sha256:b6a6143cbaebb3ee1b1496f785866164c8104a245f2423113dbd326fabae2df1

Observation ec5cef83-b4fd-464f-a38d-b37ac7d8645e · outbound

This paper cites Lan- guagebind: Extending video-language pretraining to n-modality by language-based semantic alignment.

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation Lan- guagebind: Extending video-language pretraining to n-modality by language-based semantic alignment

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:27:02.020529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:27:00.996471Z digest=sha256:2c3f58bd8e7316b32187b12ca5ec7265e3f86180c4f07a1061efdd5621a2f3d0

Pith citing papers

Observation 4929d710-214b-4f3a-b5a0-17abcb6b57a5 · inbound

CAM-VFD: Cross-Attention Multimodal Video Forgery Detection cites this paper.

CAM-VFD: Cross-Attention Multimodal Video Forgery Detection CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:13:25.042025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-20T15:08:25.309094Z digest=sha256:841495a0feb9841a0c3691f519da6c1d6e360ec266fbfa841a5811e057b896c1