REVIEW 4 cited by
Exposing Lip-syncing Deepfakes from Mouth Inconsistencies
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
A lip-syncing deepfake is a digitally manipulated video in which a person's lip movements are created convincingly using AI models to match altered or entirely new audio. Lip-syncing deepfakes are a dangerous type of deepfakes as the artifacts are limited to the lip region and more difficult to discern. In this paper, we describe a novel approach, LIP-syncing detection based on mouth INConsistency (LIPINC), for lip-syncing deepfake detection by identifying temporal inconsistencies in the mouth region. These inconsistencies are seen in the adjacent frames and throughout the video. Our model can successfully capture these irregularities and outperforms the state-of-the-art methods on several benchmark deepfake datasets. Code is available at https://github.com/skrantidatta/LIPINC
Forward citations
Cited by 4 Pith papers
-
Human Action CLIPs: Detecting AI-generated Human Motion
CLIP-based semantic embeddings, with a fine-tuned variant, detect AI-generated human-motion video with high accuracy (up to 99.2% video-level) and generalize to unseen generators.
-
KLASSify to Verify: Audio-Visual Deepfake Detection Using SSL-based Audio and Handcrafted Visual Features
A challenge entry combining Wav2Vec-AASIST audio scores with lightweight handcrafted-feature video scores via calibration and maxout reports 92.78% AUC on AV-Deepfake1M++ testA.
-
GC-ConsFlow: Leveraging Optical Flow Residuals and Global Context for Robust Deepfake Detection
A dual-stream deepfake detector using global context attention and optical flow residuals reports high accuracy on FF++ and Celeb-DF, but with no error bars or code and only marginal gains over prior work.
-
Generating and Detecting Various Types of Fake Image and Audio Content: A Review of Modern Deep Learning Technologies and Tools
A survey of deepfake generation and detection methods and tools, with no new experimental or theoretical contributions.
Discussion (0). Continue with ORCID to comment.