Pith. sign in

REVIEW 4 cited by

Exposing Lip-syncing Deepfakes from Mouth Inconsistencies

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2401.10113 v2 pith:J4F4SRCE submitted 2024-01-18 cs.CV

classification cs.CV
keywords lip-syncingdeepfakedeepfakesinconsistenciesmouthdetectionlipincregion
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

A lip-syncing deepfake is a digitally manipulated video in which a person's lip movements are created convincingly using AI models to match altered or entirely new audio. Lip-syncing deepfakes are a dangerous type of deepfakes as the artifacts are limited to the lip region and more difficult to discern. In this paper, we describe a novel approach, LIP-syncing detection based on mouth INConsistency (LIPINC), for lip-syncing deepfake detection by identifying temporal inconsistencies in the mouth region. These inconsistencies are seen in the adjacent frames and throughout the video. Our model can successfully capture these irregularities and outperforms the state-of-the-art methods on several benchmark deepfake datasets. Code is available at https://github.com/skrantidatta/LIPINC

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Human Action CLIPs: Detecting AI-generated Human Motion

    cs.CV 2024-11 conditional novelty 5.0 of 10

    CLIP-based semantic embeddings, with a fine-tuned variant, detect AI-generated human-motion video with high accuracy (up to 99.2% video-level) and generalize to unseen generators.

  2. KLASSify to Verify: Audio-Visual Deepfake Detection Using SSL-based Audio and Handcrafted Visual Features

    eess.AS 2025-08 conditional novelty 4.0 of 10

    A challenge entry combining Wav2Vec-AASIST audio scores with lightweight handcrafted-feature video scores via calibration and maxout reports 92.78% AUC on AV-Deepfake1M++ testA.

  3. GC-ConsFlow: Leveraging Optical Flow Residuals and Global Context for Robust Deepfake Detection

    cs.CV 2025-01 conditional novelty 4.0 of 10

    A dual-stream deepfake detector using global context attention and optical flow residuals reports high accuracy on FF++ and Celeb-DF, but with no error bars or code and only marginal gains over prior work.

  4. Generating and Detecting Various Types of Fake Image and Audio Content: A Review of Modern Deep Learning Technologies and Tools

    cs.CR 2025-01 unverdicted novelty 2.0 of 10

    A survey of deepfake generation and detection methods and tools, with no new experimental or theoretical contributions.

Pith tools