Pith. sign in

REVIEW

Unsupervised Audio-Visual Subspace Alignment for High-Stakes Deception Detection

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2102.03673 v1 pith:OXZZFU3O submitted 2021-02-06 cs.CV cs.HCcs.LG

classification cs.CVcs.HCcs.LG
keywords deceptionhigh-stakesmodelsreal-worldunsupervisedapproachaudio-visualdetect
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Automated systems that detect deception in high-stakes situations can enhance societal well-being across medical, social work, and legal domains. Existing models for detecting high-stakes deception in videos have been supervised, but labeled datasets to train models can rarely be collected for most real-world applications. To address this problem, we propose the first multimodal unsupervised transfer learning approach that detects real-world, high-stakes deception in videos without using high-stakes labels. Our subspace-alignment (SA) approach adapts audio-visual representations of deception in lab-controlled low-stakes scenarios to detect deception in real-world, high-stakes situations. Our best unsupervised SA models outperform models without SA, outperform human ability, and perform comparably to a number of existing supervised models. Our research demonstrates the potential for introducing subspace-based transfer learning to model high-stakes deception and other social behaviors in real-world contexts with a scarcity of labeled behavioral data.

Discussion (0). Continue with ORCID to comment.

Pith tools