REVIEW 4 major objections 6 minor 68 references
AI-guided stimuli discovery and generation to optimize facial emotion perception studies in autism
T0 review · 4 major / 6 minor · reviewed 2026-07-10 · grok-4.5
Pith's one-line read Autistic–neurotypical emotion-judgment differences are sparse at the image level, and behavior-aligned neural nets can both find and transform the faces that reveal them.
desk verdict Solid methods paper: sparsity and CLIP-guided selection are real; synthesis is weaker and the Methods/Results loss descriptions do not match. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
Population-specific ridge-regression readouts on fixed ANN visual embeddings, whose predicted autistic–neurotypical gap both ranks candidate faces for selection and supplies the loss for closed-loop GAN latent-code optimization that synthesizes gap-reduced expressions.
What would settle it
In a new independent cohort, without phenotype matching, CLIP-ranked faces fail to beat identity- and intensity-matched random faces on |ASD–NT| happy-response difference, or the synthesized faces fail to reduce that gap relative to their diagnostic bases under the same leave-one-image-out protocol.
Extended reading notes
Core claim
Autistic–neurotypical differences in facial emotion judgments are concentrated in a sparse subset of high-leverage images rather than expressed uniformly. Population-specific ANN models that predict those image-level judgments can prospectively select novel faces that produce larger group separation than random sampling in a new cohort, and the same models can guide generative transformations of diagnostic faces that reduce separation when validation participants are matched to the targeted response phenotype.
Load-bearing premise
The synthesis result assumes that matching new participants by correlation to the original group response templates is a fair test of gap reduction rather than mainly recovering the subspace the optimizer was trained to fix.
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. The paper argues that autistic–neurotypical differences in facial emotion judgments are sparse at the image level rather than uniform across stimuli, and that population-specific ANN readouts on shared visual features can both (i) prospectively select novel faces that enlarge group separation and (ii) guide GANmut transformations that reduce separation. Reanalysis of Wang & Adolphs (2017) shows diagnostic images form a high-leverage tail. Models trained on that dataset are applied to MSFDE; in an independent lab cohort (12 ASD, 13 NT), ANN-selected sets vary by architecture, with CLIP clearly outperforming matched random sets (|ASD–NT| 0.149 vs 0.091, empirical p=.006), and selection success tracking NT behavioral alignment (ρ=.82). Closed-loop synthesis then optimizes GANmut latent codes to reduce predicted group divergence; under leave-one-image-out correlation-based phenotype matching in a Prolific cohort, mean gap falls from 0.138 to 0.076 (t(14)=2.38, one-tailed p=.016). AU-region summaries describe structured but multi-region facial changes.
Significance. If the dual claim holds, the work supplies a concrete template for moving autism behavioral assays from stimulus-averaged summaries to image-computable, prospectively optimized stimulus design—an important methodological shift for heterogeneous neurodevelopmental phenotypes. Strengths include prospective testing of selection on a new stimulus set and independent lab cohort, explicit architecture comparison rather than a single black-box model, and a closed-loop synthesis intervention that goes beyond post-hoc explanation. The sparsity result (Figs. 1B, 3E) is a useful reframing of why facial-emotion findings have been inconsistent. The synthesis half, if cleaned up, would be a stronger validity test than selection alone. Code and de-identified data are promised, which supports reproducibility.
major comments (4)
- Results (Fig. 3B) report that the average selected-versus-random uplift across seven ANNs is not significant (mean 0.010 ± 0.012 SEM; paired t(6)=0.84; one-tailed p=.217); only CLIP clearly beats random (empirical p=.006). The abstract and strongest framing present “model-selected images produced larger behavioral differences than matched random images” as a general result. That claim should be restated as architecture-dependent, with CLIP (and NT-alignment) as the operative finding, or the multi-model average should be demoted from a positive result.
- Methods vs Results disagree on the synthesis objective. Results and Fig. 4A describe minimizing the predicted autistic–neurotypical difference on the candidate image (Δ(θ,ρ) between population-specific “happy” scores). Methods define the loss as (neurotypical score of the original image − autistic score of the synthesized image)². These are not the same quantity. If Methods is correct, the optimizer does not implement the closed-loop intervention claimed in Results, so the behavioral gap reduction is not a clean test of that intervention. This inconsistency must be resolved with the actual loss used, and any mismatch between claimed and implemented objective should be stated.
- Closed-loop synthesis validation (Results §Closed-loop synthesis; Fig. 4C; Methods) uses leave-one-image-out correlation matching of Prolific participants to the original Wang & Adolphs group response templates, then measures gap reduction on held-out pairs. Discussion correctly scopes this as testing attenuation within the targeted image-level phenotype, but the abstract and primary results presentation still report phenotype-matched reduction as the synthesis result without that qualifier. The claim should be limited to the matched subspace, and unmatched or randomly sampled cohort analyses (even if null or weaker) should be reported so readers can judge generalizability.
- Lab selection cohort is small (n=12 ASD, 13 NT; Table S1) and the online synthesis cohort relies on self-reported diagnosis plus SRS/AQ (Fig. S3). Trait-binned analyses (Fig. S4) help, but the synthesis effect size and selection uplift need either larger independent replication or explicit power/reliability bounds before the framework is positioned as ready for assay construction. At minimum, report image-level reliability and subject-level stability of |ASD–NT| for selected vs random sets.
minor comments (6)
- Figure 2C hypotheses (H0/H1/H2) are useful but the main text should state which architectures support H2 before the multi-model average is discussed.
- Clarify ridge regularization, cross-validation folds, and whether decoder hyperparameters were chosen on Wang & Adolphs only or tuned with any MSFDE information.
- Fig. 3C uses ΔP(happy)=Control−ASD while elsewhere Δ is autistic−NT; keep signed conventions consistent.
- AU analysis (Fig. 5) is appropriately descriptive; state explicitly that AUCanvas masks come from neutral references and that pixel-thresholding for overlays is not used in the quantitative vectors.
- Abstract says “independent cohort” for selection and “phenotype-matched validation” for synthesis; keep that distinction in the Results lead sentences as well.
- Minor typos: “mechani stic”, “netw ork”, spacing artifacts in the abstract/intro PDF text.
Circularity Check
Prospective selection is independent; synthesis success is measured only inside leave-one-out phenotype-matched subspace the optimizer targeted, not forced by construction.
-
fitted input called prediction
[Results §Closed-loop synthesis; Fig. 4C; Methods (phenotype-matched validation)]
"For each held-out image pair, neurotypical participants were selected according to the correlation between their responses to the remaining base images and the original neurotypical behavioral template. Autistic participants were selected in the same way, using the original autistic behavioral template. ... Under this correlation-based phenotype-matched validation, synthesized images reliably reduced autistic–neurotypical behavioral divergence (Figure 4C)."
Participant inclusion is conditioned on matching the same group-specific image-level response templates that defined the synthesis objective. Gap reduction is then reported only inside that reselected subspace. This is not pure tautology—the held-out image is unused for matching and the behavioral outcome can fail—but success is measured only for the fitted phenotype the optimizer targeted, so the synthesis 'prediction' is statistically favored rather than tested in unmatched samples. Abstract/strongest claim still present this as the synthesis result.
full rationale
The paper's derivation chain is largely non-circular. Population-specific ANN decoders are trained on Wang & Adolphs image-level judgments, then applied to a new stimulus set (MSFDE) and tested in an independent lab cohort; selected-versus-random separation is an empirical outcome, not a fit renamed as prediction. Closed-loop synthesis optimizes GAN latent codes under those same predictors and is then behaviorally retested. The only load-bearing soft spot is validation design: leave-one-image-out correlation matching reselects Prolific participants to the original group response templates before measuring gap reduction on held-out pairs. That scopes the claim to the image-level phenotype the models were trained to capture (as the Discussion acknowledges) and raises a mild fitted-subspace concern, but the held-out image is never used for matching and the measured |ASD–NT| change remains free empirical data—not equal to the training labels or the synthesis loss by construction. Self-citations (e.g., Kar 2022) supply related prior framing, not uniqueness theorems that force the present results. Methods/Results disagreement on the synthesis loss is a correctness inconsistency, not circularity. Overall score 3: one non-tautological but self-referential validation step; central selection claim is clean.
Assumptions & free parameters
free parameters (5)
- ridge-regression regularization strength for ASD and NT decoders
- feature layer choice (penultimate / visual embedding)
- synthesis SGD halt criteria (ε=0.0001 or 25 iterations)
- phenotype-matching correlation template and leave-one-image-out selection rule
- AU pixel-difference thresholding / 17-region masks from neutral references
assumptions (6)
- domain assumption Fixed pretrained vision backbones (AlexNet, VGG-19, ResNet-50, ConvNeXt, CORNet-S, ViT, CLIP) provide image embeddings that can be linearly decoded into population-specific emotion judgments.
- domain assumption Group-mean image-level happy probabilities are adequate targets for discovering stimuli that separate autistic and neurotypical observers.
- domain assumption GANmut’s 2D polar latent space (θ emotion category, ρ intensity) can express behaviorally relevant facial transformations while preserving identity and realism.
- domain assumption Self-reported formal autism diagnosis plus SRS/AQ on Prolific, after consistency checks, sufficiently identifies autistic vs neurotypical online groups for synthesis validation.
- ad hoc to paper Leave-one-image-out correlation matching to original group templates tests synthesis without circular use of the held-out image.
- standard math Standard ridge regression, SGD, and Pearson/Spearman statistics apply to these behavioral vectors.
invented entities (3)
-
Population-specific ANN behavioral decoders (ASD vs NT readouts on shared visual features)
independent evidence
-
Diagnostic / high-leverage facial-expression stimuli (sparse image-level phenotype)
independent evidence
-
Gap-reducing closed-loop synthesized faces
Cite this review
Pith. "Pith review of AI-guided stimuli discovery and generation to optimize facial emotion perception studies in autism." pith.science (2026). https://pith.science/paper/TLDVMFRE
@misc{pith2026260708533,
author = {Pith},
title = {Pith review of: AI-guided stimuli discovery and generation to optimize facial emotion perception studies in autism},
year = {2026},
howpublished = {\url{https://pith.science/paper/TLDVMFRE}},
note = {Machine review of arXiv:2607.08533}
}
read the original abstract
Understanding perceptual differences between autistic and neurotypical adults requires behavioral assays that are sensitive, reliable, and mechanistically informative. Facial emotion perception is a useful test case because group differences have been reported, but findings vary across studies. Here we show that this variability may reflect image-level sparsity: autistic-neurotypical differences in emotion judgments were concentrated in a small subset of diagnostic facial expressions rather than spread uniformly across stimuli. We trained population-specific artificial neural network models to predict image-level judgments for autistic and neurotypical participants, then used these models to select novel faces predicted to maximize group separation. In an independent cohort, model-selected images produced larger behavioral differences than matched random images. We then used the same models with a generative adversarial network to transform diagnostic images toward greater predicted group agreement. In phenotype-matched validation, synthesized images reduced behavioral separation relative to their matched originals. These results establish a model-guided framework for discovering and transforming stimuli that reveal population-specific perceptual differences. More broadly, they show how behavioral phenotyping can move beyond averaging across fixed stimulus sets toward optimized assays that identify the conditions under which neurodivergent perception diverges or converges.
Figures
Figures from the paper (2 more)
Reference graph
Works this paper leans on
-
[1]
Lord, C. et al. Autism spectrum disorder. Nat. Rev. Dis. Primer 6, 5 (2020)
work page 2020
-
[2]
Behrmann, M., Thomas, C. & Humphreys, K. Seeing it differently: visual processing in autism. Trends Cogn. Sci. 10, 258–264 (2006)
work page 2006
-
[3]
Robertson, C. E. & Baron-Cohen, S. Sensory perception in autism. Nat. Rev. Neurosci. 18, 671–684 (2017)
work page 2017
-
[4]
Hadad, B.-S. & Yashar, A. Sensory Perception in Autism: What Can We Learn? Annu. Rev. Vis. Sci. 8, 239–264 (2022)
work page 2022
-
[5]
Zilbovicius, M. et al. Autism, the superior temporal sulcus and social perception. Trends Neurosci. 29, 359–366 (2006)
work page 2006
-
[6]
Morrison, K. E. et al. Psychometric Evaluation of Social Cognitive Measures for Adults with Autism. Autism Res. 12, 766–778 (2019)
work page 2019
-
[7]
Kennedy, D. P. & Adolphs, R. Perception of emotions from facial expressions in high- functioning adults with autism. Neuropsychologia 50, 3313–3319 (2012)
work page 2012
-
[8]
Loth, E. et al. Facial expression recognition as a candidate marker for autism spectrum disorder: how frequent and severe are deficits? Mol. Autism 9, 7 (2018)
work page 2018
Show all 68 references
-
[9]
Sato, W. et al. Impaired detection of happy facial expressions in autism. Sci. Rep. 7, 13340 (2017)
2017
-
[10]
& Adolphs, R
Wang, S. & Adolphs, R. Reduced specificity in emotion judgment in people with autism spectrum disorder. Neuropsychologia 99, 286–295 (2017)
2017
-
[11]
M., Moody, E
Beall, P. M., Moody, E. J., McIntosh, D. N., Hepburn, S. L. & Reed, C. L. Rapid facial reactions to emotional facial expressions in typically developing children and children with autism spectrum disorder. J. Exp. Child Psychol. 101, 206–223 (2008)
2008
-
[12]
Yi, L. et al. Abnormality in face scanning by children with autism spectrum disorder is limited to the eye region: Evidence from multi-method analyses of eye tracking data. J. Vis. 13, 5– 5 (2013). 27
2013
-
[13]
L., Piven, J
Neumann, D., Spezio, M. L., Piven, J. & Adolphs, R. Looking you in the mouth: abnormal gaze in autism resulting from impaired top-down modulation of visual attention. Soc. Cogn. Affect. Neurosci. 1, 194–202 (2006)
2006
-
[14]
Dalton, K. M. et al. Gaze fixation and the neural circuitry of face processing in autism. Nat. Neurosci. 8, 519–526 (2005)
2005
-
[15]
A., Morris, J
Pelphrey, K. A., Morris, J. P., McCarthy, G. & LaBar, K. S. Perception of dynamic changes in facial affect and identity in autism. Soc. Cogn. Affect. Neurosci. 2, 140–149 (2007)
2007
-
[16]
Kleinhans, N. M. et al. fMRI evidence of neural abnormalities in the subcortical face processing system in ASD. NeuroImage 54, 697–704 (2011)
2011
-
[17]
Rutishauser, U. et al. Single-Neuron Correlates of Atypical Face Processing in Autism. Neuron 80, 887–899 (2013)
2013
-
[18]
Griffin, J. W. et al. Decoding the temporal dynamics of face-specific neural representations in autism. Nat. Ment. Health 4, 1109–1121 (2026)
2026
-
[19]
& Hamilton, A
Uljarevic, M. & Hamilton, A. Recognition of Emotions in Autism: A Formal Meta-Analysis. J. Autism Dev. Disord. 43, 1517–1526 (2013)
2013
-
[20]
López Pérez, D. et al. Visual Search Performance Does Not Relate to Autistic Traits in the General Population. J. Autism Dev. Disord. 49, 2624–2631 (2019)
2019
-
[21]
& Wagemans, J
Evers, K., Steyaert, J., Noens, I. & Wagemans, J. Reduced Recognition of Dynamic Facial Emotional Expressions and Emotion-Specific Response Bias in Children with an Autism Spectrum Disorder. J. Autism Dev. Disord. 45, 1774–1784 (2015)
2015
-
[22]
& DiCarlo, J
Kar, K. & DiCarlo, J. J. The Quest for an Integrated Set of Neural Mechanisms Underlying Object Recognition in Primates. Annu. Rev. Vis. Sci. 10, 91–121 (2024)
2024
-
[23]
Schrimpf, M. et al. Brain-Score: Which Artificial Neural Network for Object Recognition Is Most Brain-Like? http://biorxiv.org/lookup/doi/10.1101/407007 (2018) doi:10.1101/407007
2018 doi
-
[24]
Rajalingham, R. et al. Large-Scale, High-Resolution Comparison of the Core Visual Object Recognition Behavior of Humans, Monkeys, and State-of-the-Art Deep Artificial Neural Networks. J. Neurosci. 38, 7255–7269 (2018)
2018
-
[25]
Yamins, D. L. K. et al. Performance-optimized hierarchical models predict neural responses in higher visual cortex. Proc. Natl. Acad. Sci. 111, 8619–8624 (2014). 28
2014
-
[26]
& DiCarlo, J
Bashivan, P., Kar, K. & DiCarlo, J. J. Neural population control via deep image synthesis. Science 364, eaav9436 (2019)
2019
-
[27]
& Tsao, D
Chang, L. & Tsao, D. Y. The Code for Facial Identity in the Primate Brain. Cell 169, 1013- 1028.e14 (2017)
2017
-
[28]
Snoek, L. et al. Testing, explaining, and exploring models of facial expressions of emotions. Sci. Adv. 9, eabq8421 (2023)
2023
-
[29]
T., Ramezanpour, H
Wehrheim, M., Alamooti, S. T., Ramezanpour, H. & Kar, K. Facial expression discrimination emerges from neural subspaces shared with detection and identity. Preprint at https://doi.org/10.1101/2025.08.25.672186 (2025)
2025 doi
-
[30]
A Computational Probe into the Behavioral and Neural Markers of Atypical Facial Emotion Processing in Autism
Kar, K. A Computational Probe into the Behavioral and Neural Markers of Atypical Facial Emotion Processing in Autism. J. Neurosci. 42, 5115–5126 (2022)
2022
-
[31]
Yu, H. et al. Multimodal investigations of emotional face processing and social trait judgment of faces. Ann. N. Y. Acad. Sci. 1531, 29–48 (2024)
2024
-
[32]
A., Bashivan, P., Abate, A., DiCarlo, J
Ratan Murty, N. A., Bashivan, P., Abate, A., DiCarlo, J. J. & Kanwisher, N. Computational models of category-selective brain regions enable high-throughput tests of selectivity. Nat. Commun. 12, 5540 (2021)
2021
-
[33]
Ponce, C. R. et al. Evolving Images for Visual Neurons Using a Deep Generative Network Reveals Coding Principles and Neuronal Preferences. Cell 177, 999-1009.e10 (2019)
2019
-
[34]
P., Huang, Z., Romero, A
d’Apolito, S., Paudel, D. P., Huang, Z., Romero, A. & Gool, L. V. GANmut: Learning Interpretable Conditional Space for Gamut of Emotions. in 2021 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 568–577 (IEEE, Nashville, TN, USA, 2021). doi:10.1109/CVPR464...
2021 doi
- [35]
-
[36]
& Hinton, G
Krizhevsky, A., Sutskever, I. & Hinton, G. E. ImageNet Classification with Deep Convolutional Neural Networks. in Advances in Neural Information Processing Systems (eds Pereira, F., Burges, C. J. C., Bottou, L. & Weinberger, K. Q.) vol. 25 (Curran Associates, Inc., 2012). 29
2012
-
[37]
& Zisserman, A
Simonyan, K. & Zisserman, A. Very Deep Convolutional Networks for Large-Scale Image Recognition. ArXiv14091556 Cs http://arxiv.org/abs/1409.1556 (2015)
2015 arXiv
- [38]
- [39]
-
[40]
Kubilius, J. et al. Brain-Like Object Recognition with High-Performing Shallow Recurrent ANNs. in Advances in Neural Information Processing Systems (eds Wallach, H. et al.) vol. 32 (Curran Associates, Inc., 2019)
2019
- [41]
- [42]
-
[43]
Beaupré, M. G. & Hess, U. Cross-Cultural Emotion Recognition among Canadian Ethnic Groups. J. Cross-Cult. Psychol. 36, 355–370 (2005)
2005
-
[44]
Doerig, A. et al. High-level visual representations in the human brain are aligned with large language models. Nat. Mach. Intell. 7, 1220–1234 (2025)
2025
- [45]
-
[46]
Kar, K., Kubilius, J., Schmidt, K., Issa, E. B. & DiCarlo, J. J. Evidence that recurrent circuits are critical to the ventral stream’s execution of core object recognition behavior. Nat. Neurosci. 22, 974–983 (2019)
2019
-
[47]
& Kar, K
Muzellec, S. & Kar, K. Reverse predictivity for bidirectional comparison of neural networks and biological brains. Nat. Mach. Intell. 8, 474–488 (2026)
2026
-
[48]
Sörensen, L. K. A., DiCarlo, J. J. & Kar, K. Hierarchical optimization predicts plasticity in the macaque inferior temporal cortex following object training. Nat. Commun. https://doi.org/10.1038/s41467-026-74816-0 (2026) doi:10.1038/s41467-026-74816-0. 30
2026 doi
-
[49]
& Gosselin, F
Blais, C., Roy, C., Fiset, D., Arguin, M. & Gosselin, F. The eyes are not the window to basic emotions. Neuropsychologia 50, 2830–2838 (2012)
2012
-
[50]
& Plomin, R
Happé, F., Ronald, A. & Plomin, R. Time to give up on a single explanation for autism. Nat. Neurosci. 9, 1218–1220 (2006)
2006
-
[51]
Geschwind, D. H. Advances in Autism. Annu. Rev. Med. 60, 367–380 (2009)
2009
-
[52]
Lai, M.-C., Lombardo, M. V. & Baron-Cohen, S. Autism. The Lancet 383, 896–910 (2014)
2014
-
[53]
& Breakspear, M
Opel, N. & Breakspear, M. Transforming mental health research and care through artificial intelligence. Science 391, 249–258 (2026)
2026
-
[54]
K., Kornblith, S
Muzellec, S., Alghetaa, Y. K., Kornblith, S. & Kar, K. MAPS: Masked Attribution-based Probing of Strategies- A computational framework to align human and model explanations. Preprint at https://doi.org/10.48550/arXiv.2510.12141 (2025)
2025 doi
-
[55]
& Isola, P
Goetschalckx, L., Andonian, A., Oliva, A. & Isola, P. GANalyze: Toward Visual Definitions of Cognitive Image Properties. ArXiv190610112 Cs http://arxiv.org/abs/1906.10112 (2019)
1906 arXiv
-
[56]
& Mohsenzadeh, Y
Younesi, M. & Mohsenzadeh, Y. Controlling Memorability of Face Images. Preprint at http://arxiv.org/abs/2202.11896 (2022)
2022 arXiv
-
[57]
G., Skora, L
Krumhuber, E. G., Skora, L. I., Hill, H. C. H. & Lander, K. The role of facial movements in emotion recognition. Nat. Rev. Psychol. 2, 283–296 (2023)
2023
-
[58]
& Todorov, A
Aviezer, H., Trope, Y. & Todorov, A. Body Cues, Not Facial Expressions, Discriminate Between Intense Positive and Negative Emotions. Science 338, 1225–1229 (2012)
2012
-
[59]
& Gosselin, F
Belin, P., Fillion-Bilodeau, S. & Gosselin, F. The Montreal Affective Voices: A validated set of nonverbal affect bursts for research on auditory affective processing. Behav. Res. Methods 40, 531–539 (2008)
2008
-
[60]
& Lord, C
Hus, V. & Lord, C. The Autism Diagnostic Observation Schedule, Module 4: Revised Algorithm and Standardized Severity Scores. J. Autism Dev. Disord. 44, 1996–2012 (2014)
1996
-
[61]
N., Penton-Voak, I
Dalili, M. N., Penton-Voak, I. S., Harmer, C. J. & Munafò, M. R. Meta-analysis of emotion recognition deficits in major depressive disorder. Psychol. Med. 45, 1135–1144 (2015)
2015
-
[62]
& Schitter, C
Palan, S. & Schitter, C. Prolific.ac—A subject pool for online experiments. J. Behav. Exp. Finance 17, 22–27 (2018). 31
2018
-
[63]
& Shapiro, D
Chandler, J. & Shapiro, D. Conducting Clinical Research Using Crowdsourced Convenience Samples. Annu. Rev. Clin. Psychol. 12, 53–81 (2016)
2016
-
[64]
Constantino, J. N. Social Responsiveness Scale. in Encyclopedia of Autism Spectrum Disorders (ed. Volkmar, F. R.) 2919–2929 (Springer New York, New York, NY, 2013). doi:10.1007/978-1-4419-1698-3_296
2013 doi
-
[65]
& Clubley, E
Baron-Cohen, S., Wheelwright, S., Skinner, R., Martin, J. & Clubley, E. The Autism- Spectrum Quotient (AQ): Evidence from Asperger Syndrome/High-Functioning Autism, Males and Females, Scientists and Mathematicians. J. Autism Dev. Disord. 31, 5–17 (2001)
2001
-
[66]
Choi, Y. et al. StarGAN: Unified Generative Adversarial Networks for Multi-domain Image- to-Image Translation. in 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition 8789–8797 (IEEE, Salt Lake City, UT, 2018). doi:10.1109/CVPR.2018.00916
2018 doi
-
[67]
M., Sanfeliu, A
Pumarola, A., Agudo, A., Martinez, A. M., Sanfeliu, A. & Moreno-Noguer, F. GANimation: Anatomically-Aware Facial Animation from a Single Image. in Computer Vision – ECCV 2018 (eds Ferrari, V., Hebert, M., Sminchisescu, C. & Weiss, Y.) vol. 11214 835–851 (Springer International...
2018
-
[68]
& Gunes, H
Luo, C., Song, S., Xie, W., Shen, L. & Gunes, H. Learning Multi-dimensional Edge Feature- based AU Relation Graph for Facial Action Unit Recognition. in Proceedings of the Thirty-First International Joint Conference on Artificial Intelligence 1239–1246 (International Joint Con...
2022 doi
Reviewed July 10, 2026 · model on record in the stance chip above.
Discussion (0). Continue with ORCID to comment.