Pith. sign in

REVIEW 5 cited by

NECOMIMI: Neural-Cognitive Multimodal EEG-informed Image Generation with Diffusion Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2410.00712 v2 pith:PDXJRTMZ submitted 2024-10-01 q-bio.NC cs.LG

classification q-bio.NCcs.LG
keywords generationdiffusionimageimagesmodelsnecomimichallengesclassification
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

NECOMIMI (NEural-COgnitive MultImodal EEG-Informed Image Generation with Diffusion Models) introduces a novel framework for generating images directly from EEG signals using advanced diffusion models. Unlike previous works that focused solely on EEG-image classification through contrastive learning, NECOMIMI extends this task to image generation. The proposed NERV EEG encoder demonstrates state-of-the-art (SoTA) performance across multiple zero-shot classification tasks, including 2-way, 4-way, and 200-way, and achieves top results in our newly proposed Category-based Assessment Table (CAT) Score, which evaluates the quality of EEG-generated images based on semantic concepts. A key discovery of this work is that the model tends to generate abstract or generalized images, such as landscapes, rather than specific objects, highlighting the inherent challenges of translating noisy and low-resolution EEG data into detailed visual outputs. Additionally, we introduce the CAT Score as a new metric tailored for EEG-to-image evaluation and establish a benchmark on the ThingsEEG dataset. This study underscores the potential of EEG-to-image generation while revealing the complexities and challenges that remain in bridging neural activity with visual representation.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. EEG2TEXT-CN: An Exploratory Study of Open-Vocabulary Chinese Text-EEG Alignment via Large Language Model and Contrastive Learning on ChineseEEG

    cs.CL 2025-06 reject novelty 6.0 of 10

    A Chinese EEG-to-text system that aligns 128-channel EEG with per-character text embeddings achieves BLEU-1 6.38% on a held-out subject, claimed as the first open-vocabulary EEG-to-Chinese decoder.

  2. Category-aware EEG image generation based on wavelet transform and contrast semantic loss

    cs.CV 2025-05 conditional novelty 6.0 of 10

    A DWT-gated transformer EEG encoder with CLIP alignment and category-aware clustering loss generates semantic images via a pre-trained diffusion model, achieving 43% max single-subject top-1 classification accuracy an...

  3. Vision-QRWKV: Exploring Quantum-Enhanced RWKV Models for Image Classification

    cs.LG 2025-06 conditional novelty 4.0 of 10

    A quantum-enhanced RWKV with a variational circuit in its channel mixer slightly outperforms classical RWKV on a majority of small-image benchmarks, but the differences are within noise.

  4. Large Cognition Model: Towards Pretrained EEG Foundation Model

    eess.SP 2025-02 conditional novelty 4.0 of 10

    LCM, a transformer EEG model combining contrastive alignment and masked reconstruction, reports state-of-the-art balanced accuracy on BCIC-2A and BCIC-2B.

  5. Quantum-Enhanced Natural Language Generation: A Multi-Model Framework with Hybrid Quantum-Classical Architectures

    quant-ph 2025-08 reject novelty 3.0 of 10

    A benchmark of QASA, QRWKV, and QKSAN against Transformer and MLP on five tiny datasets, with results that contradict the paper's own tables.

Pith tools