REVIEW 5 cited by
DeWave: Discrete EEG Waves Encoding for Brain Dynamics to Text Translation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
The translation of brain dynamics into natural language is pivotal for brain-computer interfaces (BCIs). With the swift advancement of large language models, such as ChatGPT, the need to bridge the gap between the brain and languages becomes increasingly pressing. Current methods, however, require eye-tracking fixations or event markers to segment brain dynamics into word-level features, which can restrict the practical application of these systems. To tackle these issues, we introduce a novel framework, DeWave, that integrates discrete encoding sequences into open-vocabulary EEG-to-text translation tasks. DeWave uses a quantized variational encoder to derive discrete codex encoding and align it with pre-trained language models. This discrete codex representation brings forth two advantages: 1) it realizes translation on raw waves without marker by introducing text-EEG contrastive alignment training, and 2) it alleviates the interference caused by individual differences in EEG waves through an invariant discrete codex with or without markers. Our model surpasses the previous baseline (40.1 and 31.7) by 3.06% and 6.34%, respectively, achieving 41.35 BLEU-1 and 33.71 Rouge-F on the ZuCo Dataset. This work is the first to facilitate the translation of entire EEG signal periods without word-level order markers (e.g., eye fixations), scoring 20.5 BLEU-1 and 29.5 Rouge-1 on the ZuCo Dataset.
Forward citations
Cited by 5 Pith papers
-
Escaping the BLEU Trap: A Signal-Grounded Framework with Decoupled Semantic Guidance for EEG-to-Text Decoding
SemKey predicts four semantic attributes from EEG and conditions a frozen LLM on them, beating prior decoders on new semantic-alignment metrics while leaving true word-level accuracy low (2.7% content recall).
-
Decoding Visual Neural Representations by Multimodal with Dynamic Balancing
A multimodal EEG-image-text contrastive framework with dynamic gradient balancing and stochastic noise improves zero-shot object recognition from EEG on ThingsEEG, raising top-1 accuracy from 13.8% to 15.8%.
-
Large Language Models for EEG: A Comprehensive Survey and Taxonomy
A taxonomy and review of studies applying large language models to EEG signals, organized into four domains and three adaptation strategies.
-
ETS: Open Vocabulary Electroencephalography-To-Text Decoding and Sentiment Classification
ETS pairs EEG and eye-tracking with a CNN-transformer encoder and BART/T5 decoder to decode read sentences from brain signals and classify their sentiment in a zero-shot pipeline.
-
Bridging Brain Signals and Language: A Deep Learning Approach to EEG-to-Text Decoding
This paper reproduces an existing EEG-to-text architecture, swaps in T5 and ProphetNet, and reports lower scores than the prior work it claims to beat.
Discussion (0). Continue with ORCID to comment.