REVIEW 3 cited by
A Penny for Your (visual) Thoughts: Self-Supervised Reconstruction of Natural Movies from Brain Activity
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Reconstructing natural videos from fMRI brain recordings is very challenging, for two main reasons: (i) As fMRI data acquisition is difficult, we only have a limited amount of supervised samples, which is not enough to cover the huge space of natural videos; and (ii) The temporal resolution of fMRI recordings is much lower than the frame rate of natural videos. In this paper, we propose a self-supervised approach for natural-movie reconstruction. By employing cycle-consistency over Encoding-Decoding natural videos, we can: (i) exploit the full framerate of the training videos, and not be limited only to clips that correspond to fMRI recordings; (ii) exploit massive amounts of external natural videos which the subjects never saw inside the fMRI machine. These enable increasing the applicable training data by several orders of magnitude, introducing natural video priors to the decoding network, as well as temporal coherence. Our approach significantly outperforms competing methods, since those train only on the limited supervised data. We further introduce a new and simple temporal prior of natural videos, which - when folded into our fMRI decoder further - allows us to reconstruct videos at a higher frame-rate (HFR) of up to x8 of the original fMRI sample rate.
Forward citations
Cited by 3 Pith papers
-
Neuro-3D: Towards 3D Visual Decoding from EEG Signals
A first benchmark for decoding 3D object shape and dominant color from EEG, with above-chance but limited reconstruction performance.
-
Deep Neural Encoder-Decoder Model to Relate fMRI Brain Activity with Naturalistic Stimuli
An encoder-decoder CNN predicts visual-cortex fMRI from movie frames and reconstructs frames from the predicted fMRI, with saliency maps pointing to occipital and fusiform regions.
-
Comprehensive Review of EEG-to-Output Research: Decoding Neural Signals into Images, Videos, and Audio
A PRISMA-style review of EEG-to-output decoding claims to analyze 1,800 studies but omits the flow diagram, study list, and quantitative synthesis needed to back that claim.
Discussion (0). Continue with ORCID to comment.