Image retrieval based on closed-loop visual-semantic neural decoding
Fukuma, R.; Yanagisawa, T.; Sugano, H.; Tamura, K.; Oshino, S.; Tani, N.; Iimura, Y.; Khoo, H. M.; Suzuki, H.; Yang, H.; Iwata, T.; Nakajima, M.; Nishimoto, S.; Kamitani, Y.; Kishima, H.
Show abstract
Neural decoding via the latent space of deep neural network models can infer perceived and imagined images from neural activities, even when the image is novel for the subject and decoder. Brain-computer interfaces (BCIs) using the latent space enable a subject to retrieve intended image from a large dataset on the basis of their neural activities but have not yet been realized. Here, we used neural decoding in a closed-loop condition to retrieve images of the instructed categories from 2.3 million images on the basis of the latent vector inferred from electrocorticographic signals of visual cortices. Using a latent space of contrastive language-image pretraining (CLIP) model, two subjects retrieved images with significant accuracy exceeding 80% for two instructions. In contrast, the image retrieval failed using the latent space of another model, AlexNet. In another task to imagine an image while viewing a different image, the imagery made the inferred latent vector significantly closer to the vector of the imagined category in the CLIP latent space but significantly further away in the AlexNet latent space, although the same electrocorticographic signals from nine subjects were decoded. Humans can retrieve the intended information via a closed-loop BCI with an appropriate latent space.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- Neural Processing of Naturalistic Audiovisual Events in Space and Time 95%
- Fronto-parietal networks shape human conscious report through attention gain and reorienting 95%
- Manifold learning analysis suggests novel strategies for aligning single-cell multi-modalities and revealing functional genomics for neuronal electrophysiology 94%
Similar papers in this journal
- Intracranial electroencephalography reveals effector-independent evidence accumulation dynamics in multiple human brain regions 94%
- Driving and suppressing the human language network using large language models 93%
- Feature-based encoding of face identity by single neurons in the human amygdala and hippocampus 93%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.