Decoding of human identity by computer vision and neuronal vision
Zhang, Y.; M. Aghajan, Z.; Ison, M.; Lu, Q.; Tang, H.; Kalender, G.; Monsoor, T.; Zheng, J.; Kreiman, G.; Roychowdhury, V.; Fried, I.
Show abstract
Extracting meaning from a dynamic and variable flow of incoming information is a major goal of both natural and artificial intelligence. Computer vision (CV) guided by deep learning (DL) has made significant strides in recognizing a specific identity despite highly variable attributes1,2. This is the same challenge faced by the nervous system and partially addressed by the concept cells--neurons exhibiting selective firing in response to specific persons/places, described in the human medial temporal lobe (MTL)3-6. Yet, access to neurons representing a particular concept is limited due to these neurons sparse coding. It is conceivable, however, that the information required for such decoding is present in relatively small neuronal populations. To evaluate how well neuronal populations encode identity information in natural settings, we recorded neuronal activity from multiple brain regions of nine neurosurgical epilepsy patients implanted with depth electrodes, while the subjects watched an episode of the TV series "24". We implemented DL models that used the time-varying population neural data as inputs and decoded the visual presence of the main characters in each frame. Before training and testing the DL models, we devised a minimally supervised CV algorithm (with comparable performance against manually-labelled data7) to detect and label all the important characters in each frame. This methodology allowed us to compare "computer vision" with "neuronal vision"--footprints associated with each character present in the activity of a subset of neurons--and identify the brain regions that contributed to this decoding process. We then tested the DL models during a recognition memory task following movie viewing where subjects were asked to recognize clip segments from the presented episode. DL model activations were not only modulated by the presence of the corresponding characters but also by participants subjective memory of whether they had seen the clip segment, and by the associative strengths of the characters in the narrative plot. The described approach can offer novel ways to probe the representation of concepts in time-evolving dynamic behavioral tasks. Further, the results suggest that the information required to robustly decode concepts is present in the population activity of only tens of neurons even in brain regions beyond MTL.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Inferring Latent Behavioral Strategy from the Representational Geometry of Prefrontal Cortex Activity 97%
- Hippocampal gamma oscillations form complex ensembles modulated by behavior and learning 96%
- Automatic mapping of multiplexed social receptive fields by deep learning and GPU-accelerated 3D videography 96%
Similar papers in this journal
- A deep learning framework for inference of single-trial neural population dynamics from calcium imaging with sub-frame temporal resolution 96%
- Database and deep learning toolbox for noise-optimized, generalized spike inference from calcium imaging data 96%
- The geometry of cortical representations of touch in rodents 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.