Lightweight Reinforcement Algorithms for autonomous, scalable intra-cortical Brain Machine Interfaces
Shaikh, S.; So, R.; Sibindi, T.; Libedinsky, C.; Basu, A.
Show abstract
Intra-cortical Brain Machine Interfaces (iBMIs) with wireless capability could scale the number of recording channels by integrating an intention decoder to reduce data rates. However, the need for frequent retraining due to neural signal non-stationarity is a big impediment. This paper presents an alternate paradigm of online reinforcement learning (RL) with a binary evaluative feedback in iBMIs to tackle this issue. This paradigm eliminates time-consuming calibration procedures. Instead, it relies on updating the model on a sequential sample-by-sample basis based on an instantaneous evaluative binary feedback signal. However, batch updates of weight in popular deep networks is very resource consuming and incompatible with constraints of an implant. In this work, using offline open-loop analysis on pre-recorded data, we show application of a simple RL algorithm - Banditron -in discrete-state iBMIs and compare it against previously reported state of the art RL algorithms - Hebbian RL, Attention gated RL, deep Q-learning. Owing to its simplistic single-layer architecture, Banditron is found to yield at least two orders of magnitude of reduction in power dissipation compared to state of the art RL algorithms. At the same time, post-hoc analysis performed on four pre-recorded experimental datasets procured from the motor cortex of two non-human primates performing joystick-based movement-related tasks indicate Banditron performing significantly better than state of the art RL algorithms by at least 5%, 10%, 7% and 7% in experiments 1, 2, 3 and 4 respectively. Furthermore, we propose a non-linear variant of Banditron, Banditron-RP, which gives an average improvement of 6%, 2% in decoding accuracy in experiments 2,4 respectively with only a moderate increase in power consumption.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A Spiking Neural Network with Continuous Local Learning for Robust Online Brain Machine Interface 97%
- Firing-rate-modulated spike detection and neural decoding co-design 96%
- Speech decoding from a small set of spatially segregated minimally invasive intracranial EEG electrodes with a compact and interpretable neural network 96%
Similar papers in this journal
- Sparse Ensemble Machine Learning to improve robustness of long-term decoding in iBMIs 98%
- Optimal versus approximate channel selection methods for EEG decoding with application to topology-constrained neuro-sensor networks 96%
- ScoreNet: A Neural network-based post-processing model for identifying epileptic seizure onset and offset in EEGs 94%
Similar papers in this journal
- Hardware-Efficient Compression of Neural Multi-Unit Activity 94%
- Measuring Connectivity in Linear Multivariate Processes with Penalized Regression Techniques 94%
- HeartNet: Self Multi-Head Attention Mechanism via Convolutional Network with Adversarial Data Synthesis for ECG-based Arrhythmia Classification 94%
Similar papers in this journal
- Resource-efficient Neural Network Architectures forClassifying Nerve Cuff Recordings on Implantable Devices 97%
- Assessing the robustness of deep learning based brain age prediction models across multiple EEG datasets 94%
- Identifiability analysis and noninvasive online estimation of the first-order neural activation dynamics in the brain with closed-loop transcranial magnetic stimulation 93%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.