Towards Autonomous Intra-cortical Brain Machine Interfaces: Applying Bandit Algorithms for Online Reinforcement Learning
Shaikh, S.; So, R.; Sibindi, T.; Libedinsky, C.; Basu, A.
Show abstract
This paper presents application of Banditron - an online reinforcement learning algorithm (RL) in a discrete state intra-cortical Brain Machine Interface (iBMI) setting. We have analyzed two datasets from non-human primates (NHPs) - NHP A and NHP B each performing a 4-option discrete control task over a total of 8 days. Results show average improvements of {approx} 15%, 6% in NHP A and 15%, 21% in NHP B over state of the art algorithms - Hebbian Reinforcement Learning (HRL) and Attention Gated Reinforcement Learning (AGREL) respectively. Apart from yielding a superior decoding performance, Banditron is also the most computationally friendly as it requires two orders of magnitude less multiply-and-accumulate operations than HRL and AGREL. Furthermore, Banditron provides average improvements of at least 40%, 15% in NHPs A, B respectively compared to popularly employed supervised methods - LDA, SVM across test days. These results pave the way towards an alternate paradigm of temporally robust hardware friendly reinforcement learning based iBMIs.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A Spiking Neural Network with Continuous Local Learning for Robust Online Brain Machine Interface 96%
- Speech decoding from a small set of spatially segregated minimally invasive intracranial EEG electrodes with a compact and interpretable neural network 96%
- Scalability of Random Forest in Myoelectric Control 95%
Similar papers in this journal
Similar papers in this journal
- Sparse Ensemble Machine Learning to improve robustness of long-term decoding in iBMIs 97%
- ScoreNet: A Neural network-based post-processing model for identifying epileptic seizure onset and offset in EEGs 94%
- Optimal versus approximate channel selection methods for EEG decoding with application to topology-constrained neuro-sensor networks 94%
Similar papers in this journal
- HeartNet: Self Multi-Head Attention Mechanism via Convolutional Network with Adversarial Data Synthesis for ECG-based Arrhythmia Classification 95%
- Measuring Connectivity in Linear Multivariate Processes with Penalized Regression Techniques 94%
- Hardware-Efficient Compression of Neural Multi-Unit Activity 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.