arrow
返回

The value-complexity trade-off for reinforcement learning based brain-computer interfaces

delete2020-12-16
delete1
PRE
AI
H
Hadar Levi-Aharoni *
N
Naftali Tishby
DOI:10.1088/1741-2552/abc8d8delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Objective. One of the recent developments in the field of brain-computer interfaces (BCI) is the reinforcement learning (RL) based BCI paradigm, which uses neural error responses as the reward feedback on the agent's action. While having several advantages over motor imagery based BCI, the reliability of RL-BCI is critically dependent on the decoding accuracy of noisy neural error signals. A principled method is needed to optimally handle this inherent noise under general conditions. Approach. By determining a trade-off between the expected value and the informational cost of policies, the info-RL (IRL) algorithm provides optimal low-complexity policies, which are robust under noisy reward conditions and achieve the maximal obtainable value. In this work we utilize the IRL algorithm to characterize the maximal obtainable value under different noise levels, which in turn is used to extract the optimal robust policy for each noise level. Main results. Our simulation results of a setting with Gaussian noise show that the complexity level of the optimal policy is dependent on the reward magnitude but not on the reward variance, whereas the variance determines whether a lower complexity solution is favorable or not. We show how this analysis can be utilized to select optimal robust policies for an RL-BCI and demonstrate its use on EEG data. Significance. We propose here a principled method to determine the optimal policy complexity of an RL problem with a noisy reward, which we argue is particularly useful for RL-based BCI paradigms. This framework may be used to minimize initial training time and allow for a more dynamic and robust shared control between the agent and the operator under different conditions.
Keyword:
brain– computer interface
EEG
ErrP
noise
reinforcement learning
control information
complexity

期刊

Journal of Neural Engineering 封面图
Journal of Neural Engineering
IF:
3.8
论文数:
4.0K
被引数:
1.4W

机构

H
Hebrew University of Jerusalem
学者数:
2.8W
论文数: 2.3W
被引数: 2.7W
引用论文

引用论文

Unsupervised Learning for Brain-Computer Interfaces Based on Event-Related Potentials: Review and Online Comparison
err2018-05-01
err18
PREAI
errHuebner, David; Verhoeven, Thibault; Mueller, Klaus-Robert; Kindermans, Pieter-Jan; Tangermann, Michael
err分享
err收藏
Online use of error-related potentials in healthy users and people with severe motor impairment increases performance of a P300-BCI
err2012-07-01
err126
PREAI
errSpueler, Martin; Bensch, Michael; Kleih, Sonja; Rosenstiel, Wolfgang; Bogdan, Martin; Kuebler, Andrea
err分享
err收藏
Benefit of baseline cytometry for surveillance of patients with Barrett’s esophagus
err2009-12-08
err0
PREAI
errNicole Vogt; René Schönegg; Jürgen M. Gschossmann; Jan Borovicka
err分享
err收藏
Teaching brain-machine interfaces as an alternative paradigm to neuroprosthetics control
err2015-09-10
err124
errOAAI
errIturrate, Inaki; Chavarriaga, Ricardo; Montesano, Luis; Minguez, Javier; Millan, Jose del R.
err分享
err收藏
A Review of Error-Related Potential-Based Brain-Computer Interfaces for Motor Impaired People
err2019-01-01
err41
errOAAI
errKumar, Akshay; Gao, Lin; Pirogova, Elena; Fang, Qiang
err分享
err收藏
学者 查看更多内容