Robust Contextual Bandit via the Capped-$$\ell _{2}$$ Norm for Mobile Health Intervention

Autor:	Feiyun Zhu, Jiawen Yao, Junzhou Huang, Sheng Wang, Zhichun Xiao, Xinliang Zhu
Rok vydání:	2018
Předmět:	Computer science business.industry Gaussian Machine learning computer.software_genre Health intervention symbols.namesake Robustness (computer science) Approximation error Norm (mathematics) Outlier symbols Artificial intelligence business mHealth computer
Zdroj:	Machine Learning in Medical Imaging ISBN: 9783030009182 MLMI@MICCAI
DOI:	10.1007/978-3-030-00919-9_2
Popis:	This paper considers the actor-critic contextual bandit for the mobile health (mHealth) intervention. The state-of-the-art decision-making methods in the mHealth generally assume that the noise in the dynamic system follows the Gaussian distribution. Those methods use the least-square-based algorithm to estimate the expected reward, which is prone to the existence of outliers. To deal with the issue of outliers, we are the first to propose a novel robust actor-critic contextual bandit method for the mHealth intervention. In the critic updating, the capped-$\ell _{2}$ norm is used to measure the approximation error, which prevents outliers from dominating our objective. A set of weights could be achieved from the critic updating. Considering them gives a weighted objective for the actor updating. It provides the ineffective sample in the critic updating with zero weights for the actor updating. As a result, the robustness of both actor-critic updating is enhanced. There is a key parameter in the capped-$\ell _{2}$ norm. We provide a reliable method to properly set it by making use of one of the most fundamental definitions of outliers in statistics. Extensive experiment results demonstrate that our method can achieve almost identical results compared with the state-of-the-art methods on the dataset without outliers and dramatically outperform them on the datasets noised by outliers.
Databáze:	OpenAIRE
Externí odkaz:	https://explore.openaire.eu/search/publication?articleId=doi_________::75cda879912988dd219ec6d7e9c461df https://doi.org/10.1007/978-3-030-00919-9_2 Zobrazit plný text záznamu