Papers
Topics
Authors
Recent
Search
2000 character limit reached

Efficient and Robust Algorithms for Adversarial Linear Contextual Bandits

Published 1 Feb 2020 in cs.LG and stat.ML | (2002.00287v3)

Abstract: We consider an adversarial variant of the classic KK-armed linear contextual bandit problem where the sequence of loss functions associated with each arm are allowed to change without restriction over time. Under the assumption that the dd-dimensional contexts are generated i.i.d.~at random from a known distributions, we develop computationally efficient algorithms based on the classic Exp3 algorithm. Our first algorithm, RealLinExp3, is shown to achieve a regret guarantee of O~(KdT)\widetilde{O}(\sqrt{KdT}) over TT rounds, which matches the best available bound for this problem. Our second algorithm, RobustLinExp3, is shown to be robust to misspecification, in that it achieves a regret bound of O~((Kd)<sup>1/3T<sup>2/3)</sup></sup>+εdT\widetilde{O}((Kd)<sup>{1/3}T<sup>{2/3})</sup></sup> + \varepsilon \sqrt{d} T if the true reward function is linear up to an additive nonlinear error uniformly bounded in absolute value by ε\varepsilon. To our knowledge, our performance guarantees constitute the very first results on this problem setting.

Authors (2)
Citations (42)

Summary

No one has generated a summary of this paper yet.

Paper to Video (Beta)

No one has generated a video about this paper yet.

Whiteboard

No one has generated a whiteboard explanation for this paper yet.

Open Problems

We haven't generated a list of open problems mentioned in this paper yet.

Continue Learning

We haven't generated follow-up questions for this paper yet.