Papers
Topics
Authors
Recent
Search
2000 character limit reached

High-Dimensional Sparse Linear Bandits

Published 8 Nov 2020 in stat.ML, cs.LG, math.ST, and stat.TH | (2011.04020v2)

Abstract: Stochastic linear bandits with high-dimensional sparse features are a practical model for a variety of domains, including personalized medicine and online advertising. We derive a novel Ω(n<sup>2/3)\Omega(n<sup>{2/3}) dimension-free minimax regret lower bound for sparse linear bandits in the data-poor regime where the horizon is smaller than the ambient dimension and where the feature vectors admit a well-conditioned exploration distribution. This is complemented by a nearly matching upper bound for an explore-then-commit algorithm showing that that Θ(n<sup>2/3)\Theta(n<sup>{2/3}) is the optimal rate in the data-poor regime. The results complement existing bounds for the data-rich regime and provide another example where carefully balancing the trade-off between information and regret is necessary. Finally, we prove a dimension-free O(n)O(\sqrt{n}) regret upper bound under an additional assumption on the magnitude of the signal for relevant features.

Authors (3)
Citations (53)

Summary

No one has generated a summary of this paper yet.

Paper to Video (Beta)

No one has generated a video about this paper yet.

Whiteboard

No one has generated a whiteboard explanation for this paper yet.

Open Problems

We haven't generated a list of open problems mentioned in this paper yet.

Continue Learning

We haven't generated follow-up questions for this paper yet.