Papers
Topics
Authors
Recent
Search
2000 character limit reached

Best-item Learning in Random Utility Models with Subset Choices

Published 19 Feb 2020 in cs.LG, cs.AI, and stat.ML | (2002.07994v1)

Abstract: We consider the problem of PAC learning the most valuable item from a pool of nn items using sequential, adaptively chosen plays of subsets of kk items, when, upon playing a subset, the learner receives relative feedback sampled according to a general Random Utility Model (RUM) with independent noise perturbations to the latent item utilities. We identify a new property of such a RUM, termed the minimum advantage, that helps in characterizing the complexity of separating pairs of items based on their relative win/loss empirical counts, and can be bounded as a function of the noise distribution alone. We give a learning algorithm for general RUMs, based on pairwise relative counts of items and hierarchical elimination, along with a new PAC sample complexity guarantee of O(nc<sup>2ϵ<sup>2</sup></sup>logkδ)O(\frac{n}{c<sup>2\epsilon<sup>2}</sup></sup> \log \frac{k}{\delta}) rounds to identify an ϵ\epsilon-optimal item with confidence 1δ1-\delta, when the worst case pairwise advantage in the RUM has sensitivity at least cc to the parameter gaps of items. Fundamental lower bounds on PAC sample complexity show that this is near-optimal in terms of its dependence on n,kn,k and cc.

Authors (2)
Citations (8)

Summary

No one has generated a summary of this paper yet.

Paper to Video (Beta)

No one has generated a video about this paper yet.

Whiteboard

No one has generated a whiteboard explanation for this paper yet.

Open Problems

We haven't generated a list of open problems mentioned in this paper yet.

Continue Learning

We haven't generated follow-up questions for this paper yet.