Semi-Bandit Learning for Monotone Stochastic Optimization

Agarwal, Arpit; Ghuge, Rohan; Nagarajan, Viswanath

Computer Science > Machine Learning

arXiv:2312.15427 (cs)

[Submitted on 24 Dec 2023]

Title:Semi-Bandit Learning for Monotone Stochastic Optimization

Authors:Arpit Agarwal, Rohan Ghuge, Viswanath Nagarajan

View PDF HTML (experimental)

Abstract:Stochastic optimization is a widely used approach for optimization under uncertainty, where uncertain input parameters are modeled by random variables. Exact or approximation algorithms have been obtained for several fundamental problems in this area. However, a significant limitation of this approach is that it requires full knowledge of the underlying probability distributions. Can we still get good (approximation) algorithms if these distributions are unknown, and the algorithm needs to learn them through repeated interactions? In this paper, we resolve this question for a large class of "monotone" stochastic problems, by providing a generic online learning algorithm with $\sqrt{T \log T}$ regret relative to the best approximation algorithm (under known distributions). Importantly, our online algorithm works in a semi-bandit setting, where in each period, the algorithm only observes samples from the r.v.s that were actually probed. Our framework applies to several fundamental problems in stochastic optimization such as prophet inequality, Pandora's box, stochastic knapsack, stochastic matchings and stochastic submodular optimization.

Subjects:	Machine Learning (cs.LG); Data Structures and Algorithms (cs.DS)
Cite as:	arXiv:2312.15427 [cs.LG]
	(or arXiv:2312.15427v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2312.15427

Submission history

From: Rohan Ghuge [view email]
[v1] Sun, 24 Dec 2023 07:46:37 UTC (142 KB)

Computer Science > Machine Learning

Title:Semi-Bandit Learning for Monotone Stochastic Optimization

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Semi-Bandit Learning for Monotone Stochastic Optimization

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators