论文信息 - Stochastic Gradient Methods for Distributionally Robust Optimization with f-divergences

Stochastic Gradient Methods for Distributionally Robust Optimization with f-divergences

We develop efficient solution methods for a robust empirical risk minimization problem designed to give calibrated confidence intervals on performance and provide optimal tradeoffs between bias and variance. Our methods apply to distributionally robust optimization problems proposed by Ben-Tal et al., which put more weight on observations inducing high loss via a worst-case approach over a non-parametric uncertainty set on the underlying data distribution. Our algorithm solves the resulting minimax problems with nearly the same computational cost of stochastic gradient descent through the use of several carefully designed data structures. For a sample of size n, the per-iteration cost of our method scales as O(log n), which allows us to give optimality certificates that distributionally robust optimization provides at little extra cost compared to empirical risk minimization and stochastic gradient methods.

John C. Duchi | Hongseok Namkoong | Hongseok Namkoong

[1] Martin Zinkevich,et al. Online Convex Programming and Generalized Infinitesimal Gradient Ascent , 2003, ICML.

[2] Yonatan Wexler,et al. Minimizing the Maximal Loss: How and Why , 2016, ICML.

[3] Xin-She Yang,et al. Introduction to Algorithms , 2021, Nature-Inspired Optimization Algorithms.

[4] J. Borwein,et al. Uniformly convex functions on Banach spaces , 2008 .

[5] Timothy R. C. Read,et al. Multinomial goodness-of-fit tests , 1984 .

[6] Elad Hazan. The convex optimization approach to regret minimization , 2011 .

[7] Shie Mannor,et al. Oracle-Based Robust Optimization via Online Learning , 2014, Oper. Res..

[8] Jean-Yves Audibert,et al. Regret Bounds and Minimax Policies under Partial Monitoring , 2010, J. Mach. Learn. Res..

[9] Alexander Shapiro,et al. Stochastic Approximation approach to Stochastic Programming , 2013 .

[10] Gábor Lugosi,et al. Prediction, learning, and games , 2006 .

[11] S. Boucheron,et al. Theory of classification : a survey of some recent advances , 2005 .