Stochastic Dual Coordinate Ascent and its Proximal Extension for Regularized Loss Minimization

2 249

124.9

Microsoft Research336 тыс

Следующее

28.07.16 – 17359:32

Approximating the Expansion Profile and Almost Optimal Local Graph Clustering

Популярные

60 дней – 6722:27

Low latency carbon budget 2023

304 дня – 99 70716:22

Research Forum 2 | Keynote: The Revolution in Scientific Discovery

Опубликовано 28 июля 2016, 23:16

Stochastic Gradient Descent (SGD) has become popular for solving large scale supervised machine learning optimization problems such as SVM, due to their strong theoretical guarantees. While the closely related Dual Coordinate Ascent (DCA) method has been implemented in various software packages, it has so far lacked good convergence analysis. We present a new analysis of Stochastic Dual Coordinate Ascent (SDCA) showing that this class of methods enjoy strong theoretical guarantees that are comparable or better than SGD. This analysis justifies the effectiveness of SDCA for practical applications. Moreover, we introduce a proximal version of dual coordinate ascent method. We demonstrate how the derived algorithmic framework can be used for numerous regularized loss minimization problems, including L1 regularization and structured output SVM. The convergence rates we obtain match or improve state-of-the-art results. Joint work with Shai Shalev-Shwartz

Свежие видео