Skip to main navigation Skip to search Skip to main content

Optimal algorithms for non-smooth distributed optimization in networks

  • Kevin Scaman
  • , Francis Bach
  • , Sébastien Bubeck
  • , Yin Tat Lee
  • , Laurent Massoulié
  • Huawei Noah's Ark Lab
  • Université PSL
  • Microsoft Research
  • University of Washington
  • MSR-Inria Joint Centre

Research output: Contribution to journalConference articlepeer-review

121 Citations (Scopus)

Abstract

In this work, we consider the distributed optimization of non-smooth convex functions using a network of computing units. We investigate this problem under two regularity assumptions: (1) the Lipschitz continuity of the global objective function, and (2) the Lipschitz continuity of local individual functions. Under the local regularity assumption, we provide the first optimal first-order decentralized algorithm called multi-step primal-dual (MSPD) and its corresponding optimal convergence rate. A notable aspect of this result is that, for non-smooth functions, while the dominant term of the error is in O(1/t), the structure of the communication network only impacts a second-order term in O(1/t), where t is time. In other words, the error due to limits in communication resources decreases at a fast rate even in the case of non-strongly-convex objective functions. Under the global regularity assumption, we provide a simple yet efficient algorithm called distributed randomized smoothing (DRS) based on a local smoothing of the objective function, and show that DRS is within a d1/4 multiplicative factor of the optimal convergence rate, where d is the underlying dimension.

Original languageEnglish
Pages (from-to)2740-2749
Number of pages10
JournalAdvances in Neural Information Processing Systems
Volume2018-December
Publication statusPublished - 1 Jan 2018
Externally publishedYes
Event32nd Conference on Neural Information Processing Systems, NeurIPS 2018 - Montreal, Canada
Duration: 2 Dec 20188 Dec 2018

Fingerprint

Dive into the research topics of 'Optimal algorithms for non-smooth distributed optimization in networks'. Together they form a unique fingerprint.

Cite this