Skip to content

Parallel Restarted SGD with Faster Convergence and Less Communication: Demystifying Why Model Averaging Works for Deep Learning.

Hao Yu, Sen Yang, Shenghuo Zhu

VenueA*AAAI
Year2019
ProceedingsAAAI

Browse the full AAAI paper archive.