Neural gradients are near-lognormal: improved quantized and sparse training.
Brian Chmiel, Liad Ben-Uri, Moran Shkolnik, Elad Hoffer, Ron Banner, Daniel Soudry
Browse the full ICLR paper archive.
Brian Chmiel, Liad Ben-Uri, Moran Shkolnik, Elad Hoffer, Ron Banner, Daniel Soudry
Browse the full ICLR paper archive.