Highly Efficient 8-bit Low Precision Inference of Convolutional Neural Networks with IntelCaffe.
Jiong Gong, Haihao Shen, Guoming Zhang, Xiaoli Liu, Shane Li, Ge Jin, Niharika Maheshwari, Evarist Fomenko, Eden Segal
Browse the full ASPLOS paper archive.