Accelerating SGD for Distributed Deep-Learning Using Approximated Hessian Matrix | Read Paper on Bytez