Researcher profile

Amirkeivan Mohtashami

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. Critical Parameters for Scalable Distributed Learning with Large Batches and Asynchronous Updates

    2021 · arXiv (Cornell University)

    It has been experimentally observed that the efficiency of distributed training with stochastic gradient (SGD) depends decisively on the batch size and -- in asynchronous implementations -- on the gradient staleness. Especially, it has been …