← Back Systems — Paper Notes Distributed systems and systems for ML 2014 OSDI-2014 Scaling Distributed Machine Learning with the Parameter Server OSDI Systems 2014-10-06 distributed-dataflow 2019 NeurIPS-2019 GPipe:Efficient Training of Giant Neural Networks using Pipeline Parallelism NeurIPS Systems 2019-12-06 pipeline-parallelism 2020 SC-2020 ZeRO:Memory Optimizations Toward Training Trillion Parameter Models SC Systems 2020-11-09 distributed-dataflow GTC-2020 Megatron-LM:Training Multi-Billion Parameter Language Models Using Model Parallelism GTC Systems 2020-03-13 model-parallelism 2022 MLSys-2022 Pathways:Asynchronous Distributed Dataflow for ML MLSys Systems 2022-03-23 distributed-dataflow ← Browse all notes