Papers
Communities
Events
Blog
Pricing
Search
Open menu
Home
Papers
2112.10065
Cited By
Efficient Strong Scaling Through Burst Parallel Training
19 December 2021
S. Park
Joshua Fried
Sunghyun Kim
Mohammad Alizadeh
Adam Belay
GNN
LRM
Re-assign community
ArXiv
PDF
HTML
Papers citing
"Efficient Strong Scaling Through Burst Parallel Training"
2 / 2 papers shown
Title
MuxFlow: Efficient and Safe GPU Sharing in Large-Scale Production Deep Learning Clusters
Yihao Zhao
Xin Liu
Shufan Liu
Xiang Li
Yibo Zhu
Gang Huang
Xuanzhe Liu
Xin Jin
27
11
0
24 Mar 2023
Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism
M. Shoeybi
M. Patwary
Raul Puri
P. LeGresley
Jared Casper
Bryan Catanzaro
MoE
245
1,817
0
17 Sep 2019
1