I present a new way to parallelize the training of convolutional neural networks across multiple GPUs. The method scales significantly better than all alternatives when applied to modern convolutional neural networks.
No takes yet. Share an insight, caveat, or question.
Alex Krizhevsky (2014) studied this question.
Synapse has enriched 3 closely related papers on similar clinical questions. Consider them for comparative context: