It's worth pointing out that we have been pushing on large-scale training and asynchronous technique...

TL;DR · AI 摘要
核心要点
Jeff Dean on X: "It's worth pointing out that we have been pushing on large-scale training and asynchronous techniques for the last ~14 years. Here's our NeurIPS 2012 paper where we demonstrated that this approach could be used to train very large neural networks (for the time: 30X larger than" / X
Don’t miss what’s happening

It's worth pointing out that we have been pushing on large-scale training and asynchronous techniques for the last ~14 years. Here's our NeurIPS 2012 paper where we demonstrated that this approach could be used to train very large neural networks (for the time: 30X larger than any previous neural network), and to spread the training out across thousands of machines in a fault-tolerant manner. PDF: https://static.googleusercontent.com/media/research.google.com/en//archive/large_deep_networks_nips2012.pdf…https://static.googleusercontent.com/media/research.google.com/en//archive/large_deep_networks_nips2012.pdf… (This paper doesn't get as much attention because we neglected to put it on Arxiv at the time: oops!)
·
3
1
23
5