Loading slide

The GPU Moment

  1. 01The GPU Moment
  2. 02One thing first
  3. 03Why training is slow
  4. 04What a GPU actually does
  5. 05The realization
  6. 06The scale this enabled
  7. 07Something unexpected
  8. 08Take a breath
  9. 09Data and compute
  10. 10Memory card
  11. 11Reinforce your understanding
  12. 12Question: Why GPUs and not faster CPUs?
  13. 13Question: Three ingredients
  14. 14Question: Why the field sped up
  15. 15Quiz: answer
  16. 16The hardware was ready. The data existed.
  17. 17Want to go deeper?
6 / 16
BackNext

The scale this enabled

The speedup didn't just make the old work faster. It opened up work that had never been possible at all.

Before GPU training, you could train a network on tens of thousands of examples. Now you could train on millions. You could build networks dozens of layers deep instead of two or three. You could run in an afternoon an experiment that would once have taken months, which meant you could try ten ideas in the time it used to take to try one.

That last part matters more than it sounds. Research moves at the speed of how fast you can test an idea and see if it worked. Collapse the training time from months to hours, and the whole field speeds up with it. More people, trying more things, learning faster.

So the obstacles from the last chapter fell away. The deep networks that had been too slow to train were suddenly practical. And the moment researchers could actually build them at size, they noticed something nobody had predicted.

Citations(2)↓
  1. 1. dl.acm.org
  2. 2. ieeexplore.ieee.org
Citations(2)↓
  1. 1. dl.acm.org
  2. 2. ieeexplore.ieee.org