Comment on The Nvidia DGX-1 Deep Learning Supercomputer in a BoxComments−Coding_Cat10yWait, how many chips did they cram in there that they're getting 170 TFlops. Even at a very generous 10 TFLOP per chip that would be 17 chips.−krasin10yNVIDIA Tesla P100 has 21 TeraFLOPS of FP16 performance by their words. So they got 8 chips there.−jsheard10yYep, they showed a diagram of how it fits together: http://i.imgur.com/xk1daFG.jpg−aconz210yhttps://devblogs.nvidia.com/parallelforall/wp-content/upload...source: https://devblogs.nvidia.com/parallelforall/inside-pascal/−cptskippy10yI wish that made that information more accessible. I wasn't able to find it on the site and it was all I really cared about.−Coding_Cat10yAh, half-floats. That explains it. Still pretty high but realistic at least.
Comments
Wait, how many chips did they cram in there that they're getting 170 TFlops. Even at a very generous 10 TFLOP per chip that would be 17 chips.
NVIDIA Tesla P100 has 21 TeraFLOPS of FP16 performance by their words. So they got 8 chips there.
Yep, they showed a diagram of how it fits together: http://i.imgur.com/xk1daFG.jpg
https://devblogs.nvidia.com/parallelforall/wp-content/upload...
source: https://devblogs.nvidia.com/parallelforall/inside-pascal/
I wish that made that information more accessible. I wasn't able to find it on the site and it was all I really cared about.
Ah, half-floats. That explains it. Still pretty high but realistic at least.