Yesterday, a video was released in which Jen-Hsun Huang, CEO of NVIDIA, unveiled a new HPC platform based on Ampere. This new solution is specifically designed for AI, Deep Learning, Machine Learning, and related technologies. Now we've learned more about the silicon integrated into this platform, just hours before the start of GTC 2020.
Specifically, a rendered image of the NVIDIA Tesla A100 has been leaked, which is accompanied by 6 stack of HBM2 memory. This Tesla A100 solution has been developed for the industrial and business sector. This solution is based on the Ampere architecture, which promises a brutal leap from the current Tesla architecture.
[amazon box="B07JBQ4DBV"]First image of the NVIDIA Tesla A100
These NVIDIA solutions are typically used in data centers, high-performance computing systems, robotics, AI, and a host of specialized fields where high computing power is required. Like the solutions based on Volta, the solutions based on Ampere will be a reference in these sectors.
The Tesla A100 unit is based on the GA100 silicon chip with Ampere architecture and a 6-stack HBM2 memory chip. It appears to have been redesigned, as the mounting holes do not align with those of the Tesla V100. Furthermore, the die seems to be slightly larger than that of the GV100, which is approximately 820-840 mm².
Something very interesting that has been revealed is that the GA100 GPU will have 54.000 million transistors and has been manufactured in 7nm. What we do not know is whether this lithograph is TSMC's or Samsung's, something that has not been indicated at the moment. We also do not know if the memories are HBM2 or HBM2E and the size of each of the stacks.

Source: VZ