4x Stacks of HBM2 means 1TBps memory bandwidth at 16GB. Only 60-compute units enabled (maybe 4-CUs are expected to break during manufacturing? Its a weird number for sure...).
Since it shares dies with the MI50, the Radeon VII will have 1/2 speed double-precision, making this the cheapest high-performance double-precision card in existance.
------
FP16 compute is supported at double speed, but there are no tensor cores. So FP16 matrix multiplication / tensor ops are still a major benefit to NVidia.
But the memory size and bandwidth is quite salivating. That's a lot of bandwidth, and a number of problems are known to be memory-bound. Deep learning enthusiasts probably will stick to NVidia cards, but other compute problems may want to start playing around with this thing.
------
Video gamers seem meh about the specs. But I think anyone looking at this card for its compute performance would be impressed.