Nvidia's Blackwell worsened matrix-vector ratio to 32:1, neglecting inference efficiency
Nvidia's architectural focus on dense matrix-matrix training workloads caused the matrix-vector performance ratio to degrade from 16:1 on Hopper to 32:1 on Blackwell, even as tota…