
Community Discussion · Policy
Stacking GPUs Won't Cure AI's Anemia; What's Missing is a Storage 'Amplifier'
In 2024, in AI inference scenarios, the average GPU utilization of a single H100 was only 38%, while over 60% of the time was spent waiting for data to be moved from memory to compute units. These are real statistics from a cloud provider's internal reports, not theoretical limits. What's even more painful is that about 70% of token generation costs are consumed by the I/O path—putting these numbers together points to a simple conclusion: The GPUs you bought for hundreds of thousands are mostly "slacking off."
Physix Frontier