The company claims that what used to require 10 servers could potentially run on just one.
A 10x server reduction claim is extraordinary and will need rigorous third-party validation before any hyperscaler procurement decision. If even partially true at production scale, the TCO implications for AI inference clusters are massive — but this is precisely the kind of claim that must survive contact with real workloads.