KONST's published H100 bare-metal price is $2.00 per GPU per hour. The same GPU costs $2.96 in the cloud. These are not two prices for the same product. They assign idle-capacity risk differently.
With bare metal, you rent the entire server and billing starts by the month. You absorb the cost of idle time in exchange for the lowest unit price, exclusive PCIe and InfiniBand bandwidth, and performance unaffected by neighboring workloads. In the cloud, the platform absorbs idle-capacity risk. You pay by the second only while the instance is running, and the rate includes the platform layer, elastic resource pool, and spare capacity kept ready.
The break-even point usually depends on utilization. For teams with predictable workloads and long-term utilization above sixty percent, bare metal or a long-term contract is almost always less expensive. If demand fluctuates, the architecture is still being validated, or peak traffic needs overflow capacity, the cloud can save more in idle costs than the difference in hourly price. Bare-metal terms of three months or longer receive lower rates and can be combined with cloud resources.