GPU rental prices doubled in six months while inference costs kept falling—efficiency is the hinge
B200 GPU rental hit $8.08/hr, doubling in six months. Meanwhile Claude Opus 5.5 runs 40% cheaper than its predecessor; OpenAI slashed Luna pricing 80% in July and another 50% in September. A benchmark that cost $0.55 18 months ago now clears for $0.0015—a 377x drop. Tunguz argues efficiency gains are offsetting hardware cost inflation, with the two curves running neck and neck for now. The post doesn't predict whether efficiency can keep outpacing GPU price hikes, but says gross profit per GPU-hour is the metric to watch.
Why it matters: Tunguz lays out the parallel logic of hardware scarcity vs. software efficiency with two clean data lines: B200 rent doubling and inference cost dropping 377x. Concrete numbers plus the Oracle New Mexico force majeure anecdote ground it. Not scored higher because it's an expla...