AnalysisBusinessSeptember 22, 2026

Ornn paper: open-weight inference keeps older Nvidia GPUs earning

Across 11 open-weight and 8 closed models on the Artificial Analysis Intelligence Index, the cheapest open-weight model completes a task at roughly one fifth the cost of a comparable closed model. Self-hosting on rented hardware drops to $0.12-$0.35 per million output tokens at full utilization, and on gpt-oss-120b the A100 beats the H100 on output cost.

1 source

More stories today

Open the live feed