LaunchDevelopersJuly 8, 2026

Together AI launches Provisioned Throughput for open models

Reserved inference capacity with token-based pricing, 99% uptime SLA, and up to 90% lower cost than Claude Opus 4.8. Available today for MiniMax M3 and GLM-5.2.

2 sources

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed