AnalysisDevelopersSeptember 23, 2026

Nebius benchmarks NVIDIA Cosmos 3 Super serving via vLLM-Omni

Nebius Physical AI team members Stewart Tong and Timothy Le examine NVIDIA Cosmos 3 Super benchmarks through vLLM-Omni on NVIDIA HGX B200 and HGX systems. The livestream covers how serving world models for video generation requires a different latency-versus-throughput balance than serving large language models.

1 source

More stories today

Open the live feed