AnalysisAI ModelsAugust 18, 2026

Qwen3.8-27B hits 124 tps on RTX 3090

A developer pushed Qwen3.8-27B to 124 tokens per second on a single RTX 3090, up from 82 tps in the initial release and 99 tps in yesterday's update. The engine also reached 672 peak tps.

1 source

AI Models by email

Get an email when there's news on AI Models

No news that day, no email.

More stories today

Open the live feed
Qwen3.8-27B hits 124 tps on RTX 3090 — AIBriefs