AnalysisAI ModelsSeptember 25, 2026

Kev 4B matches Jev within 2 points on 362-question test

Read original source →opper.ai

Opper hosted Jared Palmer's Apache-2.0 Kev 4B (a Qwen3.5-4B fine-tune) behind the same endpoint as TypeSafe's Jev and found the two land within 2 points on every task across 362 questions published after both shipped. Jev counts a fixed ~257 extra input tokens per request; Kev answered in ~220 ms vs Jev's ~275 ms.

1 source

More stories today

Open the live feed