AnalysisAI ModelsJuly 8, 2026

Mimo v2.5 outperforms DeepSeek V4 Flash on Terminal Bench

Mimo v2.5 scored 55% on Terminal Bench v2.0 with the Hermes harness, beating DeepSeek V4 Flash and others that scored under 50%, according to a r/LocalLLaMA user's benchmarks. Tests ran via the OpenCode endpoint across Codex, Hermes, and Pi harnesses.

1 source

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed