Stealth model Ox Alpha matches GPT-5.6 Sol mid on DeepSWE
Ox Alpha scored 96% on SWE-bench Verified Mini (48/50 tasks) and over 80% on 10 DeepSWE tasks, beating Fable (65%) and GPT-5.6-sol (52%). It performs on par with GPT-5.6 Sol mid, possibly a Chinese model like GLM-5.3 Flash.
How this story unfolded
2 days · 0 reports · 4 community posts · from Aug 21
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Qwen releases cua-driver-rs v0.20.0 with prebuilt binaries
- AI bots flood social media with generic replies
- User connects Codex to Fusion 360 via MCP for 3D modeling
- AI companion plays Skyrim with you in real time
- Seinfeld AI video shows George in GTA 6 using Minimax H3