AppleAnalysisAI ModelsAugust 18, 2026

Apple study finds GRPO training works in non-English languages

Apple ML Research's large-scale study tests GRPO-based RLVR across many base models and languages, finding native-language reasoning training leaves only a small gap to English. It also shows strong crosslingual transfer, but warns that some languages cause severe out-of-domain regressions, requiring broad evaluation.

1 source

Apple by email

Get an email when Apple has news

No news that day, no email.

More stories today

Open the live feed