AppleAnalysisAI ModelsOctober 7, 2026

Apple researchers propose RISED rubric-based agent training

Read original source →machinelearning.apple.com

RISED repurposes rubrics as both an online data-selection signal and a supervision source for multi-environment LLM agent training, instead of relying only on scalar rewards. An LLM judge tags each rollout with a shared rubric vocabulary; positive rubrics feed an on-policy self-distillation teacher, negative ones steer later rollouts away from repeated failures.

1 source

More stories today

Open the live feed