AppleAnalysisAI ModelsSeptember 30, 2026

Apple study finds activation steering hurts fluency in LLMs

Read original source →machinelearning.apple.com

Apple researchers systematically compared LLM conditioning methods for concept injection and removal, finding efficient activation steering achieves control at a steep cost to fluency. Steering is also far less effective on instruction-tuned models than on base models, while prompting and supervised fine-tuning work for injection but not removal.

1 source

More stories today

Open the live feed