AppleAnalysisAI ModelsSeptember 24, 2026

Apple compresses streaming audio encoders via latent-space distillation

Read original source →machinelearning.apple.com

Apple's method distills the pre-quantizer latent rather than discrete tokens or output distributions, keeping the student within 1.9% relative WER of its teacher at 2.8× compression on five of six teacher–student pairs without fine-tuning. It beats an independently trained tokenizer of identical capacity by 3.9% relative.

1 source

More stories today

Open the live feed