Apple proposes integrated enlarge-and-prune pipeline for LLM pretraining

Apple researchers propose IDEA Prune, an integrated enlarge-and-prune pipeline that combines enlarged model training, pruning, and recovery under a single cosine annealing schedule. Experiments compressing 2.8B models to 1.3B with up to 2T tokens show superior pruned model performance.
1 source
Apple by email
Get an email when Apple has news
No news that day, no email.
More stories today
- Apple's Luce generates relightable 3D assets from single images
- Claude gets its own browser in Cowork
- Steve Case discusses AI buildout and Nvidia's role
- DHH discusses AI agents, vibe coding, and the future of programming on Lex Fridman Podcast
- Dwarkesh Patel interviews Ryan Greenblatt on AI and politics