AnalysisAI ModelsAugust 28, 2026

New papers and llama.cpp support advance speculative decoding

Four arXiv papers propose speculative decoding methods: TreeGraft, AgentSpec, LiLiCorr, and Self-Speculation. Meanwhile, llama.cpp merged support for DFlash2, a diffusion-style block head drafter.

How this story unfolded

4 days · 4 reports · 1 community post · from Aug 24

  1. Aug 24
  2. Aug 26
  3. Aug 27
  4. Aug 28

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed