AnalysisAI ModelsOctober 7, 2026

Benchmark tests LLMs writing in 1866 telegraphese

Read original source →fiveminutesforward.com

A 50-passage, ~1,300-question benchmark found GLM-5.3-Flash cut token use 48.4% with a lowercase instruction, and foreign models answered from compressed records at 0.99–1.10 recovery ratios. Compression varied 25%–49% across models; gpt-5-mini's writes billed about double plaintext.

1 source

More stories today

Open the live feed