AnalysisAI ModelsSeptember 28, 2026

Reddit user builds finger-based t/s counter for local LLM jokes

Read original source →reddit.com

A r/LocalLLaMA user made a typing-speed counter using real tokenizers after tiring of optimizing local model throughput. Their best is about 2 t/s, which the page says beats a 70B on a laptop CPU and is roughly 76x slower than an 8B model.

1 source

More stories today

Open the live feed