Language Models Can Control Their Own Attention

Declarative Attention lets language models declare relevant context regions during reasoning, skipping most KV cache reads and reducing attended tokens with small accuracy trade-offs. The paper is on arXiv.
2 sources
More stories today
Users report GPT-5.6 Sol quality shifts after updates
Reddit users report GPT-5.6 Sol in ChatGPT feels 'lobotomized' or 'nerfed' since the 08/06/2026 update and Astra's release, while others claim it was secretly upgraded. Complaints cite reduced reasoning depth and laziness on coding and research tasks.
r/ChatGPT·1 hour ago
Perplexity details GPU embedding stack: Ivy, Tulip, ROSE
Perplexity published research on its serving infrastructure for pplx-embed, combining Ivy, Tulip, and ROSE to lower latency and improve throughput across online and batch workloads. The stack reduces cost compared to off-the-shelf solutions.
MarkTechPost·3 hours ago

Plus users share tips to avoid burning ASTRA's 5-hour limit
A Reddit user on r/OpenAI advises Plus subscribers to use ASTRA for planning and architecture, then switch to Sol for coding, warning that ASTRA can consume the 5-hour limit in as little as one prompt and 10 minutes.
r/OpenAI·5 hours agoAI Models by email
Get an email when there's news on AI Models
No news that day, no email.
Minimax NSFW quality questioned in ComfyUI
A Reddit user asks if anyone has achieved decent NSFW content with Minimax comparable to Wan2.2, noting they've tried various LoRAs and a newer checkpoint without success.
r/ComfyUI·5 hours agoGreg Brockman touts Astra for checking scientific papers and medicine
OpenAI co-founder Greg Brockman posted two tweets highlighting Astra's use in checking scientific papers and in medical applications. No further details were provided.
Greg Brockman·6 hours agoUC Berkeley releases CUA-Lite, an open platform for computer-use agents
CUA-Lite unifies sandboxes, data, evaluation, and reinforcement learning for computer-use agents. The platform is open-source and designed to address the infrastructural challenges of training and benchmarking CUAs.
MarkTechPost·6 hours ago
Long-context retrieval improves, context rot nearly solved
A Reddit user highlights major improvements in long-context retrieval in under a year, suggesting context rot is nearly solved. The post has 44 upvotes and 6 comments.
r/Singularity·7 hours ago
Reddit users ask ChatGPT to fire them in viral prompt
A Reddit prompt asks ChatGPT to act as a boss and give one reason for firing the user, based on chat history. The thread has 30 upvotes and 33 comments.
r/ChatGPT·7 hours ago