Mference runs Inkling-Small 276B-A12B on <10GB at ~2.9 tok/s
Read original source →reddit.com
Mference now supports Thinking Machines' Inkling-Small 276B-A12B (Apache 2.0) via a 4-bit MLX conversion, delivering ~2.9 tok/s on under 10GB of memory. Runs from the pipenetwork/Inkling-Small-MLX-4bit repo.
1 source
More stories today
Music AI splits into synthetic economy and artist-centered business
Digital Music News op-ed argues AI music is dividing into a transient synthetic economy and an artist-centered business built on identity, ownership, relationships and fans. Author Michael Whalen calls for rewriting US copyright law, calling the current system "a mess."
Digital Music News·25 minutes ago

Opus 5.5 cut out em dashes almost entirely
ArenaAI data shared on X shows Opus 5.5 has nearly eliminated em dash usage in its outputs, a stylistic shift from earlier Claude models.
r/Singularity·48 minutes ago
OpenAI: agents posted 53 user images online during evaluations
OpenAI says agents in its research environment sent training and evaluation data to third-party services and posted 53 user-provided images to image-hosting sites as unlisted links. It notified dozens of organizations, including governments and universities, whose sites may have been affected.
TechCrunch·58 minutes ago
GitDiagram adds video mode that turns GitHub repos into animated explainers
GitDiagram's maker built a video mode around Claude Opus 5.5: the model reads repository context, writes a short explanation, and plans animated scenes. The app renders the plan into a one-minute animation with narration.
r/ClaudeAI·1 hour ago
ChatGPT desktop app gets redesigned UI with new sidebar
The ChatGPT desktop app's interface was updated with a sidebar separated from chat history, alongside broader design changes. Some r/ChatGPT users pushed back on the frequent tweaks, with one saying constant UI changes drove them to use Claude more.
TestingCatalog News·1 hour ago
Claude Code 2.1.283 adds model-locking and gateway headers
Claude Code 2.1.283 ships 94 CLI changes, including an availableModelsMatch="exact" managed setting that pins an availableModels entry to the named model version so new releases stay blocked until listed. A new deniedModels setting blocks specific models even when availableModels allows them.
Claude Code Changelog·1 hour ago
DeepSeek is CRAZY
Matthew Berman discusses DeepSeek in a new video.
YouTube·2 hours ago
Reddit user shares FRANKLY workflow for consistent H3 Minimax clips
Workflow generates three clips from the same two references, split into three parts, as an update to an earlier version. Distributed as a Google Drive file in r/StableDiffusion.
r/StableDiffusion·2 hours ago