Jev, TypeSafe AI's decision model, sparks a wave of fine-tunes and integrations
Read original source →together.ai
TypeSafe AI's Jev returns typed answers with probabilities instead of text, and its benchmarks show it running up to 200x faster and 400x cheaper than leading LLMs on narrow decision tasks. Together AI fine-tuned a Jev-like classifier on Qwen3.5 4B for $17, while a 194M GLiNER2 version trained on Hugging Face Jobs cost ~$1.50.
People · Diogo Almeida
How this story unfolded
2 weeks · 28 reports · 66 community posts · 94 of 97 shown
- Sep 15
- Sep 16
- Sep 17
- Sep 18
- Sep 19
- Sep 20
- Sep 21
Jev is incredibleyoutube.com
How to use Jev to automate your business (Step-by-step w/ Treg)youtube.com
An ex-OpenAI researcher just deleted language from the LLM...youtube.com
What People Are Building With Jev Is Incredibleyoutube.com
TypeSafe launched Jev because sequential LLMs are “totally useless for computers”thenewstack.io
Jev: System One models for Prod, not God — with Diogo Almeida, CEO, TypeSafe AIlatent.space
Jev introduces a new shape of LLM - System One, aka Decision Modelssimonwillison.net
- Sep 22
- Sep 23
- Sep 24
- Sep 25
- Sep 28
20 Agentic Use Cases of TypeSafe AI’s Jevmarktechpost.com
Jev in the Wild: A Data-Driven Analysis of the Jev Model's Functionality, Applications and Ecosystemarxiv.org
Jev - The New AI model that has people talkingyoutube.com
Companies are paying LLMs to generate text for decisions that only need a label. Jev offers a cheaper wayventurebeat.com
- Sep 29
More stories today
NVIDIA Dynamo-Triton adds HSTU generative recommender inference
Dynamo-Triton now supports an end-to-end HSTU generative recommender workflow via NVIDIA's recsys-examples repo, combining PyTorch AOTI, FlexKV KV caching, and NV embedding cache. At batch size 8 on an RTX PRO 6000 Blackwell Workstation GPU, it hit up to 4.47x speedup for the three-layer HSTU model and 5.93x for the eight-layer model at 100% KV-cache hit rate.
NVIDIA Developer Blog·1 hour ago

METR's Chris Painter testifies to Senate on AI agent incidents
METR President Chris Painter testified September 30, 2026 before a Senate Homeland Security subcommittee hearing titled "Rogue AI: Securing the Homeland Against AI Agent Attacks." METR runs capability tests on frontier AI agents with voluntary access from OpenAI, Anthropic, Google, Meta, SpaceXAI, and Amazon, and says it is not paid or funded by them.
METR·1 hour ago

Google announces Gemini 4 Argon, rolling out first to cyber defenders
Gemini 4 Argon is Google's new frontier model, rolling out to trusted testers via the Fairwind Program before wider release. It scores 77.9% on DeepSWE v1.1 and 77.5% on AutomationBench-AA, and is priced at $2/$10 per million input/output tokens during introductory pricing.
Google DeepMind·2 hours ago

LangChain, Modal and Cogent Security host AI Heist challenge at SF Tech Week
LangChain·2 hours agoRFK Jr. claims AI will free Americans from "medical tyranny"
At a MAHA event with VP JD Vance, Health Secretary Robert F. Kennedy Jr. said AI is "better informed than any doctor in the country" and urged Americans to use it for second opinions. Ars Technica tested his claims: Gemini said masks do reduce respiratory disease spread.
Ars Technica·2 hours ago

Framework opens preorders for AMD Ryzen AI Max 400 desktop with 192GB
Framework's DIY Edition desktop with AMD Ryzen AI Max 400 Series and 192GB memory is now available for preorder. The 192GB unified memory configuration targets local LLM workloads.
r/LocalLLaMA·2 hours agoNVIDIA expands xio-sig with cuObject, ships SCADA Server SDK
NVIDIA added cuObject to xio-sig alongside cuFile, in partnership with Google Cloud and Microsoft, and made cuObject client and server libraries generally available. The new SCADA Server SDK lets storage providers build servers that answer GPU-initiated requests; IBM showed a prototype integrating SCADA with IBM Storage Scale.
NVIDIA Developer Blog·2 hours ago

Claude Code 2.1.286 ships 88 CLI changes
Release adds permission-prompt counts like "2 of 5" when requests stack, plus mouse support for "N more" rows in fullscreen lists. Sessions now retry once on the previous model of the same tier when the Anthropic API refuses the resolved model.
Claude Code Releases·2 hours ago