AnalysisPolicySeptember 9, 2026

Goodfire uses Ai2's open post-training stack to trace unwanted model behavior

Goodfire used Ai2's fully open post-training stack to predict LLM behavioral changes, trace a safety regression back to individual preference examples, and test targeted fixes without sacrificing capability gains.

2 sources

More stories today

Open the live feed