How Goodfire used Ai2’s open post-training stack to trace unwanted model behavior
Allen Institute for AI Blog6d4 min read
Goodfire used Ai2’s fully open post-training stack to predict LLM behavioral changes, trace unwanted model behavior back to individual training examples, and test targeted fixes without sacrificing broader capability gains.
