← Back to events
ActiveAI

How Goodfire used Ai2’s open post-training stack to trace unwanted model behavior

Photo: Allen Institute for AI

What happened

Goodfire used Ai2’s fully open post-training stack to predict LLM behavioral changes, trace unwanted model behavior back to individual training examples, and test targeted fixes without sacrificing broader capability gains.

Summary assembled by rule from the sources below

Why it's spreading

Sources