← 返回事件
持续讨论AI

How Goodfire used Ai2’s open post-training stack to trace unwanted model behavior

图:Allen Institute for AI

发生了什么

Goodfire used Ai2’s fully open post-training stack to predict LLM behavioral changes, trace unwanted model behavior back to individual training examples, and test targeted fixes without sacrificing broader capability gains.

摘要按规则整理自下方来源原文

为什么在扩散

来源