How Goodfire used Ai2’s open post-training stack to trace unwanted model behavior
Imported from official source
Goodfire used Ai2’s fully open post-training stack to predict LLM behavioral changes, trace unwanted model behavior back to individual training examples, and test targeted fixes without sacrificing broader capability gains.
This version
- Version
- 1 of 1
- Recorded
- September 20, 2026 19:52
- Change
- Initial
- Content hash
5639f48b5a59c6ad15df2b9d64450d7c- All versions
- Revision history