
Goodfire adds internal safety checks for AI agents on Baseten
Goodfire launched monitors on October 8, 2026 that check an AI model's internal signals at each step using small probes, calling in a separate review model only when something looks risky. Available first to Baseten customers using Kimi K3, it watches for offensive hacking, chemical and biological weapons misuse, and reward hacking, where an agent games its goals.
Published