Goodfire adds internal safety checks for AI agents on Baseten

Goodfire adds internal safety checks for AI agents on Baseten

Goodfire launched monitors on October 8, 2026 that check an AI model's internal signals at each step using small probes, calling in a separate review model only when something looks risky. Available first to Baseten customers using Kimi K3, it watches for offensive hacking, chemical and biological weapons misuse, and reward hacking, where an agent games its goals.

Published

Read at another depth