Goodfire just launched what it says is a cheaper way to keep AI agents in check: instead of paying a second AI to read everything an agent does, its monitors peek inside the model while it works and only call in backup when something looks fishy.
Goodfire presented probes that examine internal signals from models while running AI agents. The company claims that its method detected 94% of malicious sessions in tests with Kimi K3 and reduced costs compared to other monitoring systems.