Goodfire's probes watch AI agents from the inside, about 50 times cheaper than a judge model
Goodfire has launched monitors that read an AI model's internal signals while it works instead of rereading everything it writes. In its tests on Kimi K3 they caught about 93 percent of harmful hacking sessions for under 200 dollars per million turns, roughly 50 times cheaper than a judge model on every step.