Goodfire launches inside-out monitors to spot rogue AI agents
Goodfire says its new monitors inspect an agent’s internal processes and only summon a secondary model when anomalous behavior appears. The company claims this ‘inside-out’ approach is substantially cheaper than using a second AI to review all agent actions.