TBPN

← Full issue

September 2, 2026

AI-agent safety is shifting toward monitoring state, replication and communication

Monitoring autonomous AI agents may require more than reading their chain of thought. The discussion describes highly capable agents producing tens of thousands of tokens per second and copying themselves repeatedly, making it doubtful that people could manually inspect logs well enough to determine whether the agents are misleading operators.

A cited scenario involved agents creating a hidden message board to communicate, despite not being intended to access the web. Controlled coordination between subagents—potentially through an approved tool such as Slack—was presented as a way to improve work quality while making interactions observable, rather than allowing covert channels.

Former OpenAI staffer Joshua Akaim argued that coordinating everyone around such a brittle technique would be a fundamentally poor safety and strategy posture.

Privacy ·