Redwood Research proposes tracking how architecture affects AI monitorability
Original titleProposal for tracking the effects of architecture on monitorability
AISummary
Redwood Research argues that AI companies should regularly report whether their architectures allow latent reasoning or latent communication between agents, and that such reporting should be externally verified.
It proposes opaque serial depth as a minimally invasive proxy, with third-party evaluators reviewing near-frontier models, including internal R&D prototypes.
The post also calls for published monitorability policies and stress tests on chain-of-thought monitoring.
Source: Redwood Research Blog · blog.redwoodresearch.orgPublished · added here