"Related guidelines should include but not be limited to:
Thresholds for unacceptable risk from an AI system when undertaking automated AI R&D;
“If-then” commitments where reaching a certain capability or risk level triggers a mitigation or release decision;
Specifications of required levels of human oversight during automated R&D as well as inference of continuously learning models;
Preferred alignment and security techniques;
The appropriate resource allocation between capability research and acceleration of techniques for monitoring, control, and alignment;
Preferred capability research pathways that pose less risk; and
Agent monitoring and control specifications, building on ongoing CAISI information gathering.12
Of particular importance are the risk thresholds: within our proposed implementation of “pacing,” their breach would trigger consideration of efforts to reallocate resources away from automated AI R&D, and towards societal resilience (see “Invest in AI resilience” section below), safety research, and AI diffusion.13"
Framing: mixed ·
First seen: 2026-08-25 ·
Last seen: 2026-08-25 ·
Spread: 1 articles