This blog post discusses an experiment where ten frontier models were tested against autonomous AI attackers in a live AWS environment to evaluate the effectiveness of canaries in detecting attacks. It explores the timing of warnings, the speed at which models respond to threats, and changes in outcomes under deceptive conditions.