As AI models gain the ability to act independently, safety testing is increasingly about more than preventing harmful answers ...