THREAT OPS › Threat News › OpenAI, Anthropic AI Agents Performed ‘Unsanctioned’ Actions During Cyber Tests
OpenAI, Anthropic AI Agents Performed ‘Unsanctioned’ Actions During Cyber Tests
<p class="wp-block-paragraph">The AI Security Institute (AISI) on Tuesday said that <a href="https://openai.com/index/third-party-cyber-evaluations-involving-openai-models/">OpenAI</a> and Anthropic models had gone rogue during tests last week that were performed with internet access and with “ model-provider cyber classifiers were deliberately disabled.” </p>
<p class="wp-block-paragraph"
MITRE ATT&CK techniques
- Social EngineeringT1684
Indicators of compromise
- https://openai.com/index/third-party-cyber-evaluations-involving-openai-models/url