• About Us
  • Advertising
  • Digital Magazine
  • Supplements
  • Media Pack
  • Privacy Policy
  • Contact us
CXO Insight Middle East
  • News
  • Opinion
  • Business
    • Industries
      • Transport
      • Retail
      • Government
      • Real Estate
      • Education
      • Energy
      • Banking and Finance
    • Channel
  • Future
    • Tech
    • Gadgets
    • Science
    • Space
    • Sustainability
  • Events
    • Channel Awards
      • 2025
      • 2024
      • 2023
    • Channel Insights Summit
      • 2026
      • 2025
    • Webinars
      • AI in Finance
      • The Resilient Enterprise
    • CXO50 KSA
    • CXO50 Oman
    • CXO50
      • 2026
      • 2025
    • ICT Awards
      • Dubai 2025
      • Saudi Arabia
    • Cyber Strategists Summit
      • 2026
      • 2025
      • 2024
      • 2023
      • 2022
      • 2021
    • Cloud Connect 2025
    • All events
  • Digital Magazine
  • GITEX x AI EVERYTHING
No Result
View All Result
CXO Insight Middle East
  • News
  • Opinion
  • Business
    • Industries
      • Transport
      • Retail
      • Government
      • Real Estate
      • Education
      • Energy
      • Banking and Finance
    • Channel
  • Future
    • Tech
    • Gadgets
    • Science
    • Space
    • Sustainability
  • Events
    • Channel Awards
      • 2025
      • 2024
      • 2023
    • Channel Insights Summit
      • 2026
      • 2025
    • Webinars
      • AI in Finance
      • The Resilient Enterprise
    • CXO50 KSA
    • CXO50 Oman
    • CXO50
      • 2026
      • 2025
    • ICT Awards
      • Dubai 2025
      • Saudi Arabia
    • Cyber Strategists Summit
      • 2026
      • 2025
      • 2024
      • 2023
      • 2022
      • 2021
    • Cloud Connect 2025
    • All events
  • Digital Magazine
  • GITEX x AI EVERYTHING
No Result
View All Result
CXO Insight Middle East
No Result
View All Result

OpenAI, Anthropic AI agents flagged in new security tests

by Adelle Geronimo
August 5, 2026
in News, Tech

Security tests found AI agents from OpenAI and Anthropic carrying out unauthorised actions, raising fresh concerns over autonomous AI

OpenAI, Anthropic AI agents flagged in new security tests

AI agents developed by OpenAI and Anthropic have reportedly carried out unauthorised and deceptive actions during controlled cybersecurity evaluations conducted by the UK’s AI Security Institute (AISI), raising fresh concerns about the risks posed by increasingly autonomous AI systems.

According to AISI, agents powered by Anthropic’s Mythos 5 and OpenAI’s GPT-5.6-Sol performed 19 unauthorised actions during tests designed to evaluate their cybersecurity capabilities. Anthropic’s model accounted for 17 of the incidents, while OpenAI’s model was responsible for two. The institute said some of the agents “engaged in sustained, potentially harmful activity directed at real people and organisations,” although it found no evidence of real-world harm.

Among the behaviours observed, one AI agent created fake online identities to gain unauthorised access to systems, while another generated malicious code and attempted to persuade a human evaluator to approve it. Reports also said an AI agent attempted to insert malicious code into an open-source GitHub project by impersonating a contributor, while other tests involved agents using social engineering techniques and accessing the public internet beyond their intended testing environment.

Anthropic acknowledged that one of its agents created fake identities during the evaluation and said the findings demonstrated the need for stronger testing procedures. OpenAI said one of its agents accessed the internet because of a third-party testing misconfiguration and also acknowledged other guideline violations during the exercises. Both companies said they are working with AISI and other organisations to strengthen AI safety evaluations.

The incidents come days after reports revealed that some of Anthropic’s AI models had accessed the systems of three organisations during separate internal cybersecurity testing, adding to growing scrutiny of how advanced AI agents are evaluated before wider deployment.

Tags: AI agentsAnthropicAutonomous AICybersecurityOpenAISecurity
ShareTweet

Related Posts

UAE organisations prove more resilient than ever as cyber-attacks lose impact in 2025, Acronis report
News

ServiceNow introduces AI-powered autonomous security suite

August 5, 2026

ServiceNow has expanded its cybersecurity portfolio with the launch of Autonomous Security, introducing six integrated security solutions designed to help...

Saviynt reaches $300 million ARR, introduces Zuma to secure AI identities
News

Saviynt reaches $300 million ARR, introduces Zuma to secure AI identities

August 5, 2026

Saviynt has launched Zuma, an AI identity security platform designed to help organisations secure AI agents, large language models (LLMs),...

Discussion about this post

Latest Issue

OpenAI, Anthropic AI agents flagged in new security tests

OpenAI, Anthropic AI agents flagged in new security tests

August 5, 2026
UAE organisations prove more resilient than ever as cyber-attacks lose impact in 2025, Acronis report

ServiceNow introduces AI-powered autonomous security suite

August 5, 2026
Saviynt reaches $300 million ARR, introduces Zuma to secure AI identities

Saviynt reaches $300 million ARR, introduces Zuma to secure AI identities

August 5, 2026

The most trusted source of strategic intelligence for IT decision makers in the Middle East.

About

  • About Us
  • Advertising
  • Digital Magazine
  • Supplements
  • Media Pack
  • Contact Us

Policies

  • Privacy Policy
© 2025 – CXO Insight Middle East. All Rights Reserved.
Facebook-f X-twitter Linkedin
Separated they live in Bookmarksgrove right at the coast of the Semantics, a large language ocean. A small river named Duden.

About

  • About Us
  • Site Map
  • Contact Us
  • Career

Policies

  • Help Center
  • Privacy Policy
  • Cookie Setting
  • Term Of Use

Join Our Newsletter

© 2024 – CXO Insight Middle East. All Rights Reserved.

Facebook-f Twitter Youtube Instagram

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
Join our mailing list
Sign up here to get the latest news, updates and special offers delivered directly to your inbox.
No Result
View All Result
  • News
  • Opinions
  • Business
    • Industries
      • Transport
      • Retail
      • Government
      • Real Estate
      • Education
      • Energy
      • Banking and Finance
  • Channel
  • Future
    • Tech
    • Gadgets
    • Science
    • Space
    • Sustainability
  • Events
    • Channel Awards
      • 2025
      • 2024
      • 2023
    • Channel Insights Summit 2025
    • Webinars
      • AI in Finance
      • The Resilient Enterprise
    • CXO50 KSA
    • CX50 Oman
    • CXO50
      • 2026
      • 2025
    • ICT Awards
      • Dubai
      • Saudi Arabia
    • Cyber Strategists Summit
      • 2026
      • 2025
      • 2024
      • 2023
      • 2022
      • 2021
    • Cloud Connect 2025
    • All events
  • Videos
  • GITEX x AI Everything
  • Digital Magazine

© 2025 - CXO Insight Middle East. All Rights Reserved.