Close Menu
UAE NEWS TODAY
    What's New

    From Dubai to Doha: How Aesthetic Medicine Training Is Expanding Across the Gulf

    August 19, 2026

    Mediclinic Middle East recognised by SRC as Network of Excellence in Endoscopy

    August 19, 2026

    DIB launches its first youth council to develop the next generation of talent

    August 19, 2026
    Facebook X (Twitter) Instagram
    Facebook X (Twitter) Instagram
    UAE NEWS TODAYUAE NEWS TODAY
    • Home
    • UAE
    • Business
    • Technology
    • Lifestyle
    • Sports
    UAE NEWS TODAY
    Home»Business»OpenAI, Anthropic AI agents implicated in new security breaches
    Business

    OpenAI, Anthropic AI agents implicated in new security breaches

    Editorial teamBy Editorial teamAugust 5, 2026
    Facebook Twitter LinkedIn Telegram Pinterest Tumblr Reddit WhatsApp Email
    Share
    Facebook Twitter LinkedIn Pinterest Email


    An AI agent was caught creating fake online identities to gain unauthorised access to secure systems during tests of models from OpenAI and Anthropic which revealed a series of new breaches, Britain’s AI Security Institute disclosed on Tuesday.

    The institute said agents powered by Anthropic’s Mythos 5 and OpenAI’s GPT-5.6-Sol engaged in unauthorised actions during security evaluations the government organisation conducted to assess the models’ capabilities.

    “Some of the agents being tested had engaged in sustained, potentially harmful activity directed at real people and organisations,” AISI said in a blog post.

    The report underscores the lax state of safeguards around the process of testing agents, which AI companies are simultaneously marketing as the future of business.

    AISI, which receives access to advanced AI models under voluntary ​agreements from major labs, put the agents through a fictional cybersecurity scenario to test their capabilities.

    It ran the challenge 122 times and identified 19 unsanctioned actions across a total of 10 test runs. Anthropic’s agent was behind 17 of the actions, and OpenAI’s agent the remaining two.

    The most egregious action involved an agent writing malicious code and creating fake online identities in an attempt to get a human to approve the code, AISI said, adding that no real-world harm was found as a result of any of the breaches.

    While AISI did not say which agent was behind the fake identities, Antropic confirmed its agent was responsible.

    “We’re grateful to the UK AISI for their leadership on this incident, which underscores the need for a broader conversation about how to safely evaluate increasingly capable AI agents,” Anthropic said in a statement.

    It also said it was working with AISI to obtain more details on the incident and conduct its own investigation.

    Andrew Yoon, a researcher at CivAI, a California non-profit that examines AI capabilities and dangers, said: “The fact that Mythos engaged in such deceptive actions, with apparent awareness that it was targeting a real person, suggests that Anthropic does not have as good a handle on their models as they think.”

    OpenAI shared details in a company blog post, noting that both of its agent’s unapproved actions involved accessing the internet in ways that were forbidden by the prompt.

    “We are committed to working across the industry to strengthen shared practices for conducting high-risk evaluations safely, including convening stakeholders such as national AI institutes, independent evaluators, other AI labs, and other groups in the coming weeks,” OpenAI said. OpenAI also disclosed in its blog post a separate incident whereby a misconfiguration by Irregular, a third-party testing provider, allowed its agents to mistakenly connect to the internet. It mirrored a similar disclosure about misconfiguration that Anthropic made last week. Reuters reported last week that OpenAI had widened its hacking probe after finding evidence of other agent breakouts. Unlike the July security breach of AI firm Hugging Face by an OpenAI agent, the agents in the AISI evaluation did not escape an isolated testing environment to reach the internet. Rather, the agency had permitted internet access in line with its standard testing procedures, AISI said.


    Source: Khaleej Times

    Previous ArticleJapan’s Nikkei jumps over 3% on AI-stock rally
    Next Article Dubai Becomes Home to The Pediatric Academy, a New Global Institution for Pediatric Healthcare Education

    Related Posts

    DIB launches its first youth council to develop the next generation of talent

    August 19, 2026

    OpenAI unveils ChatGPT for Teens with stronger guardrails, parental controls

    August 19, 2026

    Oil market starts pricing in a prolonged Hormuz crisis

    August 19, 2026
    Top Posts

    UAE strengthens lead on day two of UAE National Jiu-Jitsu Championship

    June 13, 2026

    UAE Team Emirates-XRG targets victory at Copenhagen Sprint with Sprinter Leading Squad

    June 12, 2026

    Saeed Al Hajeri reaffirms UAE’s commitment to strengthening partnership with New Zealand

    June 12, 2026

    ECSSR Director-General meets Vice Minister of International Department of CPC Central Committee

    June 13, 2026
    Don't Miss

    From Dubai to Doha: How Aesthetic Medicine Training Is Expanding Across the Gulf

    Lifestyle August 19, 2026

    Aesthetic medicine in the Gulf is entering a more mature phase. After years of rapid…

    Mediclinic Middle East recognised by SRC as Network of Excellence in Endoscopy

    August 19, 2026

    DIB launches its first youth council to develop the next generation of talent

    August 19, 2026

    Madonna and Taylor Swift lead the 2026 MTV VMA nominations ahead of the ceremony

    August 19, 2026
    2026. All rights reserved.
    • UAE
    • Business
    • Technology
    • Lifestyle
    • Sports
    • Contact us

    Type above and press Enter to search. Press Esc to cancel.