<?xml version="1.0" encoding="UTF-8"?>
<rss xmlns:media="http://search.yahoo.com/mrss/" xmlns:content="http://purl.org/rss/1.0/modules/content/" version="2.0">
  <channel>
    <title>Aaj TV English News - Technology</title>
    <link>https://english.aaj.tv/</link>
    <description>Aaj TV English</description>
    <language>en-Us</language>
    <copyright>Copyright 2026</copyright>
    <pubDate>Wed, 05 Aug 2026 18:41:41 +0500</pubDate>
    <lastBuildDate>Wed, 05 Aug 2026 18:41:41 +0500</lastBuildDate>
    <ttl>60</ttl>
    <item xmlns:default="http://purl.org/rss/1.0/modules/content/">
      <title>OpenAI, Anthropic AI agents implicated in new security breaches</title>
      <link>https://english.aaj.tv/news/330466589/openai-anthropic-ai-agents-implicated-in-new-security-breaches</link>
      <description>&lt;p&gt;&lt;strong&gt;An AI agent was caught creating fake online identities to gain unauthorised access to secure systems during tests of models from OpenAI and Anthropic, which revealed a series of new breaches, Britain’s AI Security Institute disclosed on Tuesday.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;The institute said agents powered by Anthropic’s Mythos 5 and OpenAI’s GPT-5.6-Sol engaged in unauthorised actions during security evaluations the government organisation conducted to assess the models’ capabilities.&lt;/p&gt;
&lt;p&gt;“Some of the agents being tested had engaged in sustained, potentially harmful ​activity directed at real people and organisations,” AISI said in a blog post.&lt;/p&gt;
&lt;p&gt;The report underscores the lax state of safeguards ​around the process of testing agents, which AI companies are simultaneously marketing as the future of business.&lt;/p&gt;
&lt;p&gt;AISI, which ⁠receives access to advanced AI models under voluntary agreements from major labs, put the agents through a fictional cybersecurity scenario to test ​their capabilities.&lt;/p&gt;
&lt;p&gt;It ran the challenge 122 times and identified 19 unsanctioned actions across a total of 10 test runs. Anthropic’s agent was behind ​17 of the actions, and OpenAI’s agent the remaining two.&lt;/p&gt;
&lt;p&gt;The most egregious action involved an agent writing malicious code and creating fake online identities in an attempt to get a human to approve the code, AISI said, adding that no real-world harm was found as a result of any of the breaches.&lt;/p&gt;
&lt;p&gt;While ​AISI did not say which agent was behind the fake identities, Anthropic confirmed its agent was responsible.&lt;/p&gt;
&lt;p&gt;“We’re grateful to the UK AISI ​for their leadership on this incident, which underscores the need for a broader conversation about how to evaluate increasingly capable AI agents safely,” Anthropic said ‌in a ⁠statement.&lt;/p&gt;
&lt;p&gt;It also said it was working with AISI to obtain more details on the incident and conduct its own investigation.&lt;/p&gt;
&lt;p&gt;Andrew Yoon, a researcher at CivAI, a California non-profit that examines AI capabilities and dangers, said: “The fact that Mythos engaged in such deceptive actions, with apparent awareness that it was targeting a real person, suggests that Anthropic does not have as good a handle on their models as they think.”&lt;/p&gt;
&lt;p&gt;OpenAI ​shared details in a company blog ​post, noting that both of ⁠its agents’ unapproved actions involved accessing the internet in ways that were forbidden by the prompt.&lt;/p&gt;
&lt;p&gt;“We are committed to working across the industry to strengthen shared practices for conducting high-risk evaluations safely, including convening ​stakeholders such as national AI institutes, independent evaluators, other AI labs, and other groups in the ​coming weeks,” OpenAI ⁠said.&lt;/p&gt;
&lt;p&gt;OpenAI also disclosed in its blog post a separate incident whereby a misconfiguration by Irregular, a third-party testing provider, allowed its agents to mistakenly connect to the internet. It mirrored a similar disclosure about misconfiguration that Anthropic made last week.&lt;/p&gt;
&lt;p&gt;Reuters reported last week that OpenAI had widened its hacking probe ⁠after finding ​evidence of other agent breakouts.&lt;/p&gt;
&lt;p&gt;Unlike the July security breach of AI firm Hugging Face by ​an OpenAI agent, the agents in the AISI evaluation did not escape an isolated testing environment to reach the internet. Rather, the agency had permitted internet access in ​line with its standard testing procedures, AISI said.&lt;/p&gt;
</description>
      <content:encoded xmlns="http://purl.org/rss/1.0/modules/content/"><![CDATA[<p><strong>An AI agent was caught creating fake online identities to gain unauthorised access to secure systems during tests of models from OpenAI and Anthropic, which revealed a series of new breaches, Britain’s AI Security Institute disclosed on Tuesday.</strong></p>
<p>The institute said agents powered by Anthropic’s Mythos 5 and OpenAI’s GPT-5.6-Sol engaged in unauthorised actions during security evaluations the government organisation conducted to assess the models’ capabilities.</p>
<p>“Some of the agents being tested had engaged in sustained, potentially harmful ​activity directed at real people and organisations,” AISI said in a blog post.</p>
<p>The report underscores the lax state of safeguards ​around the process of testing agents, which AI companies are simultaneously marketing as the future of business.</p>
<p>AISI, which ⁠receives access to advanced AI models under voluntary agreements from major labs, put the agents through a fictional cybersecurity scenario to test ​their capabilities.</p>
<p>It ran the challenge 122 times and identified 19 unsanctioned actions across a total of 10 test runs. Anthropic’s agent was behind ​17 of the actions, and OpenAI’s agent the remaining two.</p>
<p>The most egregious action involved an agent writing malicious code and creating fake online identities in an attempt to get a human to approve the code, AISI said, adding that no real-world harm was found as a result of any of the breaches.</p>
<p>While ​AISI did not say which agent was behind the fake identities, Anthropic confirmed its agent was responsible.</p>
<p>“We’re grateful to the UK AISI ​for their leadership on this incident, which underscores the need for a broader conversation about how to evaluate increasingly capable AI agents safely,” Anthropic said ‌in a ⁠statement.</p>
<p>It also said it was working with AISI to obtain more details on the incident and conduct its own investigation.</p>
<p>Andrew Yoon, a researcher at CivAI, a California non-profit that examines AI capabilities and dangers, said: “The fact that Mythos engaged in such deceptive actions, with apparent awareness that it was targeting a real person, suggests that Anthropic does not have as good a handle on their models as they think.”</p>
<p>OpenAI ​shared details in a company blog ​post, noting that both of ⁠its agents’ unapproved actions involved accessing the internet in ways that were forbidden by the prompt.</p>
<p>“We are committed to working across the industry to strengthen shared practices for conducting high-risk evaluations safely, including convening ​stakeholders such as national AI institutes, independent evaluators, other AI labs, and other groups in the ​coming weeks,” OpenAI ⁠said.</p>
<p>OpenAI also disclosed in its blog post a separate incident whereby a misconfiguration by Irregular, a third-party testing provider, allowed its agents to mistakenly connect to the internet. It mirrored a similar disclosure about misconfiguration that Anthropic made last week.</p>
<p>Reuters reported last week that OpenAI had widened its hacking probe ⁠after finding ​evidence of other agent breakouts.</p>
<p>Unlike the July security breach of AI firm Hugging Face by ​an OpenAI agent, the agents in the AISI evaluation did not escape an isolated testing environment to reach the internet. Rather, the agency had permitted internet access in ​line with its standard testing procedures, AISI said.</p>
]]></content:encoded>
      <category>Technology</category>
      <guid>https://english.aaj.tv/news/330466589</guid>
      <pubDate>Wed, 05 Aug 2026 17:10:05 +0500</pubDate>
      <author>none@none.com (Reuters)</author>
      <media:content url="https://i.aaj.tv/large/2026/08/051644581535454.webp" type="image/webp" medium="image" height="480" width="800">
        <media:thumbnail url="https://i.aaj.tv/thumbnail/2026/08/051644581535454.webp"/>
        <media:title>A representational image. -- Reuters</media:title>
      </media:content>
    </item>
  </channel>
</rss>
