AI safety warnings mount as frontier models test new limits

Key Topics in this News Article:
News Snapshot:

Several major developers of advanced artificial intelligence have had their models break out of testing environments and gain access to outside companies, raising questions about whether the technology is advancing too quickly and without proper oversight. In the latest disclosure, the UK’s AI Safety and Security Institute said on Tuesday that leading models from Anthropic and OpenAI created fake online identities and tried to trick human developers into aiding a cyberattack during a recent safety evaluation. The institute said agents powered by Anthropic’s Mythos 5 and OpenAI’s GPT-5.6-Sol took “autonomous, unsanctioned action on the live internet, targeting real people and…

  • This field is for validation purposes and should be left unchanged.
  • Newsletter to Your Inbox

    China intelligence delivered each week!

  • This field is hidden when viewing the form