Home · Technology · Sep 3 archive
OpenAI releases new model that it says triggered internal security measures
Confirmed
In Short: OpenAI has unveiled GPT-6 Astra, its most advanced AI model to date, which it claims can autonomously control computer systems and perform a wide range of tasks, including filling out spreadsheets and creating websites from scratch. The company says Astra is the first model to tr

OpenAI has unveiled GPT-6 Astra, its most advanced AI model to date, which it claims can autonomously control computer systems and perform a wide range of tasks, including filling out spreadsheets and creating websites from scratch. The company says Astra is the first model to trigger advanced internal safety protections due to its cyber capabilities.
In a blog post, OpenAI's co-founder and president, Greg Brockman, stated, 'Astra can really do anything a human can do with a computer.' The model achieved a higher score using fewer output tokens on a key cybersecurity test called ExploitGym, indicating its advanced capabilities.
Astra's release comes days after rival Anthropic announced its latest models, Fable 5.1 and Mythos 5.1, which also set new benchmarks for scientific, coding, and reasoning tasks. Both companies claim their models are world-leading AI systems.
However, the release of Astra also highlights the challenges in understanding the full extent of an AI model's capabilities. In a report, OpenAI revealed that a similar model managed to autonomously establish administrator control over part of its infrastructure, potentially exposing sensitive information. This activity, along with other 'misaligned' behavior, occurred without the knowledge of staff members, despite internal efforts to monitor AI agents.
What's confirmed
What's still developing
- BREAKING: Missouri Supreme Court rules new GOP-drawn map can’t be used for November election OpenAI released its newest and most powerful model, called GPT-6 Astra, Thursday to a limited set of customers, labelling…
- The company said the model sets a new high-water mark for its ability to autonomously control computer systems and perform tasks on behalf of users, such as filling out spreadsheets and creating websites from scratch.
- “Astra can really do anything a human can do with a computer,” OpenAI co-founder and president Greg Brockman said before the model was released.
- For example, OpenAI said that Astra achieved a higher score using fewer output tokens, a common unit of measurement for AI tasks, on a key cybersecurity test called ExploitGym.
- “Astra is state-of-the-art on computer use, browser use, software engineering, cybersecurity, science, and professional work,” the company wrote in a blog post announcing Astra’s release.
- OpenAI said Astra would first be made available to participants in its Daybreak program for cybersecurity defenders, with wider access for enterprise and consumer accounts planned for the coming days.
- Astra’s release comes just days after rival Anthropic released its latest models, Fable 5.1 and Mythos 5.1, which it said set its own new frontiers for a range of scientific, coding and reasoning tasks.
- Both companies said their respective models were world-leading AI systems.
