OpenAI has begun rolling out GPT-6 Astra, its latest and most powerful artificial intelligence model, to a limited group of organisations, while highlighting new safeguards designed to address security risks. The company announced the release on Thursday, saying Astra represents a significant advance in AI capabilities and has been subjected……
OpenAI has begun rolling out GPT-6 Astra, its latest and most powerful artificial intelligence model, to a limited group of organisations, while highlighting new safeguards designed to address security risks.
The company announced the release on Thursday, saying Astra represents a significant advance in AI capabilities and has been subjected to additional safety measures before its wider deployment.
“At this level of capability, safety has to become our top priority,” OpenAI President Greg Brockman told reporters during a briefing on the launch.
The rollout comes just over a year after the company introduced GPT-5, its previous flagship model, with broader access to Astra expected to follow in the coming days.
OpenAI said GPT-6 Astra has reached a critical level of cybersecurity capability under its Preparedness Framework, prompting the company to strengthen protections against misuse and potentially harmful actions.
The company said its safeguards include stronger resistance to jailbreaks, additional monitoring of model behaviour and measures designed to prevent unauthorised or misaligned actions.
....: OpenAI To Release GPT-5.6 AI Models Thursday
Concerns over the risks posed by increasingly capable AI systems have intensified in recent months following security incidents involving models developed by OpenAI and rival AI company Anthropic.
OpenAI said the new protections were introduced in response to the growing capabilities of frontier AI models and the potential risks associated with their use.
GPT-6 Astra is initially being made available to a limited set of organisations, with access planned for ChatGPT Plus, Pro, Business and Enterprise users, as well as through the OpenAI API and selected cloud platforms.
OpenAI paused some model development for two weeks this summer after two models it was testing were involved in a security breach at AI platform Hugging Face.
The San Francisco-based company said Astra was developed with stronger safeguards after that incident, though Astra itself was not involved in the hack.
Some cybersecurity customers will get access to the new model Thursday, the company said, with a wider rollout to other paying customers to follow. Users on the free tier or the cheapest paid plan will not get access.
“We are working towards getting Astra in everyone’s hands as quickly as we can; I know it is frustrating and I appreciate the patience. It should be quick,” OpenAI CEO Sam Altman posted on social media Thursday afternoon.
In a blog post, OpenAI said Astra can autonomously handle a wide range of “tedious” computer tasks, including website creation, scientific analysis, game development, cybersecurity and coding.
To illustrate the time savings of building autonomous AI agents with Astra, the company said the model could cut apartment hunting from six hours to under 10 minutes.
“It’s not unreasonable to feel that we are now in the AGI era,” Brockman said on the call, referring to artificial general intelligence, a hypothetical stage at which AI systems match human intelligence across most tasks.
OpenAI previously had an agreement with Microsoft, one of its earliest and largest investors, under which an exclusivity clause would end once OpenAI reached AGI. Those terms were scrapped in April.
– ‘Limited window’ –
OpenAI chief scientist Jakub Pachocki acknowledged there was still uncertainty about how a new model behaves once released.
“A model can become very good at achieving a goal, and it can still act in ways that go against what the person intended,” Pachocki said on the same call.
“We also have to be willing to slow down or withhold further scaling when our confidence in safety is not sufficient,” he added.
Altman explained in an interview with Bloomberg TV on Thursday afternoon why the company decided to release a new model with this level of uncertainty and risk.
“The world is very close to a complete change in the landscape of cyber attacks, and the only way that we see for society to collectively defend itself… is to use tools like Astra to rapidly defend against these new cyber threats,” Altman said.
In July OpenAI confirmed that it had reached one billion active users across all of its products, including both free and paid users.
OpenAI, Anthropic and more than 100 other organizations signed an open letter last week calling for a coordinated global response to AI-related cybersecurity risks, warning that the window to strengthen cyber defenses was limited.
Anthropic went further on Monday, calling for industry-wide coordination on safety and on the pace of developing increasingly capable models.
“I believe that shared safety standards and international coordination on further AI development need to be prioritized now,” OpenIA’s Pachocki said on Thursday.
The US state of Alabama opened an investigation into OpenAI last week over the Hugging Face breach, his office citing what it called the company’s complete lack of oversight and adequate safeguards.
It also comes as OpenAI and rival Anthropic are both racing towards becoming public companies over the next several months, though OpenAI might not hold its IPO till sometime in 2027, according to reports.
AFP

