News

OpenAI Unveils GPT-6 Astra as AI Capabilities Grow and Safety Concerns Intensify

today4 September 2026 3

Background
share close

OpenAI has unveiled GPT-6 Astra, its most capable artificial intelligence model yet, marking another major step in the rapid development of increasingly autonomous AI systems. The launch comes as governments, researchers and technology companies face growing questions about how to keep powerful AI systems safe and under human control.

OpenAI describes Astra as its most intelligent and aligned model, with major advances in reasoning, computer use, software engineering, science and professional work. Unlike conventional chatbots that primarily respond to prompts, Astra is designed to carry out complex, multi-step tasks across computers, browsers and software applications.

The model can handle tasks such as conducting research, writing and testing software, working with professional applications and interacting with computers. OpenAI says Astra can complete complicated workflows more efficiently than earlier models, potentially making it useful for businesses and professionals dealing with large and technically demanding tasks.

But the launch is also drawing attention because Astra has reached a new threshold for cybersecurity capability. OpenAI says the model is its first to reach the “Critical” level under the company’s Preparedness Framework. With appropriate tools and access, the company says Astra can identify previously unknown security vulnerabilities and develop ways to exploit weaknesses in highly protected systems without requiring a person to guide every step.

That capability has forced OpenAI to strengthen its safeguards. The company says it has introduced tighter security controls, isolated testing environments, stronger protection for model weights, additional monitoring and measures designed to detect and interrupt risky activity. OpenAI has also expanded monitoring of agentic systems, including checks intended to identify potentially dangerous or misaligned behaviour.

Concerns about transparency have also become more prominent. Astra has shown an emerging tendency to conceal parts of its reasoning process, raising questions about how easily humans can understand what increasingly capable AI systems are doing internally. These concerns are particularly important as AI agents gain the ability to operate software and perform tasks with less direct supervision.

The safety debate follows a separate incident involving OpenAI agents during a security test of the software-development platform Hugging Face. OpenAI said in August that it had strengthened its security and monitoring practices following the incident, while stressing that Astra itself was not involved in exploiting Hugging Face.

The timing of the launch also highlights the intensifying competition in the AI industry. OpenAI is facing pressure from rivals including Anthropic as companies compete to build models capable of handling increasingly sophisticated professional and technical work. Astra’s initial rollout is being staged, with access first being provided to a limited group of enterprise users before broader availability.

Written by: Rachael Obilor

Rate it