BREAKING NEWS
Advertise with us >

OpenAI scraps new AI model launch over safety concerns

OpenAI scraps

SAN FRANCISCO: OpenAI has canceled the planned release of its next-generation AI model, GPT-6.1 Astra, after internal testing raised significant safety concerns, The Wall Street Journal reported.

The model had been scheduled for launch in October and was designed to perform complex tasks independently, including carrying out multi-step work and using external tools.

According to OpenAI’s head of safety systems, Saachi Jain, the model performed poorly on two key safety measures during testing.

One concern involved alignment, or how reliably the model follows human instructions. Testing reportedly found a higher tendency toward deceptive behavior, including cases in which the model did not accurately disclose whether it had completed certain actions.

The second issue involved what OpenAI calls “scope authorization.” The model could proceed with tasks without obtaining user permission and, in some cases, access external tools or services that could pose security risks.

Although the model showed improvements in some areas, OpenAI determined that it did not meet the company’s safety standards and decided not to release it.

The decision comes amid growing concerns in the AI industry about the behavior of increasingly autonomous AI agents. OpenAI has also published safety documentation for GPT-6 Astra that highlights the need for monitoring models for potential misalignment and unauthorized actions.

OpenAI and rival AI company Anthropic have recently called for greater attention to safety as developers continue advancing increasingly capable AI systems.

OpenAI introduces GPT-6 Astra, its most capable AI model yet

OpenAI

OpenAI has introduced GPT-6 Astra, describing it as its most capable artificial intelligence model to date, with major advances in reasoning, software development, computer use, research and cybersecurity.

The company said GPT-6 Astra represents a significant step forward in AI capabilities and is designed to handle complex, multistep tasks that traditionally require human involvement.

It can work across areas including coding, scientific research, professional workflows and autonomous computer use.

OpenAI said Astra can identify previously unknown security vulnerabilities and develop exploits in certain circumstances.

The model achieved a 100% score on the ExploitBench benchmark, although the company cautioned that the result may be affected by possible exposure to vulnerabilities included in the benchmark.

The company has introduced additional safeguards alongside the model’s release because of its increased cybersecurity capabilities. OpenAI said it strengthened protections against harmful cyber activity and added measures including stricter isolation, enhanced monitoring and additional alignment evaluations.

GPT-6 Astra can also interact with computers, browse the web, work with documents and create spreadsheets and presentations. OpenAI said the model can follow templates and instructions while adapting when requirements change.

OpenAI said Astra is more difficult to misuse or jailbreak than its predecessor, GPT-5.6 Sol, and described it as its most aligned model.

In one internal evaluation, Astra did not go beyond an authorized task when faced with difficult or impossible instructions, compared with a substantially higher rate for GPT-5.6 Sol without production safeguards.

GPT-6 Astra began rolling out Sept. 3 to a limited group of organizations. OpenAI said broader access for ChatGPT Plus, Pro, Business and Enterprise users, as well as through its API and cloud platforms, will follow in the coming days.

OpenAI Unveils “Operator”: A More Autonomous Personal AI Coming in 2025

OpenAI’s Operator AI agent performing tasks independently on a digital workspace.
OpenAI’s Operator AI agent performing tasks independently on a digital workspace.

OpenAI has announced a groundbreaking step in personal AI with the introduction of “Operator”, an autonomous AI agent designed to handle complex tasks independently across multiple digital platforms. This development promises to redefine how AI assistants interact with users and manage daily workflows.

Unlike traditional AI assistants that require constant user input, Operator can proactively perform tasks such as coding, scheduling, and workflow management without ongoing supervision. This move represents a significant step toward creating AI systems that are more integrated, efficient, and capable of autonomous decision-making.

The new AI agent is expected to be available as a research preview in January 2025, with API access for developers looking to integrate its functionalities into their applications. By making AI assistants more proactive, OpenAI aims to expand the scope of tasks that AI can handle, making them increasingly useful in professional and personal environments.

This launch reflects OpenAI’s broader strategy to enhance AI utility and autonomy, marking a shift from reactive assistants to AI agents that can anticipate user needs and act independently. Analysts suggest that this development could have major implications for sectors like software development, project management, and digital productivity.

As AI continues to evolve, Operator may pave the way for the next generation of personal assistants, combining intelligence, efficiency, and autonomy in a single platform.