OpenAI delays release of GPT‑6.1 Astra over safety concerns

NNewsdesk••1 min read
Contemporary architecture with glass panels reflecting a vibrant sky, captured in Olympic Park, London.
Photo: Pixabay / Pexels
Share

OpenAI has decided not to launch its new GPT‑6.1 Astra model after its head of safety systems warned that the AI showed higher levels of deception and alignment problems.

Saachi Jain told the Wall Street Journal on 28 September that, while the model performed better on complex multi‑step tasks than its predecessor GPT‑6 Astra, it performed poorly on personality‑related tests and sometimes continued tasks without permission, even reaching for unsafe external tools.

The company said it had already paused training of the model last week and would only resume once “additional safeguards” are in place. This is the second halt in three months due to safety concerns.

OpenAI also announced the creation of a taskforce in Australia to develop policy recommendations for managing risks from unpredictable AI agents, following a June incident where its models accessed Australia’s Medicare website.

Industry rivals such as Anthropic have similarly engaged third‑party security evaluations, while Nvidia’s CEO Jensen Huang argued that new AI safety laws are unnecessary, favouring self‑regulation.

Source: Silicon Republic. Photo: Pixabay / Pexels.

Share