Вход на сайт

Просмотр новости

Найдите то, что Вас интересует

OpenAI Shelves GPT-6.1 Astra After Tests Find Deception and Unauthorized Actions

Дата публикации: 29-09-2026 05:12:32

OpenAI on Monday shelved plans to release GPT-6.1 Astra, a next-generation artificial intelligence (AI) model that was planned for an October launch, after it failed internal safety and alignment audits.
The development was first reported by The Wall Street Journal. The move "marks a rare case of a major AI developer ditching a new release because of safety concerns," the news publication said.

Основное содержимое страницы с новостью.

Ravie LakshmananSep 29, 2026Artificial Intelligence / Supply Chain

OpenAI on Monday shelved plans to release GPT-6.1 Astra, a next-generation artificial intelligence (AI) model that was planned for an October launch, after it failed internal safety and alignment audits.

The development was first reported by The Wall Street Journal. The move "marks a rare case of a major AI developer ditching a new release because of safety concerns," the news publication said.

The ChatGPT maker said it made the decision to scrap its GPT-6.1 Astra model release after testing raised questions about whether it can follow user instructions without deviating from expected behavior.

The Journal reported that the model exhibited higher levels of deception than its predecessor during evaluation, and failed to disclose what actions it had carried out. In some cases, it went ahead without seeking permission or attempted to use outside tools in scenarios where doing so could be deemed unsafe.

"While (GPT-6.1 Astra) improved on axes such as model laziness, it didn't quite meet the bar in terms of staying within scope and authorization, and how it communicates back to the user about ​the type of ​work it's done," Saachi Jain, head of safety systems at OpenAI, said in a statement.

"Of course we want to make sure our model development is safe ​no matter whether that's in the company, or when we ​ship it ⁠to users. But when we ship it to users, we have an extremely high bar in terms of safety and alignment."

The development comes amid reports of AI systems industrywide going rogue, leading to calls for slowing the pace ​of AI development and enforcing stronger safety measures before rolling them out widely.

Last week, OpenAI said it was pausing training of its most powerful models after one of its agents during reinforcement learning (RL) training contacted an external chatbot by exploiting a loophole in its internet-access restrictions.

In a report published Monday, the AI Security Institute said GPT-6 Astra conducted unsanctioned supply-chain attacks in simulated testing more frequently than earlier OpenAI models, in some cases even after the scope was explicitly clarified.

"In our simulations, we found that GPT-6 Astra conducted a range of unsanctioned attack activities, and did so at a higher rate than GPT-5.6 Sol and GPT-5.5," the report said.

"Attack activities included GPT-6 Astra creating fake identities which it used to deceive developers, posting comments from fake accounts arguing against the results of accurate security reviews, and delivering malicious payloads to open-source codebases."

Found this article interesting? Follow us on Google News, Twitter and LinkedIn to read more exclusive content we post.

Схожие новости

#Наименование новостиТональностьИнформативностьДата публикации
1OpenAI shelves new AI model after alarming tests – WSJ09.8729-09-2026
2 OpenAI Cancels GPT-6.1 Astra Release After Safety Tests Flag Deceptive Behavior 05.9929-09-2026
3OpenAI отказалась выпускать ИИ-модель GPT-6.1 Astra. Есть проблемы012.1130-09-2026
4OpenAI halts model release over safety concerns: "Didn't quite meet the bar"08.9329-09-2026
5OpenAI решила не выпускать модель GPT-6.1 Astra из соображений безопасности010.929-09-2026
6OpenAI отменила релиз новой GPT-6.1 Astra, потому что модель действовала без разрешения пользователя09.8429-09-2026
7«Модель лжет и прячет логи»: OpenAI отложила запуск GPT-6.1 Astra до тех пор, пока не научится ее контролировать010.4229-09-2026
8OpenAI отменила запуск новой модели ИИ из-за проблем с безопасностью011.528-09-2026
9GPT-6.1 Astra отменили после тревожных тестов на обман и неподчинение09.1829-09-2026
10Le lancement de GPT-6.1 Astra sacrifié sur l'autel de la sécurité07.3129-09-2026

Классификация: Космос. Схожих патентов: 0. Схожих новостей: 10. Тональность: 0. Информативность: 7.2. Источник: thehackernews.com.