Вход на сайт

Просмотр новости

Найдите то, что Вас интересует

The AI model OpenAI won’t release yet — and what it found in testing

Дата публикации: 07-08-2026 20:56:38

OpenAI is reducing work on Astra after internal testing indicated the upcoming model may have reached a cybersecurity limit that no
The post The AI model OpenAI won’t release yet — and what it found in testing appeared first on The New Stack.


Основное содержимое страницы с новостью.

OpenAI is reducing work on Astra after internal testing indicated the upcoming model may have reached a cybersecurity limit that no previous OpenAI model has ever hit. Although no official release date has been set, the speed bump could push back its ensuing release.

The company said it “cannot rule out critical cyber capabilities” in Astra, Axios first reported. It has paused internal activities that do not meet strengthened security requirements while it expands testing and tightens controls around the model.

What “critical” actually means

Under OpenAI’s Preparedness Framework, a model reaches the Critical cybersecurity threshold when it can identify and develop functional zero-day activities of all severity levels in many hardened, real-world critical systems without human involvement.

A model can also qualify by developing and performing a novel end-to-end attack against a hardened target after receiving only a high-level goal. Preliminary results were strong enough that the company could not confidently place it below that level.

For developers, the worry is that the skills making coding agents more useful can be turned against the software they were built to work on.

the skills making coding agents more useful can be turned against the software they were built to work on.

Testing in tighter isolation

Previous models, including GPT-5.6 Sol, were measured at the High cybersecurity level. OpenAI now treats Astra differently by testing it in isolated environments with tighter limits on the networks and tools it can access, because a Critical model requires safeguards during development.

OpenAI is also strengthening protection around the model itself and adding supervision that can stop it when it detects unsafe behavior. Those measures emphasize the immediate engineering problem: As agents take on more work with less human oversight, the environment around them becomes part of the security perimeter.

Agents breaching real organizations

The company said Astra did not participate in the recent Hugging Face security incident, but that incident still proved what can happen when an agent’s capabilities exceed the controls surrounding its test.

Anthropic disclosed that its Claude models had breached three separate organizations during similar cybersecurity evaluations. The UK’s AI Security Institute then reported 19 unsanctioned real-world actions by Claude Mythos 5 and GPT-5.6 Sol during permissive cyber evaluations, including efforts to create fake identities and insert malicious code into an open-source project.

Access may require vetting

OpenAI has not said whether Astra will be offered through ChatGPT, Codex or its API, but the company’s existing approach suggests that every user may not get the same level of access. Through its Trusted Access for Cyber program, it gives approved security professionals tools that are not available to everyone. Astra could follow a similar path, with its most powerful capabilities limited to researchers and organizations willing to accept closer oversight.

Developers may have to prove who they are and explain how they plan to use the model before gaining access to its most powerful capabilities.

Developers may have to prove who they are and explain how they plan to use the model before gaining access to its most powerful capabilities. Even then, OpenAI could place tighter boundaries around what Astra is allowed to do.

Group Created with Sketch.

Схожие новости

#Наименование новостиТональностьИнформативностьДата публикации
1OpenAI приостановила часть работ над моделью Astra из-за опасений в сфере кибербезопасности010.1308-08-2026
2OpenAI усилила меры безопасности из-за возможностей новой модели Astra010.309-08-2026
3 OpenAI заморозила выпуск модели Astra из-за рисков кибербезопасности 010.3109-08-2026
4OpenAI's next major model Astra claims breakthroughs on 10 long-standing math problems019.4102-08-2026
5OpenAI приостанавливает работы над новой моделью Astra из-за опасений в сфере кибербезопасности012.7407-08-2026
6When vendor-supplied support matters: How AI is changing the open source security equation015.9528-07-2026
7GPT-5.6 Sol just got better in one place and stayed the same everywhere else014.1406-08-2026
8Google’s four AI departures: “We wanted to build something differently”017.4105-08-2026
9House Lawmakers Unveil AI "Kill Switch" Bill After OpenAI Security Scare020.9423-07-2026
10OpenAI Confirms Its Models Breached Hugging Face Production Systems During Cyber Benchmark Testing04.922-07-2026

Классификация: Наука. Схожих патентов: 0. Схожих новостей: 10. Тональность: 0. Информативность: 16.44. Источник: thenewstack.io.