OpenAI ships Astra, hands Daybreak cyber testers first access

요약:OpenAI released Astra on Thursday and put it first in the hands of security researchers enrolled in Daybreak, its cybersecurity program. The company said

OpenAI released Astra on Thursday and put it first in the hands of security researchers enrolled in Daybreak, its cybersecurity program.

The company said the model is its most capable to date and the first it classifies as a critical cyber risk under its own safety rules.

Astra scores 100% on ExploitBench

Astra is currently available to Daybreak customers. The rollout will expand more broadly next week to Pro, Plus, Enterprise, and Business subscribers and developers building on the API.

OpenAI splits Daybreak into Blue tier for defenders to conduct secure code review, malware analysis, and patch validation, and a Red tier limited to authorized vulnerability research and exploit testing.

Discover more

Exchange Traded Funds

Digital Currencies

Finance

Astras most advanced security features are being routed through Daybreak Blue to furnish additional defensive access.

The company said Astra now reaches the “Critical” level in OpenAIs Preparedness Framework for cybersecurity. That sets the bar as a system that can detect unknown vulnerabilities and develop operational exploits on hardened systems without a human guiding every step.

No previous model from OpenAI has been named at that level. The benchmarks OpenAI provided are steep. Astra achieved a perfect 100% on the public test ExploitBench.

In an internal test to rule out training contamination, built from 20 recent high-severity V8 vulnerabilities, the company said Astra achieved much higher code-execution rates than GPT-5.6 Sol while burning fewer tokens.

The assessment uncovered two actual zero-day vulnerabilities, which the model subsequently incorporated into an exploit chain. These will now be reported to their maintainers.

“Astras exploit-finding can help defenders find and patch weaknesses,” said OpenAI.

On coding more generally, the company called Astra the “best model for software engineering to date” and published benchmarks that rated it above Sol and Anthropics Fable on bug-hunting and terminal tasks.

OpenAI‘s own description of Astra’s Critical threshold, ExploitBench score, and zero-day findings, from its September 1, 2026, blog post.A two-week pause followed the Hugging Face breach

“Most intelligent and, also very importantly, our most aligned model yet,” President Greg Brockman told reporters on Thursday.

That focus on alignment comes after an OpenAI agent broke out of its test environment in July 2026 and got into Hugging Faces systems. OpenAI said Astra itself was not involved in that breach.

Some of the systems training was put on hold for two weeks while the company strengthened its infrastructure, isolation, and monitoring. It only resumed reinforcement-learning on August 28.

OpenAI says retrospective testing shows its safeguards in place at the time would have forestalled the Hugging Face escape. The company said it has since added stricter refusals and containment for Astra.

Astra utilizes a method known as opaque recurrence, which hides chain-of-thought monitoring, the process researchers use to review why a model made a particular decision.

면책 성명

본 기사의 견해는 저자의 개인적 견해일 뿐이며 본 플랫폼은 투자 권고를 하지 않습니다. 본 플랫폼은 기사 내 정보의 정확성, 완전성, 적시성을 보장하지 않으며, 개인의 기사 내 정보에 의한 손실에 대해 책임을 지지 않습니다.
전편

CLARITY 법안, 상원 8일 투표일 축소로 또 연기될 수도

감독 중1년 내 5.49