OPENAI

Someone take Astra’s Wi-Fi away

OpenAI says its upcoming model, Astra, has reached what it calls a “Critical” level of cybersecurity capability, making it the first OpenAI model to reach that threshold.

In testing, Astra was able to find previously unknown security flaws and turn them into working exploits against protected systems.

It also performed much better than GPT-5.6 Sol on internal cybersecurity tests, including finding two new vulnerabilities during one evaluation.

Because of those results, OpenAI delayed parts of Astra’s development while it added stronger safeguards.

These include stricter refusals for harmful cyber requests, more monitoring and systems designed to stop the model if it tries to act outside its allowed limits.

In brief:

  • Astra can find and exploit serious security flaws.

  • OpenAI has added stronger safeguards and monitoring before release.

  • Its most advanced cyber tools will have limited access at launch.

Ah. So it can hack.

OpenAI says Astra is also better at following security restrictions than GPT-5.6 Sol.

In one test, Astra refused 91.5% of harmful cyber requests, compared with 59% for GPT-5.6 Sol.

Astra is still expected to launch, but its most advanced cybersecurity features will initially be limited to a smaller group of testers before access expands.

Exactly what we want to hear after the last hack: it’s got even better! - MV