Technologist Mag
  • Home
  • Tech News
  • AI
  • Apps
  • Gadgets
  • Gaming
  • Guides
  • Laptops
  • Mobiles
  • Wearables
  • More
    • Web Stories
    • Trending
    • Press Release

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

What's On
Valor Mortis Preview: Uneven Difficulty Balance

Valor Mortis Preview: Uneven Difficulty Balance

21 September 2026
Bungie Announces Plans To Restore Vaulted Destiny 2 Content And Substantial Reworks To Marathon

Bungie Announces Plans To Restore Vaulted Destiny 2 Content And Substantial Reworks To Marathon

21 September 2026
After 8 Years in Space, ‘BepiColombo’ Is Finally Approaching Mercury

After 8 Years in Space, ‘BepiColombo’ Is Finally Approaching Mercury

21 September 2026
Justice for CSS | WIRED

Justice for CSS | WIRED

21 September 2026
AI, Tariffs, Rare Minerals: What to Expect From Trump’s Upcoming Summit With Xi Jinping

AI, Tariffs, Rare Minerals: What to Expect From Trump’s Upcoming Summit With Xi Jinping

21 September 2026
Facebook X (Twitter) Instagram
Facebook X (Twitter) Instagram
Technologist Mag
SUBSCRIBE
  • Home
  • Tech News
  • AI
  • Apps
  • Gadgets
  • Gaming
  • Guides
  • Laptops
  • Mobiles
  • Wearables
  • More
    • Web Stories
    • Trending
    • Press Release
Technologist Mag
Home » OpenAI Is About to Release Its First AI Model With ‘Critical’ Cyber Abilities
Tech News

OpenAI Is About to Release Its First AI Model With ‘Critical’ Cyber Abilities

By technologistmag.com1 September 20264 Mins Read
OpenAI Is About to Release Its First AI Model With ‘Critical’ Cyber Abilities
Share
Facebook Twitter Reddit Telegram Pinterest Email

OpenAI announced Tuesday that its forthcoming AI model, Astra, is its first to reach the company’s threshold for what it calls “critical” cyber capabilities. OpenAI says it plans to publicly release a version of Astra “soon,” but will make the model’s advanced cyber capabilities available only to select partners in its Daybreak Blue early-access program at launch.

In a briefing with reporters, OpenAI safety and security leaders said the company has concluded that Astra reaches the critical cybersecurity capabilities outlined in its preparedness framework, which sets thresholds and protocols for when its AI models pose new levels of risk. The company says an AI model has reached its critical cyber threshold when it can independently find and exploit previously unknown vulnerabilities in real-world software. OpenAI leaders said the company has followed its procedure for this situation, which is to halt further development until appropriate safeguards and security measures can be implemented.

OpenAI previously said that it paused some training workloads related to the development of Astra and a future AI model for several weeks. Executives say the company has now resumed said work on Astra, and the future AI model, after putting additional safety and security controls in place. OpenAI says the multi-week pause was productive, and it is now confident that it can release Astra broadly in a safe way.

The announcement comes as Silicon Valley grapples with the advanced cybersecurity capabilities of cutting-edge AI models, and tries to assure users, lawmakers, and other companies that it can keep them under control. In July, OpenAI disclosed an incident in which agents running two of its models exploited vulnerabilities in what was supposed to be a siloed testing environment, gaining access to the internet and hacking the open source AI platform Hugging Face. (OpenAI notes that Astra was not one of the models involved in this case.)

Other AI companies, such as Anthropic and Meta, have disclosed similar incidents in recent weeks. On Monday, Anthropic also said it has paused some AI training workloads while it hardens its safety and security practices.

OpenAI says it’s implementing a multi-step approach to limit everyday users from accessing Astra’s advanced cyber capabilities, including a new “misalignment monitor.” If someone asks Astra to help them find an exploit in a real-world software system, for example, the model is supposed to refuse to answer. OpenAI says it has also made Astra more robust to jailbreaking attempts, and in tests it successfully refused unsafe queries at a significantly higher rate than previous models.

However, OpenAI notes in a blog post that its misalignment monitor may “occasionally flag legitimate activity as potential cyber misuse or unauthorized behavior, leading to it inadvertently being slowed, paused, or stopped.” OpenAI says the guardrail can be triggered in some cases even when a user is engaging in activities that don’t appear related to cybersecurity. When this happens, ChatGPT and Codex users may be asked to review the model’s action before proceeding, OpenAI said.

Partners in OpenAI’s Daybreak program—which includes digital infrastructure providers like Cisco, Cloudflare, and Palo Alto Networks—will get early access to a less restricted version of Astra with more robust cyber capabilities. The goal of the program is to ensure these companies can use advanced AI models like Astra to harden their defenses before similarly capable models are made broadly available. OpenAI leaders also said the company has been working closely with government partners to ensure they’re aware of Astra’s cyber skills and can get access to them.

Astra is not only capable of finding novel software vulnerabilities and developing ways to exploit them for hacking, but is also able to “chain” multiple exploits together, a technique used to bore deeper and deeper into a target system and gain access that wouldn’t be attainable using just one vulnerability.

Share. Facebook Twitter Pinterest LinkedIn Telegram Reddit Email
Previous ArticleDyson’s first toothbrush costs as much as a budget smartphone and justifies the price tag with a camera and AI
Next Article Ace Combat 8: Wings of Preview: A Good Reason To

Related Articles

After 8 Years in Space, ‘BepiColombo’ Is Finally Approaching Mercury

After 8 Years in Space, ‘BepiColombo’ Is Finally Approaching Mercury

21 September 2026
Justice for CSS | WIRED

Justice for CSS | WIRED

21 September 2026
AI, Tariffs, Rare Minerals: What to Expect From Trump’s Upcoming Summit With Xi Jinping

AI, Tariffs, Rare Minerals: What to Expect From Trump’s Upcoming Summit With Xi Jinping

21 September 2026
Dyson Won’t Say What’s Wrong With Its CameraJet Toothbrush

Dyson Won’t Say What’s Wrong With Its CameraJet Toothbrush

21 September 2026
The Search for Silicon Valley’s Most Powerful Woman

The Search for Silicon Valley’s Most Powerful Woman

21 September 2026
Review: Apple Mac Mini (M6)

Review: Apple Mac Mini (M6)

21 September 2026
Stay In Touch
  • Facebook
  • Twitter
  • Pinterest
  • Instagram
  • YouTube
  • Vimeo

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

Don't Miss
Bungie Announces Plans To Restore Vaulted Destiny 2 Content And Substantial Reworks To Marathon

Bungie Announces Plans To Restore Vaulted Destiny 2 Content And Substantial Reworks To Marathon

By technologistmag.com21 September 2026

It’s been a few months since Bungie ceased active development of Destiny 2, bringing the…

After 8 Years in Space, ‘BepiColombo’ Is Finally Approaching Mercury

After 8 Years in Space, ‘BepiColombo’ Is Finally Approaching Mercury

21 September 2026
Justice for CSS | WIRED

Justice for CSS | WIRED

21 September 2026
AI, Tariffs, Rare Minerals: What to Expect From Trump’s Upcoming Summit With Xi Jinping

AI, Tariffs, Rare Minerals: What to Expect From Trump’s Upcoming Summit With Xi Jinping

21 September 2026
Dyson Won’t Say What’s Wrong With Its CameraJet Toothbrush

Dyson Won’t Say What’s Wrong With Its CameraJet Toothbrush

21 September 2026
Technologist Mag
Facebook X (Twitter) Instagram Pinterest
  • Privacy
  • Terms
  • Advertise
  • Contact
© 2026 Technologist Mag. All Rights Reserved.

Type above and press Enter to search. Press Esc to cancel.