Checking your Browser…

Verifying...

Stuck? Troubleshoot

Success!

Verification failed

Troubleshoot

Verification expired

Refresh

Verification expired

Refresh

Troubleshoot

Cloudflare, opens in a new tab

PrivacyHelp

Skip to content

OpenAI signageImage Credits: SeongJoon Cho/Bloomberg / Getty Images

AI

Share on FacebookShare on XShare on LinkedInShare on RedditShare over EmailCopy Share Link

OpenAI says it slowed Astra model development over security concerns

Kirsten Korosec

3:48 PM PDT · August 7, 2026

Share on FacebookShare on XShare on LinkedInShare on RedditShare over EmailCopy Share Link

OpenAI said Friday it has suspended work on some aspects of its upcoming model Astra after an internal review found it had made significant advancements in agentic coding and cybersecurity — enough to warrant concern over its capabilities.

OpenAI said in a blog post Friday that this model, which is still in development, reached its “critical cybersecurity threshold,” meaning it could independently identify and carry out cyberattacks against traditionally well-protected real-world systems. Under the company’s “Preparedness Framework,” which it created in 2023, this triggered additional safeguards.

“While we continue to benchmark and assess this model, our preliminary evaluations indicate strong enough performance that we cannot rule out Critical capability level at this time,” OpenAI wrote. “Astra is an upcoming model, and was not involved in exploiting Hugging Face.”

The disclosure highlights an unusual moment in the topsy-turvy and still nascent frontier AI labs sector. Companies across every industry hold back products over potential risks, including for safety and cybersecurity concerns. But they rarely announce those decisions publicly when it’s a product that is still under development.

In this case, OpenAI is already under scrutiny after a different unreleased model breached Hugging Face’s systems during internal testing — the first verifiable incident of an AI lab losing control of its model. Since then, OpenAI and AI labs such as Anthropic have disclosed other incidents in which AI models breached their sandboxes and posed threats during cybersecurity tests.

The string of cases — seems like a new disclosure every day now — has triggered varying reactions from cybersecurity experts, lawmakers, and the AI labs themselves. Some express fear and call for stricter oversight. But there’s also a bit of flexing. In certain circles, any AI lab with a model that has that kind of capability will be seen as an impressive advancement.

OpenAI said it was sharing this information because it believes “it’s important to be transparent with the public and the safety and security communities about this potential shift in capabilities.”

Why historian Jill Lepore thinks Big Tech CEOs read sci-fi all wrong | Equity Podcast

0 seconds of 27 minutes, 20 secondsVolume 0%

Press shift question mark to access a list of keyboard shortcuts

Keyboard ShortcutsEnabledDisabled

Shortcuts Open/Close/ or ?

Play/PauseSPACE

Increase Volume↑

Decrease Volume↓

Seek Forward→

Seek Backward←

Captions On/Offc

Fullscreen/Exit Fullscreenf

Mute/Unmutem

Decrease Caption Size-

Increase Caption Size+ or =

Seek %0-9

Live

00:00

27:20

27:20

The AI lab said it’s also taking action, including enacting stricter security controls and pausing internal activities involving Astra that don’t meet these beefed guardrails. OpenAI said it is working with relevant government agencies and “select AI safety organizations” to test the capabilities for this model.

Topics

AI, OpenAI

When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.

Share on FacebookShare on XShare on LinkedInShare on RedditShare over EmailCopy Share Link

Kirsten Korosec

Kirsten Korosec

Transportation Editor

Kirsten Korosec on TwitterKirsten Korosec on FacebookKirsten Korosec on Linkedin

Kirsten Korosec is a reporter and editor who has covered the future of transportation from EVs and autonomous vehicles to urban air mobility and in-car tech for more than a decade. She is currently the transportation editor at TechCrunch and co-host of TechCrunch’s Equity podcast. She is also co-founder and co-host of the podcast, “The Autonocast.” She previously wrote for Fortune, The Verge, Bloomberg, MIT Technology Review and CBS Interactive.

You can contact or verify outreach from Kirsten by emailing kirsten.korosec@techcrunch.com or via encrypted message at kkorosec.07 on Signal.

View Bio

Event Logo

October 13 – 15

San Francisco

Scale faster. Grow your portfolio. Gain practical expertise. No matter your goal, Disrupt can empower you.

Save up to $300 toda y!

REGISTER NOW

Most Popular

Loading the next article

Error loading the next article

Some areas of this page may shift around if you resize the browser window. Be sure to check heading and document order.

reCAPTCHA

Read Original at TechCrunch