Skip to main content
Kaino.dev
Discover
Evals
News
Academics
Kaino.dev

Discover, evaluate, and compare AI tools, models, and agents.

Explore

  • Discover
  • Evaluations
  • News
  • Academics

Community

  • Twitter
  • YouTube
  • Instagram
Privacy PolicyTerms of Service

© 2026 Kaino.dev. All rights reserved.

Version 1.1.0
OpenAI Hit Pause on Its Largest Frontier AI Run After Astra Flagged 'Critical' Cyber Risk · News · Kaino
OpenAI Hit Pause on Its Largest Frontier AI Run After Astra Flagged 'Critical' Cyber Risk
Kaino
8h agoSep 15, 2026, 12:00 AM3 views

OpenAI Hit Pause on Its Largest Frontier AI Run After Astra Flagged 'Critical' Cyber Risk

OpenAI paused its largest frontier RL run after Astra's cyber evaluations couldn't rule out 'Critical' hacking capability. Here's what's confirmed — and what isn't.

OpenAIAstraAI cyber capabilityreinforcement learningOpenAI pausefrontier AI safetyPreparedness FrameworkAI training haltCritical cyber riskOpenAI RL training

OpenAI has paused a deployment-focused reinforcement-learning phase for two weeks and is holding back its largest planned frontier RL run — after evaluations of its upcoming Astra system couldn't rule out "Critical" cybersecurity capabilities under the company's own risk framework. OpenAI says it's using the time to strengthen security, monitoring, and alignment safeguards.

This is a real, consequential pause — but a narrow one. OpenAI hasn't stopped all training, cancelled Astra, or set a restart date. What it has done is let an internal capability evaluation directly delay a training schedule, rather than just shape release documentation after the fact. Axios independently confirmed the hold and reported OpenAI is also reconsidering parts of its preparedness framework as a result.

The trigger is specific: Astra's cyber-capability evaluations reached a point where OpenAI couldn't confidently rule out Critical-tier hacking ability. In response, the company moved related work into more secure research environments, restricted tool and network access, and added monitoring — meaning the risk assessment changed how researchers are allowed to work with the system, not just what gets published about it later.

What's still missing is the substance behind the headline. OpenAI hasn't disclosed the actual benchmark results, which specific cyber tasks triggered the concern, or the threshold Astra needs to clear before the held run resumes — and there's no public timeline for that decision. That means outsiders can't yet independently judge whether OpenAI's "Critical" classification is well-calibrated, or whether the new safeguards are actually sufficient.

Bottom line: This is a genuinely notable step in AI safety governance — a capability evaluation now has the power to gate a frontier training run, not just annotate a model card after launch. OpenAI deserves credit for disclosing it publicly. But the real test comes next: what evidence OpenAI accepts before resuming, and whether that bar can be scrutinized by anyone outside the company.

Key takeaways
  • 1

    OpenAI says it's using the time to strengthen security, monitoring, and alignment safeguards.

  • 2

    OpenAI hasn't stopped all training, cancelled Astra, or set a restart date.

  • 3

    What it has done is let an internal capability evaluation directly delay a training schedule, rather than just shape release documentation after the fact.

Continue reading

Latest from Kaino News

Story pulse

Freshness

8h ago

Views

3

Reading

2 min

Byline

Kainotomic Team

Utilities

Topics

OpenAIAstraAI cyber capabilityreinforcement learningOpenAI pausefrontier AI safetyPreparedness FrameworkAI training haltCritical cyber riskOpenAI RL training

Sources

Reference material and original reporting used in this story.

OpenAI

Published Sep 15, 2026, 12:00 AM

View source