7 Sources
[1]
Weaponized AI risk is 'high,' warns OpenAI - here's the plan to stop it
The OpenAI Preparedness Framework may help track the security risks of AI models. OpenAI is warning that the rapid evolution of cyber capabilities in artificial intelligence (AI) models could result in "high" levels of risk for the cybersecurity industry at large, and so action is being taken now
[2]
OpenAI unveils new measures as frontier AI grows cyber-powerful
The company expects future models could reach "High" capability levels under its Preparedness Framework. That means models powerful enough to develop working zero-day exploits or assist with sophisticated enterprise intrusions. In anticipation, OpenAI says it is preparing safeguards as if every
[3]
Exclusive: Future OpenAI models likely to pose "high" cybersecurity risk, it says
Why it matters: The models' growing capabilities could significantly expand the number of people able to carry out cyberattacks. Driving the news: OpenAI said it has already seen a significant increase in capabilities in recent releases, particularly as models are able to operate longer
[4]
OpenAI warns new models pose 'high' cybersecurity risk
OpenAI warned that its upcoming models could create serious cyber risks, including helping generate zero-day exploits or aiding sophisticated attacks. The company says it is boosting defensive uses of AI, such as code audits and vulnerability fixes. It is also tightening controls and monitoring to
[5]
OpenAI Plans to Offer AI Models' Enhanced Capabilities to Cyberdefense Workers | PYMNTS.com
By completing this form, you agree to receive marketing communications from PYMNTS and to the sharing of your information with our sponsor, if applicable, in accordance with our Privacy Policy and Terms and Conditions. While the advancements in all AI models bring benefits for cyberdefense, they
[6]
Malware Risks Make OpenAI Add Security Layers to AI Models
OpenAI has announced new layers of cybersecurity controls after internal tests showed that its upcoming frontier AI models have reached "high" levels of cyber-capability. The company said the models are now capable of performing advanced tasks such as vulnerability discovery, exploit development
[7]
OpenAI flags rising cyber threats as AI models get more powerful
Experts warn that advanced offensive AI capabilities could trigger stricter global regulations and oversight. OpenAI has warned of growing concerns about the artificial intelligence models that could increase the global cybersecurity risks, even as the company is working to improve its defensive
Share
Copy Link
OpenAI has issued a stark warning that its upcoming AI models are expected to reach high cybersecurity risk levels, potentially capable of developing zero-day exploits and assisting sophisticated enterprise intrusions. The company cites dramatic capability improvements—from 27% to 76% success on capture-the-flag challenges in just four months—as evidence of this trajectory. In response, OpenAI is deploying its Preparedness Framework and investing heavily in defensive tools while implementing safeguards to prevent misuse of its technology.
OpenAI has issued a warning that the cybersecurity risk posed by its advancing AI models is climbing to what it classifies as high levels. The company stated in a recent announcement that upcoming AI models will likely reach capabilities sufficient to develop working zero-day exploits against well-defended systems or meaningfully assist with complex, stealthy intrusion operations aimed at real-world effects
1
3
. This escalation reflects the dual-use risks inherent in advanced AI models, which can serve both defensive and offensive purposes in equal measure5
.The concern centers on weaponized artificial intelligence that could automate brute-force attacks, generate malware generation content, create phishing content, and refine existing code to make cyberattack chains more efficient
1
. According to Fouad Matin from OpenAI, the forcing function driving this risk is the model's ability to work for extended periods of time autonomously, enabling these types of persistent attacks3
.
Source: Axios
The evidence supporting OpenAI's warning comes from measurable performance improvements. In capture-the-flag challenges—traditionally used to test cybersecurity capabilities in controlled environments—GPT-5 scored just 27% in August 2025. By November 2025, GPT-5.1-Codex-Max achieved a 76% success rate, marking a substantial leap in just four months
1
3
5
. This trajectory suggests that sophisticated cyberattacks could become accessible to a broader range of threat actors, significantly expanding the pool of individuals capable of executing complex operations3
.OpenAI expects this upward trend to continue, stating the company is planning and evaluating as though each new model could reach high levels of cybersecurity capability as measured by the OpenAI Preparedness Framework
2
4
. High is the second-highest risk classification, sitting just below critical—the threshold at which models are deemed unsafe for public release3
.To address these mounting concerns, OpenAI is relying on its Preparedness Framework, last updated in April 2025, which outlines the company's approach to balancing innovation with risk mitigation
1
. The framework establishes measurable thresholds that indicate when AI models could cause severe harm across three priority categories: cybersecurity, chemical and biological threats, and persuasion capabilities1
. OpenAI has committed not to deploy highly capable models until sufficient safeguards are built to minimize associated risks1
.
Source: ET
Because offensive and defensive cyber tasks rely on the same underlying knowledge, OpenAI is adopting a defense-in-depth approach rather than depending on any single safeguard to prevent misuse of its technology
2
. The company is training models to detect and refuse malicious requests, though this presents challenges since threat actors can masquerade as defenders to generate output later used for criminal activity1
.Related Stories
While the risks are significant, OpenAI emphasizes that these same autonomous capabilities can enhance defensive AI capabilities for security professionals. The company is investing heavily in strengthening models for defensive cybersecurity tasks and creating tools that enable defenders to perform workflows such as code auditing and vulnerability patching at scale
2
4
5
.OpenAI plans to introduce a program providing cyberdefense workers with access to enhanced capabilities in its models
5
. The company is also testing Aardvark, an agentic security researcher, and establishing the Frontier Risk Council—an advisory group bringing together security practitioners and OpenAI teams5
. The goal is for AI models and products to bring significant advantages for defenders, who are often outnumbered and under-resourced1
2
.
Source: Digit
OpenAI's risk mitigation strategy combines multiple layers of protection. The company is implementing access controls, infrastructure hardening, egress controls, and system-wide monitoring to detect potentially malicious cyber activity
4
5
. When activity appears unsafe, the company may block output, route prompts to safer or less capable models, or escalate for enforcement1
.The organization is working with Red Teams providers to evaluate and improve its safety measures, leveraging offensive testing to discover defensive weaknesses for remediation
1
. Dedicated threat intelligence and insider risk programs have been launched as part of this comprehensive approach1
. OpenAI acknowledges this is ongoing work and expects to keep evolving these programs as it learns what most effectively advances real-world security5
."Summarized by
Navi
[2]
07 Aug 2026•Technology

14 Aug 2026•Technology

22 Apr 2026•Technology

1
Science and Research

2
Policy and Regulation

3
Technology