OpenAI follows Anthropic's lead in limited release of GPT‑5.4‑Cyber

The cybersecurity-focused AI model is less resistant to seemingly malicious actions, such as finding security vulnerabilities.
 By 
Amanda Yeo
 on 
The OpenAI logo appears on the screen of a smartphone placed on a laptop keyboard illuminated by green light.
Credit: Samuel Boivin / NurPhoto via Getty Images

OpenAI has unveiled GPT-5.4-Cyber, a new AI model that may be willing to accept seemingly malicious prompts in the name of cybersecurity. Fortunately, the ChatGPT developer won't let just anyone play with its less restrictive, more freewheeling AI.

Announced via a blog post on Tuesday, GPT-5.4-Cyber is a variant of OpenAI's publicly available GPT-5.4 large language model. According to OpenAI, its frontier AI models such as GPT-5.4 have safeguards against clearly malicious use, making them refuse harmful user requests such as stealing credentials or finding vulnerabilities in code. In contrast, the company's new GPT-5.4-Cyber model is trained to be more lenient, and potentially accept these prompts instead. 

Describing GPT-5.4-Cyber as "cyber-permissive," OpenAI states that this change is to allow the AI to be used for defensive cybersecurity measures, such as helping researchers find vulnerabilities to be addressed.


You May Also Like

"We want to empower defenders by giving broad access to frontier capabilities, including models which have been tailor-made for cybersecurity," wrote OpenAI. "This is a version of GPT‑5.4 which lowers the refusal boundary for legitimate cybersecurity work and enables new capabilities for advanced defensive workflows."

Given the potential danger posed by GPT-5.4-Cyber's lowered safeguards, not everyone will be able to immediately dive in to push the AI's arguably flexible ethical limits even further. OpenAI states that it is starting with "limited, iterative deployment to vetted security vendors, organizations, and researchers." As such, only members of its Trusted Access for Cyber⁠ (TAC) program will be given access to GPT-5.4-Cyber at present, and only those at its highest tiers. 

Introduced in February, TAC is a network of users who have been through OpenAI's automated identity verification process, including completing a government ID check. Once approved, users in OpenAI's TAC program are allowed access to versions of its AI models with fewer safeguards, such as GPT‑5.4‑Cyber. OpenAI states that this is intended to enable cybersecurity research, education, and programming. 

Not every TAC-approved user will immediately get their hands on GPT-5.4-Cyber, however. OpenAI states that users who aren't already part of TAC's higher tiers may request access to it, which will require going through further authentication to verify themselves as "legitimate cyber defenders." 

GPT-5.4-Cyber's reveal comes just one week after OpenAI competitor Anthropic announced Project Glasswing. Like TAC, Project Glasswing is an initiative that restricts Anthropic's cybersecurity-focused Claude Mythos Preview AI model to select approved organisations. Claiming that Claude Mythos Preview "has already found thousands of high-severity vulnerabilities," Anthropic stated that Project Glasswing was an effort to ensure its AI model was used for solely defensive cybersecurity purposes.

"Given the rate of AI progress, it will not be long before such capabilities proliferate, potentially beyond actors who are committed to deploying them safely," Anthropic wrote.


Disclosure: Ziff Davis, Mashable’s parent company, in April 2025 filed a lawsuit against OpenAI, alleging it infringed Ziff Davis copyrights in training and operating its AI systems.

Amanda Yeo
Amanda Yeo
Assistant Editor

Amanda Yeo is an Assistant Editor at Mashable, covering entertainment, culture, tech, science, and social good. Based in Australia, she writes about everything from video games and K-pop to movies and gadgets.

Mashable Potato

Recommended For You
OpenAI's GPT-5.6 finally set for public release after delays
The GPT-5.6 presentation page is displayed on a smartphone screen.

Anthropic's Claude Mythos, or a model like it, to get public release
Anthropic: AI safety company developing Claude language models. Founded by ex-OpenAI members, focusing on safe AI systems.

Anthropic CEO says AI growth is exponential. Anthropic research says otherwise.
Dario Amodei, CEO of Anthropic, looking perplexed in an interview

Anthropic pulls Claude Fable 5, Mythos 5 after Trump admin order
Anthropic Claude Fable logo on mobile device

Anthropic overtakes OpenAI to become world's most valuable AI company
Claude logo on mobile device in front of OpenAI on computer screen

Trending on Mashable
NYT Connections hints today: Clues, answers for July 10, 2026
Connections game on a smartphone

Wordle today: Answer, hints for July 10, 2026
Wordle game on a smartphone

NYT Connections hints today: Clues, answers for July 9, 2026
Connections game on a smartphone

NYT Strands hints, answers for July 10, 2026
A game being played on a smartphone.

FIFA World Cup schedule today: Games, kickoff times, livestream info for July 10
Lamine Yamal holds the ball
The biggest stories of the day delivered to your inbox.
These newsletters may contain advertising, deals, or affiliate links. By clicking Subscribe, you confirm you are 16+ and agree to our Terms of Use and Privacy Policy.
Thanks for signing up. See you at your inbox!