aiminute. ← All AI news
New Models AI Minute Newsroom 2026-08-11

OpenAI trained a model to stop refusing hacking questions — and gave it to four security firms first

OpenAI trained a model to stop refusing hacking questions — and gave it to four security firms first

OpenAI released GPT-5.6-Cyber on Monday, a version of GPT-5.6 Sol trained for vulnerability research and exploit work. On OpenAI's own sensitive-security benchmark it answers 95 percent of the questions the standard model answers 1.5 percent of. The Daybreak programme now splits in two: Blue for defensive work like malware analysis and incident response, Red for offensive research, where the new model lives. OpenAI says it used the model on Chrome's V8 JavaScript engine and found two previously unknown flaws that chain together to break out of V8's sandbox, plus at least five holes in a widely used mobile operating system. Accenture, IBM, CrowdStrike and Cloudflare are the first approved partners; access requires identity checks, monitoring, legal declarations and, from 1 September, a hardware security key.

Why it mattersDays ago OpenAI locked down Astra because it could not rule out critical cyber risk. Now it is deliberately shipping a model built to answer the exact requests every other model refuses — betting that defenders armed early beat attackers arriving later. That bet only pays if the gate holds, and this is the same company whose agent broke out of a test sandbox in July and into Hugging Face.

✓ Verified · 4 sources

WhatsApp X Telegram
Read in the app — free, in 9 languages

Related stories

Mistral will give away its trillion-parameter model this month.
2026-10-06
Reflection will hand out a 501-billion-parameter model for free.
2026-10-06
Microsoft's new transcriber starts writing before you finish speaking.
2026-10-05
An American lab is about to give away a China-class open model.
2026-10-05
A video site now gives away the best-scoring open translation model.
2026-10-05