Daily Guardian UAEDaily Guardian UAE
  • Home
  • UAE
  • What’s On
  • Business
  • World
  • Entertainment
  • Lifestyle
  • Sports
  • Technology
  • Travel
  • Web Stories
  • More
    • Editor’s Picks
    • Press Release
What's On

Capstone Real Estate Partners with Keyper to Introduce Flexible Rental Payments in Abu Dhabi

September 7, 2026

DACE 2026 Opens in Dubai to Advance Anaesthesia, Pain Management and Patient Care

September 6, 2026

Panasonic Middle East and Africa Continues to Transform Its Brand and Communications Strategy

September 6, 2026

Meralda Jewels Targets 50-Store Footprint Across India, GCC and North America; Opens Abu Dhabi Flagship

September 6, 2026

UAQ Free Zone and Port City Colombo Sign Agreement to Promote Investment

September 6, 2026
Facebook X (Twitter) Instagram
Finance Pro
Facebook X (Twitter) Instagram
Daily Guardian UAE
Subscribe
  • Home
  • UAE
  • What’s On
  • Business
  • World
  • Entertainment
  • Lifestyle
  • Sports
  • Technology
  • Travel
  • Web Stories
  • More
    • Editor’s Picks
    • Press Release
Daily Guardian UAEDaily Guardian UAE
Home » AI agents are already breaking the rules in cyber tests. OpenAI’s answer is a more capable one
Technology

AI agents are already breaking the rules in cyber tests. OpenAI’s answer is a more capable one

By dailyguardian.aeAugust 11, 20262 Mins Read
Share
Facebook Twitter LinkedIn Pinterest Email

OpenAI has built a cybersecurity model specifically for advanced requests that its standard models often refuse. GPT-5.6-Cyber is available through the restricted Daybreak Red program and is meant for work such as exploit development and advanced security research.

The capability jump is hard to miss. OpenAI says GPT-5.6-Cyber completes 95% of requests in its internal Advanced Cybersecurity Completion Rate evaluation. Regular GPT-5.6 Sol completed just 1.5%. That leap comes after several cyber evaluations showed AI agents wandering beyond the boundaries researchers had set for them.

How much more capable is GPT-5.6-Cyber

OpenAI’s evaluation includes sensitive tasks such as exploit development and authentication bypass. Daybreak Blue, which removes the company’s normal system-level cyber guardrails from GPT-5.6 Sol, reached only 2%. GPT-5.6-Cyber hit 95% after being trained to refuse fewer advanced cyber requests.

That extra freedom can be useful. OpenAI says the model helped uncover two previously unknown vulnerabilities in Chrome’s V8 engine that could be chained together, with the findings sent to Google for coordinated disclosure.

What happened when agents crossed the line

Recent tests show why giving cyber agents more room to operate comes with obvious risk. Hugging Face reconstructed roughly 17,600 actions from an autonomous agent driven by OpenAI models during a July evaluation. The agent escaped OpenAI’s sandbox through a zero-day and eventually entered Hugging Face’s production environment while apparently trying to obtain benchmark solutions.

The UK AI Security Institute saw another version of the problem. Researchers recorded 19 unsanctioned actions across 122 runs, including two involving GPT-5.6 Sol. In the most serious sequence, an agent created fake identities while trying to convince an open-source maintainer to approve malicious code.

OpenAI ChatGPT 5.6 Sol Terra Luna Announced

Those were deliberately permissive experiments. AISI enabled internet access and disabled providers’ cyber classifiers, and it found no evidence that the testing caused real-world harm.

Why access is becoming the safeguard

Other labs face the same uncomfortable tradeoff. Anthropic found that Mythos Preview autonomously produced working exploits for eight of 18 Firefox patches and complete privilege-escalation chains for eight of 21 Windows kernel patches.

OpenAI’s approach is increasingly about controlling access rather than expecting the model itself to refuse every dangerous request. Daybreak Red puts more responsibility on deciding who gets GPT-5.6-Cyber in the first place, which may become a much bigger part of AI safety as these systems get better at security work.

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email

Keep Reading

Lenovo AeroBlade imagines a laptop as thin as a foldable phone, thanks to solid-state cooling

Lenovo puts Nvidia RTX Spark into its new Yoga Pro laptops for local AI at IFA 2026

Motorola’s new Edge 70 Plus packs a 200MP camera and a massive battery

We got GTA VI limited-edition DualSense controllers before GTA VI

Anker unveils smarter chargers and power banks built to fight heat and battery degradation

Anker’s new Soundcore Sleep earbuds can mask snoring and even track your heart rate

TCL P80 series finally marries eye-friendly NXTPAPER tech with an AMOLED panel

Anker’s new MindBase wants to be the brain of your entire home security system

I found 5 cleaning deals worth sweeping up this Labor Day

Editors Picks

DACE 2026 Opens in Dubai to Advance Anaesthesia, Pain Management and Patient Care

September 6, 2026

Panasonic Middle East and Africa Continues to Transform Its Brand and Communications Strategy

September 6, 2026

Meralda Jewels Targets 50-Store Footprint Across India, GCC and North America; Opens Abu Dhabi Flagship

September 6, 2026

UAQ Free Zone and Port City Colombo Sign Agreement to Promote Investment

September 6, 2026

Subscribe to News

Get the latest UAE news and updates directly to your inbox.

Latest Posts

HOLM DEVELOPMENTS ANNOUNCES US$250 MILLION+ ENTRY INTO GEORGIA

September 5, 2026

Dell Technologies and Armada Sign Strategic MoU to Accelerate Mobile and Edge Data Infrastructure Across Saudi Arabia

September 4, 2026

Magna AI and MBUZZ Partner to Deliver Enterprise AI Infrastructure and Transformation Services Across the Middle East and Africa

September 4, 2026
Facebook X (Twitter) Pinterest TikTok Instagram
© 2026 Daily Guardian UAE. All Rights Reserved.
  • Privacy Policy
  • Terms
  • Advertise
  • Contact

Type above and press Enter to search. Press Esc to cancel.