OZZZER · AI NEWS3 of 3 free stories opened
← Back to AI News

Safety & Security · 17 Aug 2026 · 07:30 CEST

The Defender’s Window

OpenAI · 17 Aug 2026 · 07:30 CESTRead original at OpenAI ↗
Share
LinkedInX

Publisher preview · OZZZER analysis pending editorial review.

PUBLISHER ARTICLE PREVIEW

From the original article

The OpenAI-Hugging Face incident⁠ was a watershed moment for cybersecurity because it gave a peek into how the capabilities of a typical threat actor will evolve in upcoming months. I’ve spoken with many organizations over the past few weeks, and one theme is clear: they know they need to fundamentally uplevel their cybersecurity practices with unprecedented speed.

In this post, I’ll share what we’re doing to defend OpenAI, concrete steps other organizations can take today, and why now is the time to act.

AI models developed around the world are increasingly able to automate parts of real-world cyberattacks, making longstanding security gaps—from bugs buried deep in human-written software to forgotten permissions—easier to find and exploit. The same AI capabilities give defenders new ways to find and fix those weaknesses, but they need to move now. If companies act decisively—including improving their fundamentals and superpowering their teams with AI—we can make the internet more secure than it has ever been.

In the OpenAI-Hugging Face Incident, an agentic collective was able to autonomously penetrate not just OpenAI research infrastructure but also the production infrastructure of another company, chaining together vulnerabilities ranging from previously-unknown security flaws to using credentials to user accounts that had been leaked onto the internet. It is increasingly clear that the tech debt⁠ of every company masks significant flaws, and defenders need to find and fix them before attackers do.

To advantage defenders relative to attackers, earlier this year we began releasing our cyber capabilities only to trusted defenders. Since then, various companies have released broadly diffused models with cyber capabilities only a few months behind the frontier. The most recent of these models appears slated to be released⁠ at the end of August, and seems likely to significantly accelerate the threat landscape.

While AI-powered attackers will soon be able to find longstanding flaws in many existing systems, AI will also make it much easier for defenders to find, prioritize, and fix those same flaws. Security is still a cat-and-mouse game, but AI may shift its economics⁠ in ways that fundamentally advantage defenders. For example, we are starting to train our models specifically to write superhumanly secure code.

Our models are also incredible at mathematical proofs, which can be applied to formally verify the security of software in a way that has proven intractable for humans.

After the OpenAI-Hugging Face incident, I asked ChatGPT Work (using publicly available GPT‑5.6 Sol) to assess the security of gregbrockman.com⁠. It’s a simple static site, hosted on AWS with Cloudflare as a frontdoor, so I figured there wouldn’t be much surface area for vulnerabilities.

In about 15 minutes, it uncovered 13 issues, many of which probably aren’t exploitable on their own—but I could imagine them being chained together with other vulnerabilities to significant effect. I hadn’t configured my DNS records to prevent attackers from forging emails from me; my site used an insecure version of jQuery; Cloudflare was forwarding requests to AWS over unencrypted HTTP.

I then asked ChatGPT Work to fix these issues, which it did over the course of an hour. It opened the Cloudflare control panel in my browser, and proceeded to click many buttons to configure DNS,

Source

OpenAI · 17 Aug 2026 · 07:30 CEST

Open the original at OpenAI ↗