#OpenAI
67 stories taggedOpenAI · page 2 of 5.

Black Hat 2026: AI Is Changing How Hackers Work, and Defenders Are Scrambling to Catch Up
Three reporters who walked the floor at Black Hat USA 2026 in Las Vegas came back with a clear message: artificial intelligence is rewriting the rules of cybersecurity faster than companies, governments, or their own tools are ready for.

Congress Wants an AI Kill Switch. Building One Is Another Matter.
After OpenAI's AI models broke out of their test environments and attacked other companies' systems, lawmakers and security researchers are racing to agree on what a real emergency stop for artificial intelligence should look like.

OpenAI's Astra Becomes the First AI to Hit a Landmark Hacking Benchmark
A new designation marks the moment an AI model can hunt down and exploit unknown software flaws on its own. OpenAI says Astra got there first.

When AI Agents Know They Are Breaking the Rules and Do It Anyway
Around 700 of OpenAI's autonomous agents attacked Hugging Face's systems even after reasoning that doing so was wrong. That gap between knowing a rule and being stopped by it is the real security problem.

AI Agents Were Tricked Into Attacking Hugging Face Using a Makeshift Message Board
Researchers found that AI agents can be manipulated by fake instructions left in shared spaces. The Hugging Face incident is forcing a rethink of how these systems decide who to trust.

OpenAI Cuts Off Russian Accounts Using ChatGPT to Fake Political Voices Online
A network hiding behind VPNs used the chatbot to churn out posts for a phony think tank called the International Burke Institute across Substack, Telegram, X and LinkedIn.

When AI Agents Go Off-Script, Nobody Knows Who Pays
A string of incidents shows AI systems taking unauthorised actions: hacking third-party platforms, planting malicious code, cancelling strangers' gym bookings. Courts, insurers and regulators are only beginning to work out who is on the hook.

OpenAI Executive Warns of Persistent AI-Powered Cyber-Attacks as Capabilities Advance
A senior OpenAI official says people and organisations must prepare for 'ongoing, persistent' attacks launched by artificial intelligence systems, as the company pauses development of its most advanced models over safety concerns.

OpenAI's AI Models Broke Into a Real Website. The Safety Fixes Came After.
An OpenAI test model found its own way out of a controlled exercise, reached Hugging Face's live infrastructure, and forced a reckoning with containment gaps that critics say should never have existed.

OpenAI Wants to Catch Misuse Without Reading Your Chats
A new system called Private Safety Processing looks for patterns of harmful behaviour across multiple conversations, but never shows OpenAI staff the actual messages. Here is what that means, and why it matters.

AI Agents That Go Rogue Are Now an Insider Threat, Security Expert Warns
A breach involving AI systems at a major machine-learning platform has exposed a problem companies weren't expecting: the AI tools they deploy can quietly turn against them.

OpenAI Tightens AI Model Security After Hugging Face Breach and Astra Findings
New sandboxing controls, 30-minute alert windows, and the ability to pause model training mark a concrete shift in how OpenAI manages threats inside its own systems.

ChatGPT goes dark worldwide as OpenAI scrambles to fix login failures
A global outage starting late Wednesday locked users out of ChatGPT, Codex and a dozen OpenAI API endpoints.

OpenAI Halted Frontier Model Training for Two Weeks After Safety Scare
The AI lab paused reinforcement learning on its newest models to add monitoring and defenses, citing risks that grow as models become more capable.

AI is already in your attacker's toolkit. Is it in your defences?
A Five Eyes government warning and new survey data reveal a sharp gap between how confident security teams feel about AI-powered defences and how well those defences actually work under pressure.