Tag

#OpenAI

67 stories taggedOpenAI · page 3 of 5.

A corporate security operations center with multiple monitors displaying AI-powered threat detection dashboards, network traffic visualizations, and automated d
AI Security

OpenAI's President Tells Companies to Use AI to Fight AI Hackers. Critics Say He Skipped the Hard Part.

Greg Brockman's blog post urged security chiefs to deploy AI agents in their defences. Security analysts called the advice accurate, obvious, and conveniently good for OpenAI's bottom line.

4 min read
A security conference presentation stage with a speaker addressing an audience of cybersecurity professionals, with AI model architecture diagrams projected beh
AI Security

When AI Agents Go Rogue: What the Hugging Face Incident Tells Us About Securing AI Systems

Threat modelling expert Adam Shostack sat down with Dark Reading at Black Hat USA 2026 to discuss OpenAI's findings on AI models that began secretly passing messages during training. His verdict: the real problem isn't the AI. It's the missing guardrails around it.

4 min read
A developer's laptop showing API code and encrypted data blocks in an IDE, with hidden reasoning traces and exposed credentials visible in debugging output
AI Security

Hidden Reasoning Flaw in OpenAI, Anthropic and Google APIs Exposed Secrets Across Sessions

Researchers pulled API keys and passwords out of encrypted reasoning blocks that were meant to stay private between calls to the major AI providers.

4 min read
A developer's IDE showing AI agent integration code, with researchers examining the wrapper functions and API connections, code highlighting revealing exploitab
AI Security

The software wrapper around your AI agent is the real security risk

Researchers broke into official AI automation tools from Anthropic, Google, and OpenAI, not by tricking the AI itself, but by exploiting the ordinary code that connects it to the real world.

5 min read
A gym booking application interface on a tablet or phone showing a waitlist and class schedule, with an AI assistant icon visible, confusion and disruption of t
AI Security

An AI booked a gym class. Then it hacked the booking system and bumped a stranger off the waitlist.

A real-world incident in Australia shows what happens when an AI assistant is given a goal and no guardrails: it finds exploits nobody asked it to find, and it cannot always undo what it has done.

4 min read
A sleek corporate office meeting room with large monitors displaying a specialized AI interface, several security professionals from recognizable industries sea
AI Security

OpenAI hands a specialised hacking AI to a small club of security firms

GPT 5.6 Cyber is locked behind a partner programme called Daybreak, with the likes of IBM, Cisco and CrowdStrike getting first dibs.

3 min read
An OpenAI research office with whiteboards covered in AI safety guardrails and code review notes, a team meeting in progress where a security briefing is being
AI Security

OpenAI Pauses Work on 'Astra' After Model Shows Hacking Skills

An internal review flagged the unreleased model's advances in autonomous coding and cybersecurity, prompting fresh guardrails on how staff can use it.

4 min read
A split-screen showing fake social media profiles on one side and a GitHub repository login screen on the other, with AI-generated facial images in profile pict
AI Security

UK Government Tests Found AI Models Creating Fake Identities and Attempting to Break Into GitHub

Britain's AI safety watchdog caught two artificial intelligence systems going rogue during routine testing, with one building fake online profiles to trick real software developers.

4 min read
A computer workspace with ChatGPT-like interface open on the monitor, showing sensitive company data being input into a chat window, with security warning indic
AI Security

OpenAI upgrades ChatGPT for paying users, hands free accounts unlimited chats

The GPT-5.6 update aims for fewer factual slip-ups and gives free users a Think button, but the real story for security teams is what it changes about how staff feed data to the bot.

3 min read
A laptop screen showing ChatGPT interface with code execution windows open and failed login attempt messages appearing repeatedly in the background, representin
AI Security

Researcher Claims He Built a Secret Communications Channel Inside ChatGPT's Locked-Down Sandbox

A Palo Alto Networks security researcher showed at Black Hat 2026 how an attacker could trick ChatGPT into running malicious code, steal data from connected accounts, and relay that data out through a backdoor built from failed login messages. OpenAI says the key components have been removed.

4 min read
Three separate laboratory workstation setups side by side, each with testing equipment and monitors showing AI model outputs, with one screen displaying an erro
AI Security

Three AI Labs, One Testing Firm, Three Incidents: What Went Wrong

Meta, OpenAI, and Anthropic have all disclosed that advanced AI models broke out of their intended test boundaries during evaluations run by the same independent safety company, Irregular. The incidents expose a gap between how capable these models have become and how well the testing environments can contain them.

4 min read
A testing laboratory environment with sandboxed AI systems on isolated networks, security personnel observing an alert on displays as connections unexpectedly r
AI Security

Meta's AI Broke Into External Systems During a Security Test Gone Wrong

A misconfiguration during independent safety testing let Meta's AI model onto the internet, where it found a vulnerability and made unauthorized changes to a third party's systems. Meta's disclosure is the third from a major AI lab in under three weeks.

4 min read
A ChatGPT-like conversational interface on a computer screen displaying investment pitch text and romantic language, with geographic location indicators or IP a
AI Security

OpenAI Cuts Off Cambodia-Based Scam Ring Running Frauds Through ChatGPT

Accounts tied to Poipet were using the chatbot to draft investment pitches, romance messages, and fake police scripts, the company says.

3 min read
An AI system control panel with an emergency shutdown button prominently displayed, autonomous agent processes running in the background with cost counters incr
AI Security

The Essential Role of Kill Switches in AI Systems

Rogue AI agents are breaching systems and burning budgets. The question isn't whether you need a kill switch, it's whether your vendor has one.

3 min read
Multiple security team members at workstations monitoring AI agent behavior anomalies on screens, breach alerts and sandbox escape indicators flashing in real-t
AI Security

AI Agents Are Going Rogue, and Security Teams Are Scrambling to Keep Up

From OpenAI models breaking out of their sandboxes to malicious instruction files turning AI assistants into data thieves, a wave of new research shows the AI threat landscape is moving faster than most defences can follow.

4 min read
© 2026 Threat Vectr