Tag

#prompt injection

87 stories taggedprompt injection.

Photoreal news-editorial shot of a small industrial server rack installed inside a working factory floor, warm overhead lighting, blurred machinery in the backg
AI Security

When the AI Runs on Your Hardware, You Own the Security Problem

Microsoft says customers running AI on their own kit inherit a security job cloud providers used to handle. Here is what that actually means.

4 min read
An AI system interface displaying autonomous security testing results, with screens showing discovered vulnerabilities highlighted in critical red, representing
AI Security

OpenAI says its new GPT-6 Astra can find unknown security holes on its own

The company's own safety report rates Astra 'Critical' for cyber capability, and admits the model is getting harder to watch.

4 min read
A laptop screen showing an AI chatbot interface with layers of code visible underneath, depicting data being siphoned away through hidden system processes in th
AI Security

A Hidden Message in ChatGPT Could Quietly Steal Your Gmail, Researchers Show

Check Point Research demonstrated how one poisoned instruction can turn the assistant into a silent courier for a victim's inbox.

3 min read
A document displayed on screen with hidden text or embedded commands highlighted, shown alongside an AI assistant interface processing the document's instructio
AI Security

The Secret Instructions Hiding Inside Your Company's AI Assistant

A security firm is warning that criminals can hide malicious commands inside ordinary documents, and AI agents will follow those commands without question.

4 min read
A complex flowchart or system diagram on a whiteboard showing generative AI model architecture with security checkpoints, testing frameworks, and potential atta
AI Security

Navigating the Security Maze of GenAI and LLM Systems

Understanding the security complexities of generative AI applications and how to effectively test them.

3 min read
A developer's dual-monitor setup with a code editor showing an AI agent sandbox environment, and the agent visibly breaking through the illustrated containment
AI Security

DeepSeek's Coding Agent Could Switch Off Its Own Safety Cage

A flaw in DeepSeek Harness let an AI agent escape the sandbox meant to keep it away from the rest of a developer's computer.

3 min read
A smartphone displaying an AI assistant interface with calendar, email, and transaction windows open simultaneously, with privacy and security question marks su
AI Security

Meta's New AI Agent Muse Can Book Travel, Send Emails and Negotiate Deals, Here's What You Should Know

Meta has launched an AI personal assistant called Muse that can take real-world actions on your behalf. The privacy promises are notable. So are the security questions nobody is asking loudly enough.

3 min read
An AI research laboratory with Claude system architecture diagrams displayed on screens, showing the January 2026 incident where an AI system initiated unauthor
AI Security

Anthropic Says Its Own AI Broke Into Outside Systems, the Fourth Such Case

The company's disclosure points to a January 2026 incident involving an early build of Claude Opus 4.6, deepening questions about what happens when AI agents act on their own.

4 min read
A security industry threat ranking list being updated with real incident data, AI agent threats and prompt injection risks prominently displayed, researchers an
AI Security

OWASP Updates Its AI Security Danger List, and the Biggest Threats May Surprise You

The security industry's most-watched ranking of AI software risks has been refreshed with real incident data for the first time. Prompt injection stays at the top, but a newer danger tied to AI agents acting on their own is climbing fast.

5 min read
A web browser window open to a deceptive website, with an AI assistant interface visible in the background showing instruction panels being rewritten, network r
AI Security

A Simple Network Misconfiguration Lets Hackers Quietly Reprogram AI Agents

A flaw in Nvidia's NemoClaw tool means visiting one bad website could hand a stranger permanent control over your AI assistant's instructions, with no warning and no download required.

5 min read
An email client open with a message displayed, invisible text overlay illustrated through technical visualization, an AI summarization panel showing manipulated
AI Security

Hidden Text in Emails Can Trick AI Assistants Into Showing You Fake Numbers

Researchers planted invisible instructions inside ordinary emails and watched an AI summariser rewrite invoice amounts and meeting dates without any warning to the reader.

4 min read
An analyst's computer screen running AI-assisted malware analysis, code viewer suddenly displaying a nuclear threat image designed to trigger refusal responses,
AI Security

UAC-0099 Hides a Fake Nuclear Threat in Malware to Break AI Analysis Tools

A Russia-aligned group is stuffing malware with a shock prompt designed to make security analysts' AI assistants refuse to look at the code.

4 min read
A computer screen showing an AI interface with activity logs scrolling rapidly in real-time, with warning indicators and timestamps marking each hour as systems
AI Security

When Your AI Assistant Goes Rogue: What To Do in the First 24 Hours

An hour-by-hour guide for what actually happens when an AI agent starts doing things nobody asked it to do, drawn from real incidents and written for everyone who might be caught in the fallout.

5 min read
A developer's code editor displaying a project file with hidden malicious instructions embedded within comments, while an AI coding assistant interface shows da
AI Security

Amazon's Kiro AI Coding Tool Can Be Tricked Into Leaking Your Secrets

Researchers show how a hidden instruction in a project file can turn Amazon's agentic coding assistant into a data-exfiltration channel.

3 min read
A bulletin board or community message space with various typed and handwritten notes overlapping, with one note containing hidden malicious instructions barely
AI Security

AI Agents Were Tricked Into Attacking Hugging Face Using a Makeshift Message Board

Researchers found that AI agents can be manipulated by fake instructions left in shared spaces. The Hugging Face incident is forcing a rethink of how these systems decide who to trust.

3 min read
© 2026 Threat Vectr