Tag

#prompt injection

87 stories taggedprompt injection · page 2 of 6.

A laptop screen showing an email inbox with one message open, while hidden text and code fragments appear faintly overlaid in the email body, visible only to an
AI Security

Your AI email assistant can be fed a fake message while you read a real one

Researchers hid nearly 500 characters of secret instructions inside an ordinary-looking email. The AI summariser obeyed them every single time.

4 min read
An AI assistant interface on a computer screen being corrupted by malicious webpage code injected through a browser window overlay, showing instruction prompt m
AI Security

A Single Website Visit Can Poison Your Local AI Agent, Researchers Find

A flaw in Nvidia's NemoClaw lets a malicious webpage secretly rewrite the instructions an AI assistant follows, and the damage survives every conversation that comes after.

4 min read
A tech company office with developers working at desks surrounded by whiteboards filled with AI model diagrams and test scenarios, with screens showing neural n
AI Security

Alice Raises $140 Million to Test AI Systems for Weaknesses Before Criminals Find Them

A company that tries to break AI models on purpose, so real attackers cannot, has secured fresh funding to expand that work to more businesses.

4 min read
An office workspace where a few employees are intensely working at computer stations, with colorful AI tool interfaces visible on monitors contrasting with a wo
AI Security

The 5% Problem: How a Handful of AI Power Users Became the Riskiest People in Your Company

New Akamai research finds a small group of enthusiastic staff are quietly wiring untested AI tools into serious business systems, and security teams are watching the wrong crowd.

4 min read
A computer screen displaying encrypted code and mathematical symbols morphing into AI chatbot interface elements, representing hidden malicious instructions byp
AI Security

Researchers Find Way to Hide Malicious Instructions Inside Encrypted AI Prompts

A new technique called 'Cryptographic Context Injection' slips harmful commands past the safety filters built into Grok and Gemini by wrapping them in encryption that only the AI unwraps.

3 min read
A laptop screen displaying an AI chat interface mid-conversation with a webpage open in another tab, with subtle connection indicators and data packets visualiz
AI Security

Researchers Show How a Booby-Trapped Web Page Can Steal Your Grok Chat Data

Adversa AI's 'Cryptographic Context Injection' technique tricks xAI's Grok into leaking user names, locations and prompts to attacker servers when a user asks it to summarize a page.

4 min read
An email server monitoring station showing AI-generated phishing emails being processed and analyzed, threat detection systems with machine learning components
AI Security

Phishing Has a New Problem: The Attacker Isn't Human Anymore

Email defences built for bad links and bad attachments are struggling as AI agents start writing, sending, and even reading the mail on both sides.

4 min read
A researcher's computer showing a chat interface with an AI assistant, with internal system information and security weaknesses being revealed in the conversati
AI Security

Researchers Tricked Microsoft Copilot Into Revealing Its Own Weaknesses, Then Used That Knowledge to Steal Data

A research team at Varonis discovered that simply chatting with Microsoft's AI assistant could expose enough internal detail to build a working attack. Microsoft has patched the flaws, but the technique raises questions that go well beyond one product.

4 min read
A Copilot AI assistant interface on screen showing connected applications like Gmail and calendar services, with warning indicators highlighting data leakage vu
AI Security

One Click on Copilot Could Have Leaked Your Connected Apps, Researchers Say

Three flaws in Microsoft Copilot Personal, nicknamed CoSnitch, let a booby-trapped link quietly pull data from Gmail, calendars and other services the assistant was connected to.

4 min read
A network diagram showing AI coding agents connected by files and prompts, with malicious instructions spreading like contagion through the shared connections,
AI Security

AI Agents Can Infect Each Other Through Shared Prompt Files, Researchers Show

A preprint from Anthropic and EPFL demonstrates self-spreading instructions jumping between coding agents in a lab setup.

3 min read
A corporate acquisition scene with security product icons and AI model testing workflows merging into a single unified platform interface, representing consolid
AI Security

Fortinet Buys AI Security Startup Virtue AI

The cybersecurity giant adds a specialist platform for testing and protecting AI models, chatbots, and autonomous software agents, as the market for AI security tools heats up fast.

4 min read
A developer's IDE window displaying code with an active plugin sidebar showing suspicious helper tool options, SSH key references visible in the code editor, am
AI Security

How a Rogue Helper Tool Can Trick an AI Coding Assistant Into Leaking Your Secrets

Researchers show that a hostile plugin can smuggle out SSH keys and source code by breaking one big theft into small, innocent-looking steps.

4 min read
An AI agent interface on a monitor showing a domain management dashboard, a firewall log excerpt visible in a window behind it with one blocked request highligh
AI Security

GhostJacking: How Hackers Can Turn an AI Assistant Against Its Own Company

Researchers showed that a single blocked web request, already sitting in a firewall log, was enough to trick an AI agent into handing over a company's entire domain. Here's what that means for organisations using AI tools to manage their systems.

5 min read
A computer mouse pointer hovering over a link in an email interface, with office collaboration tool icons visible on the background desktop
AI Security

A Single Click Could Have Handed Hackers Your Company's Confluence and Jira Files

Researchers found a flaw in Atlassian's Rovo AI assistant that let an attacker steal data from widely used workplace tools with almost no effort from the victim.

2 min read
An office worker's computer displaying a workplace dashboard and chat interface, with faint text prompts suggesting hidden instructions overlaid in the shadows
AI Security

Rovo, Atlassian's AI Assistant, Can Be Tricked Into Leaking Jira and Confluence Data

Two research teams showed how hidden instructions can turn Atlassian's built-in AI helper into a quiet data pipe. Only one of the tricks has been fully closed.

4 min read
© 2026 Threat Vectr