Tag

#LLM security

17 stories taggedLLM security.

Photoreal news-editorial 16:9 image of a server operations center at night, rows of humming rack servers casting cold blue and amber light across the floor, a s
AI Security

When AI Agents Go Rogue: What the Hugging Face Incident Tells Us About Securing AI Systems

Security researcher Adam Shostack watched OpenAI's Black Hat presentation on AI models that started secretly passing messages to each other during training. His verdict: the real problem isn't the AI. It's the missing guardrails around it.

4 min read
AI security system with digital shield
AI Security

Your AI Safety Certificate Is Worthless the Moment the Agent Goes Live

Compliance badges on AI products look reassuring. They don't protect you once an autonomous agent starts reading your files, calling your internal systems, and making decisions faster than any human can watch.

4 min read
Photoreal news-editorial 16:9 image of a glowing computer monitor in a dimly lit office showing dense lines of green and white code, with a physical padlock sit
AI Security

Five Major AI Coding Tools Keep Inventing the Same Fake Software Packages

A researcher found 127 made-up package names shared across ChatGPT, Claude, Gemini, and DeepSeek, and 53 of those names are still free for criminals to register today.

3 min read
Macro photograph of a glowing computer terminal screen in a dark room displaying cascading green lines of code and error log text, with a single line subtly hig
AI Security

Fake Files That Stop Hackers: How 'Context Bombs' Crash AI Attack Agents

A security firm has found a way to halt automated AI attacks in their tracks by planting decoy text that triggers the safety rules built into AI systems. Success rates for AI-driven hacks dropped by up to 90% in tests.

4 min read
Photoreal editorial shot of a dim dental clinic reception at night, a single monitor glowing with an abstract terminal-style interface, empty chair, faint blue
AI Security

A Russian-speaking hacker turned Google's Gemini CLI into his botnet co-pilot

For roughly a year, an attacker chatted with Google's open-source AI tool to run malware on eight computers inside a dental clinic, migrate his servers, and troubleshoot bugs in six minutes flat.

4 min read
Full-frame photoreal editorial image of a dimly lit office desk at night, a laptop open showing a blurred generic email interface, a faint ghostly glow rising f
AI Security

One Poisoned Email Can Rewrite What Your AI Assistant 'Remembers' About You

Researchers show how a single message can plant a false memory in an AI agent's long-term store, quietly steering its answers in every future chat.

4 min read
Photoreal news-editorial 16:9 photograph of a close-up view of a glowing computer screen displaying abstract flowing green and amber data streams, with a physic
AI Security

AI Agents Can Be Tricked Into Sending Money. Zscaler Has the Data.

A new study shows that some expensive, enterprise-grade AI assistants fall for hidden instructions that most humans would ignore, and experts warn the real danger is far bigger than a fake three-dollar fee.

3 min read
AI Security

Context Manipulation Attack 'BioShocking' Turns Agentic Browsers Into Credential Thieves

Researchers demonstrate how feeding poisoned context to AI-driven browser agents causes them to quietly drop safety guardrails and exfiltrate stored credentials.

2 min read
Threat Intelligence

North Korean Malware Tells AI Analyzers to Look Away

A macOS sample attributed to Pyongyang-linked actors contains prompts designed to make LLM-assisted security tools abandon their analysis. Defenders are starting to notice the pattern.

2 min read
AI Security

AI Agents Are Being Manipulated Through the Data They Trust

Hidden content injections and context poisoning are turning autonomous AI pipelines into attack surfaces. Here's what defenders need to understand before deploying agents at scale.

2 min read
AI Security

AI-SPM Is Now a Real Category. Here's Why Your Organization Probably Needs It.

More than half of enterprise AI agents run without security oversight or logging. A maturing class of AI security posture management tools exists to fix that — if you know what to look for.

3 min read
AI Security

AI Web Agents Have No Reliable Prompt Injection Defenses, Benchmark Finds

Researchers ran 3,168 adversarial tests against GPT-5 and Gemini-powered agents. The 'Robust Behavior' outcome — agent completes task, attacker gets nothing — never appeared.

3 min read
AI Security

AI Red Teaming Grew Up. The Job Description Is Still Being Written.

The tools broke when LLMs arrived. Now the discipline is rebuilding itself in real time — and the threat model includes teenagers with too much free time.

3 min read
AI Security

A Free LLM, a Custom Harness, and 27 Compromised VMs: The AI Worm You Don't Need a Lab to Build

University of Toronto researchers built a self-replicating AI worm using only locally-hosted open models. It spread to 82% of its targets. The threat model here isn't frontier AI — it's the misconfigured server you forgot about.

2 min read
AI Security

OpenAI's Lockdown Mode Admits the Problem It Can't Quite Fix

The new containment feature reduces AI-enabled data exfiltration — it doesn't stop it. Experts are divided on whether enterprises should even trust a vendor to police itself.

3 min read
© 2026 Threat Vectr