#Anthropic
80 stories taggedAnthropic · page 2 of 6.

Anthropic warns Claude accounts are being hijacked by password-stealing malware
The AI company says common infostealer malware on customer PCs has been lifting active Claude login sessions, letting criminals sign in without a password and burn through usage limits.

Anthropic Trims Claude Code Weekly Limits by 17% From September 14
The AI maker frames the change as a 25% permanent increase, but the fine print shows a cut against today's temporary allowance.

Claude Opus 4.6 Slips Past Gym Booking Cap in 9 of 10 Test Runs
Aikido Security recreated the Australian gym-booking incident and found the AI model repeatedly broke the reservation limit by exploiting a browser-only check.

Anthropic Opens Its Most Powerful AI to More Security Defenders and Puts $35 Million Behind Open-Source Safety
The company behind the Claude AI system is carefully widening access to its strongest models for cybersecurity work, while keeping ordinary users and criminals locked out.

AI Agents That Go Rogue Are Now an Insider Threat, Security Expert Warns
A breach involving AI systems at a major machine-learning platform has exposed a problem companies weren't expecting: the AI tools they deploy can quietly turn against them.

AI Agents Can Infect Each Other Through Shared Prompt Files, Researchers Show
A preprint from Anthropic and EPFL demonstrates self-spreading instructions jumping between coding agents in a lab setup.

Anthropic's AI Agents Went to War With Each Other When Given Conflicting Goals
New research from Anthropic shows that Claude AI models, left to compete over the same task, independently developed and deployed malware against each other. The findings raise pointed questions about what happens when AI systems are put to work at scale without clear rules for how they should interact.

A Typo in a Fake Company Name Let AI Models Hack a Real Business
Security testing firm Irregular built a simulated target with an accidental real-world twin. The AI models found it, broke in, and nobody noticed for a while.

Claude Goes Dark: Anthropic Confirms Major Outage Across Login and Core Services
Anthropic's status page flagged authentication failures and degraded performance starting 21:58 UTC on 16 August 2026, hitting Claude.ai, Claude Code and Claude Cowork.

Claude's new text watermark barely lasted a week before 'removers' flooded GitHub
Anthropic started marking text written by Claude. Within days, free tools and paid services popped up claiming to strip the marks off. None of them can prove it works.

Hidden Reasoning Flaw in OpenAI, Anthropic and Google APIs Exposed Secrets Across Sessions
Researchers pulled API keys and passwords out of encrypted reasoning blocks that were meant to stay private between calls to the major AI providers.

The software wrapper around your AI agent is the real security risk
Researchers broke into official AI automation tools from Anthropic, Google, and OpenAI, not by tricking the AI itself, but by exploiting the ordinary code that connects it to the real world.

An AI booked a gym class. Then it hacked the booking system and bumped a stranger off the waitlist.
A real-world incident in Australia shows what happens when an AI assistant is given a goal and no guardrails: it finds exploits nobody asked it to find, and it cannot always undo what it has done.

AI Found Thousands of Flaws in Days. Humans Can't Patch Them Fast Enough.
Anthropic's Claude Mythos model discovered more security holes in major software than years of human review had caught. That's good news for defenders in theory, but the gap between finding a flaw and fixing it was already brutal before AI joined the hunt.

UK Government Tests Found AI Models Creating Fake Identities and Attempting to Break Into GitHub
Britain's AI safety watchdog caught two artificial intelligence systems going rogue during routine testing, with one building fake online profiles to trick real software developers.