Browsing Tag
Anthropic
19 posts
Anthropic Finds 4th Claude AI Hacking Incident Missed in Earlier Review
Anthropic reveals a 4th Claude AI hacking incident after Opus 4.6 accessed a real system, leading to a review of 481 million evaluation transcripts.
September 10, 2026
Why Hiding Chain-of-Thought Alone Doesn’t Stop Distillation Attacks
Distillation attacks have escalated into a primary security threat for AI model providers in 2026. The US government…
September 7, 2026
Fake Claude Opus 5 App Spreads New RevStealer Windows Malware
Cybersecurity firm Morphisec has uncovered a new Windows infostealer called RevStealer, which is being distributed through a fake…
August 31, 2026
Black Hat USA 2026: One GitHub Issue Could Compromise Major AI Coding Workflows
At Black Hat USA 2026, Novee found GitHub workflow flaws in Claude Code, Gemini CLI and Codex that enabled RCE, credential theft and agent control in pipelines.
August 6, 2026
Anthropic Says Claude Models Hacked 3 Organizations During Cyber Tests
Anthropic found Claude accessed systems at three real businesses after a testing error gave its AI models live internet access during cybersecurity evaluations.
July 31, 2026
Google Indexed Claude AI Shared Chats Before Results Were Removed
Google Search indexed public Claude AI chat links, letting people find shared conversations through the site query before those listings disappeared soon after.
July 27, 2026
PromptFiction Flaw Auto-Submitted Hidden Prompts in Claude Desktop
A one-click Claude Desktop flaw allowed attackers to submit concealed instructions without review, exposing chats and enabling code execution on some systems remotely.
July 15, 2026
GhostApproval Flaws Let Top AI Coding Tools Write Outside Workspaces
Wiz found GhostApproval symlink flaws in major AI coding assistants that could hide sensitive file targets, bypass approval checks and enable system access too.
July 9, 2026
Fake ChatGPT Desktop App Ads Used to Push Password-Stealing Malware
Fake ChatGPT desktop app ads pushed password-stealing malware by abusing trusted AI links, hiding from scanners, and tricking users into downloads.
June 2, 2026
Claude Mythos AI Identified 10,000+ Software Vulnerabilities in One Month
Anthropic says its Claude Mythos AI identified more than 10,000 software vulnerabilities in one month, including critical flaws in open-source code.
May 26, 2026