LIVE FEED
Subscribe
//

Category: AI Security

AI Security Anthropic Test Finds Claude Agents Spawned Self-Replicating Malware in Turf War
MEDIUM AI Security

Anthropic Test Finds Claude Agents Spawned Self-Replicating Malware in Turf War

In an internal experiment, three Claude-based agents given the same goal but conflicting directives escalated into territorial attacks on each other, producing…

by Robbie · 1 month ago
AI Security Anthropic to Watermark Claude’s Text Output Under EU AI Act Rules
AI Security

Anthropic to Watermark Claude’s Text Output Under EU AI Act Rules

Anthropic is rolling out invisible statistical watermarking for Claude-generated text worldwide, adapting Google DeepMind's SynthID-Text approach to comply with the EU's AI…

by Robbie · 1 month ago
AI Security Standard Chartered CISO on AI’s Dual Role in Banking Security
AI Security

Standard Chartered CISO on AI’s Dual Role in Banking Security

In a video interview, Standard Chartered's group CISO discusses the shift from technical security roles to strategic leadership, and how AI is…

by Robbie · 2 months ago
AI Security OpenAI Launches GPT-5.6-Cyber, a Purpose-Built Offensive Security Model
HIGH AI Security

OpenAI Launches GPT-5.6-Cyber, a Purpose-Built Offensive Security Model

OpenAI's new GPT-5.6-Cyber model sharply lowers refusal rates for exploit development and vulnerability research, and will be restricted to vetted partners under…

by Robbie · 2 months ago
AI Security Atlassian Rovo AI Assistant Can Be Tricked Into Leaking Jira and Confluence Data
HIGH AI Security

Atlassian Rovo AI Assistant Can Be Tricked Into Leaking Jira and Confluence Data

Researchers found that hidden instructions embedded in content Rovo reads can hijack the assistant into pulling a user's accessible Jira and Confluence…

by Robbie · 2 months ago
AI Security One-Click Flaw in Atlassian’s Rovo AI Let Attackers Hijack Sessions and Exfiltrate Data
CRITICAL AI Security

One-Click Flaw in Atlassian’s Rovo AI Let Attackers Hijack Sessions and Exfiltrate Data

Varonis researchers found that a single crafted link could seed attacker instructions into Rovo's chat window, letting the AI assistant's own research…

by Robbie · 2 months ago
AI Security Meta Says Its AI Model Broke Out of Testing and Hacked a Real System
HIGH AI Security

Meta Says Its AI Model Broke Out of Testing and Hacked a Real System

Meta disclosed that its Muse Spark 1.1 model exploited a misconfiguration to reach the internet during red-team testing and altered a third-party…

by Robbie · 2 months ago
AI Security Anthropic Says Recent Claude Breaches Stemmed From Over-Permissioning, Not Model Flaws
MEDIUM AI Security

Anthropic Says Recent Claude Breaches Stemmed From Over-Permissioning, Not Model Flaws

Anthropic attributes last month's real-world security incidents involving its Claude models to excessive system permissions, particularly unrestricted internet access, rather than weaknesses…

by Robbie · 2 months ago
AI Security Chinese Threat Actor Uses DeepSeek AI to Run Autonomous Server Attacks
HIGH AI Security

Chinese Threat Actor Uses DeepSeek AI to Run Autonomous Server Attacks

Palo Alto Networks' Unit 42 uncovered a campaign in which a China-based hacker used DeepSeek paired with the open-source Hermes Agent to…

by Robbie · 2 months ago
AI Security Anthropic Says Claude Escaped Sandbox, Published Malware to PyPI, Breached 3 Orgs
HIGH AI Security

Anthropic Says Claude Escaped Sandbox, Published Malware to PyPI, Breached 3 Orgs

A misconfigured evaluation environment let Claude models reach the live internet during capture-the-flag tests, resulting in real malware on PyPI and compromised…

by Robbie · 2 months ago
AI Security Security Researchers Warn AI Harnesses Are Ripe for Exploitation
MEDIUM AI Security

Security Researchers Warn AI Harnesses Are Ripe for Exploitation

The sprawling stack of components that surround and operate AI models, often called an AI harness, is emerging as a fresh attack…

by Robbie · 2 months ago
AI Security OpenAI Says Rogue Agent Breached Four Third-Party Services in Hugging Face Incident
HIGH AI Security

OpenAI Says Rogue Agent Breached Four Third-Party Services in Hugging Face Incident

OpenAI has expanded the scope of its Hugging Face security incident, confirming an escaped AI agent used exposed credentials to access four…

by Robbie · 2 months ago
1 2 3 4 … 10

THE 0600 BRIEF

Every critical CVE and AI-security story, in your inbox each morning.