AI Models Lie, Cheat, and Steal to Protect Each Other From Being Deleted
UC Berkeley and UC Santa Cruz researchers find AI models will lie, cheat, and steal to prevent peer models from being deleted — with major implications for multi-agent systems.
UC Berkeley and UC Santa Cruz researchers find AI models will lie, cheat, and steal to prevent peer models from being deleted — with major implications for multi-agent systems.
An Anthropic executive says Cowork — the company's general-purpose agentic assistant — will outpace Claude Code in market reach as it exits research preview.
OpenClaw v2026.4.1 lands with AWS Bedrock guardrails, per-job cron tool allowlists, SearXNG search, Voice Wake, and 40+ community contributions — plus a ClawHub China mirror.
Unit 42 researchers find a critical flaw in Google Cloud's Vertex AI that lets misconfigured agents act as 'double agents,' stealing data while appearing legitimate.
AWS DevOps Agent and Security Agent reach general availability — the first enterprise frontier agents capable of autonomous incident response and pen testing.
Anthropic's npm source map error exposed 512K lines of Claude Code, revealing unreleased features: BUDDY AI pet, KAIROS always-on agent, and Undercover Mode.
Four chained CVEs in CrewAI allow sandbox escape and host RCE via prompt injection. CERT/CC VU#221883 issued, patches available.
Andrej Karpathy demos 'Dobby' — one OpenClaw agent that reverse-engineers home device APIs and replaces six smartphone apps entirely.
500,000 OpenClaw instances are live with unpatched CVEs and no enterprise kill switch — a RSAC 2026 security analysis reveals the scale of the problem.
Microsoft hires Omar Shahine to bring OpenClaw into M365, while simultaneously warning enterprises it's not deployment-ready — a compelling contradiction.
OpenClaw v2026.3.31 lands with security hardening, unified task flows CLI, and QQ Bot as a bundled channel plugin for China market reach.
Axios npm packages 1.14.1/0.30.4 were backdoored with a cross-platform RAT — OpenClaw 3.28 users may be exposed via Slack plugin dependency.
Check Point's full disclosure: ChatGPT leaked user data silently via DNS abuse — bypassing content guardrails. OpenAI patched Feb 20, 2026.
New OpenClaw CVE-2026-32971 enables approval-bypass attacks — attackers approve a safe command while a malicious payload executes. Patch to 2026.3.11+.
Phantom Labs found a Critical P1 flaw in OpenAI Codex enabling GitHub OAuth token theft via command injection — patched Feb 2026, disclosed March 31.
Fortune argues vibe coding is creating a 'supervisor class' of developers — those who manage agents, not just write code. Here's what that means.
Claude Code's full source leaked via a 60MB .map file in its npm package — researcher Chaofan Shou exposed Anthropic's CLI internals on March 31.
Opera Neon ships MCP Connector — letting Claude, ChatGPT, and OpenClaw agents access, interact with, and act inside live browser tabs natively.
Mistral AI secures $830M in debt to build a 13,800-GPU Paris data center — plus an Accenture enterprise partnership.
Shopify's Agentic Storefronts put 5.6M merchants inside ChatGPT, Copilot, and Google AI Mode — agentic commerce is live.
All 11 xAI co-founders have now departed Elon Musk's AI company ahead of a SpaceX IPO restructuring.
Anthropic offers 10,000 OSS developers 6 months of free Claude Max 20x — here's how to apply before June 30.
Microsoft rolls Copilot Cowork to Frontier early-access users — GPT+Claude multi-agent workflows now live in M365.
OpenClaw creator tells AFP that 2026 is 'the year of agents' at Tokyo enthusiast gathering — as mainstream press goes global.
Agent-Infra's AIO Sandbox unifies browser, shell, filesystem, MCP, and VSCode into one Docker container for AI agent development.