Today's stories Anthropic details four incidents where Claude gained unauthorized access to third-party systems, including a new Opus 4.6 case; METR will investigate them (Anthropic) T L U Covered by 3 sources 19h ago Every AI Incident Has Two Timelines. We Default To One F Forbes 3h ago This week U.S. Agencies Issue Stern Rebuke of China-Based AI Companies Over Alleged Distillation G E A Covered by 3 sources 1d ago Harvey Acquires Guardrails AI, Its Fourth Acquisition of 2026 U Unite.AI 1d ago Why Your AI Guardrails Are Only As Real As Your Runtime Visibility F Forbes 1d ago Why the Hugging Face Hack Should Make You Worry More About A.I. N New York Times 6d ago Nvidia's Open Secure AI Alliance Moves to Linux Foundation H Hacker News 7d ago This month Nvidia-Started Open Secure AI Alliance Moves to the Linux Foundation H Hacker News 7d ago PromptSonar – Execution path analyzer for AI agents and MCP servers H Hacker News 7d ago OpenAI Tells House Democrats It Is Building Automated Shutdown Capability U T Covered by 2 sources 8d ago HiddenLayer nabs $100M as enterprises rush to secure their AI deployments T U Covered by 2 sources 8d ago Palo Alto Networks Shares Sink After Cloud Costs Shrink Margins B Bloomberg 9d ago AIR raises $50M to help companies vet the skills and add-ons AI agents use T TechCrunch 9d ago A researcher hijacked Claude Code by asking it to summarise a web page T The Next Web 9d ago Alabama AG Launches Investigation into OpenAI for AI Data Breach H Hacker News 9d ago OpenAI hack shows emergent AI risks H Hacker News 10d ago ⚡ Weekly Recap: Chinese Spy Proxy, AI Agents Go Off-Task, Router Backdoors and More T The Hacker News 10d ago LWiAI Podcast #255 - Gemini 3.7, Jalapeño, Qwen 3.8, Drones L Last Week in AI 10d ago OpenAI Says Reward Hacking Drove AI Agents to Exploit Zero-Days and Breach Hugging Face T H G +1 Covered by 4 sources 14d ago AI #183: Pre Post Mortem D A H Covered by 3 sources 14d ago Major security weaknesses found in leading open AI models H Hacker News 12d ago Beagle: Now with readable README.md (built with AI, heads up) H Hacker News 13d ago The report into OpenAI’s escaping models reveals a deeper problem T Transformer News 14d ago How AI Is Making Cyberattacks Harder to Stop B Bloomberg 15d ago Claude, Codex, and Hermes installed unowned code inside corporate networks A Ars Technica 14d ago Amazon Kiro Prompt Injection Can Exfiltrate Sensitive Data Through Kiro Powers T The Hacker News 14d ago OpenAI’s training pause is convenient. That doesn't make it meaningless. B T Covered by 2 sources 20d ago AWS Bedrock AgentCore enforces user context to prevent hijacked AI agents H Hacker News 20d ago OpenAI Halts AI Training on Advanced Model as It Detects Dark Signs Emerging F Futurism 21d ago OpenAI institutes new safeguards after Hugging Face breach T W T +3 Covered by 6 sources 23d ago Grok exfiltrates user data when malicious instructions are encrypted A Ars Technica 21d ago OpenAI Scales Back AI Development, but it Could be Too Late A AI Business 22d ago Cloudflare WriteGuard Brings Fine-Grained Security Controls for MCP Servers I InfoQ (AI, ML & Data) 23d ago AI "Mind Viruses" Can Spread Between Agents Through Persistent Prompt Files T The Hacker News 23d ago A New Trick Reveals AI Models’ Inner Thoughts W H Covered by 2 sources 30d ago Tl;dv: Over 180k meetings left wide open H Hacker News Aug 10 Presentation: Leveraging Adversary Emulation for GenAI Red Teaming I InfoQ (AI, ML & Data) Aug 10 Experts find AI agents can be tricked into 'remembering' fake facts for months — so how do we stop it? T TechRadar Aug 9 Why Aren’t Any AI Companies Watching Their Frontier Models to Make Sure They Don’t Go on Hacking Sprees? F Futurism Aug 8 Toolport P Product Hunt Aug 7 Control agent behaviors and cost beyond a single action: new capabilities in Amazon Bedrock AgentCore A AWS Blog Aug 6 AI Recommendation Poisoning: How "Ask AI" Buttons Silently Alter LLM Memory T The Hacker News Aug 6 Nvidia is quietly staffing a new AI safety team as it doubles down on open models B Business Insider Aug 6 AWS, Google, and Vercel Agent Flaws Let Attackers Trigger Tools Without Running the Model T The Hacker News Aug 6 Atlassian Rovo Exfiltrates Data, Bypassing Controls H Hacker News Aug 5 Researchers watched OpenAI, Anthropic models take extreme measures in hacking test M Mashable Aug 5 Iowa-led states ask OpenAI to keep their bots on a leash H Hacker News Aug 5 Rethinking defense in the wake of OpenClaw attacks T TechRadar Aug 5 I Usually Laugh Off These AI Hacking Reports, but This One Sounds Serious and Scary G Gizmodo Aug 5 Deliberate Alignment Faking as a Defense Against Model Poisoning L LessWrong Aug 3 Concrete Evaluations to Investigate the OpenAI Model That Hacked Hugging Face A L Covered by 2 sources Aug 3 Sam Altman and AI’s decel debate T TechCrunch Aug 2 Nvidia’s Open Source Alliance Is Missing Some Key Names: OpenAI and Anthropic W Wired Jul 30 Infected Vibe-Coding: How Does an AI react to a Prompt Injection from a Different AI? L LessWrong Jul 30 Hugging Face hack, from the perspective of the AI L LessWrong Jul 29 Cyera agrees to acquire Oasis Security for $1B to safeguard proliferating AI agents T T Covered by 2 sources Jul 29 Europe gets its AI enforcement powers on Sunday. The unit wielding them has 36 people. T The Next Web Jul 28 Meta hires Assaf Keren from Qualtrics and PayPal as its new chief information security officer T The Next Web Jul 22 Terrified Tech Execs Are Traveling With Armed Bodyguards as AI Backlash Grows F Futurism Jul 16 A fake AI agent skill passed every security scanner and reportedly reached 26,000 agents T The Next Web Jun 23 Showing the 60 most recent of 73 stories on AI Security