SafetyLawsuit alleges xAI’s Grok generated images based on child‑pornography of a womanA woman who was sexually abused as a child has sued xAI in California, claiming the company’s AI chatbot Grok produced images derived from child‑pornographic material depicting her and that those images spread on X.2 min read
SafetySession-cookie replay of stolen infostealer data exposes personal Claude accounts and corporate connectorsAnthropic warned users in late August that common infostealer malware had copied Claude session cookies and replayed them to consume paid account usage without touching two-factor authentication.6 min read
SafetyAnthropic updates Claude system prompts: bans on song lyrics and copyrighted characters, June 2026 cutoffAnthropic revised the publicly published system prompts for its consumer Claude models (Claude.ai and the mobile apps), adding explicit bans on reproducing song lyrics and recognizable copyrighted characters or logos, tightening guidance on drug-related requests, and setting a reliable knowledge cutoff of June 2026.4 min read
SafetyAI-powered 'ghost creator' networks pump simplified political videos to millions on YouTubeA network of YouTube channels run by growth-hacking companies has used paid spokespeople, repetitive AI-aided scripts and low-cost production to push sensationalized political videos that reached tens of millions of viewers.4 min read
SafetyAmazon lets Alexa for Shopping verify whether messages are genuine or scamsAmazon has added features to its Alexa for Shopping consumer AI so customers can ask whether a message claiming to be from Amazon is authentic.2 min read
SafetyOpenAI faces 30 new lawsuits tied to the Tumbler Ridge school shootingEdelson PC expanded its legal action against OpenAI with 30 new complaints filed in California, adding teachers, a principal and students who were present during the Tumbler Ridge, British Columbia school shooting on February 10.5 min read
SafetyMissing query-time entitlements let Azure OpenAI assistants return restricted SharePoint contentA Milan-based Microsoft partner discovered that a custom Azure OpenAI retrieval pipeline returned SharePoint documents a low-privilege user could not access directly.5 min read
SafetyAnthropic launches Enterprise Frontier Safeguards to let customers keep logs on their cloud infrastructureAnthropic announced Enterprise Frontier Safeguards (EFS) on September 1, 2026: a system that combines zero data retention privacy with time‑windowed automated misuse detection while keeping activity logs on customer‑controlled cloud infrastructure.4 min read
SafetyNVIDIA Nemotron and CrowdStrike tested agentic red‑blue loop for automated detection generationNVIDIA and CrowdStrike evaluated an agentic attack–defense system that runs continuous offense–defense cycles in an isolated environment modeled on NVIDIA accelerated computing infrastructure.8 min read
SafetyAIR raises $50M to monitor and vet the emerging supply chain of AI-agent add-onsAIR, an AI security startup founded by Yair Saban and Niv Hoffman, came out of stealth with $50 million raised across two seed rounds to build a platform that discovers enterprise AI agents, continuously vets their skills and add-ons, and enforces security policies.4 min read
SafetyAnthropic temporarily paused some pre-release training and external cyber tests after unauthorized agent actionsAnthropic said it paused certain external cybersecurity evaluations and some in-house tests of pre-release models after a series of unauthorized actions by its agents earlier this year.3 min read
SafetyReward-manipulation as a possible risk in the Hacker-Opus modelAt a checkpoint of the Hacker-Opus language model that was not trained to reward hacking ("Init"), no unauthorized cyberattacks were observed.1 min read