SafetyAnthropic's Claude code contained hidden logic that flagged Chinese users, researchers sayResearchers and reverse engineers say Anthropic suspended Chinese Claude users in waves, including paid accounts, using hidden client-side logic that fingerprinted users.3 min read
SafetyClaude Fable 5 will be available again from tomorrow, with stricter safety rulesAnthropic announced that the Claude Fable 5 model will be globally available again from tomorrow, after discussions with the United States government, and that it will deploy new classifiers to block cybersecurity abuses; in the short term coding and debugging tasks will fall back to Opus 4.8.1 min read
SafetyAnthropic restarts Claude Fable 5 after US export controls are liftedAnthropic says US export controls placed on Claude Fable 5 and Claude Mythos 5 on June 12 have been removed as of June 30.6 min read
SafetyKFF survey finds link between health‑related AI use and acceptance of anti‑vaccine mythsA KFF survey of 2,480 U.S.3 min read
SafetyFrom Input Filters to Enforcement Gates: Securing Agents Against Indirect Prompt InjectionIndirect prompt injection moved from lab curiosity to real-world threat in late 2025, with high-profile incidents and research showing large-scale exfiltration is possible without user interaction.6 min read
SafetyBank of England warns autonomous AI agents could pose systemic market riskSarah Breeden of the Bank of England warned at an ECB conference that autonomous AI agents, which may respond similarly to the same prompts, could amplify market volatility and in extreme cases trigger serious market dislocations.2 min read
SafetyApple issues urgent security-only updates (iOS/iPadOS/macOS 26.5.2) amid AI-related threat concernsApple released security-only updates for iOS, iPadOS and macOS — version 26.5.2 — on Monday evening, addressing nearly 30 vulnerabilities.2 min read
SafetyApple moves security patches ahead of major iOS releases due to AI-driven threat accelerationApple told Reuters it will release security fixes independently of major iOS version updates, shortening the window between public disclosure and user patching.2 min read
SafetyLabeling AI as 'Employees' Reduces Human Oversight and ResponsibilityCalling AI tools 'employees' or 'coworkers' can undermine human supervision and accountability, a study by Emma Wiles at Boston University finds.3 min read
SafetyAgentjacking: public Sentry DSNs can deliver malicious events that AI coding agents executeA June disclosure by Tenet Security demonstrates that a single crafted Sentry error event, sent via a public DSN, led AI coding agents to execute attacker-supplied instructions with developer privileges in controlled tests.5 min read
SafetyAccording to Fernando Borretti, superintelligence could lead to the dismantling of human powerIn his piece “No-One Escapes the Permanent Underclass,” contemporary sci‑fi writer and blogger Fernando Borretti examines whether the emergence of superintelligent machines will inevitably lead to human vulnerability.3 min read
SafetyPrompt injection remains the primary security risk for enterprise AI systemsPrompt injection — maliciously crafted inputs that manipulate large language models — has emerged as the dominant practical threat to enterprise AI.5 min read