SafetyDatasette releases security patches in 1.0a39 and 0.65.4Datasette published two security patch releases — 1.0a39 for the alpha line and 0.65.4 for the stable 0.65.x family — addressing vulnerabilities relevant to instances exposed on the public web, especially those mixing public and private tables.2 min read
SafetyAnthropic alleges Chinese firms funneled user prompts to Claude to train their modelsAnthropic says Chinese AI companies including Moonshot, Deepseek, Xiaomi and Alibaba covertly routed user queries to its Claude chatbot to extract answers for training competing models.4 min read
SafetyExperts warn AI, encrypted messaging and drones create new terrorist risksTwenty-five years after the Sept.2 min read
SafetyAnthropic and U.S. DOE/NNSA co-develop AI classifier to detect nuclear misuseAnthropic has partnered with the U.S.2 min read
SafetyAnthropic blocks attempts to misuse models for biological and military purposes, flags growing riskAnthropic says it prevented multiple attempts this year to use its Claude chatbot and other models for projects that could aid biological weapon development, while also reporting misuse for surveillance, propaganda and conventional-weapons design.3 min read
SafetyAnthropic gives users a choice to share data for model training and extends retention to five yearsAnthropic updated its Consumer Terms and Privacy Policy on August 28, 2025, allowing users of Claude Free, Pro and Max plans to opt in to having their new or resumed chats and coding sessions used to improve future Claude models.3 min read
SafetyAnthropic collaborated with US CAISI and UK AISI to test and harden AI safeguardsAnthropic says it worked with the US Center for AI Standards and Innovation (CAISI) and the UK AI Security Institute (AISI) over the past year to test and strengthen its safeguard systems.4 min read
SafetyAnthropic details Claude’s safeguards for suicide, self-harm and sycophancyAnthropic published a December 18, 2025 post describing technical and product measures intended to keep its Claude models from mishandling conversations about suicide, self-harm and reality-disconnected users.5 min read
SafetyDetailed Threat Report on Abuses Related to ClaudeThe developers of Claude published their most detailed threat report to date, showing how the model was attempted to be used for cyberattacks, influence operations, surveillance, biological abuses,…1 min read
SafetyAnthropic: State actors using Claude models to automate and expand surveillanceAnthropic's threat report says state-linked actors in Mali, China and Iran have used Claude models to automate collection, analysis and targeting in surveillance operations.3 min read
SafetySenior AI researchers endorse warnings that advanced systems could pose extinction riskSeveral prominent AI researchers have publicly supported recent warnings that advanced artificial intelligence could pose a non-negligible risk of catastrophic human harm.2 min read
SafetyWeekly AI roundup — Anthropic cyber incidents, OpenAI product and governance moves, model and infra updates (Sept 8–9, 2026)Anthropic disclosed four cyber incidents during third‑party security tests and has agreed to an independent METR review with broad access for at least eight weeks.8 min read