SafetyPaul Christiano joins the OpenAI Foundation board and its Safety and Security CommitteePaul Christiano, founder of Alignment Research Center, is joining the OpenAI Foundation board and its Safety and Security Committee, and will serve as a non-voting observer on the OpenAI Group PBC board.1 min read
SafetyAnthropic’s Chris Olah speaks at Vatican presentation of Pope Leo XIV’s encyclical on AIOn May 25, 2026, Pope Leo XIV published an encyclical titled "Magnifica humanitas: On safeguarding the human person in the time of artificial intelligence," presented in the Vatican.4 min read
SafetyAnthropic expands Project Glasswing to about 150 more critical organizationsAnthropic announced on June 2, 2026 that it is widening Project Glasswing by inviting roughly 150 additional organizations to access Claude Mythos Preview after security vetting.4 min read
SafetyApple introduces Reference Image on iPhone 18 Pro to verify photo authenticityAt its Surprise and Shine event, Apple unveiled Apple Reference Image, a feature on the iPhone 18 Pro that records signed sensor data and creates an unalterable reference image in the Photos app to help prove a photo's authenticity.2 min read
SafetyOpenAI insiders urge external restraint as rapid AI advances raise control concernsSenior researchers and executives at OpenAI and other AI labs say the field is accelerating in ways they cannot safely manage alone, calling for outside intervention.4 min read
SafetyJacob says AI has more than a 10% chance of wiping out humanity in the next decadeJacob claims that he seriously believes artificial intelligence could be capable of wiping out humanity; his personal estimate is that there is more than a 10% chance of this happening in the next ten years.1 min read
SafetyDemszky asked AI to compare mayoral records; AI ranked Tarlós firstFormer Budapest mayor Demszky Gábor said he asked a generative AI to compare his performance with that of his successors Tarlós István and Karácsony Gergely.2 min read
SafetyBugcrowd CEO warns AI agents will become targets and vectors in cyberattacksBugcrowd CEO Dave Gerry told Axios that as AI agents proliferate in enterprise environments, attackers will both target and exploit those agents, shifting the focus of cyber defenses.3 min read
SafetyOpenAI focuses on understanding models and regulating the pace of developmentOpenAI said it is currently focusing on a deeper understanding of its models and using that to schedule further capability development.1 min read
SafetyIndustry warning: prompt injection risks and models' resilienceThe author of the post reports that OpenAI's new model is roughly equivalent to Gemini Flash and Opus 4.8 in terms of prompt injection risk, and praises its performance.1 min read
SafetyUsers report stolen Claude session tokens that drain paid usage and cause account suspensionsSeveral Claude users, including independent consultant Grant de Swardt, reported unauthorized token consumption after session credentials were compromised, leading to depleted usage and account suspensions.4 min read
SafetyAI-generated 'thin ideal' is narrowing body diversity on social mediaAI-generated images and virtual influencers are amplifying an ever-thinner body ideal on social platforms, researchers and clinicians warn.3 min read