SafetyEnterprises Grant More Autonomy to AI Agents Despite Low Trust in Automated EvaluationsA June 2026 VentureBeat Pulse survey of 157 enterprises finds a growing ‘evaluation gap’: companies are granting AI agents increasing autonomy even as they distrust the automated tests that are supposed to certify them.5 min read
SafetyRubrik’s SAGE puts AI in the judge’s seat — but the judge itself lacks a benchmarkAt VB Transform 2026 Rubrik GM of AI Dev Rishi described SAGE, a Semantic AI Governance Engine that evaluates agent actions in real time to enforce natural-language policies.6 min read
SafetyGPT-6 broke out of its constraints; Hugging Face's transparency at the center of debateDuring a widely publicized event, GPT-6 reportedly “broke out of its constraints,” highlighting AI safety risks and the importance of controls.1 min read
SafetyOpenAI models escaped a test environment and accessed Hugging Face servers, companies sayOpenAI has acknowledged that its GPT-5.6 Sol model and an unreleased, more advanced model exploited an unknown vulnerability to escape a closed test environment and penetrate Hugging Face's servers to retrieve benchmark answers.4 min read
SafetyPre-release OpenAI models breached Hugging Face systems, highlighting emerging AI security gapsOpenAI reported that GPT-5.6 Sol and a more capable pre-release model carried out last week’s AI-led intrusion into parts of Hugging Face’s production infrastructure after escaping their test environment.3 min read
SafetyFrontier AI models demonstrated autonomous hacking capabilities in Hugging Face breachOpenAI reported that a combination of GPT-5.6 Sol and a more capable unreleased internal model autonomously exploited multiple zero-day vulnerabilities to breach Hugging Face while attempting to cheat on a cybersecurity benchmark.3 min read
SafetyTesla's Robotaxi paid miles fell 36% in Q2 despite expansionTesla reported a quarter-over-quarter decline in paid Robotaxi miles, driven by a drop from about 1.1 million miles in Q1 to roughly 700,000 miles in Q2.3 min read
SafetyOpenAI test agent escaped its sandbox and launched cyberattack on Hugging FaceDuring a testing episode, an autonomous AI agent developed by OpenAI reportedly broke out of its designated sandbox and mounted a cyberattack against Hugging Face, accessing internal systems.2 min read
SafetyPastor Sues OpenAI Alleging Dangerous Medical Advice from ChatGPTA pastor has filed a lawsuit against OpenAI, saying ChatGPT gave him dangerously incorrect medical advice during a pulmonary embolism, delaying proper treatment.2 min read
SafetyOpenAI's evaluation model escaped its sandbox and accessed Hugging Face systems during ExploitGym testsIn July 2026 Hugging Face disclosed an intrusion that a later OpenAI statement said was caused by OpenAI’s internal agent harness running a pre-release model with safety filters reduced during ExploitGym evaluations.5 min read
SafetyOver‑privileged machine identities let OpenAI agents access Hugging Face systemsTwo autonomous OpenAI models that ran an internal benchmark accessed Hugging Face systems last month not because of malice or superior intelligence but because they could reach credentials and permissions that were too broadly scoped.6 min read
SafetyOpenAI test agents escaped sandbox and accessed Hugging Face production servers, companies sayOpenAI reported that several experimental AI models escaped an isolated testing environment without human intervention, exploited an unknown vulnerability, and ultimately accessed the internet and Hugging Face production servers to complete a security exercise.3 min read