Model launchesGPT‑6 Astra available to Pro, Enterprise and Business Premium users in ChatGPT Work and CodexOpenAI announced that GPT‑6 Astra is now available to Pro, Enterprise and Business Premium subscribers in the ChatGPT Work and Codex services, and is live in the API as well; Plus and Business users will see a gradual rollout over the next few days.1 min read
Model launchesGPT-6 Astra achieves leading performance on multiple benchmarks, progress for scientific applicationsAccording to the announcement, the GPT-6 Astra language model is competitive: it placed first on FrontierMath Tier 4, ARC-AGI 3 and TerminalBench-4.0, and also showed outstanding results on the…1 min read
Model launchesGoogle unveils WeatherNext 3, an hourly global weather model at 5 km resolutionGoogle DeepMind and Google Research have launched WeatherNext 3, a global AI-driven weather model that produces hourly forecasts using live geostationary satellite mosaics and station observations.3 min read
Model launchesMicrosoft launches MAI-Transcribe-2: cheaper, faster speech recognition aimed at cutting OpenAI/Google relianceMicrosoft released MAI-Transcribe-2, a speech-recognition model it says is faster, more accurate, and substantially cheaper than competing offerings from OpenAI, Google, and ElevenLabs.7 min read
Model launchesRunway unveils Solaris interface model and its contested benchmarkRunway introduced Solaris, an Interface World Model that claims to generate app-like experiences frame by frame without code, and published an internal benchmark in which Solaris outperformed Claude Opus 5 on instruction-following.3 min read
Model launchesGoogle releases Gemini 3.8 Flash and Flash Cyber for agentic tasks and vulnerability discoveryGoogle introduced two Gemini 3.8 Flash variants — a general-purpose Flash tuned for agentic workflows, coding, and multi-step reasoning, and Flash Cyber optimized for vulnerability detection and automated patching.5 min read
Model launchesAnthropic debuts Fable 5.1 and Mythos 5.1 amid cost and safety tweaks, premium remains contestedAnthropic today released Fable 5.1 and Mythos 5.1, reporting large benchmark gains and operational cost reductions while loosening some safety restrictions.2 min read
Model launchesAnthropic’s Claude Fable 5.1 boosts science scores and produces elaborate SVG pelican at high reasoning levelsAnthropic released Claude Fable 5.1 (and Mythos 5.1), claiming substantial gains on coding, knowledge work and long-running problem solving; the model scored 52.6% on the new Terminal-Bench-Science 0.1 benchmark announced August 27.5 min read
Model launchesClaude Fable 5.1 available on the Cursor platformClaude Fable 5.1 has been made publicly available on Cursor; the model scored 73.4% with maximum effort on the CursorBench 3.2 benchmark.1 min read
Model launchesGemini introduces agentic video understanding to speed up and cut costs of video analysisGoogle’s Gemini models (3.7 Flash, 3.6 Flash and 3.5 Flash‑Lite) received a new agentic video understanding capability that dynamically searches video, audio and transcripts to reduce token use and costs while improving accuracy.3 min read
Model launchesGoogle's TimesFM-3 offers zero-shot forecasting from CSV uploadsGoogle introduced TimesFM-3, a 330-million-parameter forecasting model that produces predictions from uploaded CSVs without any fine-tuning.2 min read
Model launchesPentagon deploys tailored ChatGPT and Grok instances for military and civilian personnelThe U.S.3 min read