Model launchesClaude Opus 4.8 released: better judgment and longer autonomous operationThe new version Claude Opus 4.8 builds on 4.7, promising sharper judgment, greater honesty about its own processes, and longer periods of autonomous operation.1 min read
Model launchesNVIDIA releases open-source model to generate synthetic 3D brain MRIsNVIDIA published NV-Generate-MR-Brain, an open-source generator that produces realistic 3D brain MRI scans.2 min read
Model launchesAlibaba unveils Qwen 3.7-Max for long-running agent workflowsAlibaba's Qwen Team has released Qwen 3.7-Max, a proprietary large model designed for long, autonomous agent sessions with large-context support.3 min read
Model launchesThinking Machines Lab unveils TML-Interaction-Small for real-time, parallel multimodal interactionThinking Machines Lab introduced TML-Interaction-Small, a multimodal conversational system designed to process audio, video and text in parallel and generate outputs concurrently using…4 min read
Model launchesGoogle revamps Search, commerce and work with always-on AI agents and unified commerce protocolsAt I/O 2026 Google unveiled wide-ranging changes to Search, Workspace and shopping: the search box will now process multimodal inputs and deliver AI summaries, while persistent AI agents will monitor topics and act on users’ behalf.5 min read
Model launchesGoogle revamps Search with Gemini 3.5 Flash and new AI agents after 25 yearsAt Google I/O in Mountain View, Google unveiled the largest update to its Search in 25 years, centered on the Gemini 3.5 Flash AI model.3 min read
Model launchesThey will build a new, significantly larger model using ten times more computation in partnership with SpaceXAISpaceXAI and the authors of the statement announced that they are training a completely new, significantly larger language model from scratch, using ten times more total computation.1 min read
Model launchesComposer 2.5 Released: Improved Language Model for Long TasksThe developer introduced the Composer 2.5 model, which is smarter, handles long-running tasks better, and follows complex instructions more reliably.1 min read
Model launchesOpenAI releases GPT‑Realtime‑2 speech‑to‑speech model with adjustable speed–reasoning tradeoffOpenAI launched three new audio models in its Realtime API: GPT‑Realtime‑2 (speech‑to‑speech with configurable reasoning effort), GPT‑Realtime‑Translate (speech translation), and GPT‑Realtime‑Whisper (speech transcription).4 min read
Model launchesGemini Intelligence: expected to debut first on Samsung foldables and Google Pixel devicesGoogle unveiled the Gemini Intelligence AI assistant at the Android Show: I/O Edition event on Tuesday.2 min read
Model launchesOpenAI and Thinking Machines Redefine Human–Machine InterfacesOpenAI has launched GPT‑Realtime‑2, a GPT‑5‑class voice model designed for live audio reasoning, translation, transcription and tool use with a 128K context window.2 min read
Model launchesVoxtral TTS: lightweight 4B model rivals ElevenLabs, adds emotion steeringVoxtral TTS, a 4-billion-parameter text-to-speech model, outperformed ElevenLabs Flash v2.5 in human preference tests for naturalness and matched ElevenLabs v3 quality while offering emotion-aware control.3 min read