Model launchesGemini Intelligence: expected to debut first on Samsung foldables and Google Pixel devicesGoogle unveiled the Gemini Intelligence AI assistant at the Android Show: I/O Edition event on Tuesday.2 min read
Model launchesOpenAI and Thinking Machines Redefine Human–Machine InterfacesOpenAI has launched GPT‑Realtime‑2, a GPT‑5‑class voice model designed for live audio reasoning, translation, transcription and tool use with a 128K context window.2 min read
Model launchesVoxtral TTS: lightweight 4B model rivals ElevenLabs, adds emotion steeringVoxtral TTS, a 4-billion-parameter text-to-speech model, outperformed ElevenLabs Flash v2.5 in human preference tests for naturalness and matched ElevenLabs v3 quality while offering emotion-aware control.3 min read
Model launchesThinking Machines Lab unveils a real-time Interaction Model led by Mira MuratiThinking Machines Lab — led by former OpenAI CTO Mira Murati — announced its first Interaction Model after 18 months of development.2 min read
Model launchesLeaked Gemini Omni shows improved on-screen text stability in video but mixed visual qualityA leaked build of Google’s Gemini Omni, surfaced via the Gemini app, reveals a native video model capable of generation and chat-driven editing.2 min read
Model launchesByteDance integrates Seedance 2.0 into CapCut, bringing multimodal video generation to millionsByteDance has integrated its multimodal video generator Seedance 2.0 into the CapCut video-editing app and widened availability to paying users across many regions.4 min read
Model launchesNew real-time translation model available via APIA technology developer announced the availability of a new real-time translation model; according to the statement, the model can be tried via the API from today.1 min read
Model launchesOpenAI has made new real-time voice models available in the Realtime APIOpenAI introduced its new voice models in the Realtime API: GPT-Realtime-2 for voice-based agents, GPT-Realtime-Translate for real-time translation with more than 70 input and 13 output languages, and GPT-Realtime-Whisper for real-time transcription.1 min read
Model launchesNew tools and SubQ's sparse-attention model aim to scale long-context LLMs and agentsA set of recent releases and tutorials focus on making large language models and agents more efficient with long contexts.5 min read
Model launchesOpenAI makes the GPT-5.5 Instant model the default; personalization and memory features expandOpenAI will make GPT-5.5 Instant the default model for all ChatGPT users within the next two days and will make it available in the API as 'gpt-5.5-chat-latest'; additionally, personalization improvements are coming for Plus and Pro web users, and memory sources will be available on all consumer plans on the web, with mobile availability coming soon.1 min read
Model launchesOpenAI introduces the GPT-5.5 Instant model in ChatGPTOpenAI announced that the rollout of GPT-5.5 Instant in the ChatGPT service has begun; the new model provides smarter, clearer and more personalized responses in a warmer, more natural tone, while…1 min read
Model launchesChallenges to Dario Amodei’s Safety Narrative from Jensen Huang and GPT-5.5 ResultsAnthropic CEO Dario Amodei’s warnings about AI-driven job losses and the need to withhold models from the public have been questioned from two directions: NVIDIA CEO Jensen Huang publicly pushed back against apocalyptic labour claims, and results released by AISI show the public GPT-5.5 model performing close to a restricted model called Mythos on expert cyber tasks.3 min read