Model launchesGemini introduces agentic video understanding to speed up and cut costs of video analysisGoogle’s Gemini models (3.7 Flash, 3.6 Flash and 3.5 Flash‑Lite) received a new agentic video understanding capability that dynamically searches video, audio and transcripts to reduce token use and costs while improving accuracy.3 min read
Model launchesGoogle's TimesFM-3 offers zero-shot forecasting from CSV uploadsGoogle introduced TimesFM-3, a 330-million-parameter forecasting model that produces predictions from uploaded CSVs without any fine-tuning.2 min read
Model launchesPentagon deploys tailored ChatGPT and Grok instances for military and civilian personnelThe U.S.3 min read
Model launchesZhipu’s GLM-5.3 Flash („Ox Alpha”) tops developer rankings after free test periodA previously anonymous model known as “Ox Alpha” has been revealed as GLM-5.3 Flash, a new release from Chinese AI lab Zhipu.2 min read
Model launchesTencent unveils Hy4 preview: 770B parameters and 1M-token context windowChinese tech firm Tencent released a preview of its new Hy4 large language model with 770 billion parameters (49B active) and a 1,000,000 token context window.2 min read
Model launchesGoogle releases Gemini Omni 1.1 Flash with expanded generative video controls for developersGoogle announced Gemini Omni 1.1 Flash, an update to its generative video model that adds scene extension, finer camera control between keyframes, faster low-resolution drafts, and upscaling to 4K.4 min read
Model launchesAnthropic expands Claude for healthcare and life sciences with Opus 4.5, new connectors and HIPAA‑ready toolsAnthropic announced on Jan 11, 2026 that it is extending Claude’s capabilities for healthcare and life sciences with a new Claude for Healthcare suite, expanded life‑sciences connectors, and model improvements in Claude Opus 4.5.5 min read
Model launchesGoogle introduces Gemini 3.5 Transcribe, an AI model for structured transcriptions and on‑page dictationGoogle unveiled Gemini 3.5 Transcribe, an audio-focused AI model designed to convert unstructured spoken language into formatted text and to support voice-driven editing.2 min read
Model launchesQwen3.8-Flash-Next: open multimodal Mixture-of-Experts model previewing Qwen4 architectureQwen3.8-Flash-Next is an open-weights, multimodal Mixture-of-Experts (MoE) model that serves as an early preview of the architecture planned for Qwen4.2 min read
Model launchesAnthropic expands Claude's context window to 100,000 tokensOn May 11, 2023 Anthropic announced that its Claude model's context window has been increased from 9,000 to 100,000 tokens (around 75,000 words).2 min read
Model launchesAlibaba previews Qwen3.8-Flash-Next: a long-context multimodal MoE model with 262k native context and 1M-token scalingAlibaba released weights for Qwen3.8-Flash-Next as a developer preview of the upcoming Qwen4 family.3 min read
Model launchesFormer Meta Researchers Launch Vision AI for Industrial RobotsPerceptron, a startup founded in November 2024 by former Meta FAIR researchers Armen Aghajanyan and Akshat Shrivastava, released Isaac 0.5, a vision model intended to help robots perceive, reason and act in industrial environments.3 min read