ToolsNunchaku Lite: Native 4-bit SVDQuant support in Diffusers for lower VRAM and faster samplingDiffusers now supports Nunchaku-style SVDQuant checkpoints natively via a “Nunchaku Lite” path, allowing models with 4-bit weights and activations (W4A4) to be loaded with Diffusers’ from_pretrained() without a separate inference engine or local CUDA compilation.4 min read
ToolsChatGPT Health: secure US rollout lets users link Apple Health and medical recordsOpenAI has begun rolling out Health in ChatGPT to logged-in users aged 18+ in the United States on web and iOS across Free, Go, Plus and Pro plans.4 min read
ToolsWhen AI Agents Go Confidently Wrong: The Data Observability GapMany production failures in retrieval-based AI stem from stale, incomplete, or inconsistently validated data rather than the model or retrieval layer.6 min read
ToolsShopify and Vercel collaborate to provide shared infrastructure for storesAccording to announcements from Shopify and Vercel, the two companies will connect their infrastructures for applications and agents, with the goal of making these solutions available to millions of businesses.1 min read
ToolsCursor Router available on Microsoft Teams Enterprise plans with administrative controlsCursor Router is now available across all Microsoft Teams and Enterprise subscriptions.1 min read
ToolsCursor Router: intelligent model router for more cost-effective AI resultsCursor Router is an intelligent model router that automatically selects the model best suited for the task.1 min read
ToolsAnthropic makes Economic Index data queryable in Claude via new connectorAnthropic has released a connector for Claude that lets users query the Anthropic Economic Index directly from conversations.2 min read
ToolsAnthropic publishes the Anthropic Economic Index, Claude now directly queryable about itAnthropic introduced a public dataset called the Anthropic Economic Index that measures the use of artificial intelligence in the economy; the system named Claude can now directly answer questions based on the index.1 min read
ToolsMaking TensorRT Builds Observable and Cancelable with IProgressMonitorNVIDIA TensorRT exposes IProgressMonitor, an API available in NvInfer.h, that lets applications observe nested, live progress during engine builds and request cancellation.6 min read
ToolsOpenAI launches Presence for enterprise real-time voice and chat agents with high-touch deploymentsOpenAI announced Presence, a managed enterprise product for deploying and governing real-time voice and chat AI agents in customer-facing and internal workflows.7 min read
ToolsVercel AI Gateway: fast token retrieval for cold and hot tokensVercel announced the AI Gateway service, which it claims is the fastest solution for retrieving tokens (cold and hot).1 min read
ToolsOpenAI Presence: corporate voice and chat agents available with limited accessOpenAI introduced the OpenAI Presence service, which enables companies to deploy reliable voice and chat agents for customer and internal workflows; the agents can connect to systems, perform approved actions, escalate requests to humans, and learn over time.1 min read