Anthropic released Claude Haiku 5.5, its fastest and most capable small model for high-volume, cost-sensitive workloads. Anthropic says it costs about 75% less per task on average than Haiku 4.5, supports adjustable effort, and is available across the Claude Platform, AWS, Google Cloud, and Microsoft Azure. · 5.5
AWS open-sources Strands Box for AI agent sandboxing
AWS released Strands Box, an open-source sandbox for AI agents powered by Dogwood policies. The initial release supports macOS and is designed to package agent runtimes with enforceable controls that can later travel with deployed agents.
AWS introduced the publicly available Physical AI Toolchain, a curated collection of reference architectures, Infrastructure as Code and deployment automation for training and deploying physical AI policies at scale across robotics, autonomous vehicles and smart-factory workloads.
Microsoft opens preorders for Surface Laptop Ultra and Surface RTX Spark Dev Box
Microsoft opened preorders for Surface Laptop Ultra and Surface RTX Spark Dev Box, both built around NVIDIA RTX Spark for local AI development and agent workloads. Surface Laptop Ultra availability begins October 16, while the Dev Box is scheduled to ship in the United States in November.
Microsoft Execution Containers reaches general availability
Microsoft made Microsoft Execution Containers generally available on Windows 11, providing policy-driven containment for AI agents with controls over file and network access and integrations spanning Agent 365, Intune and Windows security.
NVIDIA and Microsoft open preorders for RTX Spark Windows PCs
NVIDIA and Microsoft opened preorders for Windows PCs powered by NVIDIA RTX Spark, including systems from ASUS, Dell, HP, Lenovo, MSI and Microsoft Surface. RTX Spark laptops are scheduled to become available on October 16, with compact desktops following in November.
OpenAI and Atlassian expanded their partnership to deploy GPT-6-family frontier models across Atlassian's platform and Rovo. The agreement also broadens Atlassian's use of Codex and ChatGPT Enterprise and supports deeper integrations with Jira and Atlassian's Teamwork Graph.
Google DeepMind released EmbeddingGemma 2, an Apache 2.0-licensed open-weight multimodal embedding model that maps text, code, images, video and audio into a shared 768-dimensional space. The modular model scales from 270 million parameters for text and code to 740 million parameters for all modalities. · 2
OpenAI says billing for the trusted-access GPT-Rosalind research model begins on October 5, 2026, at $5 per million input tokens, $0.50 per million cached input tokens, and $25 per million output tokens.