Skip to content

Source

AWS What's New — Machine Learning

5 items

Dated, terse notices of what Amazon shipped, including Bedrock model availability and region rollouts.

aws.amazon.com ↗

AWSBusiness+the lowest Support plan with access to the age
Agentsmedium signal

AWS previews Well-Architected Agent, limited to Business+ Support plans and above

AWS announced a preview of AWS Well-Architected Agent on October 1, 2026. The service scans your AWS accounts on a schedule and returns recommendations for cost, security, resilience, and performance, ranked against business goals that you write yourself. Remediation scripts come attached, and the agent also reviews infrastructure-as-code templates on demand. Access requires an AWS Support plan at the Business+ tier or higher, and agent profiles are hosted in three US Regions.

AWSverified

DevDay 2026 Recap artwork with colorful illustrated circles on a black background.
Agentsstrong signal

Codex gets reusable cloud environments, voice control in the CLI, and repository security scans

OpenAI turned Codex cloud tasks into published environments that a team can reuse, with each task in its own workspace and work continuing while your laptop sleeps. The CLI takes spoken instructions and shows parallel tasks in a new view. Codex Security Cloud scans whole GitHub repositories and prepares fixes in the cloud. The Agents API now drives a computer, and a Bedrock version runs the same agents inside AWS.

OpenAIverified

Anthropic logo
Modelsstrong signal

Anthropic ships Claude Opus 5.5 and cuts token prices by a fifth

Opus 5.5 charges $4 per million input tokens and $20 per million output tokens, 20% below Opus 5. Cache reads, which dominate the bill for agentic and coding work, fall 60% to $0.20 per million. Anthropic says the model works at the level of Claude Fable 5.1 on most tasks while costing 40% less to run than Opus 5. It also arrives with safeguards that reroute most cybersecurity work to Opus 4.8.

Anthropicverified

Modelsmedium signal

Amazon Bedrock added Kimi K3, and its agent runtime now starts in about two seconds

Two Bedrock announcements landed on September 18, 2026. Kimi K3, which Moonshot AI describes as the first open model with 2.8 trillion parameters, is generally available, and it is the first open-weight model on Bedrock that supports explicit prompt caching. The runtime agents run inside was replaced at the same time: AWS measured a P75 cold start of 1.9 to 2.0 seconds for container images from 200 MB to 2 GB, against 5.4 to 30 seconds on the previous version. Memory is now released during a session instead of being held at the peak.

Amazon Web Servicesverified

Modelsmedium signal

AWS added model caching to SageMaker HyperPod and measured about 60% faster scale-out

SageMaker HyperPod can now keep model weights on a node's local NVMe and pre-pull the container image, so a pod that used to wait minutes for downloads starts in seconds. AWS reports benchmarks across models from 57 GB to 145 GB showing around 60% faster scale-out, and an image cache that removes over two minutes of pull time, a 97% reduction. Those figures are the vendor's own and arrived without a description of the test. On the same day AWS made TwelveLabs Marengo 3.0 available as an embedding model in Amazon Bedrock Managed Knowledge Base, so a knowledge base can index what a video shows rather than only what was said in it.

AWSverified