Deploy a multimodal WhatsApp ordering assistant with Amazon Bedrock AgentCore

This post shows how to deploy a multimodal WhatsApp ordering assistant built with Amazon Bedrock AgentCore and Amazon Nova 2. Many quick-service restaurants spread ordering across an app, a website, a phone line, and the counter. Each of those is a separate system to build and run. Each one also fragments the customer’s history, making … Read more

Designing lifecycle policies for AgentCore memory

Memory lifecycle policies help long-running agents on Amazon Bedrock AgentCore stay effective by systematically managing what they remember and forget. Your agent generates memories from every conversation it conducts. If you don’t actively manage these memories, your agents will accumulate outdated context, which can degrade response quality and create compliance risks for your deployment. After … Read more

Build a Physical AI model factory with NVIDIA Cosmos 3 on SageMaker HyperPod

A Physical AI system, such as a robot or autonomous vehicle (AV) that translates real-world data into physical actions, can’t be built in a single training job. Instead, it takes a continuous pipeline: a loop of generating synthetic data, post-training perception and policy models, so the system understands its surroundings and can act, and evaluating … Read more

Run agent-driven Amazon SageMaker HyperPod operations with InstantStart

If you run foundation model (FM) workloads on Amazon SageMaker HyperPod, you know the work is rarely a single task. It is a chain of dependent ones. An infrastructure team creates the network and control plane, attaches accelerator capacity, and installs cluster dependencies in the right order. It also prepares storage and identity, keeps distributed … Read more

Customizing your knowledge base on Amazon Bedrock for large and complex documents using Amazon Textract

For customer service teams handling thousands of utility bills each month, accurately parsing and analyzing complex, multi-page documents is a persistent challenge. Inconsistent formats, dense tables, and varied layouts make it difficult to extract the right information quickly. This leads to delayed responses, billing errors, and frustrated customers. As document volumes grow, these inefficiencies compound, … Read more

How Intuit built an agentic disaster recovery assistant with Amazon Bedrock

Disaster recovery (DR) at scale is hard. When thousands of microservices span multiple AWS Regions, coordinating a reliable failover becomes a major operational challenge. At Intuit, we operate at this scale. We support products that millions of people rely on to run their businesses and manage their finances. These include TurboTax, QuickBooks, Mailchimp, and Credit … Read more

AI-driven development lifecycle using Amazon Bedrock AgentCore

Engineering teams adopting the AI-Driven Development Lifecycle (AI-DLC) with Amazon Bedrock AgentCore and coding agents like Kiro often struggle with the gap between conceptual frameworks and working code. Amazon Bedrock AgentCore is a service for building, connecting, and optimizing agents at scale with any framework or model. AI-DLC positions AI as a central collaborator across … Read more

Set up OpenAI ChatGPT Codex with LiteLLM on Amazon ECS and Amazon Bedrock

OpenAI ChatGPT Codex with LiteLLM can provide centralized enterprise controls for generative AI coding agents. These agents help developers understand repositories, write code, run tests, and complete multi-step engineering tasks. As organizations move from individual experimentation to managed adoption, teams need a consistent way to control model access and attribute consumption. They must also apply … Read more

Best practices for building agentic automations with Amazon Quick Automate

Agentic automations are transforming how enterprises run their business processes. Instead of following rigid scripts, AI agents reason about context, adapt to variation, and collaborate with people and other agents to move work forward. Amazon Quick Automate is a multi-agent automation capability within Amazon Quick that helps organizations build, deploy, and maintain these automations at … Read more

Accessing OpenAI models on Amazon Bedrock from Australia with global cross-Region inference

Australian teams working with OpenAI models can now access the latest OpenAI models through Amazon Bedrock. Amazon Bedrock offers OpenAI GPT-5.6 Sol, Terra, and Luna with global cross-Region inference from both Asia Pacific (Sydney) and Asia Pacific (Melbourne) AWS Regions in Australia. Your application calls the Amazon Bedrock Runtime endpoint in Asia Pacific (Sydney) or … Read more

Modernizing and scaling support operations with generative AI on AWS

Scaling support operations requires handling rising ticket volumes, meeting strict Service Level Agreements (SLAs), adapting to evolving compliance requirements, and maintaining documentation that quickly becomes outdated, all without proportional increases in headcount. In many teams, the knowledge required to resolve tickets is fragmented across SOPs, recordings, and tribal expertise, forcing analysts to spend significant time … Read more

How an AWS team detects dashboard content failures at scale using Amazon Bedrock

Picture a scenario familiar to any organization running business intelligence (BI) at scale: A user opens a dashboard minutes before an important meeting and finds a blank chart. Every infrastructure monitor reports healthy. Servers are up, APIs respond, and the data pipeline completed on schedule. Yet the content on screen is broken, and no monitoring … Read more

From code to diagrams: Agentic architecture documentation with Amazon Bedrock AgentCore

Architecture documentation remains one of the most persistent challenges in software development as code bases evolve rapidly. Development teams often spend hours manually creating architecture diagrams, only to watch them become outdated within weeks of deployment. This documentation gap creates knowledge silos, slows developer onboarding, and complicates compliance audits. Amazon Bedrock AgentCore is the platform … Read more

Trinity: Agentic AI-powered transition planning for students with disabilities

This post was co-authored with Marc Steren, Odina Salihbaeva, and Aashrit Surapaneni from University Startups, a partnership between University Startups, g/d/n/a, and AWS. Trinity is a conversational AI solution that helps students with disabilities take ownership of their postsecondary planning. It was developed by University Startups, which was founded in 2020 on a straightforward belief: … Read more

Introducing Claude Fable 5.1 on AWS

Today, we’re excited to announce the availability of Claude Fable 5.1 on Amazon Bedrock and Claude Platform on AWS. Claude Fable 5.1 delivers frontier intelligence for ambitious tasks across coding, scientific research, and enterprise workflows. Given its capabilities, Anthropic has designated Fable 5.1 a Covered Model, a category of Claude models that carry additional data … Read more

From theory to delivery: How Atos upskilled 400 engineers in agentic AI

When Atos set out to upskill 400 engineers from theory to delivery in agentic AI, the team faced a familiar challenge: how to build real-world capability, not only theoretical knowledge. Online courses and classroom-based instruction build foundations, but they do not always give teams the confidence or practical experience needed to apply AI effectively to … Read more

Securing Amazon Quick from POC to production: Agents, Flows, and Spaces

Amazon Quick proof of concept (POC) projects often succeed with a small pilot team, then stall when security and compliance teams review the production plan. A permission model that works for ten pilot users often breaks when you add five departments. Agents can return data outside their intended scope, and compliance teams struggle to audit … Read more

How ZS democratized secure ad-hoc analytics with Amazon SageMaker

This blog post is co-written with Kiran Dhamane, Abhishek I S, and Mayur Ghodekar from ZS Associates Organizations in regulated industries face a persistent tension: give developers the agility they need for ad-hoc analytics, or lock down the environment to meet compliance requirements. In this post, we explore how ZS built a security-hardened Amazon SageMaker … Read more

How Boomi Scribe streamlines documentation using AWS

Boomi Scribe alleviates documentation, one of the most persistent sources of technical debt for enterprise development teams. Enterprise developers often struggle to create and maintain documentation, especially when workflows (automated business processes) involve integrations with multiple enterprise applications and data sources. Boomi recognized this challenge and built Boomi Scribe, an AI-powered agent running on AWS … Read more

Connect an AgentCore Runtime hosted MCP server to Amazon Quick

Model Context Protocol (MCP) servers allow foundation models to access external data and tools, supporting standardized, secure access to files, databases, and APIs. They give AI agents the ability to interact with real-world applications, reduce hallucinations with accurate context, and offer stateful, multi-turn capabilities. Industry-standard architectures quickly evolved and adopted MCP to power agentic AI … Read more

AWS recognized as a Leader in The Forrester Wave: AI Infrastructure Solutions, Q4 2025

We’re excited to share that AWS has been recognized as a Leader in The Forrester Wave: AI Infrastructure Solutions, Q4 2025. In this evaluation of 13 providers, AWS received the highest score in the Strategy category. We believe this recognition reflects our commitment to delivering flexible, cost-efficient AI infrastructure that helps you move from experimentation … Read more

Build observable enterprise agentic retrieval using Managed Amazon Bedrock Knowledge Base with AWS CloudFormation

Teams that add Retrieval Augmented Generation (RAG) to a foundation model usually start with a single retrieval step against a single knowledge base. That works until the questions get harder, when the answer spans several sources, or the system has to decide which source to consult before it can respond. Enterprise agentic retrieval solves that: … Read more

Build multi-tenant agentic chat applications on enterprise data with Amazon Bedrock Managed Knowledge Base

Multi-tenant agentic chat assistants have become a frequent request for large-scale customers, and document chat sits at the top of the list. A user uploads a contract, a report, or a product manual, and then researches or asks questions about it immediately or in the future. The conversational interface is straightforward to build, but the … Read more

Batch write and discover records in Amazon SageMaker Feature Store

Amazon SageMaker Feature Store is a fully managed, purpose-built repository to store, share, and manage features for machine learning (ML) models. It provides low-latency online serving for real-time inference, an offline store for historical retention and training feature data, and supports both streaming and batch ingestion patterns. As ML platforms mature, two operational gaps surface … Read more

How Decathlon runs demand forecasting at scale with Chronos-2

This post is co-written with Vianney Bruned, Filippo Giruzzi, Belkiss Saidi, and Carlos Ramirez from Decathlon. Decathlon is one of the world’s largest sporting goods retailers, with more than 100,000 teammates and 400 million users worldwide. The company relies on accurate demand forecasting at scale to support the availability of the appropriate products in each … Read more