Artificial Intelligence

Safely Releasing Frontier Models to Customers

Safely Releasing Frontier Models to Customers

It’s our goal for AWS to be the most secure place to run any workload, and in support of that we’ve been deeply investing in security across our services since AWS’s inception more than two decades ago. Our AI services like Amazon Bedrock are built on this foundation and with the same focus. 

Pay-per-inference for AI agents: How BlockRun and Incarna use Amazon Bedrock AgentCore payments

Pay-per-inference for AI agents: How BlockRun and Incarna use Amazon Bedrock AgentCore payments

Amazon Bedrock AgentCore payments gives AI agents a managed way to pay for services on demand, with spending limits enforced by the infrastructure. See how Incarna’s agents pay BlockRun for model inference one request at a time over x402, cutting the work of adding x402 payment support from months to days.

Share GPU clusters across teams with isolation and fairness using Amazon SageMaker HyperPod

Share GPU clusters across teams with isolation and fairness using Amazon SageMaker HyperPod

A reference architecture for securely sharing one Amazon SageMaker HyperPod EKS cluster across multiple teams, using AWS IAM Identity Center for authentication, per-team SageMaker Domains and Kubernetes namespaces for isolation, HyperPod Task Governance for fairness, and namespace-level cost allocation for chargeback.

Introducing Claude Haiku 5.5 on AWS

Introducing Claude Haiku 5.5 on AWS

Claude Haiku 5.5 is now available on Amazon Bedrock and Claude Platform on AWS. According to Anthropic, it is the fastest, most efficient model in the Claude 5.5 family, built for subagents and high-volume, cost-sensitive work, and costs around 75% less than Claude Haiku 4.5 for most tasks. This post covers its improvements and how to get started.

Rethinking access control for RAG with Amazon Quick and Amazon Bedrock

Rethinking access control for RAG with Amazon Quick and Amazon Bedrock

Enterprise RAG unlocks insights from knowledge sources like SharePoint, Google Drive, and Confluence, but those sources carry complex permissions. Learn how Amazon Quick and Amazon Bedrock Knowledge Bases enforce document-level access controls in real time, verifying permissions directly with authoritative sources at query time.

Beyond hours saved: Building the business case for agentic automation

Beyond hours saved: Building the business case for agentic automation

The RPA-era ROI model misses most of the value agentic automation creates. This post gives AI center of excellence leaders a framework to size the full value of agents across time savings, exception handling, decision quality, and maintenance economics, and to prioritize which workflows to automate first.

How Cornerstone OnDemand cut database diagnosis by 78% with Amazon Bedrock

How Cornerstone OnDemand cut database diagnosis by 78% with Amazon Bedrock

Cornerstone OnDemand built Orion AI, a multi-agent system on Amazon Bedrock and Strands Agents, to turn database operations from reactive firefighting into proactive automation. A three-person team cut database diagnosis from 45 minutes to 10, a 78% reduction, in six months. See the design decisions other teams can reuse.

Building a context-aware AI assistant on AgentCore and OpenClaw

Building a context-aware AI assistant on AgentCore and OpenClaw

Off-the-shelf AI assistants forget you between conversations. This post shows how to build a personal assistant that accumulates context using OpenClaw on Amazon Bedrock AgentCore runtime, with AgentCore memory turning disposable chats into durable, structured knowledge you can retrieve with metadata filters.