AI

Inference meta-monitoring for Amazon SageMaker AI endpoints with Amazon Quick

Without the ability to track machine learning (ML) model prediction quality, organizations only realize they have issues when their customers complain or when they conduct spot checks, which jeopardizes customer trust. This post introduces inference meta-monitoring for Amazon SageMaker AI endpoints. It provides a governance layer that sits above production ML inference pipelines to continuously […]

Inference meta-monitoring for Amazon SageMaker AI endpoints with Amazon Quick Read More »

Introducing explicit prompt caching for OpenAI GPT-5.6 models on Amazon Bedrock

This post is co-written with Chris Dickens from OpenAI. OpenAI GPT-5.6 Sol, Terra, and Luna are now generally available on Amazon Bedrock. With GPT-5.6 on Amazon Bedrock, you get the newest generation of OpenAI frontier models with pay-per-token pricing, AWS security and governance controls, and usage that counts toward your existing AWS commitments. The family

Introducing explicit prompt caching for OpenAI GPT-5.6 models on Amazon Bedrock Read More »