Spreading the load: How Salesforce met Multi-AZ HA with SageMaker Inference Components
AWS Machine Learningen

Learn how Salesforce used Amazon SageMaker AI Inference Component placement (the SchedulingConfig parameter) to distribute model copies across multiple Availability Zones, meeting their Multi-AZ high availability compliance requirements without sacrificing the cost efficiency of multi-model co-hosting.
This is a short summary published by AI Global Wire. The full article is owned and hosted by AWS Machine Learning — open it there to read it in full.
Read the full story at AWS Machine LearningRelated AI news
- How Decathlon runs demand forecasting at scale with Chronos-2AWS Machine Learning · August 28, 2026
- Build agentic creative workflows with Amazon Quick and falAWS Machine Learning · August 27, 2026
- Introducing India cross-Region inference for OpenAI GPT-5.6 models on Amazon BedrockAWS Machine Learning · August 27, 2026
- Introducing OpenAI models on Amazon Bedrock for in-country inferencing in IndiaAWS Machine Learning · August 27, 2026
- Deepgram deepens Amazon SageMaker AI observability with Enhanced MetricsAWS Machine Learning · August 27, 2026
- Reduce ASR inference costs by 75% with NVIDIA MPS on Amazon EC2AWS Machine Learning · August 27, 2026