Amazon SageMaker Expands AI Inference Capabilities
Amazon is enhancing its SageMaker platform with new tools to improve AI performance and accessibility. The latest development introduces an inference gateway to significantly reduce latency, following the launch of Positron for data science workflows.
-
Posit launches Positron on Amazon SageMaker AI for data science workflows
Posit’s IDE, Positron, is now offered as a custom container that runs inside Amazon SageMaker Studio Spaces. The image, built on the SageMaker Distribution base, lets a Space execute Athena queries,…
1 source primary source -
Amazon announces SageMaker HyperPod Inference Gateway to cut first-token latency up to 82%
Amazon Web Services introduced the SageMaker HyperPod Inference Gateway, a Kubernetes‑native add‑on for EKS that routes large‑language‑model inference requests using real‑time GPU signals. The addon…
3 sources primary source