Z.ai's GLM 5.3 model becomes available on Amazon Bedrock
Z.ai (Zhipu AI) has made its GLM 5.3 model available on Amazon Bedrock, targeting enterprise customers who need advanced coding and agentic capabilities without managing their own inference infrastructure. GLM 5.3 is a 753B-parameter mixture-of-experts model that Z.ai reports shows notable cyber security capabilities, including a score of 84.5 on the CyberGym benchmark. The model builds on the…
Key points
- GLM 5.3 is a 753B-parameter mixture-of-experts model from Z.ai now on Amazon Bedrock.
- Z.ai reports a score of 84.5 on the CyberGym benchmark for GLM 5.3.
- The model supports prompt caching and cross-Region inference for enterprise customers.
On Amazon Bedrock, GLM 5.3 is accessible via fully managed APIs that support cross-Region inference, prompt caching, and various service tiers. The model supports both OpenAI-compatible APIs and native Bedrock endpoints. A key feature is prompt caching, which reduces latency and input costs for workloads that resend large context blocks, such as agentic workflows. AWS provides a walkthrough demonstrating how to use GLM 5.3 with Strix, an open-source AI penetration testing agent, to perform authorized security tests on applications like OWASP Juice Shop. This setup allows inference to run under the user's AWS account controls, avoiding third-party providers.
Access is currently limited to eligible enterprise customers. The model is available through US and Global cross-Region inference profiles. While the article highlights the model's strengths in long-horizon agentic tasks and complex systems engineering, it notes that users must have appropriate IAM permissions and, for the security demo, specific software like Docker and Strix installed. AWS emphasizes that unauthorized security testing is illegal and violates its Acceptable Use Policy.
Model page: GLM 5.3 →
Introducing GLM 5.3 on Amazon Bedrock
AWS Machine Learning Blog · 5 October 2026
Loading the full article…
This text was published by AWS Machine Learning Blog and written by Alex Thewsey. It is reproduced here with attribution so you can read it in full; the rights remain with the publisher. Read it at the source ↗
The headline, key points and digest above were generated by Digest AI's editorial model from the linked sources. Automated summaries can contain errors: the sources are the record. Spotted a mistake? Tell us. Published by Martin K., who runs Digest AI and handles corrections.
More in Generative AI & Models
All →- Reflection AI releases Beam, a 501B open-weight model claiming parity with Chinese rivals at lower compute cost · 4 src
- Anthropic prompts users to share voice data to enhance AI models · 6 src
- Reka AI releases research preview of Rho-1 omni-model · 1 src
- Claude explains AI context windows and their limits · 1 src
- OpenAI, Google, and Anthropic launch new AI models in late September · 4 src
Comments
via GitHub Discussions