Day 0 support for GLM-5.3 on DeepInfra. 🚀
@Zai_org's latest flagship model brings major advances in complex coding, long-horizon agents, and cyber defense.
GLM-5.3 is running in the US on @nvidia Blackwell GPUs with ZDR.
$1.40 input · $4.40 output · $0.26 cached / 1M tokens
The @Zai_org team is on a roll. Another open-weight drop, and a seriously good one — congrats to everyone who shipped it.
Live on DeepInfra now -> deepinfra.com/zai-org/GLM-5.3
GLM-5.3 is now open-weight.
Our most capable model for agentic coding and cyber defense is now available to download, run, and customize.
Weights: huggingface.co/zai-org/GLM-5.3
Tech blog: z.ai/blog/glm-5.3
Video and audio, from one model.
Wan3.0-Video from @Alibaba_Wan is now on DeepInfra: 30-second clips at 1080P, omni-modal reference (images, files, even web pages), and characters that stay the same character.
$0.20/sec →
Introducing GLM-5.3-Flash
- Leading capabilities at a highly competitive price
- Natively multimodal with a 1M-token context window
- A 320B-A18B model released under the MIT License
- Previously previewed as Ox Alpha, running entirely on Chinese AI chips
Blog:
Day 0 support for GLM-5.3-Flash on DeepInfra.
$0.15 in / $0.50 out per 1M. $0.03 cached.
@Zai_org's 320B-A18B multimodal model — 1M-token context, built for coding and long-horizon agents. Running in US on NVIDIA Blackwell.