Glm 53 Fast Available On Baseten
Captured source
source ↗GLM 5.3 Fast available on Baseten Try the new DeepSeek V4 Pro 0813 today. Frontier intelligence at a fraction of the cost. Here
changelog / post
GLM 5.3 Fast available on Baseten
Aug 26, 2026 Go back
GLM 5.3 Flash is now available through Baseten Model APIs. Send requests to zai-org/GLM-5.3-Flash through our OpenAI-compatible endpoint with your Baseten API key. GLM 5.3 Flash is Z.ai's latest model with a 1M-token context window, vision, and reasoning. Reasoning defaults to high and can be steered with reasoning_effort ( low , high , max) . curl https: //inference.baseten.co/v1/chat/completions \ -H "Content-Type: application/json" \ -H "Authorization: Bearer $BASETEN_API_KEY" \ -d '{ "model" : "zai-org/GLM-5.3-Flash" , "messages" : [{ "role" : "user" , "content" : "What is gradient descent?" }] }' For more information, see our docs or for dedicated inference see the GLM 5.3 Fast deployment .
Explore Baseten today Start deploying Talk to an engineer
Popular models GLM-5.3-Flash
DeepSeek V4 Pro 0813
Kimi K3
GLM-5.2 Fast
Whisper Large V3
Qwen3.8-27B
Explore all
Popular models GLM-5.3-Flash
DeepSeek V4 Pro 0813
Kimi K3
GLM-5.2 Fast
Whisper Large V3
Qwen3.8-27B
Explore all
Notability
notability 3.0/10Routine platform availability announcement, not original release.