WritingBasetenBasetenpublished Aug 26, 2026seen 2d

Glm 53 Fast Available On Baseten

Open original ↗

Captured source

source ↗
published Aug 26, 2026seen 2dcaptured 2dhttp 200method plain

GLM 5.3 Fast available on Baseten Try the new DeepSeek V4 Pro 0813 today. Frontier intelligence at a fraction of the cost. Here

changelog / post

GLM 5.3 Fast available on Baseten

Aug 26, 2026 Go back

GLM 5.3 Flash is now available through Baseten Model APIs. Send requests to zai-org/GLM-5.3-Flash through our OpenAI-compatible endpoint with your Baseten API key. GLM 5.3 Flash is Z.ai's latest model with a 1M-token context window, vision, and reasoning. Reasoning defaults to high and can be steered with reasoning_effort ( low , high , max) . curl https: //inference.baseten.co/v1/chat/completions \ -H "Content-Type: application/json" \ -H "Authorization: Bearer $BASETEN_API_KEY" \ -d '{ "model" : "zai-org/GLM-5.3-Flash" , "messages" : [{ "role" : "user" , "content" : "What is gradient descent?" }] }' For more information, see our docs or for dedicated inference see the GLM 5.3 Fast deployment .

Explore Baseten today Start deploying Talk to an engineer

Popular models GLM-5.3-Flash

DeepSeek V4 Pro 0813

Kimi K3

GLM-5.2 Fast

Whisper Large V3

Qwen3.8-27B

Explore all

Popular models GLM-5.3-Flash

DeepSeek V4 Pro 0813

Kimi K3

GLM-5.2 Fast

Whisper Large V3

Qwen3.8-27B

Explore all

Notability

notability 3.0/10

Routine platform availability announcement, not original release.