Training Is Now Ga
Captured source
source ↗Training is now GA! Announcing our Series F . Learn more
changelog / post
Training is now GA!
Oct 31, 2025 Go back
Since launching the beta of Baseten Training in May, we’ve introduced a ton of improvements, including; A more robust ML Cookbook , great starting points for: Training a coding model with GRPO
Long-context training using multi-node with Qwen3 30B A3B
A variety of examples with Qwen3, gpt-oss, Gemma3, and Llama
Resume from checkpoint : Launch jobs that pick up right where you left off
A ton of other improvements, including: Broader checkpoint recognition across FSDP, VeRL, and Megatron checkpointing formats
More availability for InfiniBand-backed multi-node training runs
Improved management and handling of the training cache
Per-GPU metric visibility and improved logs
Quality of life improvements around Training Cache
And much more!
After months of positive feedback from early users and thousands of training runs completed, Baseten Training is now immediately available for anyone on Baseten. Get started here .
Explore Baseten today Start deploying Talk to an engineer
Popular models GLM 5.2
Kimi K2.7 Code
DeepSeek V4
GPT OSS 120B
Whisper Large V3
NVIDIA Nemotron 3 Ultra
Explore all
Popular models GLM 5.2
Kimi K2.7 Code
DeepSeek V4
GPT OSS 120B
Whisper Large V3
NVIDIA Nemotron 3 Ultra
Explore all
Notability
notability 3.0/10Routine GA announcement, no traction context