WritingBasetenBasetenpublished Oct 31, 2025seen Jun 26

Training Is Now Ga

Open original ↗

Captured source

source ↗
published Oct 31, 2025seen Jun 26captured Jun 27http 200method plain

Training is now GA! Announcing our Series F . Learn more

changelog / post

Training is now GA!

Oct 31, 2025 Go back

Since launching the beta of Baseten Training in May, we’ve introduced a ton of improvements, including; A more robust ML Cookbook , great starting points for: Training a coding model with GRPO

Long-context training using multi-node with Qwen3 30B A3B

A variety of examples with Qwen3, gpt-oss, Gemma3, and Llama

Resume from checkpoint : Launch jobs that pick up right where you left off

A ton of other improvements, including: Broader checkpoint recognition across FSDP, VeRL, and Megatron checkpointing formats

More availability for InfiniBand-backed multi-node training runs

Improved management and handling of the training cache

Per-GPU metric visibility and improved logs

Quality of life improvements around Training Cache

And much more!

After months of positive feedback from early users and thousands of training runs completed, Baseten Training is now immediately available for anyone on Baseten. Get started here .

Explore Baseten today Start deploying Talk to an engineer

Popular models GLM 5.2

Kimi K2.7 Code

DeepSeek V4

GPT OSS 120B

Whisper Large V3

NVIDIA Nemotron 3 Ultra

Explore all

Popular models GLM 5.2

Kimi K2.7 Code

DeepSeek V4

GPT OSS 120B

Whisper Large V3

NVIDIA Nemotron 3 Ultra

Explore all

Notability

notability 3.0/10

Routine GA announcement, no traction context