Custom Training: RL and SFT for Open Models | Together AI

Reinforcement Learning - now in beta

Post-training, all the way to production

RL and SFT, full-weight and LoRA, on frontier open models. The highest-quality custom models through fast, large-scale experimentation.

Request access

Why Together Custom Training

One platform that takes a model from first experiment to production, without leaving the stack.

Choose your training method

Both run through a granular Python SDK, with high configurability over every training job.

Everything you need to reach production

Your code defines the run. Full configurability, dedicated capacity, and one path from training to serving.

Run many LoRA adapter experiments in parallel inside one training deployment. Converge on your best model in one fast test loop.

Backed by frontier research

Every run sits on Together's own training and inference research.

Accuracy (%) on DeepMath reasoning benchmark

UPipe vs other SOTA Approaches

82.5% less memory

FFT Optimizer results

25% less memory

security and data privacy
We take security and compliance seriously, with strict data privacy controls to keep your information protected. Your data and models remain fully under your ownership, safeguarded by robust security measures.

preferred partner