models.
Fine-tune and serve open models.
Bring a checkpoint from anywhere, or pick one from the catalog. The mesh handles sharding, checkpointing, and the GPUs — you handle the prompt.
frameworksllama · mistral · qwen
fine-tuneLoRA, full, DPO
servingvLLM · sglang
checkpointsautomatic
mesh finetune --model llama-3.1-8b --data ./instruct.jsonl --gpu h100