View a markdown version of this page

Open weight model customization - Amazon SageMaker AI

Open weight model customization

This section walks you through the process to get started with open weight model customization.

Supported models and customization types

The following table shows the supported fine-tuning recipes for each model, including SFT, DPO, RLVR, and RLAIF with LoRA or full fine-tuning (FFT).

Provider Model Model ID SFT (LoRA) SFT (FFT) DPO (LoRA) DPO (FFT) RLVR (LoRA) RLVR (FFT) RLAIF (LoRA) RLAIF (FFT)
AlibabaQwen3.6 27Bhuggingface-vlm-qwen3-6-27b
AlibabaQwen3.5 27Bhuggingface-vlm-qwen3-5-27b
AlibabaQwen3.5 9Bhuggingface-vlm-qwen3-5-9b
AlibabaQwen3.5 4Bhuggingface-vlm-qwen3-5-4b
AlibabaQwen3 32Bhuggingface-reasoning-qwen3-32b
AlibabaQwen3 14Bhuggingface-reasoning-qwen3-14b
AlibabaQwen3 8Bhuggingface-reasoning-qwen3-8b
AlibabaQwen3 4Bhuggingface-reasoning-qwen3-4b
AlibabaQwen3 1.7Bhuggingface-reasoning-qwen3-1-7b
AlibabaQwen3 0.6Bhuggingface-reasoning-qwen3-06b
AlibabaQwen2.5 Instruct 72Bhuggingface-llm-qwen2-5-72b-instruct
AlibabaQwen2.5 Instruct 32Bhuggingface-llm-qwen2-5-32b-instruct
AlibabaQwen2.5 Instruct 14Bhuggingface-llm-qwen2-5-14b-instruct
AlibabaQwen2.5 Instruct 7Bhuggingface-llm-qwen2-5-7b-instruct
DeepSeekR1 Distill Qwen 32Bdeepseek-llm-r1-distill-qwen-32b
DeepSeekR1 Distill Qwen 14Bdeepseek-llm-r1-distill-qwen-14b
DeepSeekR1 Distill Qwen 7Bdeepseek-llm-r1-distill-qwen-7b
DeepSeekR1 Distill Qwen 1.5Bdeepseek-llm-r1-distill-qwen-1-5b
DeepSeekR1 Distill Llama 70Bdeepseek-llm-r1-distill-llama-70b
DeepSeekR1 Distill Llama 8Bdeepseek-llm-r1-distill-llama-8b
GoogleGemma 4 31Bhuggingface-vlm-gemma-4-31b-it
GoogleGemma 4 26B A4Bhuggingface-vlm-gemma-4-26b-a4b-it
GoogleGemma 4 E4Bhuggingface-vlm-gemma-4-e4b-it
MetaLlama 3.3 Instruct 70Bmeta-textgeneration-llama-3-3-70b-instruct
MetaLlama 3.2 Instruct 3Bmeta-textgeneration-llama-3-2-3b-instruct
MetaLlama 3.2 Instruct 1Bmeta-textgeneration-llama-3-2-1b-instruct
MetaLlama 3.1 Instruct 8Bmeta-textgeneration-llama-3-1-8b-instruct
NVIDIANemotron 3 Super 120B (A12B)huggingface-llm-nvidia-nemotron-3-super-120b-a12b-bf16
NVIDIANemotron 3 Nano 30B (A3B)huggingface-reasoning-nvidia-nemotron-3-nano-30b-a3b-bf16
OpenAIGPT OSS 120Bopenai-reasoning-gpt-oss-120b
OpenAIGPT OSS 20Bopenai-reasoning-gpt-oss-20b