View a markdown version of this page

Personnalisation du modèle Open Weight - Amazon SageMaker AI

Les traductions sont fournies par des outils de traduction automatique. En cas de conflit entre le contenu d'une traduction et celui de la version originale en anglais, la version anglaise prévaudra.

Personnalisation du modèle Open Weight

Cette section vous explique comment démarrer avec la personnalisation des modèles de poids ouverts.

Modèles et types de personnalisation pris en charge

Le tableau suivant présente les recettes de réglage prises en charge pour chaque modèle, notamment SFT, DPO, RLVR et RLAIF avec LoRA ou réglage fin complet (FFT).

Fournisseur Modèle ID du modèle DOUX (1) LoRA DOUX (PIEDS) DPO () LoRA DPO (PIEDS) RLVR () LoRA RLVR (FFT) RELIEF () LoRA RLAIF (PIEDS)
AlibabaQwen3.6 27Bhuggingface-vlm-qwen3-6-27b
AlibabaQwen3.5 27Bhuggingface-vlm-qwen3-5-27b
AlibabaQwen3.5 9Bhuggingface-vlm-qwen3-5-9b
AlibabaQwen3.5 4Bhuggingface-vlm-qwen3-5-4b
AlibabaQwen3 32 Gohuggingface-reasoning-qwen3-32b
AlibabaQwen3 14Bhuggingface-reasoning-qwen3-14b
AlibabaQwen3 8Bhuggingface-reasoning-qwen3-8b
AlibabaQwen3 4Ghuggingface-reasoning-qwen3-4b
AlibabaQwen3 1,7 Gohuggingface-reasoning-qwen3-1-7b
AlibabaQwen3 0,6 Vhuggingface-reasoning-qwen3-06b
AlibabaQwen2.5 Instruction 72Bhuggingface-llm-qwen2-5-72b-instruct
AlibabaQwen2.5 Instruire 32Bhuggingface-llm-qwen2-5-32b-instruct
AlibabaQwen2.5 Instruction 14Bhuggingface-llm-qwen2-5-14b-instruct
AlibabaQwen2.5 Instruire 7Bhuggingface-llm-qwen2-5-7b-instruct
DeepSeekD1 Distillateur Qwen 32 Vdeepseek-llm-r1-distill-qwen-32b
DeepSeekD1 Distillateur Qwen 14Bdeepseek-llm-r1-distill-qwen-14b
DeepSeekD1 Distillateur Qwen 7Bdeepseek-llm-r1-distill-qwen-7b
DeepSeekD1 Distillateur Qwen 1,5 Vdeepseek-llm-r1-distill-qwen-1-5b
DeepSeekR1 Distill Lama 70Bdeepseek-llm-r1-distill-llama-70b
DeepSeekLama distillé R1 8Bdeepseek-llm-r1-distill-llama-8b
GoogleGemma 4 31Bhuggingface-vlm-gemma-4-31b-it
GoogleGemma 4 26B A4Bhuggingface-vlm-gemma-4-26b-a4b-it
GoogleGemma 4 E4Bhuggingface-vlm-gemma-4-e4b-it
MetaLlama 3.3 Instruct 70Bmeta-textgeneration-llama-3-3-70b-instruct
MetaLlama 3.2 Instruct 3Bmeta-textgeneration-llama-3-2-3b-instruct
MetaLlama 3.2 Instruct 1Bmeta-textgeneration-llama-3-2-1b-instruct
MetaLlama 3.1 Instruct 8Bmeta-textgeneration-llama-3-1-8b-instruct
NVIDIANémotron 3 Super 120B (A12B)huggingface-llm-nvidia-nemotron-3-super-120b-a12b-bf16
NVIDIANémotron 3 Nano 30B (A3B)huggingface-reasoning-nvidia-nemotron-3-nano-30b-a3b-bf16
IA ouverteGPT OSS 120Bopenai-reasoning-gpt-oss-120b
IA ouverteGPT OSS 20Bopenai-reasoning-gpt-oss-20b