Note
The context length for all fine-tuning models is 8k.
The context length for all fine-tuning models is 8k.
Documentation Index
Fetch the complete documentation index at: /llms.txt
Use this file to discover all available pages before exploring further.
There are different models for fine-tuning and inference. The list of models for fine-tuning includes both instruct and non-instruct models, and is categorized by provider. Deployment in Nebius Token Factory does not support all models for fine-tuning. For more information, see a list of available models.
| Name | Supported fine-tuning type | License |
|---|---|---|
| deepseek-ai/DeepSeek-V3-0324 (Model card) | Full fine-tuning | MIT License |
| Name | Supported fine-tuning type | License |
|---|---|---|
| meta-llama/Llama-3.2-1B-Instruct (Model card) | LoRA and full fine-tuning | Llama 3.2 Community License Agreement |
| meta-llama/Llama-3.2-3B-Instruct (Model card) | LoRA and full fine-tuning | Llama 3.2 Community License Agreement |
| meta-llama/Llama-3.1-8B-Instruct (Model card) | LoRA and full fine-tuning | Llama 3.1 Community License Agreement |
| meta-llama/Llama-3.1-70B (Model card) | LoRA and full fine-tuning | Llama 3.1 Community License Agreement |
| meta-llama/Llama-3.3-70B-Instruct (Model card) | LoRA and full fine-tuning | Llama 3.3 Community License Agreement |
| Name | Supported fine-tuning type | License |
|---|---|---|
| unsloth/gpt-oss-20b-BF16 (Model card) | LoRA and full fine-tuning | Apache License 2.0 |
| unsloth/gpt-oss-120b-BF16 (Model card) | LoRA and full fine-tuning | Apache License 2.0 |
| Name | Supported fine-tuning type | License |
|---|---|---|
| Qwen/Qwen3-14B (Model card) | LoRA and full fine-tuning | Apache License 2.0 |
| Qwen/Qwen3-32B (Model card) | LoRA and full fine-tuning | Apache License 2.0 |
| Name | Supported fine-tuning type | License |
|---|---|---|
| meta-llama/Llama-3.1-8B-Instruct (Model card) | LoRA and full fine-tuning | Llama 3.1 Community License Agreement |
| meta-llama/Llama-3.3-70B-Instruct (Model card) | LoRA | Llama 3.3 Community License Agreement |
Was this page helpful?