inference cloud
Nebius
Models, pricing and feature compatibility available through Nebius.
Deployments
30
Models
30
Labs
6
With pricing
30
Models on Nebius
Follow any model to compare this deployment with the same model through other providers.
| Model | Provider model ID | Mode | Region | Input | Output | Context | Tools |
|---|---|---|---|---|---|---|---|
| Baai Bge En Icl | BAAI/bge-en-icl | Embedding | global | $0.01 / 1M | $0.0000 / 1M | 32.8K | not_applicable |
| Baai Bge Multilingual Gemma2 | BAAI/bge-multilingual-gemma2 | Embedding | global | $0.01 / 1M | $0.0000 / 1M | 8.2K | not_applicable |
| DeepSeek-Ai Deepseek R1 | deepseek-ai/DeepSeek-R1 | Chat | global | $0.80 / 1M | $2.40 / 1M | 128K | supported |
| DeepSeek-Ai Deepseek R1 0528 | deepseek-ai/DeepSeek-R1-0528 | Chat | global | $0.80 / 1M | $2.40 / 1M | 164K | supported |
| Deepseek Ai Deepseek R1 Distill Llama 70B | deepseek-ai/DeepSeek-R1-Distill-Llama-70B | Chat | global | $0.25 / 1M | $0.75 / 1M | 128K | supported |
| DeepSeek-Ai Deepseek | deepseek-ai/DeepSeek-V3 | Chat | global | $0.50 / 1M | $1.50 / 1M | 128K | supported |
| DeepSeek-Ai Deepseek V3 0324 | deepseek-ai/DeepSeek-V3-0324 | Chat | global | $0.50 / 1M | $1.50 / 1M | 128K | supported |
| Gemma 3 27B IT | google/gemma-3-27b-it | Chat | global | $0.06 / 1M | $0.20 / 1M | 128K | supported |
| Intfloat E5 Mistral 7B Instruct | intfloat/e5-mistral-7b-instruct | Embedding | global | $0.01 / 1M | $0.0000 / 1M | 32.8K | not_applicable |
| Meta Llama Llama 3 3 70B Instruct | meta-llama/Llama-3.3-70B-Instruct | Chat | global | $0.13 / 1M | $0.40 / 1M | 128K | supported |
| Meta Llama Llama Guard 3 8B | meta-llama/Llama-Guard-3-8B | Chat | global | $0.02 / 1M | $0.06 / 1M | 128K | unknown |
| Meta Llama Meta Llama 3 1 405B Instruct | meta-llama/Meta-Llama-3.1-405B-Instruct | Chat | global | $1.00 / 1M | $3.00 / 1M | 128K | supported |
| Meta Llama Meta Llama 3 1 70B Instruct | meta-llama/Meta-Llama-3.1-70B-Instruct | Chat | global | $0.13 / 1M | $0.40 / 1M | 128K | supported |
| Meta Llama Meta Llama 3 1 8B Instruct | meta-llama/Meta-Llama-3.1-8B-Instruct | Chat | global | $0.02 / 1M | $0.06 / 1M | 128K | supported |
| Mistralai Mistral Nemo Instruct 2407 | mistralai/Mistral-Nemo-Instruct-2407 | Chat | global | $0.04 / 1M | $0.12 / 1M | 128K | supported |
| Nousresearch Hermes 3 Llama 3 1 405B | NousResearch/Hermes-3-Llama-3.1-405B | Chat | global | $1.00 / 1M | $3.00 / 1M | 128K | supported |
| Nvidia Llama 3 1 Nemotron Ultra 253B | nvidia/Llama-3.1-Nemotron-Ultra-253B-v1 | Chat | global | $0.60 / 1M | $1.80 / 1M | 128K | supported |
| Nvidia Llama 3 3 Nemotron Super 49B | nvidia/Llama-3.3-Nemotron-Super-49B-v1 | Chat | global | $0.10 / 1M | $0.40 / 1M | 131.1K | supported |
| Qwen Qwen2 5 32B Instruct | Qwen/Qwen2.5-32B-Instruct | Chat | global | $0.06 / 1M | $0.20 / 1M | 128K | supported |
| Qwen Qwen2 5 72B Instruct | Qwen/Qwen2.5-72B-Instruct | Chat | global | $0.13 / 1M | $0.40 / 1M | 128K | supported |
| Qwen Qwen2 5 Coder 7B | Qwen/Qwen2.5-Coder-7B | Chat | global | $0.01 / 1M | $0.03 / 1M | 32.8K | supported |
| Qwen Qwen2 5 VL 72B Instruct | Qwen/Qwen2.5-VL-72B-Instruct | Chat | global | $0.13 / 1M | $0.40 / 1M | 131.1K | supported |
| Qwen Qwen2 VL 72B Instruct | Qwen/Qwen2-VL-72B-Instruct | Chat | global | $0.13 / 1M | $0.40 / 1M | 131.1K | supported |
| Qwen Qwen2 VL 7B Instruct | Qwen/Qwen2-VL-7B-Instruct | Chat | global | $0.02 / 1M | $0.06 / 1M | 131.1K | unknown |
| Qwen Qwen3 14B | Qwen/Qwen3-14B | Chat | global | $0.08 / 1M | $0.24 / 1M | 32.8K | supported |
| Qwen Qwen3 235B A22B | Qwen/Qwen3-235B-A22B | Chat | global | $0.20 / 1M | $0.60 / 1M | 262.1K | supported |
| Qwen Qwen3 30B A3B | Qwen/Qwen3-30B-A3B | Chat | global | $0.10 / 1M | $0.30 / 1M | 32.8K | supported |
| Qwen Qwen3 32B | Qwen/Qwen3-32B | Chat | global | $0.10 / 1M | $0.30 / 1M | 32.8K | supported |
| Qwen Qwen3 4B | Qwen/Qwen3-4B | Chat | global | $0.08 / 1M | $0.24 / 1M | 32.8K | supported |
| Qwen Qwq 32B | Qwen/QwQ-32B | Chat | global | $0.15 / 1M | $0.45 / 1M | 32.8K | supported |