inference cloud

Nebius

Models, pricing and feature compatibility available through Nebius.

Deployments

30

Models

30

Labs

6

With pricing

30

Models on Nebius

Follow any model to compare this deployment with the same model through other providers.

ModelProvider model IDModeRegionInputOutputContextTools
Baai Bge En IclBAAI/bge-en-iclEmbeddingglobal$0.01 / 1M$0.0000 / 1M32.8Knot_applicable
Baai Bge Multilingual Gemma2BAAI/bge-multilingual-gemma2Embeddingglobal$0.01 / 1M$0.0000 / 1M8.2Knot_applicable
DeepSeek-Ai Deepseek R1deepseek-ai/DeepSeek-R1Chatglobal$0.80 / 1M$2.40 / 1M128Ksupported
DeepSeek-Ai Deepseek R1 0528deepseek-ai/DeepSeek-R1-0528Chatglobal$0.80 / 1M$2.40 / 1M164Ksupported
Deepseek Ai Deepseek R1 Distill Llama 70Bdeepseek-ai/DeepSeek-R1-Distill-Llama-70BChatglobal$0.25 / 1M$0.75 / 1M128Ksupported
DeepSeek-Ai Deepseekdeepseek-ai/DeepSeek-V3Chatglobal$0.50 / 1M$1.50 / 1M128Ksupported
DeepSeek-Ai Deepseek V3 0324deepseek-ai/DeepSeek-V3-0324Chatglobal$0.50 / 1M$1.50 / 1M128Ksupported
Gemma 3 27B ITgoogle/gemma-3-27b-itChatglobal$0.06 / 1M$0.20 / 1M128Ksupported
Intfloat E5 Mistral 7B Instructintfloat/e5-mistral-7b-instructEmbeddingglobal$0.01 / 1M$0.0000 / 1M32.8Knot_applicable
Meta Llama Llama 3 3 70B Instructmeta-llama/Llama-3.3-70B-InstructChatglobal$0.13 / 1M$0.40 / 1M128Ksupported
Meta Llama Llama Guard 3 8Bmeta-llama/Llama-Guard-3-8BChatglobal$0.02 / 1M$0.06 / 1M128Kunknown
Meta Llama Meta Llama 3 1 405B Instructmeta-llama/Meta-Llama-3.1-405B-InstructChatglobal$1.00 / 1M$3.00 / 1M128Ksupported
Meta Llama Meta Llama 3 1 70B Instructmeta-llama/Meta-Llama-3.1-70B-InstructChatglobal$0.13 / 1M$0.40 / 1M128Ksupported
Meta Llama Meta Llama 3 1 8B Instructmeta-llama/Meta-Llama-3.1-8B-InstructChatglobal$0.02 / 1M$0.06 / 1M128Ksupported
Mistralai Mistral Nemo Instruct 2407mistralai/Mistral-Nemo-Instruct-2407Chatglobal$0.04 / 1M$0.12 / 1M128Ksupported
Nousresearch Hermes 3 Llama 3 1 405BNousResearch/Hermes-3-Llama-3.1-405BChatglobal$1.00 / 1M$3.00 / 1M128Ksupported
Nvidia Llama 3 1 Nemotron Ultra 253Bnvidia/Llama-3.1-Nemotron-Ultra-253B-v1Chatglobal$0.60 / 1M$1.80 / 1M128Ksupported
Nvidia Llama 3 3 Nemotron Super 49Bnvidia/Llama-3.3-Nemotron-Super-49B-v1Chatglobal$0.10 / 1M$0.40 / 1M131.1Ksupported
Qwen Qwen2 5 32B InstructQwen/Qwen2.5-32B-InstructChatglobal$0.06 / 1M$0.20 / 1M128Ksupported
Qwen Qwen2 5 72B InstructQwen/Qwen2.5-72B-InstructChatglobal$0.13 / 1M$0.40 / 1M128Ksupported
Qwen Qwen2 5 Coder 7BQwen/Qwen2.5-Coder-7BChatglobal$0.01 / 1M$0.03 / 1M32.8Ksupported
Qwen Qwen2 5 VL 72B InstructQwen/Qwen2.5-VL-72B-InstructChatglobal$0.13 / 1M$0.40 / 1M131.1Ksupported
Qwen Qwen2 VL 72B InstructQwen/Qwen2-VL-72B-InstructChatglobal$0.13 / 1M$0.40 / 1M131.1Ksupported
Qwen Qwen2 VL 7B InstructQwen/Qwen2-VL-7B-InstructChatglobal$0.02 / 1M$0.06 / 1M131.1Kunknown
Qwen Qwen3 14BQwen/Qwen3-14BChatglobal$0.08 / 1M$0.24 / 1M32.8Ksupported
Qwen Qwen3 235B A22BQwen/Qwen3-235B-A22BChatglobal$0.20 / 1M$0.60 / 1M262.1Ksupported
Qwen Qwen3 30B A3BQwen/Qwen3-30B-A3BChatglobal$0.10 / 1M$0.30 / 1M32.8Ksupported
Qwen Qwen3 32BQwen/Qwen3-32BChatglobal$0.10 / 1M$0.30 / 1M32.8Ksupported
Qwen Qwen3 4BQwen/Qwen3-4BChatglobal$0.08 / 1M$0.24 / 1M32.8Ksupported
Qwen Qwq 32BQwen/QwQ-32BChatglobal$0.15 / 1M$0.45 / 1M32.8Ksupported

Labs available through Nebius