Skip to main content

Nebius

Cloud-based AI inference through the Token Factory platform with access to popular language models.

Nebius

Cloud-based AI inference through the Token Factory platform with access to popular language models.

Authentication

Auth Type
API Key
Platform Key
Not configured
Self-serve
Available

Add your Nebius API key from Organization → Integrations → AI Vendors. Keys are encrypted at rest and scoped to your organization.

Available Models

49 enabled models across 2 categories.

Text47

BGE-ICLBAAI/bge-en-icl

chatModel, languageModel, completionModel

bge-multilingual-gemma2BAAI/bge-multilingual-gemma2

chatModel, languageModel, completionModel

DeepSeek-R1-0528deepseek-ai/DeepSeek-R1-0528

chatModel, languageModel, completionModel

functionCalling, reasoning

DeepSeek R1 0528 Fastdeepseek-ai/DeepSeek-R1-0528-fast

chatModel, languageModel, completionModel

functionCalling, reasoning

DeepSeek-V3-0324deepseek-ai/DeepSeek-V3-0324

chatModel, languageModel, completionModel

functionCalling

DeepSeek-V3-0324 (Fast)deepseek-ai/DeepSeek-V3-0324-fast

chatModel, languageModel, completionModel

functionCalling

DeepSeek-V3.2deepseek-ai/DeepSeek-V3.2

chatModel, languageModel, completionModel

functionCalling, reasoning

e5-mistral-7b-instructintfloat/e5-mistral-7b-instruct

chatModel, languageModel, completionModel

Gemma-2-2b-itgoogle/gemma-2-2b-it

chatModel, languageModel, completionModel

Gemma-2-9b-it (Fast)google/gemma-2-9b-it-fast

chatModel, languageModel, completionModel

Gemma-3-27b-itgoogle/gemma-3-27b-it

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

Gemma-3-27b-it (Fast)google/gemma-3-27b-it-fast

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

GLM-4.5zai-org/GLM-4.5

chatModel, languageModel, completionModel

functionCalling

GLM-4.5-Airzai-org/GLM-4.5-Air

chatModel, languageModel, completionModel

functionCalling

GLM-4.7 (FP8)zai-org/GLM-4.7-FP8

chatModel, languageModel, completionModel

functionCalling

GLM-5zai-org/GLM-5

chatModel, languageModel, completionModel

functionCalling, reasoning

gpt-oss-120bopenai/gpt-oss-120b

chatModel, languageModel, completionModel

functionCalling, reasoning

gpt-oss-20bopenai/gpt-oss-20b

chatModel, languageModel, completionModel

functionCalling

Hermes-4-405BNousResearch/Hermes-4-405B

chatModel, languageModel, completionModel

functionCalling, reasoning

Hermes-4-70BNousResearch/Hermes-4-70B

chatModel, languageModel, completionModel

functionCalling, reasoning

INTELLECT-3PrimeIntellect/INTELLECT-3

chatModel, languageModel, completionModel

functionCalling

Kimi-K2.5moonshotai/Kimi-K2.5

chatModel, languageModel, completionModel

functionCalling, reasoning, imageInput, vision

Kimi-K2.5-fastmoonshotai/Kimi-K2.5-fast

chatModel, languageModel, completionModel

functionCalling, reasoning, imageInput, vision

Kimi-K2-Instructmoonshotai/Kimi-K2-Instruct

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

Kimi-K2-Thinkingmoonshotai/Kimi-K2-Thinking

chatModel, languageModel, completionModel

functionCalling, reasoning

Llama-3.1-Nemotron-Ultra-253B-v1nvidia/Llama-3_1-Nemotron-Ultra-253B-v1

chatModel, languageModel, completionModel

functionCalling

Llama-3.3-70B-Instructmeta-llama/Llama-3.3-70B-Instruct

chatModel, languageModel, completionModel

functionCalling

Llama-3.3-70B-Instruct (Fast)meta-llama/Llama-3.3-70B-Instruct-fast

chatModel, languageModel, completionModel

functionCalling

Llama-Guard-3-8Bmeta-llama/Llama-Guard-3-8B

chatModel, languageModel, completionModel

Meta-Llama-3.1-8B-Instructmeta-llama/Meta-Llama-3.1-8B-Instruct

chatModel, languageModel, completionModel

functionCalling

Meta-Llama-3.1-8B-Instruct (Fast)meta-llama/Meta-Llama-3.1-8B-Instruct-fast

chatModel, languageModel, completionModel

functionCalling

MiniMax-M2.1MiniMaxAI/MiniMax-M2.1

chatModel, languageModel, completionModel

functionCalling, reasoning

Nemotron-3-Nano-30B-A3Bnvidia/NVIDIA-Nemotron-3-Nano-30B-A3B

chatModel, languageModel, completionModel

functionCalling

Nemotron-3-Super-120B-A12Bnvidia/nemotron-3-super-120b-a12b

chatModel, languageModel, completionModel

functionCalling, reasoning

Nemotron-Nano-V2-12bnvidia/Nemotron-Nano-V2-12b

chatModel, languageModel, completionModel

functionCalling

Qwen2.5-Coder-7B (Fast)Qwen/Qwen2.5-Coder-7B-fast

chatModel, languageModel, completionModel

functionCalling

Qwen2.5-VL-72B-InstructQwen/Qwen2.5-VL-72B-Instruct

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

Qwen3 235B A22B Instruct 2507Qwen/Qwen3-235B-A22B-Instruct-2507

chatModel, languageModel, completionModel

functionCalling, reasoning

Qwen3 235B A22B Thinking 2507Qwen/Qwen3-235B-A22B-Thinking-2507

chatModel, languageModel, completionModel

functionCalling, reasoning

Qwen3-30B-A3B-Instruct-2507Qwen/Qwen3-30B-A3B-Instruct-2507

chatModel, languageModel, completionModel

functionCalling

Qwen3-30B-A3B-Thinking-2507Qwen/Qwen3-30B-A3B-Thinking-2507

chatModel, languageModel, completionModel

functionCalling, reasoning

Qwen3-32BQwen/Qwen3-32B

chatModel, languageModel, completionModel

functionCalling

Qwen3-32B (Fast)Qwen/Qwen3-32B-fast

chatModel, languageModel, completionModel

functionCalling

Qwen3-Coder-30B-A3B-InstructQwen/Qwen3-Coder-30B-A3B-Instruct

chatModel, languageModel, completionModel

functionCalling

Qwen3 Coder 480B A35B InstructQwen/Qwen3-Coder-480B-A35B-Instruct

chatModel, languageModel, completionModel

functionCalling

Qwen3-Embedding-8BQwen/Qwen3-Embedding-8B

chatModel, languageModel, completionModel

Qwen3-Next-80B-A3B-ThinkingQwen/Qwen3-Next-80B-A3B-Thinking

chatModel, languageModel, completionModel

functionCalling, reasoning

Image2

FLUX.1-devblack-forest-labs/flux-dev

image, imageModel

FLUX.1-schnellblack-forest-labs/flux-schnell

image, imageModel