Skip to main content

NVIDIA

NVIDIA NIM provides optimized inference microservices for AI models with enterprise-grade performance and scalability.

NVIDIA

NVIDIA NIM provides optimized inference microservices for AI models with enterprise-grade performance and scalability.

Authentication

Auth Type
API Key
Platform Key
Not configured
Self-serve
Available

Add your NVIDIA API key from Organization → Integrations → AI Vendors. Keys are encrypted at rest and scoped to your organization.

Available Models

79 enabled models across 3 categories.

Text75

Codegemma 1.1 7bgoogle/codegemma-1.1-7b

chatModel, languageModel, completionModel

Codegemma 7bgoogle/codegemma-7b

chatModel, languageModel, completionModel

Codellama 70bmeta/codellama-70b

chatModel, languageModel, completionModel

Codestral 22b Instruct V0.1mistralai/codestral-22b-instruct-v0.1

chatModel, languageModel, completionModel

functionCalling

Cosmos Nemotron 34Bnvidia/cosmos-nemotron-34b

chatModel, languageModel, completionModel

reasoning, imageInput, vision

Deepseek Coder 6.7b Instructdeepseek-ai/deepseek-coder-6.7b-instruct

chatModel, languageModel, completionModel

functionCalling

Deepseek R1deepseek-ai/deepseek-r1

chatModel, languageModel, completionModel

reasoning

Deepseek R1 0528deepseek-ai/deepseek-r1-0528

chatModel, languageModel, completionModel

functionCalling, reasoning

DeepSeek V3.1deepseek-ai/deepseek-v3.1

chatModel, languageModel, completionModel

functionCalling, reasoning

DeepSeek V3.1 Terminusdeepseek-ai/deepseek-v3.1-terminus

chatModel, languageModel, completionModel

functionCalling, reasoning

DeepSeek V3.2deepseek-ai/deepseek-v3.2

chatModel, languageModel, completionModel

functionCalling, reasoning

DeepSeek V4 Flashdeepseek-ai/deepseek-v4-flash

chatModel, languageModel, completionModel

functionCalling, reasoning

DeepSeek V4 Prodeepseek-ai/deepseek-v4-pro

chatModel, languageModel, completionModel

functionCalling, reasoning

Devstral-2-123B-Instruct-2512mistralai/devstral-2-123b-instruct-2512

chatModel, languageModel, completionModel

functionCalling, reasoning

Gemma 2 27b Itgoogle/gemma-2-27b-it

chatModel, languageModel, completionModel

functionCalling

Gemma 2 2b Itgoogle/gemma-2-2b-it

chatModel, languageModel, completionModel

functionCalling

Gemma 3 12b Itgoogle/gemma-3-12b-it

chatModel, languageModel, completionModel

functionCalling

Gemma 3 1b Itgoogle/gemma-3-1b-it

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

Gemma-3-27B-ITgoogle/gemma-3-27b-it

chatModel, languageModel, completionModel

functionCalling, reasoning, imageInput, vision

Gemma 3n E2b Itgoogle/gemma-3n-e2b-it

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

Gemma 3n E4b Itgoogle/gemma-3n-e4b-it

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

Gemma-4-31B-ITgoogle/gemma-4-31b-it

chatModel, languageModel, completionModel

functionCalling, reasoning, imageInput, vision

GLM-4.7z-ai/glm4.7

chatModel, languageModel, completionModel

functionCalling, reasoning

GLM5z-ai/glm5

chatModel, languageModel, completionModel

functionCalling, reasoning

GLM-5.1z-ai/glm-5.1

chatModel, languageModel, completionModel

functionCalling, reasoning

GPT-OSS-120Bopenai/gpt-oss-120b

chatModel, languageModel, completionModel

reasoning

Kimi K2 0905moonshotai/kimi-k2-instruct-0905

chatModel, languageModel, completionModel

functionCalling

Kimi K2.5moonshotai/kimi-k2.5

chatModel, languageModel, completionModel

functionCalling, reasoning, imageInput, vision

Kimi K2 Instructmoonshotai/kimi-k2-instruct

chatModel, languageModel, completionModel

functionCalling, reasoning

Kimi K2 Thinkingmoonshotai/kimi-k2-thinking

chatModel, languageModel, completionModel

functionCalling, reasoning

Llama 3.1 405b Instructmeta/llama-3.1-405b-instruct

chatModel, languageModel, completionModel

functionCalling

Llama 3.1 70b Instructmeta/llama-3.1-70b-instruct

chatModel, languageModel, completionModel

functionCalling

Llama 3.1 Nemotron 51b Instructnvidia/llama-3.1-nemotron-51b-instruct

chatModel, languageModel, completionModel

functionCalling

Llama 3.1 Nemotron 70b Instructnvidia/llama-3.1-nemotron-70b-instruct

chatModel, languageModel, completionModel

functionCalling

Llama-3.1-Nemotron-Ultra-253B-v1nvidia/llama-3.1-nemotron-ultra-253b-v1

chatModel, languageModel, completionModel

functionCalling, reasoning

Llama 3.2 11b Vision Instructmeta/llama-3.2-11b-vision-instruct

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

Llama 3.2 1b Instructmeta/llama-3.2-1b-instruct

chatModel, languageModel, completionModel

functionCalling

Llama 3.3 70b Instructmeta/llama-3.3-70b-instruct

chatModel, languageModel, completionModel

functionCalling

Llama 3.3 Nemotron Super 49b V1nvidia/llama-3.3-nemotron-super-49b-v1

chatModel, languageModel, completionModel

Llama 3.3 Nemotron Super 49b V1.5nvidia/llama-3.3-nemotron-super-49b-v1.5

chatModel, languageModel, completionModel

Llama3 70b Instructmeta/llama3-70b-instruct

chatModel, languageModel, completionModel

functionCalling

Llama3 8b Instructmeta/llama3-8b-instruct

chatModel, languageModel, completionModel

functionCalling

Llama3 Chatqa 1.5 70bnvidia/llama3-chatqa-1.5-70b

chatModel, languageModel, completionModel

functionCalling

Llama 4 Maverick 17b 128e Instructmeta/llama-4-maverick-17b-128e-instruct

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

Llama 4 Scout 17b 16e Instructmeta/llama-4-scout-17b-16e-instruct

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

Llama Embed Nemotron 8Bnvidia/llama-embed-nemotron-8b

chatModel, languageModel, completionModel

Mamba Codestral 7b V0.1mistralai/mamba-codestral-7b-v0.1

chatModel, languageModel, completionModel

MiniMax-M2.1minimaxai/minimax-m2.1

chatModel, languageModel, completionModel

functionCalling, reasoning

MiniMax-M2.5minimaxai/minimax-m2.5

chatModel, languageModel, completionModel

functionCalling, reasoning

MiniMax-M2.7minimaxai/minimax-m2.7

chatModel, languageModel, completionModel

functionCalling, reasoning

Ministral 3 14B Instruct 2512mistralai/ministral-14b-instruct-2512

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

Mistral Large 2 Instructmistralai/mistral-large-2-instruct

chatModel, languageModel, completionModel

functionCalling

Mistral Large 3 675B Instruct 2512mistralai/mistral-large-3-675b-instruct-2512

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

Mistral Small 3.1 24b Instruct 2503mistralai/mistral-small-3.1-24b-instruct-2503

chatModel, languageModel, completionModel

functionCalling

NeMo Retriever OCR v1nvidia/nemoretriever-ocr-v1

chatModel, languageModel, completionModel

imageInput, vision

nemotron-3-nano-30b-a3bnvidia/nemotron-3-nano-30b-a3b

chatModel, languageModel, completionModel

functionCalling, reasoning

Nemotron 3 Supernvidia/nemotron-3-super-120b-a12b

chatModel, languageModel, completionModel

functionCalling, reasoning

Nemotron 4 340b Instructnvidia/nemotron-4-340b-instruct

chatModel, languageModel, completionModel

functionCalling

nvidia-nemotron-nano-9b-v2nvidia/nvidia-nemotron-nano-9b-v2

chatModel, languageModel, completionModel

functionCalling, reasoning

Phi 3.5 Moe Instructmicrosoft/phi-3.5-moe-instruct

chatModel, languageModel, completionModel

functionCalling

Phi 3.5 Vision Instructmicrosoft/phi-3.5-vision-instruct

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

Phi 3 Medium 128k Instructmicrosoft/phi-3-medium-128k-instruct

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

Phi 3 Medium 4k Instructmicrosoft/phi-3-medium-4k-instruct

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

Phi 3 Small 128k Instructmicrosoft/phi-3-small-128k-instruct

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

Phi 3 Small 8k Instructmicrosoft/phi-3-small-8k-instruct

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

Phi 3 Vision 128k Instructmicrosoft/phi-3-vision-128k-instruct

chatModel, languageModel, completionModel

functionCalling, imageInput, vision

Qwen2.5 Coder 32b Instructqwen/qwen2.5-coder-32b-instruct

chatModel, languageModel, completionModel

functionCalling

Qwen2.5 Coder 7b Instructqwen/qwen2.5-coder-7b-instruct

chatModel, languageModel, completionModel

functionCalling

Qwen3-235B-A22Bqwen/qwen3-235b-a22b

chatModel, languageModel, completionModel

functionCalling, reasoning

Qwen3.5-397B-A17Bqwen/qwen3.5-397b-a17b

chatModel, languageModel, completionModel

functionCalling, reasoning, imageInput, vision

Qwen3 Coder 480B A35B Instructqwen/qwen3-coder-480b-a35b-instruct

chatModel, languageModel, completionModel

functionCalling

Qwen3-Next-80B-A3B-Instructqwen/qwen3-next-80b-a3b-instruct

chatModel, languageModel, completionModel

functionCalling

Qwen3-Next-80B-A3B-Thinkingqwen/qwen3-next-80b-a3b-thinking

chatModel, languageModel, completionModel

functionCalling, reasoning

Qwq 32bqwen/qwq-32b

chatModel, languageModel, completionModel

reasoning

Step 3.5 Flashstepfun-ai/step-3.5-flash

chatModel, languageModel, completionModel

functionCalling, reasoning

Image1

FLUX.1-devblack-forest-labs/flux.1-dev

image, imageModel

Sound3

Parakeet TDT 0.6B v2nvidia/parakeet-tdt-0.6b-v2

transcription

Phi-4-Minimicrosoft/phi-4-mini-instruct

transcription

functionCalling, reasoning, imageInput, vision

Whisper Large v3openai/whisper-large-v3

transcription