Nvidia
nvidia
Updated 1 hour ago
NVIDIA NIM (NVIDIA Inference Microservices) is a platform that provides optimized AI model inference containers featuring industry-leading APIs for running AI models across NVIDIA's accelerated infrastructure. NIM supports models from major providers including Meta (Llama), Google (Gemma), Mistral, xAI (Grok), DeepSeek, Microsoft (Phi), Qwen, and NVIDIA's own Nemotron family. The platform offers standard APIs across multiple deployment options including cloud, on-premises, and local workstations, with microservices optimized for NVIDIA GPUs. NIM provides an OpenAI-compatible API endpoint at integrate.api.nvidia.com for easy integration, featuring over 180 models from various AI companies hosted on NVIDIA's inference infrastructure.
Browse 37 LLM models available from Nvidia. Compare prices and features.
Models (37)
| Organization | Model Name | Original Model | Input | Output | Free | |||
|---|---|---|---|---|---|---|---|---|
|
|
Meta | Muse Glimmer-30B |
meta/muse-glimmer-30b
|
- | - | |||
|
|
Z.ai | GLM-5.2 |
z-ai/glm-5.2
|
$0.00 | $0.00 | Free |
|
|
|
|
Nvidia | Nemotron 3 Ultra (550B A55B) |
nvidia/nemotron-3-ultra-550b-a55b
|
$0.50 | $2.50 | |||
|
|
Nvidia | Nemotron 3 Ultra 550B A55B |
nemotron-3-ultra-nvfp4
|
$0.60 | $2.40 |
|
||
|
|
Minimax | MiniMax M3 |
minimaxai/minimax-m3
|
$0.00 | $0.00 | Free |
|
|
|
|
DiffusionGemma 26B-A4B |
google/diffusiongemma-26b-a4b-it
|
- | - | ||||
|
|
Gemma 4 31B |
google/gemma-4-31b-it
|
$0.00 | $0.00 | Free |
|
||
|
|
Nvidia | Nemotron 3 Super (120B A12B) |
nvidia/nemotron-3-super-120b-a12b
|
$0.20 | $0.80 | |||
|
|
Nvidia | Nemotron 3 Nano (30B A3B) |
nvidia/nemotron-3-nano-30b-a3b
|
$0.00 | $0.00 | Free | ||
|
|
OpenAI | GPT OSS 120B |
openai/gpt-oss-120b
|
$0.00 | $0.00 | Free |
|
|
|
|
Nvidia | NVIDIA Nemotron Nano 9B V2 |
nvidia/nvidia-nemotron-nano-9b-v2
|
$0.00 | $0.00 | Free | ||
|
|
OpenAI | GPT OSS 20B |
openai/gpt-oss-20b
|
$0.00 | $0.00 | Free |
|
|
|
|
Nvidia | Llama-3.3 Nemotron Super 49B v1 |
nvidia/llama-3.3-nemotron-super-49b-v1
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Llama 3.1 Nemotron Nano 8B V1 |
nvidia/llama-3.1-nemotron-nano-8b-v1
|
$0.00 | $0.00 | Free | ||
|
|
Meta | Llama 3.3 70B Instruct |
meta/llama-3.3-70b-instruct
|
$0.00 | $0.00 | Free | ||
|
|
Meta | Llama 3.1 8B Instruct |
meta/llama-3.1-8b-instruct
|
$0.00 | $0.00 | Free |
|
|
|
|
Meta | Llama 3.2 3B Instruct |
meta/llama-3.2-3b-instruct
|
$0.00 | $0.00 | Free | ||
|
|
Meta | Llama 3.1 70B Instruct |
meta/llama-3.1-70b-instruct
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Nemotron 3.5 Lightning (30B A3B) |
nvidia/nemotron-3.5-lightning-30b-a3b
|
$0.00 | $0.00 | Free | ||
|
|
BaseTen | Inkling |
thinkingmachines/inkling
|
$0.00 | $0.00 | Free | ||
|
|
Poolside | Poolside: Laguna XS 2.1 |
poolside/laguna-xs-2.1
|
$0.00 | $0.00 | Free | ||
|
|
StepFun | Step 3.7 Flash |
stepfun-ai/step-3.7-flash
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Whisper Large v3 |
openai/whisper-large-v3
|
$0.00 | $0.00 | Free | ||
|
|
Black Forest Labs | flux-2-klein-4b |
black-forest-labs/flux.2-klein-4b
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Nemotron Nano 12B 2 VL (free) |
nvidia/nemotron-nano-12b-v2-vl
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Llama 3.3 Nemotron Super 49b V1.5 |
nvidia/llama-3.3-nemotron-super-49b-v1.5
|
$0.00 | $0.00 | Free | ||
|
|
Groq | Llama Guard 4 12B |
meta/llama-guard-4-12b
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Llama 3.2 11b Vision Instruct |
meta/llama-3.2-11b-vision-instruct
|
$0.00 | $0.00 | Free | ||
|
|
Meta | llama-3.2-1b-instruct |
meta/llama-3.2-1b-instruct
|
$0.00 | $0.00 | Free | ||
|
|
Azure | Llama-3.2-90B-Vision-Instruct |
meta/llama-3.2-90b-vision-instruct
|
$0.00 | $0.00 | Free | ||
|
|
Black Forest Labs | flux-1-kontext-dev |
black-forest-labs/FLUX.1-Kontext-dev
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | FLUX.1-dev |
black-forest-labs/FLUX.1-dev
|
$0.00 | $0.00 | Free | ||
|
|
Black Forest Labs | FLUX.1-schnell |
black-forest-labs/FLUX.1-schnell
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Nemotron 3 Nano Omni 30B A3B Reasoning |
nvidia/nemotron-3-nano-omni-30b-a3b-reasoning
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Parakeet TDT 0.6B v2 |
nvidia/parakeet-tdt-0.6b-v2
|
- | - | |||
|
|
qwen | Qwen Image |
qwen/qwen-image
|
$0.00 | $0.00 | Free | ||
|
|
qwen | Qwen Image Edit |
qwen/qwen-image-edit
|
$0.00 | $0.00 | Free |