Nvidia
nvidia
Updated 7 minutes ago
NVIDIA NIM (NVIDIA Inference Microservices) is a platform that provides optimized AI model inference containers featuring industry-leading APIs for running AI models across NVIDIA's accelerated infrastructure. NIM supports models from major providers including Meta (Llama), Google (Gemma), Mistral, xAI (Grok), DeepSeek, Microsoft (Phi), Qwen, and NVIDIA's own Nemotron family. The platform offers standard APIs across multiple deployment options including cloud, on-premises, and local workstations, with microservices optimized for NVIDIA GPUs. NIM provides an OpenAI-compatible API endpoint at integrate.api.nvidia.com for easy integration, featuring over 180 models from various AI companies hosted on NVIDIA's inference infrastructure.
Browse 29 LLM models available from Nvidia. Compare prices and features.
Models (29)
| Organization | Model Name | Original Model | Input | Output | Free | |||
|---|---|---|---|---|---|---|---|---|
|
|
Meta | Muse Glimmer-30B |
meta/muse-glimmer-30b
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Nemotron 3.5 Lightning (30B A3B) |
nvidia/nemotron-3.5-lightning-30b-a3b
|
$0.00 | $0.00 | Free | ||
|
|
Poolside | Laguna XS 2.1 |
poolside/laguna-xs-2.1
|
$0.00 | $0.00 | Free | ||
|
|
Minimax | MiniMax M3 |
minimaxai/minimax-m3
|
$0.00 | $0.00 | Free |
|
|
|
|
Nvidia | Nemotron 3 Ultra (550B A55B) |
nvidia/nemotron-3-ultra-550b-a55b
|
$0.50 | $2.50 | |||
|
|
DeepSeek | DeepSeek V4 Pro |
deepseek-ai/deepseek-v4-pro-0813
|
$0.44 | $0.87 |
|
||
|
|
DeepSeek | DeepSeek V4 Flash |
deepseek-ai/deepseek-v4-flash-0731
|
$0.00 | $0.00 | Free |
|
|
|
|
DiffusionGemma 26B-A4B |
google/diffusiongemma-26b-a4b-it
|
- | - | ||||
|
|
Gemma 4 31B |
google/gemma-4-31b-it
|
$0.00 | $0.00 | Free |
|
||
|
|
Nvidia | Nemotron 3 Super (120B A12B) |
nvidia/nemotron-3-super-120b-a12b
|
$0.20 | $0.80 | |||
|
|
Nvidia | Nemotron 3 Nano (30B A3B) |
nvidia/nemotron-3-nano-30b-a3b
|
$0.00 | $0.00 | Free | ||
|
|
OpenAI | GPT OSS 120B |
openai/gpt-oss-120b
|
$0.00 | $0.00 | Free |
|
|
|
|
Nvidia | NVIDIA Nemotron Nano 9B V2 |
nvidia/nvidia-nemotron-nano-9b-v2
|
$0.00 | $0.00 | Free | ||
|
|
OpenAI | GPT OSS 20B |
openai/gpt-oss-20b
|
$0.00 | $0.00 | Free |
|
|
|
|
StepFun | Step 3.7 Flash |
stepfun-ai/step-3.7-flash
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Whisper Large v3 |
openai/whisper-large-v3
|
$0.00 | $0.00 | Free | ||
|
|
Black Forest Labs | flux-2-klein-4b |
black-forest-labs/flux.2-klein-4b
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Nemotron Nano 12B 2 VL (free) |
nvidia/nemotron-nano-12b-v2-vl
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Llama 3.3 Nemotron Super 49b V1.5 |
nvidia/llama-3.3-nemotron-super-49b-v1.5
|
$0.00 | $0.00 | Free | ||
|
|
Groq | Llama Guard 4 12B |
meta/llama-guard-4-12b
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Llama 3.2 11b Vision Instruct |
meta/llama-3.2-11b-vision-instruct
|
$0.00 | $0.00 | Free | ||
|
|
Azure | Llama-3.2-90B-Vision-Instruct |
meta/llama-3.2-90b-vision-instruct
|
$0.00 | $0.00 | Free | ||
|
|
Black Forest Labs | flux-1-kontext-dev |
black-forest-labs/FLUX.1-Kontext-dev
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | FLUX.1-dev |
black-forest-labs/FLUX.1-dev
|
$0.00 | $0.00 | Free | ||
|
|
Black Forest Labs | FLUX.1-schnell |
black-forest-labs/FLUX.1-schnell
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Nemotron 3 Nano Omni 30B A3B Reasoning |
nvidia/nemotron-3-nano-omni-30b-a3b-reasoning
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Parakeet TDT 0.6B v2 |
nvidia/parakeet-tdt-0.6b-v2
|
- | - | |||
|
|
qwen | Qwen Image |
qwen/qwen-image
|
$0.00 | $0.00 | Free | ||
|
|
qwen | Qwen Image Edit |
qwen/qwen-image-edit
|
$0.00 | $0.00 | Free |