Nvidia
nvidia
Updated 1 hour ago
NVIDIA NIM (NVIDIA Inference Microservices) is a platform that provides optimized AI model inference containers featuring industry-leading APIs for running AI models across NVIDIA's accelerated infrastructure. NIM supports models from major providers including Meta (Llama), Google (Gemma), Mistral, xAI (Grok), DeepSeek, Microsoft (Phi), Qwen, and NVIDIA's own Nemotron family. The platform offers standard APIs across multiple deployment options including cloud, on-premises, and local workstations, with microservices optimized for NVIDIA GPUs. NIM provides an OpenAI-compatible API endpoint at integrate.api.nvidia.com for easy integration, featuring over 180 models from various AI companies hosted on NVIDIA's inference infrastructure.
Browse 25 LLM models available from Nvidia. Compare prices and features.
Models (25)
| Organization | Model Name | Original Model | Input | Output | Free | |||
|---|---|---|---|---|---|---|---|---|
|
|
DeepSeek | DeepSeek-V4.1-Flash |
deepseek-ai/deepseek-v4.1-flash
|
$0.00 | $0.00 | Free |
|
|
|
|
Z.ai | GLM-5.3-Flash |
z-ai/glm-5-3-flash
|
$0.00 | $0.00 | Free |
|
|
|
|
Z.ai | GLM-5.3 |
z-ai/glm-5-3
|
$0.00 | $0.00 | Free |
|
|
|
|
Moonshot AI | Kimi K3 |
moonshotai/kimi-k3
|
$0.00 | $0.00 | Free | ||
|
|
Meta | Muse Glimmer-30B |
meta/muse-glimmer-30b
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Nemotron 3.5 Lightning (30B A3B) |
nvidia/nemotron-3.5-lightning-30b-a3b
|
$0.00 | $0.00 | Free | ||
|
|
Poolside | Laguna XS 2.1 |
poolside/laguna-xs-2.1
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Nemotron 3 Ultra (550B A55B) |
nvidia/nemotron-3-ultra-550b-a55b
|
$0.50 | $2.50 | |||
|
|
DiffusionGemma 26B-A4B |
google/diffusiongemma-26b-a4b-it
|
$0.00 | $0.00 | Free | |||
|
|
Gemma 4 31B |
google/gemma-4-31b-it
|
$0.00 | $0.00 | Free |
|
||
|
|
Nvidia | Nemotron 3 Super (120B A12B) |
nvidia/nemotron-3-super-120b-a12b
|
$0.20 | $0.80 | |||
|
|
OpenAI | GPT OSS 20B |
openai/gpt-oss-20b
|
$0.00 | $0.00 | Free |
|
|
|
|
Nvidia | NVIDIA: Nemotron 3.5 Content Safety |
nvidia/nemotron-3.5-content-safety
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Whisper Large v3 |
openai/whisper-large-v3
|
$0.00 | $0.00 | Free | ||
|
|
Black Forest Labs | flux-2-klein-4b |
black-forest-labs/flux.2-klein-4b
|
$0.00 | $0.00 | Free | ||
|
|
Groq | Llama Guard 4 12B |
meta/llama-guard-4-12b
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Llama 3.2 11b Vision Instruct |
meta/llama-3.2-11b-vision-instruct
|
$0.00 | $0.00 | Free | ||
|
|
Azure | Llama-3.2-90B-Vision-Instruct |
meta/llama-3.2-90b-vision-instruct
|
$0.00 | $0.00 | Free | ||
|
|
Black Forest Labs | flux-1-kontext-dev |
black-forest-labs/FLUX.1-Kontext-dev
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | FLUX.1-dev |
black-forest-labs/FLUX.1-dev
|
$0.00 | $0.00 | Free | ||
|
|
Black Forest Labs | FLUX.1-schnell |
black-forest-labs/FLUX.1-schnell
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Nemotron 3 Nano Omni 30B A3B Reasoning |
nvidia/nemotron-3-nano-omni-30b-a3b-reasoning
|
$0.00 | $0.00 | Free | ||
|
|
Nvidia | Parakeet TDT 0.6B v2 |
nvidia/parakeet-tdt-0.6b-v2
|
- | - | |||
|
|
qwen | Qwen Image |
qwen/qwen-image
|
$0.00 | $0.00 | Free | ||
|
|
qwen | Qwen Image Edit |
qwen/qwen-image-edit
|
$0.00 | $0.00 | Free |