Weights & Biases
wandb
Updated 49 minutes ago
Weights & Biases (W&B) is an AI/ML platform that provides model training tracking, experiment management, and inference services. Their W&B Inference API offers an OpenAI-compatible interface to access curated open-source language models including DeepSeek, Qwen, Meta Llama, Google Gemma, MiniMax, Moonshot AI (Kimi), NVIDIA Nemotron, Microsoft Phi, and others. The platform is known for its MLOps tooling and has expanded into hosted inference with competitive pricing.
Browse 40 LLM models available from Weights & Biases. Compare prices and features.
Models (40)
| Organization | Model Name | Original Model | Input | Output | Free | |||
|---|---|---|---|---|---|---|---|---|
|
|
Z.ai | GLM-5.3-Flash |
GLM-5.3-Flash
|
- | - |
|
||
|
|
Z.ai | GLM-5.3-Flash |
zai-org/GLM-5.3-Flash
|
$0.15 | $0.50 |
|
||
|
|
DeepSeek | DeepSeek-V4.1-Flash |
DeepSeek-V4.1-Flash
|
- | - |
|
||
|
|
DeepSeek | DeepSeek-V4.1-Flash |
deepseek-ai/DeepSeek-V4.1-Flash
|
$0.20 | $0.65 |
|
||
|
|
qwen | Qwen3.8-27B |
Qwen/Qwen3.8-27B
|
$0.40 | $3.00 | |||
|
|
qwen | Qwen3.8-27B |
Qwen3.8-27B
|
- | - | |||
|
|
Z.ai | GLM-5.2 |
GLM-5.2
|
- | - |
|
||
|
|
Z.ai | GLM-5.2 |
zai-org/GLM-5.2
|
$0.76 | $2.42 |
|
||
|
|
WandB | IBM Granite 4.2 8B |
ibm-granite/granite-4.2-8b
|
$0.10 | $0.15 | |||
|
|
WandB | IBM Granite 4.2 8B |
granite-4.2-8b
|
- | - | |||
|
|
Moonshot AI | Kimi K2.7 Code |
Kimi-K2.7-Code
|
- | - | |||
|
|
Moonshot AI | Kimi K2.7 Code |
moonshotai/Kimi-K2.7-Code
|
$0.71 | $3.50 | |||
|
|
Minimax | MiniMax M3 |
MiniMaxAI/MiniMax-M3
|
$0.23 | $0.96 | |||
|
|
Minimax | MiniMax M3 |
MiniMax-M3
|
- | - | |||
|
|
Nvidia | Nemotron 3 Ultra (550B A55B) |
NVIDIA-Nemotron-3-Ultra-550B-A55B
|
- | - |
|
||
|
|
Nvidia | Nemotron 3 Ultra (550B A55B) |
nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B
|
$0.50 | $2.15 |
|
||
|
|
DeepSeek | DeepSeek-V4-Pro-Max |
DeepSeek-V4-Pro-0813
|
- | - |
|
||
|
|
DeepSeek | DeepSeek-V4-Pro-Max |
deepseek-ai/DeepSeek-V4-Pro-0813
|
$1.31 | $3.96 |
|
||
|
|
DeepSeek | DeepSeek-V4-Flash-Max |
deepseek-ai/DeepSeek-V4-Flash-0731
|
$0.13 | $0.28 |
|
||
|
|
DeepSeek | DeepSeek-V4-Flash-Max |
DeepSeek-V4-Flash-0731
|
- | - |
|
||
|
|
Moonshot AI | Kimi K2.6 |
Kimi-K2.6
|
- | - | |||
|
|
Moonshot AI | Kimi K2.6 |
moonshotai/Kimi-K2.6
|
$0.65 | $3.41 | |||
|
|
qwen | Qwen3.6-35B-A3B |
Qwen3.6-35B-A3B
|
- | - | |||
|
|
qwen | Qwen3.6-35B-A3B |
Qwen/Qwen3.6-35B-A3B
|
$0.25 | $1.25 | |||
|
|
Gemma 4 31B |
gemma-4-31B-it
|
- | - |
|
|||
|
|
Gemma 4 31B |
google/gemma-4-31B-it
|
$0.10 | $0.34 |
|
|||
|
|
Gemma 4 26B-A4B |
gemma-4-26B-A4B-it
|
- | - |
|
|||
|
|
Gemma 4 26B-A4B |
google/gemma-4-26B-A4B-it
|
$0.10 | $0.30 |
|
|||
|
|
OpenAI | GPT OSS 120B |
openai/gpt-oss-120b
|
$0.03 | $0.17 |
|
||
|
|
OpenAI | GPT OSS 120B |
gpt-oss-120b
|
- | - |
|
||
|
|
OpenAI | GPT OSS 20B |
openai/gpt-oss-20b
|
$0.03 | $0.13 |
|
||
|
|
OpenAI | GPT OSS 20B |
gpt-oss-20b
|
- | - |
|
||
|
|
DeepSeek | DeepSeek-V3.1 |
deepseek-ai/DeepSeek-V3.1
|
$0.55 | $1.65 | |||
|
|
DeepSeek | DeepSeek-V3.1 |
DeepSeek-V3.1
|
- | - | |||
|
|
Meta | Llama 3.3 70B Instruct |
meta-llama/Llama-3.3-70B-Instruct
|
$0.71 | $0.71 | |||
|
|
Meta | Llama 3.3 70B Instruct |
Llama-3.3-70B-Instruct
|
- | - | |||
|
|
Meta | Llama 3.1 8B Instruct |
meta-llama/Llama-3.1-8B-Instruct
|
$0.22 | $0.22 | |||
|
|
Meta | Llama 3.1 8B Instruct |
Llama-3.1-8B-Instruct
|
- | - | |||
|
|
WandB | Nemotron 3.5 Lightning |
nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B
|
$0.07 | $0.20 | |||
|
|
WandB | Nemotron 3.5 Lightning |
NVIDIA-Nemotron-3.5-Lightning-30B-A3B
|
- | - |