Weights & Biases icon

Weights & Biases

wandb

Updated 51 minutes ago

Weights & Biases (W&B) is an AI/ML platform that provides model training tracking, experiment management, and inference services. Their W&B Inference API offers an OpenAI-compatible interface to access curated open-source language models including DeepSeek, Qwen, Meta Llama, Google Gemma, MiniMax, Moonshot AI (Kimi), NVIDIA Nemotron, Microsoft Phi, and others. The platform is known for its MLOps tooling and has expanded into hosted inference with competitive pricing.

Browse 56 LLM models available from Weights & Biases. Compare prices and features.

Models (56)

Organization Model Name Original Model Input Output Free
Moonshot AI
Moonshot AI Kimi K3 moonshotai/Kimi-K3 $3.00 $15.00
qwen
qwen Qwen3.8-27B Qwen/Qwen3.8-27B $0.40 $3.00
qwen
qwen Qwen3.8-27B Qwen3.8-27B - -
Z.ai
Z.ai GLM-5.2 GLM-5.2 - -
Z.ai
Z.ai GLM-5.2 zai-org/GLM-5.2 $0.76 $2.42
Moonshot AI
Moonshot AI Kimi K2.7 Code Kimi-K2.7-Code - -
Moonshot AI
Moonshot AI Kimi K2.7 Code moonshotai/Kimi-K2.7-Code $0.71 $3.50
Minimax
Minimax MiniMax M3 MiniMaxAI/MiniMax-M3 $0.23 $0.96
Minimax
Minimax MiniMax M3 MiniMax-M3 - -
Nvidia
Nvidia Nemotron 3 Ultra 550B A55B NVIDIA-Nemotron-3-Ultra-550B-A55B - -
Nvidia
Nvidia Nemotron 3 Ultra 550B A55B nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B $0.75 $2.75
DeepSeek
DeepSeek DeepSeek V4 Pro DeepSeek-V4-Pro - -
DeepSeek
DeepSeek DeepSeek V4 Pro deepseek-ai/DeepSeek-V4-Pro $1.15 $2.55
DeepSeek
DeepSeek DeepSeek V4 Flash DeepSeek-V4-Flash - -
DeepSeek
DeepSeek DeepSeek V4 Flash deepseek-ai/DeepSeek-V4-Flash $0.14 $0.28
DeepSeek
DeepSeek DeepSeek V4 Flash deepseek-ai/DeepSeek-V4-Flash-0731 $0.13 $0.28
DeepSeek
DeepSeek DeepSeek V4 Flash DeepSeek-V4-Flash-0731 - -
Moonshot AI
Moonshot AI Kimi K2.6 Kimi-K2.6 - -
Moonshot AI
Moonshot AI Kimi K2.6 moonshotai/Kimi-K2.6 $0.65 $3.41
Alibaba
Alibaba Qwen3.6 27B Qwen3.6-27B - -
Alibaba
Alibaba Qwen3.6 27B Qwen/Qwen3.6-27B $0.60 $3.60
Z.ai
Z.ai GLM-5.1 GLM-5.1 - -
qwen
qwen Qwen3.6 35B A3B Qwen3.6-35B-A3B - -
qwen
qwen Qwen3.6 35B A3B Qwen/Qwen3.6-35B-A3B $0.25 $1.25
Minimax
Minimax MiniMax M2.5 MiniMax-M2.5 - -
google
google Gemma 4 31B gemma-4-31B-it - -
google
google Gemma 4 31B google/gemma-4-31B-it $0.10 $0.34
qwen
qwen Qwen3.5-35B-A3B Qwen3.5-35B-A3B - -
qwen
qwen Qwen3.5-35B-A3B Qwen/Qwen3.5-35B-A3B $0.25 $1.25
OpenAI
OpenAI GPT OSS 120B openai/gpt-oss-120b $0.03 $0.17
OpenAI
OpenAI GPT OSS 120B gpt-oss-120b - -
qwen
qwen Qwen3-Coder 480B A35B Instruct Qwen3-Coder-480B-A35B-Instruct - -
OpenAI
OpenAI GPT OSS 20B openai/gpt-oss-20b $0.03 $0.13
OpenAI
OpenAI GPT OSS 20B gpt-oss-20b - -
qwen
qwen Qwen3-235B-A22B-Instruct-2507 Qwen3-235B-A22B-Instruct-2507 - -
DeepSeek
DeepSeek DeepSeek-V3.1 deepseek-ai/DeepSeek-V3.1 $0.55 $1.65
DeepSeek
DeepSeek DeepSeek-V3.1 DeepSeek-V3.1 - -
Meta
Meta Llama 3.1 70B Instruct meta-llama/Llama-3.1-70B-Instruct $0.80 $0.80
Meta
Meta Llama 3.1 70B Instruct Llama-3.1-70B-Instruct - -
Meta
Meta Llama 3.1 8B Instruct meta-llama/Llama-3.1-8B-Instruct $0.22 $0.22
Meta
Meta Llama 3.1 8B Instruct Llama-3.1-8B-Instruct - -
Meta
Meta Llama 3.3 70B Instruct meta-llama/Llama-3.3-70B-Instruct $0.71 $0.71
Meta
Meta Llama 3.3 70B Instruct Llama-3.3-70B-Instruct - -
WandB
WandB Granite 4.2 8B ibm-granite/granite-4.2-8b $0.10 $0.15
WandB
WandB Granite 4.2 8B granite-4.2-8b - -
WandB
WandB Nemotron 3.5 Lightning nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B $0.10 $0.25
WandB
WandB Nemotron 3.5 Lightning NVIDIA-Nemotron-3.5-Lightning-30B-A3B - -
WandB
WandB Mellum2 12B A2.5B JetBrains/Mellum2-12B-A2.5B-Instruct $0.05 $0.10
WandB
WandB Mellum2 12B A2.5B Mellum2-12B-A2.5B-Instruct - -
IBM Granite 4.1 8B granite-4.1-8b - -
IBM Granite 4.1 8B ibm-granite/granite-4.1-8b $0.05 $0.10
WandB
WandB NVIDIA Nemotron 3 Super 120B NVIDIA-Nemotron-3-Super-120B-A12B-FP8 - -
Alibaba
Alibaba qwen3-30b-a3b-instruct-2507 Qwen/Qwen3-30B-A3B-Instruct-2507 $0.10 $0.30
Alibaba
Alibaba qwen3-30b-a3b-instruct-2507 Qwen3-30B-A3B-Instruct-2507 - -
WandB
WandB OpenPipe Qwen3 14B Instruct OpenPipe/Qwen3-14B-Instruct $0.05 $0.22
WandB
WandB OpenPipe Qwen3 14B Instruct Qwen3-14B-Instruct - -