Weights & Biases

[PROVIDER]
id: wandb
npm: @ai-sdk/openai-compatible
env: WANDB_API_KEY
api: https://api.inference.wandb.ai/v1

Models

DeepSeek V3.1

deepseek-ai/DeepSeek-V3.1
in $0.55/M
out $1.65/M
cache read $0.55/M
ctx: 161,000 max out: 161,000 in: text out: text
reasoning tools vision structured temp open weights

DeepSeek V4 Flash

deepseek-ai/DeepSeek-V4-Flash
in $0.14/M
out $0.28/M
cache read $0.07/M
ctx: 1,048,576 max out: 1,048,576 in: text out: text
reasoning tools vision structured temp open weights

DeepSeek V4 Flash 0731

deepseek-ai/DeepSeek-V4-Flash-0731
in $0.13/M
out $0.28/M
cache read $0.07/M
ctx: 262,144 max out: 262,144 in: text out: text
reasoning tools vision structured temp open weights

DeepSeek V4 Pro

deepseek-ai/DeepSeek-V4-Pro
in $1.15/M
out $2.55/M
cache read $0.20/M
ctx: 1,048,576 max out: 1,048,576 in: text out: text
reasoning tools vision structured temp open weights

Gemma 4 31B

google/gemma-4-31B-it
in $0.10/M
out $0.34/M
cache read $0.10/M
ctx: 262,144 max out: 262,144 in: text, image out: text
reasoning tools vision structured temp open weights

GLM 5.1

zai-org/GLM-5.1
in $1.40/M
out $4.40/M
cache read $0.26/M
ctx: 202,752 max out: 202,752 in: text out: text
reasoning tools vision structured temp open weights

GLM 5.2

zai-org/GLM-5.2
in $0.76/M
out $2.42/M
cache read $0.14/M
ctx: 262,144 max out: 262,144 in: text out: text
reasoning tools vision structured temp open weights

gpt-oss-120b

openai/gpt-oss-120b
in $0.03/M
out $0.17/M
cache read $0.03/M
ctx: 131,072 max out: 131,072 in: text out: text
reasoning tools vision structured temp open weights

gpt-oss-20b

openai/gpt-oss-20b
in $0.03/M
out $0.13/M
cache read $0.03/M
ctx: 131,072 max out: 131,072 in: text out: text
reasoning tools vision structured temp open weights

Granite 4.1 8B

ibm-granite/granite-4.1-8b
in $0.05/M
out $0.10/M
cache read $0.05/M
ctx: 131,072 max out: 131,072 in: text out: text
reasoning tools vision structured temp open weights

Kimi K2.6

moonshotai/Kimi-K2.6
in $0.65/M
out $3.41/M
cache read $0.15/M
ctx: 262,144 max out: 262,144 in: text, image out: text
reasoning tools vision structured temp open weights

Kimi K2.7 Code

moonshotai/Kimi-K2.7-Code
in $0.71/M
out $3.50/M
cache read $0.15/M
ctx: 262,144 max out: 262,144 in: text, image out: text
reasoning tools vision structured temp open weights

Kimi K3

moonshotai/Kimi-K3
in $3.00/M
out $15.00/M
cache read $0.30/M
ctx: 1,048,576 max out: 1,048,576 in: text, image out: text
reasoning tools vision structured temp open weights

Llama 3.1 70B

meta-llama/Llama-3.1-70B-Instruct
in $0.80/M
out $0.80/M
cache read $0.80/M
ctx: 128,000 max out: 128,000 in: text out: text
reasoning tools vision structured temp open weights

Llama 3.1 8B

meta-llama/Llama-3.1-8B-Instruct
in $0.22/M
out $0.22/M
cache read $0.22/M
ctx: 128,000 max out: 128,000 in: text out: text
reasoning tools vision structured temp open weights

Llama 3.3 70B

meta-llama/Llama-3.3-70B-Instruct
in $0.71/M
out $0.71/M
cache read $0.71/M
ctx: 128,000 max out: 128,000 in: text out: text
reasoning tools vision structured temp open weights

Mellum2 12B A2.5B

JetBrains/Mellum2-12B-A2.5B-Instruct
in $0.05/M
out $0.10/M
cache read $0.05/M
ctx: 131,072 max out: 131,072 in: text out: text
reasoning tools vision structured temp open weights

MiniMax M2.5

MiniMaxAI/MiniMax-M2.5
in $0.30/M
out $1.20/M
cache read $0.30/M
ctx: 196,608 max out: 196,608 in: text out: text
reasoning tools vision structured temp open weights

MiniMax M3

MiniMaxAI/MiniMax-M3
in $0.23/M
out $0.96/M
cache read $0.05/M
ctx: 262,144 max out: 262,144 in: text, image out: text
reasoning tools vision structured temp open weights

Nemotron 3 Super

nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-FP8
in $0.20/M
out $0.80/M
cache read $0.20/M
ctx: 262,144 max out: 262,144 in: text out: text
reasoning tools vision structured temp open weights

Nemotron 3 Ultra

nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B
in $0.75/M
out $2.75/M
cache read $0.15/M
ctx: 262,144 max out: 262,144 in: text out: text
reasoning tools vision structured temp open weights

Nemotron 3.5 Lightning

nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B
in $0.10/M
out $0.25/M
cache read $0.05/M
ctx: 262,144 max out: 262,144 in: text out: text
reasoning tools vision structured temp open weights

Qwen3 14B Instruct

OpenPipe/Qwen3-14B-Instruct
in $0.05/M
out $0.22/M
cache read $0.05/M
ctx: 32,768 max out: 32,768 in: text out: text
reasoning tools vision structured temp open weights

Qwen3 30B A3B Instruct 2507

Qwen/Qwen3-30B-A3B-Instruct-2507
in $0.10/M
out $0.30/M
cache read $0.10/M
ctx: 262,144 max out: 262,144 in: text out: text
reasoning tools vision structured temp open weights

Qwen3 Coder 480B A35B

Qwen/Qwen3-Coder-480B-A35B-Instruct
in $1.00/M
out $1.50/M
cache read $1.00/M
ctx: 262,144 max out: 262,144 in: text out: text
reasoning tools vision structured temp open weights

Qwen3.5-35B-A3B

Qwen/Qwen3.5-35B-A3B
in $0.25/M
out $1.25/M
cache read $0.25/M
ctx: 262,144 max out: 262,144 in: text, image out: text
reasoning tools vision structured temp open weights

Qwen3.6 27B

Qwen/Qwen3.6-27B
in $0.60/M
out $3.60/M
cache read $0.12/M
ctx: 262,144 max out: 262,144 in: text, image out: text
reasoning tools vision structured temp open weights

Qwen3.6 35B A3B

Qwen/Qwen3.6-35B-A3B
in $0.25/M
out $1.25/M
cache read $0.25/M
ctx: 262,144 max out: 262,144 in: text, image out: text
reasoning tools vision structured temp open weights

Qwen3.8 27B

Qwen/Qwen3.8-27B
in $0.40/M
out $3.00/M
cache read $0.15/M
ctx: 262,144 max out: 262,144 in: text, image out: text
reasoning tools vision structured temp open weights