wandb No lender online

nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B on Wandb.

Wandb serves nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B at $0.500 per million input tokens and $2.15 per million output. Context window: 262K tokens. No lender is online for it at the moment.

Input /Mtok

$0.500

Output /Mtok

$2.15

Context

262K

Lenders online

0

01Overview

Wandb serves nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B at $0.500 per million input tokens and $2.15 per million output. Context window: 262K tokens. No lender is online for it at the moment.

02Pricing

Per million tokens, in USD. This is the rate you are charged.
Input$0.500
Output$2.15
Cache read$0.100

03Capabilities

This model can reason at length before it answers.This model can call tools and functions you define.

Not supported: image input.

Context window262K tokens
Maximum output262K tokens

Anything not listed here, we don't publish for this model yet.

05Usage

This model has not been served through Aile yet.

06Call it

Any OpenAI- or Anthropic-compatible SDK works. Use this exact model id:

curlTerminal
curl https://aile.sh/v1/chat/completions \
  -H "Authorization: Bearer $AILE_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "wandb/nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

Quickstart · API reference

07Alternatives

Other providers serve nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B too, at their own rates — compare every provider for nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B.

Other models from Wandb:

ModelProviderInput /MtokOutput /MtokStatus
zai-org/GLM-5.3-Flashwandb$0.150$0.500Offline
zai-org/GLM-5.2wandb$0.760$2.42Offline
moonshotai/Kimi-K2.7-Codewandb$0.710$3.50Offline
openai/gpt-oss-120bwandb$0.0300$0.170Offline
Qwen/Qwen3-Coder-480B-A35B-Instructwandb——Offline
deepseek-ai/DeepSeek-V3.1wandb$0.550$1.65Offline
Qwen/Qwen3.8-27Bwandb$0.400$3.00Offline
deepseek-ai/DeepSeek-V4-Pro-0813wandb$1.31$3.96Offline
deepseek-ai/DeepSeek-V4-Flash-0731wandb$0.130$0.280Offline
MiniMaxAI/MiniMax-M3wandb$0.230$0.960Offline
nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3Bwandb$0.0700$0.200Offline
ibm-granite/granite-4.2-8bwandb$0.100$0.150Offline

Similar published rates elsewhere:

ModelProviderInput /MtokOutput /MtokStatus
deepseek-v4-flashbazaarlink$0.200$0.400Offline
deepseek-v4-flashcharm-hyper$0.200$0.400Offline
deepseek-v4-flashdeepseek$0.300$1.20Offline
deepseek-v4-flashinference-net$0.230$0.620Offline
deepseek-v4-flashllmgateway$0.440$1.32Offline
deepseek-v4-flashopencode-go$0.300$1.20Online