venice No lender online

nvidia-nemotron-3-ultra-550b-a55b on Venice.

Venice serves nvidia-nemotron-3-ultra-550b-a55b at $0.625 per million input tokens and $3.13 per million output. Context window: 256K tokens. No lender is online for it at the moment.

Input /Mtok

$0.625

Output /Mtok

$3.13

Context

256K

Lenders online

0

01Overview

Venice serves nvidia-nemotron-3-ultra-550b-a55b at $0.625 per million input tokens and $3.13 per million output. Context window: 256K tokens. No lender is online for it at the moment.

02Pricing

Per million tokens, in USD. This is the rate you are charged.
Input$0.625
Output$3.13
Cache read$0.188

03Capabilities

This model can reason at length before it answers.This model can call tools and functions you define.

Not supported: image input.

Context window256K tokens
Maximum output33K tokens

Anything not listed here, we don't publish for this model yet.

05Usage

This model has not been served through Aile yet.

06Call it

Any OpenAI- or Anthropic-compatible SDK works. Use this exact model id:

curlTerminal
curl https://aile.sh/v1/chat/completions \
  -H "Authorization: Bearer $AILE_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "venice/nvidia-nemotron-3-ultra-550b-a55b",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

Quickstart · API reference

07Alternatives

Other providers serve nvidia-nemotron-3-ultra-550b-a55b too, at their own rates — compare every provider for nvidia-nemotron-3-ultra-550b-a55b.

Other models from Venice:

ModelProviderInput /MtokOutput /MtokStatus
deepseek-v4-flashvenice$0.138$0.275Offline
claude-opus-5-5venice$4.80$24.00Offline
deepseek-v4-provenice$1.65$3.30Offline
claude-sonnet-5venice$3.00$15.00Offline
claude-sonnet-4-6venice$3.60$18.00Offline
claude-sonnet-5-5venice$3.75$18.75Offline
claude-opus-4-8venice$6.00$30.00Offline
claude-opus-5venice$6.00$30.00Offline
claude-opus-4-7venice$6.00$30.00Offline
venice-latestvenice——Offline
gemini-3-6-flashvenice$0.938$4.69Offline
gemini-3-7-flashvenice$0.938$4.69Offline

Similar published rates elsewhere:

ModelProviderInput /MtokOutput /MtokStatus
deepseek-v4-flashdeepseek$0.300$1.20Offline
deepseek-v4-flashinference-net$0.230$0.620Offline
deepseek-v4-flashllmgateway$0.440$1.32Offline
deepseek-v4-flashopencode-go$0.300$1.20Online
deepseek/deepseek-v4-flashorcarouter$0.220$0.660Offline
deepseek-v4-flashpoe$0.444$1.33Offline