Endpoint / venice-nvidia-nemotron-3-ultra-550b-a55b

Nvidia Nemotron 3 Ultra 550b A55b

served by Venice

Default workload estimate$3.751M input + 1M output tokens
Provider model ID
nvidia-nemotron-3-ultra-550b-a55b
Protocol
openai chat completions
Region
global
Status
active
Base URL
https://api.venice.ai/api/v1

Current offers

Price components

Exact source values; no hidden currency conversion.

On demandon demand
$3.75
input tokens$0.63per million tokens
output tokens$3.13per million tokens
cached input tokens$0.19per million tokens

Provider assertions

Claims

capability · code optimized

unsupported · 1 source assertion

capability · reasoning

supported · 1 source assertion

compatibility · structured outputs

supported · 1 source assertion

compatibility · tools

supported · 1 source assertion

limits · context tokens

reported · 1 source assertion

privacy · e2ee

unsupported · 1 source assertion

privacy · mode

reported · 1 source assertion

privacy · tee attestation

unsupported · 1 source assertion

Independent checks

Observations

0

independent observations · probes begin in Stage 2

Immutable record

Pricing history

Newest observed values first.

ObservedOfferComponentPriceSnapshot
On demandinput tokens$0.63f24aab2471
On demandoutput tokens$3.13f24aab2471
On demandcached input tokens$0.19f24aab2471
On demandinput tokens$0.63f24aab2471
On demandoutput tokens$3.13f24aab2471
On demandcached input tokens$0.19f24aab2471
On demandinput tokens$0.63f24aab2471
On demandoutput tokens$3.13f24aab2471
On demandcached input tokens$0.19f24aab2471
On demandinput tokens$0.63f24aab2471
On demandoutput tokens$3.13f24aab2471
On demandcached input tokens$0.19f24aab2471