Endpoint / venice-llama-3-3-70b

Llama 3 3 70b Instruct

served by Venice

Default workload estimate$3.501M input + 1M output tokens
Provider model ID
llama-3.3-70b
Protocol
openai chat completions
Region
global
Status
active
Base URL
https://api.venice.ai/api/v1

Current offers

Price components

Exact source values; no hidden currency conversion.

On demandon demand
$3.50
input tokens$0.70per million tokens
output tokens$2.80per million tokens

Provider assertions

Claims

capability · code optimized

unsupported · 1 source assertion

capability · reasoning

unsupported · 1 source assertion

compatibility · structured outputs

unsupported · 1 source assertion

compatibility · tools

supported · 1 source assertion

limits · context tokens

reported · 1 source assertion

privacy · e2ee

unsupported · 1 source assertion

privacy · mode

reported · 1 source assertion

privacy · tee attestation

unsupported · 1 source assertion

Independent checks

Observations

0

independent observations · probes begin in Stage 2

Immutable record

Pricing history

Newest observed values first.

ObservedOfferComponentPriceSnapshot
On demandinput tokens$0.70f24aab2471
On demandoutput tokens$2.80f24aab2471
On demandinput tokens$0.70f24aab2471
On demandoutput tokens$2.80f24aab2471
On demandinput tokens$0.70f24aab2471
On demandoutput tokens$2.80f24aab2471
On demandinput tokens$0.70f24aab2471
On demandoutput tokens$2.80f24aab2471