Endpoint / venice-e2ee-qwen3-6-35b-a3b

Qwen3 6 35b A3b fp8

served by Venice

Default workload estimate$1.361M input + 1M output tokens
Provider model ID
e2ee-qwen3-6-35b-a3b
Protocol
openai chat completions
Region
global
Status
active
Base URL
https://api.venice.ai/api/v1

Current offers

Price components

Exact source values; no hidden currency conversion.

On demandon demand
$1.36
input tokens$0.18per million tokens
output tokens$1.18per million tokens
cached input tokens$0.06per million tokens

Provider assertions

Claims

capability · code optimized

supported · 1 source assertion

capability · reasoning

supported · 1 source assertion

compatibility · structured outputs

unsupported · 1 source assertion

compatibility · tools

supported · 1 source assertion

limits · context tokens

reported · 1 source assertion

privacy · e2ee

supported · 1 source assertion

privacy · mode

reported · 1 source assertion

privacy · tee attestation

supported · 1 source assertion

Independent checks

Observations

0

independent observations · probes begin in Stage 2

Immutable record

Pricing history

Newest observed values first.

ObservedOfferComponentPriceSnapshot
On demandinput tokens$0.18f24aab2471
On demandoutput tokens$1.18f24aab2471
On demandcached input tokens$0.06f24aab2471
On demandinput tokens$0.18f24aab2471
On demandoutput tokens$1.18f24aab2471
On demandcached input tokens$0.06f24aab2471
On demandinput tokens$0.18f24aab2471
On demandoutput tokens$1.18f24aab2471
On demandcached input tokens$0.06f24aab2471
On demandinput tokens$0.18f24aab2471
On demandoutput tokens$1.18f24aab2471
On demandcached input tokens$0.06f24aab2471