Endpoint / venice-e2ee-qwen3-6-27b

Qwen3 6 27b fp8

served by Venice

Default workload estimate$3.811M input + 1M output tokens
Provider model ID
e2ee-qwen3-6-27b
Protocol
openai chat completions
Region
global
Status
active
Base URL
https://api.venice.ai/api/v1

Current offers

Price components

Exact source values; no hidden currency conversion.

On demandon demand
$3.81
input tokens$0.35per million tokens
output tokens$3.46per million tokens
cached input tokens$0.17per million tokens

Provider assertions

Claims

capability · code optimized

supported · 1 source assertion

capability · reasoning

supported · 1 source assertion

compatibility · structured outputs

unsupported · 1 source assertion

compatibility · tools

supported · 1 source assertion

limits · context tokens

reported · 1 source assertion

privacy · e2ee

supported · 1 source assertion

privacy · mode

reported · 1 source assertion

privacy · tee attestation

supported · 1 source assertion

Independent checks

Observations

0

independent observations · probes begin in Stage 2

Immutable record

Pricing history

Newest observed values first.

ObservedOfferComponentPriceSnapshot
On demandinput tokens$0.35f24aab2471
On demandoutput tokens$3.46f24aab2471
On demandcached input tokens$0.17f24aab2471
On demandinput tokens$0.35f24aab2471
On demandoutput tokens$3.46f24aab2471
On demandcached input tokens$0.17f24aab2471
On demandinput tokens$0.35f24aab2471
On demandoutput tokens$3.46f24aab2471
On demandcached input tokens$0.17f24aab2471
On demandinput tokens$0.35f24aab2471
On demandoutput tokens$3.46f24aab2471
On demandcached input tokens$0.17f24aab2471