Endpoint / venice-e2ee-deepseek-v4-flash

DeepSeek V4 Flash

served by Venice

Default workload estimate$0.561M input + 1M output tokens
Provider model ID
e2ee-deepseek-v4-flash
Protocol
openai chat completions
Region
global
Status
active
Base URL
https://api.venice.ai/api/v1

Current offers

Price components

Exact source values; no hidden currency conversion.

On demandon demand
$0.56
input tokens$0.18per million tokens
output tokens$0.37per million tokens
cached input tokens$0.04per million tokens

Provider assertions

Claims

capability · code optimized

supported · 1 source assertion

capability · reasoning

supported · 1 source assertion

compatibility · structured outputs

unsupported · 1 source assertion

compatibility · tools

supported · 1 source assertion

limits · context tokens

reported · 1 source assertion

privacy · e2ee

supported · 1 source assertion

privacy · mode

reported · 1 source assertion

privacy · tee attestation

supported · 1 source assertion

Independent checks

Observations

0

independent observations · probes begin in Stage 2

Immutable record

Pricing history

Newest observed values first.

ObservedOfferComponentPriceSnapshot
On demandinput tokens$0.18f24aab2471
On demandoutput tokens$0.37f24aab2471
On demandcached input tokens$0.04f24aab2471
On demandinput tokens$0.18f24aab2471
On demandoutput tokens$0.37f24aab2471
On demandcached input tokens$0.04f24aab2471
On demandinput tokens$0.18f24aab2471
On demandoutput tokens$0.37f24aab2471
On demandcached input tokens$0.04f24aab2471
On demandinput tokens$0.18f24aab2471
On demandoutput tokens$0.37f24aab2471
On demandcached input tokens$0.04f24aab2471