Endpoint / deepinfra-meta-llama-llama-3-3-70b-instruct-turbo

Llama 3 3 70b Instruct Turbo

served by DeepInfra

Default workload estimate$0.421M input + 1M output tokens
Provider model ID
meta-llama/Llama-3.3-70B-Instruct-Turbo
Protocol
openai chat completions
Region
global
Status
active
Base URL
https://api.deepinfra.com/v1/openai

Current offers

Price components

Exact source values; no hidden currency conversion.

Standard on demandon demand
$0.42
input tokens$0.10per million tokens
output tokens$0.32per million tokens

Provider assertions

Claims

capability · reasoning

unsupported · 1 source assertion

compatibility · structured outputs

supported · 1 source assertion

compatibility · tools

supported · 1 source assertion

limits · context tokens

reported · 1 source assertion

runtime · quantization

reported · 1 source assertion

Independent checks

Observations

0

independent observations · probes begin in Stage 2

Immutable record

Pricing history

Newest observed values first.

ObservedOfferComponentPriceSnapshot
Standard on demandinput tokens$0.10bb193c809d
Standard on demandoutput tokens$0.32bb193c809d
Standard on demandinput tokens$0.1018b5cfc8cd
Standard on demandoutput tokens$0.3218b5cfc8cd
Standard on demandinput tokens$0.106a13d395e4
Standard on demandoutput tokens$0.326a13d395e4
Standard on demandinput tokens$0.10a0c2739383
Standard on demandoutput tokens$0.32a0c2739383