Llama 3.3 70B Instruct Turbo API
Pricing & Specs

Meta · $1.04 input / $1.04 output per 1M tokens · 131.072K context window

$1.04
per 1M input tokens
$1.04
per 1M output tokens
131.072K
context window (tokens)

The per-token price never tells the full story. A typical task (1K input + 500 output tokens) on Llama 3.3 70B Instruct Turbo API costs about $0.0016 — roughly $1.56 per 1,000 runs. But if it needs 3x the tokens of a cheaper model to match quality on YOUR task, the economics flip. The only way to know is to benchmark it on your actual workload.

Llama 3.3 70B Instruct Turbo API API Pricing

TokensPrice
Input$1.04 / 1M tokens
Output$1.04 / 1M tokens

50% discount available for batch API requests. Batch API pricing is available via Together at 50%; cache pricing is not published.

What that means in practice

1x One typical task (1K input + 500 output tokens): $0.0016
1K 1,000 runs of that task: $1.56
💡 Real cost depends on how verbose the model is on YOUR prompts — benchmark to measure actual tokens, not estimates.

Llama 3.3 70B Instruct Turbo API Specs

SpecValue
Context window131.072K tokens
Input modalitiestext
Output modalitiestext
Latency (measured by OpenMark)~1.3s median response
Reasoning modelNo
Tool / function callingYes
JSON modeYes
StreamingYes
Prompt cachingNo
Batch APIYes

Capabilities

chat multilingual streaming structured outputs function calling

FAQ

How much does Llama 3.3 70B Instruct Turbo API cost?

Llama 3.3 70B Instruct Turbo API costs $1.04 per 1M input tokens and $1.04 per 1M output tokens.

What is Llama 3.3 70B Instruct Turbo API's context window?

Llama 3.3 70B Instruct Turbo API has a 131.072K-token context window.

Can I test Llama 3.3 70B Instruct Turbo API on my own task?

Yes. OpenMark lets you benchmark Llama 3.3 70B Instruct Turbo API against 100+ models on your own task with real API calls — no API keys needed, free tier available.

Is Llama 3.3 70B Instruct Turbo API the right model for YOUR task?

Pricing tables can't answer that. Benchmark it against 100+ models
on your actual workload — real API calls, real costs. Free tier available.

Benchmark Llama 3.3 70B Instruct Turbo API — Free →