Qwen3.5 9B FP8
Pricing & Specs

Qwen (Alibaba) · $0.17 input / $0.25 output per 1M tokens · 262.144K context window

$0.17
per 1M input tokens
$0.25
per 1M output tokens
262.144K
context window (tokens)

The per-token price never tells the full story. A typical task (1K input + 500 output tokens) on Qwen3.5 9B FP8 costs about $0.0003 — roughly $0.30 per 1,000 runs. But if it needs 3x the tokens of a cheaper model to match quality on YOUR task, the economics flip. The only way to know is to benchmark it on your actual workload.

Qwen3.5 9B FP8 API Pricing

TokensPrice
Input$0.17 / 1M tokens
Output$0.25 / 1M tokens

50% discount available for batch API requests.

What that means in practice

1x One typical task (1K input + 500 output tokens): $0.0003
1K 1,000 runs of that task: $0.30
💡 Real cost depends on how verbose the model is on YOUR prompts — benchmark to measure actual tokens, not estimates.

Qwen3.5 9B FP8 Specs

SpecValue
Context window262.144K tokens
Input modalitiestext, image
Output modalitiestext
Reasoning modelYes
Tool / function callingYes
JSON modeYes
StreamingYes
Batch APIYes

FAQ

How much does Qwen3.5 9B FP8 cost?

Qwen3.5 9B FP8 costs $0.17 per 1M input tokens and $0.25 per 1M output tokens.

What is Qwen3.5 9B FP8's context window?

Qwen3.5 9B FP8 has a 262.144K-token context window.

Can I test Qwen3.5 9B FP8 on my own task?

Yes. OpenMark lets you benchmark Qwen3.5 9B FP8 against 100+ models on your own task with real API calls — no API keys needed, free tier available.

Is Qwen3.5 9B FP8 the right model for YOUR task?

Pricing tables can't answer that. Benchmark it against 100+ models
on your actual workload — real API calls, real costs. Free tier available.

Benchmark Qwen3.5 9B FP8 — Free →