CheapestCoding Long-contextBest value Get an API key
๐ŸŽฏ Don't want to pick? Send model: "appmo/auto" and Appmo routes each request to the cheapest model that's good enough โ€” cutting your bill up to 40%. See how โ†’
Models โ€บ qwen โ€บ Qwen: Qwen3.8 2.4T A95B

Qwen: Qwen3.8 2.4T A95B โ€” pricing & benchmarks

By qwen ยท Updated 2026-07-24 ยท long-context / RAGstep-by-step reasoning

Qwen: Qwen3.8 2.4T A95B costs $2 per 1M input tokens and $6 per 1M output tokens, with a 1M context window. It is best suited for long-context / RAG, step-by-step reasoning, and is available through one OpenAI-compatible Appmo API key.
$2Input / 1M
$6Output / 1M
1MContext
โ€”Intelligence
โ€”Coding
โ€”GPQA %
โ€”Speed tok/s

What Qwen: Qwen3.8 2.4T A95B is used for

Qwen: Qwen3.8 2.4T A95B is best suited for long-context / RAG, step-by-step reasoning. Access it through one Appmo API key alongside 390 other models โ€” Appmo's optimizer can route to it automatically when it is the best value for a request.

Use it via the Appmo API (OpenAI-compatible)

curl https://appmo.com/v1/chat/completions \
  -H "Authorization: Bearer $APPMO_KEY" \
  -d '{"model":"qwen/qwen3.8-2.4t-a95b","messages":[{"role":"user","content":"Hello"}]}'

FAQ

How much does Qwen: Qwen3.8 2.4T A95B cost?
Qwen: Qwen3.8 2.4T A95B costs $2 per 1M input tokens and $6 per 1M output tokens ($3 blended), as of 2026-07-24. Prices are expected to fall as competition increases.

What is Qwen: Qwen3.8 2.4T A95B best for?
Qwen: Qwen3.8 2.4T A95B is best suited for long-context / RAG, step-by-step reasoning.

How do I use Qwen: Qwen3.8 2.4T A95B?
Send requests to Appmo's OpenAI-compatible endpoint with model id "qwen/qwen3.8-2.4t-a95b" and one API key โ€” no separate provider account needed.