All AI models · LLM benchmarks · Methodology

Alibaba

Qwen3 235B-A22B

Qwen instruction model for multilingual chat, reasoning, and tool use

Model facts

Context window
131K tokens
Maximum output
16K tokens
Cheapest paid input
$0.13 per 1M tokens
Cheapest paid output
$0.6 per 1M tokens
Open weights
yes
Providers
16

Capabilities and modalities

reasoning, tool calling, structured output, attachments, temperature control, open weights, text, pdf.

Local hardware estimate

At 4K context: Q4_K_M 127.0 GiB, Q8_0 236.5 GiB, F16 441.7 GiB working memory.

Parameters
235.1 billion
Architecture
qwen3moe
Model license
apache-2.0

Planning estimate adapted from the MIT-licensed whichllm estimator using metadata from Hugging Face; not a vendor minimum requirement.

Artificial Analysis benchmarks

Compare this model on the full LLM benchmark leaderboard.

ConfigurationAAICodingMathOutput tokens/s
Reasoning13.5not measured82.0not measured
Non-reasoning10.8not measured23.7not measured

Observed price history

  • 2026-08-24: $0.13 input / $0.6 output per 1M tokens

API providers

  • iFlow — model id qwen3-235b; free input / free output per 1M tokens; provider documentation
  • iFlow — model id qwen3-235b-a22b-instruct; free input / free output per 1M tokens; provider documentation
  • Abacus — model id Qwen/Qwen3-235B-A22B-Instruct-2507; $0.13 input / $0.6 output per 1M tokens; provider documentation
  • NanoGPT — model id qwen/qwen3-235b-a22b; $0.3 input / $0.5 output per 1M tokens; provider documentation
  • Hugging Face — model id Qwen/Qwen3-235B-A22B; $0.2 input / $0.8 output per 1M tokens; provider documentation
  • Jiekou.AI — model id qwen/qwen3-235b-a22b-fp8; $0.2 input / $0.8 output per 1M tokens; provider documentation
  • NovitaAI — model id qwen/qwen3-235b-a22b-fp8; $0.2 input / $0.8 output per 1M tokens; provider documentation
  • Vertex — model id qwen/qwen3-235b-a22b-instruct-2507-maas; $0.22 input / $0.88 output per 1M tokens; provider documentation
  • Alibaba (China) — model id qwen3-235b-a22b; $0.287 input / $1.15 output per 1M tokens; provider documentation
  • Merge Gateway — model id qwen/qwen3-235b-a22b; $0.287 input / $1.15 output per 1M tokens; provider documentation
  • Kilo Gateway — model id qwen/qwen3-235b-a22b; $0.455 input / $1.82 output per 1M tokens; provider documentation
  • OpenRouter — model id qwen/qwen3-235b-a22b; $0.455 input / $1.82 output per 1M tokens; provider documentation
  • 302.AI — model id qwen3-235b-a22b; $0.29 input / $2.86 output per 1M tokens; provider documentation
  • Alibaba — model id qwen3-235b-a22b; $0.7 input / $2.8 output per 1M tokens; provider documentation
  • Arena — model id qwen3-235b-a22b; not published input / not published output per 1M tokens; provider documentation
  • Qiniu — model id qwen3-235b-a22b; not published input / not published output per 1M tokens; provider documentation