LLM List: AI models compared
Browse all LLMs in one comprehensive list: 1,781 large language models and AI models across 214 API providers. Compare token prices, context windows, capabilities, benchmark scores, open-weight availability, and local hardware requirements. Updated every six hours.
Explore the LLM benchmark leaderboard · How prices, benchmarks and model identities are calculated
- qwen/qwen3-coder-next by Alibaba — 262K context; $0.2 input / $1.5 output per 1M tokens
- Qwen/Qwen3-Next-80B-A3B-Instruct — 262K context; $0.14 input / $1.4 output per 1M tokens
- Qwen/Qwen3-VL-235B-A22B-Instruct by Alibaba — 262K context; $0.3 input / $1.5 output per 1M tokens
- Qwen/Qwen3-VL-235B-A22B-Thinking by Alibaba — 262K context; $0.45 input / $3.5 output per 1M tokens
- Qwen/Qwen3-VL-30B-A3B-Instruct by Alibaba — 262K context; $0.13 input / $0.52 output per 1M tokens
- Qwen/Qwen3-VL-30B-A3B-Thinking by Alibaba — 262K context; $0.2 input / $1 output per 1M tokens
- Qwen/Qwen3-VL-32B-Instruct by Alibaba — 262K context; $0.104 input / $0.416 output per 1M tokens
- Qwen/Qwen3-VL-32B-Thinking by Alibaba — 262K context; $0.2 input / $1.5 output per 1M tokens
- Qwen/Qwen3-VL-8B-Instruct by Alibaba — 262K context; $0.117 input / $0.455 output per 1M tokens
- Qwen/Qwen3.5-122B-A10B by Alibaba — 262K context; $0.29 input / $2.32 output per 1M tokens
- Qwen/Qwen3.5-27B by Alibaba — 262K context; $0.26 input / $2.09 output per 1M tokens
- Qwen/Qwen3.5-35B-A3B by Alibaba — 262K context; $0.23 input / $1.86 output per 1M tokens
- Qwen/Qwen3.5-397B-A17B by Alibaba — 262K context; $0.29 input / $1.74 output per 1M tokens
- Qwen/Qwen3.5-4B by Alibaba — 262K context; free input / free output per 1M tokens
- Qwen/Qwen3.5-9B by Alibaba — 262K context; $0.1 input / $0.15 output per 1M tokens
- Qwen/Qwen3.6-35B-A3B by Alibaba — 262K context; $0.23 input / $1.86 output per 1M tokens
- Qwen2.5 14B Instruct — 131K context; $0.144 input / $0.431 output per 1M tokens
- Qwen2.5 32B Instruct by Alibaba — 131K context; $0.287 input / $0.861 output per 1M tokens
- Qwen2.5 72B by Alibaba — 131K context; $0.11 input / $0.38 output per 1M tokens
- Qwen2.5 7B Instruct by Alibaba — 131K context; $0.07 input / $0.07 output per 1M tokens
- Qwen2.5 Coder 32B by Alibaba — 33K context; $0.06 input / $0.2 output per 1M tokens
- Qwen2.5 Coder 7B fast — 32K context; $0.03 input / $0.09 output per 1M tokens
- Qwen2.5 VL 72B by Alibaba — 128K context; $0.25 input / $0.747 output per 1M tokens
- Qwen2.5 VL 72B TEE — 66K context; $0.7 input / $0.7 output per 1M tokens
- Qwen2.5-Coder 7B Instruct by Alibaba — 131K context; $0.144 input / $0.287 output per 1M tokens
- Qwen2.5-Coder-0.5B by Alibaba — 33K context; $0.1 input / $0.1 output per 1M tokens
- Qwen2.5-Math 72B Instruct — 4K context; $0.574 input / $1.72 output per 1M tokens
- Qwen2.5-Math 7B Instruct — 4K context; $0.144 input / $0.287 output per 1M tokens
- Qwen2.5-Max-2025-01-25 — 128K context; not published input / not published output per 1M tokens
- Qwen2.5-Omni 7B — 33K context; $0.087 input / $0.345 output per 1M tokens
- Qwen2.5-VL 7B Instruct — 131K context; $0.287 input / $0.717 output per 1M tokens
- Qwen3 1.7B Base by Alibaba — 33K context; $0.1 input / $0.1 output per 1M tokens
- Qwen3 14B by Alibaba — 131K context; $0.05 input / $0.22 output per 1M tokens
- Qwen3 235B A22B FP8 — 41K context; $0.2 input / $0.8 output per 1M tokens
- Qwen3 235B A22B Instruct 2507 by Alibaba — 262K context; $0.087 input / $0.35 output per 1M tokens
- Qwen3 235B A22B Instruct 2507 FP8 by Alibaba — 262K context; $0.2 input / $0.6 output per 1M tokens
- Qwen3 235B A22B Thinking — 262K context; $0.3 input / $2.9 output per 1M tokens
- Qwen3 235B A22B Thinking 2507 by Alibaba — 131K context; $0.3 input / $0.5 output per 1M tokens
- Qwen3 235B A22B Thinking 2507 TEE by Alibaba — 262K context; $0.299 input / $1.2 output per 1M tokens
- Qwen3 235B-A22B by Alibaba — 131K context; $0.13 input / $0.6 output per 1M tokens
- Qwen3 30B A3B by Alibaba — 41K context; $0.08 input / $0.29 output per 1M tokens
- Qwen3 30B A3B 2507 by Alibaba — 262K context; free input / free output per 1M tokens
- Qwen3 30B A3b fp8 by Alibaba — 33K context; $0.051 input / $0.335 output per 1M tokens
- Qwen3 30B A3B Instruct 2507 by Alibaba — 262K context; $0.048 input / $0.193 output per 1M tokens
- Qwen3 30B A3B Thinking 2507 by Alibaba — 262K context; $0.36 input / $1.3 output per 1M tokens
- Qwen3 32B by Alibaba — 131K context; $0.09 input / $0.25 output per 1M tokens
- Qwen3 32B (dense) — 16K context; $0.15 input / $0.6 output per 1M tokens
- Qwen3 32B TEE by Alibaba — 41K context; $0.104 input / $0.416 output per 1M tokens
- Qwen3 4B by Alibaba — 128K context; $0.03 input / $0.03 output per 1M tokens
- Qwen3 4B Base by Alibaba — 33K context; $0.15 input / $0.15 output per 1M tokens
- Qwen3 8B by Alibaba — 131K context; $0.035 input / $0.138 output per 1M tokens
- Qwen3 Coder by Alibaba — 262K context; $0.3 input / $1.2 output per 1M tokens
- Qwen3 Coder 30B by Alibaba — 262K context; free input / free output per 1M tokens
- Qwen3 Coder 480B A35B by Alibaba — 262K context; $0.2 input / $0.8 output per 1M tokens
- Qwen3 Coder 480B A35B Instruct by Alibaba — 262K context; $0.38 input / $1.55 output per 1M tokens
- Qwen3 Coder 480B A35B Instruct Turbo by Alibaba — 262K context; $0.22 input / $0.95 output per 1M tokens
- Qwen3 Coder Flash by Alibaba — 1M context; $0.144 input / $0.574 output per 1M tokens
- Qwen3 Coder Next by Alibaba — 262K context; $0.108 input / $0.675 output per 1M tokens
- Qwen3 Coder Next FP8 by Alibaba — 262K context; $0.5 input / $1.2 output per 1M tokens
- Qwen3 Coder Plus by Alibaba — 1M context; $0.574 input / $2.29 output per 1M tokens
- Qwen3 Embedding 0.6B by Alibaba — 41K context; $0.01 input / free output per 1M tokens
- Qwen3 Embedding 8B by Alibaba — 33K context; $0.01 input / free output per 1M tokens
- Qwen3 Max by Alibaba — 262K context; $0.36 input / $1.43 output per 1M tokens
- Qwen3 Max 2026-01-23 — 256K context; $1.2 input / $6 output per 1M tokens
- Qwen3 Max Preview by Alibaba — 256K context; $1.2 input / $6 output per 1M tokens
- Qwen3 Omni 30B A3B Instruct by Alibaba — 66K context; $0.25 input / $0.97 output per 1M tokens
- Qwen3 Omni 30B A3B Thinking by Alibaba — 66K context; $0.25 input / $0.97 output per 1M tokens
- Qwen3 Reranker 0.6B by Alibaba — 41K context; $0.01 input / $0.01 output per 1M tokens
- Qwen3 Reranker 4B by Alibaba — 33K context; $0.058 input / free output per 1M tokens
- Qwen3 VL 235B by Alibaba — 218K context; $0.21 input / $1.9 output per 1M tokens
- Qwen3 VL 235B A22B by Alibaba — 131K context; $0.2 input / $0.88 output per 1M tokens
- Qwen3 VL 235B A22B Instruct by Alibaba — 262K context; $0.2 input / $0.88 output per 1M tokens
- Qwen3 VL 235B A22B Instruct Original — 33K context; $0.5 input / $1.2 output per 1M tokens
- Qwen3 VL 235B A22B Thinking by Alibaba — 131K context; $0.287 input / $2.87 output per 1M tokens
- Qwen3 VL 30B A3B Instruct by Alibaba — 262K context; $0.15 input / $0.6 output per 1M tokens
- Qwen3 VL 30B A3B Thinking by Alibaba — 262K context; $0.2 input / $2.4 output per 1M tokens
- Qwen3 VL 32B Instruct by Alibaba — 131K context; $0.104 input / $0.416 output per 1M tokens
- Qwen3 VL 8B Instruct by Alibaba — 262K context; $0.117 input / $0.455 output per 1M tokens
- Qwen3 VL 8B Thinking by Alibaba — 131K context; $0.18 input / $2.1 output per 1M tokens
- Qwen3 VL Flash — 262K context; $0.05 input / $0.4 output per 1M tokens
- Qwen3 VL Flash (Alibaba Cloud) by Alibaba — 262K context; $0.05 input / $0.4 output per 1M tokens
- Qwen3 VL Instruct by Alibaba — 131K context; $0.4 input / $1.6 output per 1M tokens
- Qwen3 VL Thinking by Alibaba — 131K context; $0.4 input / $4 output per 1M tokens
- qwen3-235b-2507-cs — not published context; not published input / not published output per 1M tokens
- qwen3-235b-a22b-no-thinking — not published context; not published input / not published output per 1M tokens
- qwen3-32b-cs — not published context; not published input / not published output per 1M tokens
- Qwen3-ASR Flash — 53K context; $0.032 input / $0.032 output per 1M tokens
- Qwen3-Coder 30B-A3B Instruct by Alibaba — 262K context; $0.06 input / $0.25 output per 1M tokens
- Qwen3-Coder-Next-FP8-no-thinking — 26K context; free input / free output per 1M tokens
- Qwen3-LiveTranslate Flash Realtime — 53K context; $10 input / $10 output per 1M tokens
- qwen3-max-2025-09-23 — 258K context; $0.86 input / $3.43 output per 1M tokens
- qwen3-max-2025-09-26 — not published context; not published input / not published output per 1M tokens
- Qwen3-Next 80B-A3B (Thinking) by Alibaba — 131K context; $0.15 input / $0.65 output per 1M tokens
- Qwen3-Next 80B-A3B Instruct by Alibaba — 131K context; $0.144 input / $0.574 output per 1M tokens
- Qwen3-Omni Flash — 66K context; $0.058 input / $0.23 output per 1M tokens
- Qwen3-Omni Flash Realtime — 66K context; $0.23 input / $0.918 output per 1M tokens
- Qwen3-VL 30B-A3B by Alibaba — 262K context; $0.108 input / $0.431 output per 1M tokens
- Qwen3-VL Embedding 8B by Alibaba — 32K context; $0.09 input / $0.09 output per 1M tokens
- Qwen3-VL Plus by Alibaba — 262K context; $0.143 input / $1.43 output per 1M tokens
- Qwen3.5 0.8B by Alibaba — 262K context; $0.06 input / $0.12 output per 1M tokens