LLM List: AI models compared
Browse all LLMs in one comprehensive list: 1,781 large language models and AI models across 214 API providers. Compare token prices, context windows, capabilities, benchmark scores, open-weight availability, and local hardware requirements. Updated every six hours.
Explore the LLM benchmark leaderboard · How prices, benchmarks and model identities are calculated
- Ox Alpha Free (Unlimited) — 1M context; free input / free output per 1M tokens
- PaddleOCR-VL — 16K context; $0.02 input / $0.02 output per 1M tokens
- PaddlePaddle/PaddleOCR-VL-1.5 — 16K context; free input / free output per 1M tokens
- paligemma by Google — 128K context; free input / free output per 1M tokens
- Palmyra X4 — 128K context; $2.5 input / $10 output per 1M tokens
- Palmyra X5 — 1.04M context; $0.6 input / $6 output per 1M tokens
- Pareto Code Router — 2M context; free input / free output per 1M tokens
- Pearl AI Gemma 4 31B Instruct — 32K context; $0.28 input / $0.86 output per 1M tokens
- Perceptron Mk1 — 33K context; $0.15 input / $1.5 output per 1M tokens
- Perceptron: Perceptron Mk1 — 33K context; $0.15 input / $1.5 output per 1M tokens
- Perplexity Academic Researcher — 128K context; $2 input / $8 output per 1M tokens
- Perplexity Deep Research — 128K context; $3.4 input / $13.6 output per 1M tokens
- Perplexity Pro — 2K context; $3 input / $15 output per 1M tokens
- Perplexity Reasoning Pro — 128K context; $2 input / $8 output per 1M tokens
- Perplexity Simple — 127K context; $1 input / $1 output per 1M tokens
- Perplexity Sonar Reasoning — 127K context; $1 input / $5 output per 1M tokens
- Phi 4 by Microsoft — 128K context; $0.07 input / $0.14 output per 1M tokens
- Phi 4 Multimodal by Microsoft — 128K context; $0.07 input / $0.11 output per 1M tokens
- Phi-4-mini by Microsoft — 128K context; $0.075 input / $0.3 output per 1M tokens
- Phi-4-mini-reasoning — 128K context; $0.075 input / $0.3 output per 1M tokens
- Phi-4-reasoning — 32K context; $0.125 input / $0.5 output per 1M tokens
- Phi-4-reasoning-plus — 32K context; $0.125 input / $0.5 output per 1M tokens
- Pioneer Auto — 1.05M context; not published input / not published output per 1M tokens
- Pixtral 12B by Mistral — 128K context; $0.15 input / $0.15 output per 1M tokens
- Pixtral 12B 2409 — 128K context; $0.2 input / $0.2 output per 1M tokens
- Pixtral Large (25.02) by Mistral — 128K context; $1.99 input / $5.98 output per 1M tokens
- Pixtral Large (latest) by Mistral — 128K context; $2 input / $6 output per 1M tokens
- Pokee-Isaac 28B — 10M context; $0.15 input / $1 output per 1M tokens
- Poolside: Laguna S 2.1 by Poolside — 262K context; free input / free output per 1M tokens
- Poolside: Laguna XS 2.1 by Poolside — 262K context; free input / free output per 1M tokens
- ppl-sonar-reasoning-pro-high — not published context; not published input / not published output per 1M tokens
- Pro/deepseek-ai/DeepSeek-R1 by DeepSeek — 164K context; $0.5 input / $2.18 output per 1M tokens
- Pro/deepseek-ai/DeepSeek-V3 by DeepSeek — 164K context; $0.25 input / $1 output per 1M tokens
- Pro/deepseek-ai/DeepSeek-V3.1-Terminus by DeepSeek — 164K context; $0.27 input / $1 output per 1M tokens
- Pro/deepseek-ai/DeepSeek-V3.2 by DeepSeek — 164K context; $0.27 input / $0.42 output per 1M tokens
- Pro/MiniMaxAI/MiniMax-M2.5 by MiniMax — 192K context; $0.3 input / $1.22 output per 1M tokens
- Pro/moonshotai/Kimi-K2.5 by Moonshot AI — 262K context; $0.45 input / $2.25 output per 1M tokens
- Pro/moonshotai/Kimi-K2.6 by Moonshot AI — 262K context; $0.95 input / $4 output per 1M tokens
- Pro/zai-org/GLM-5 by Z.AI — 205K context; $1 input / $3.2 output per 1M tokens
- Pro/zai-org/GLM-5.1 by Z.AI — 205K context; $1.4 input / $4.4 output per 1M tokens
- Prompt Guard 2 86M by Meta — 512 context; $0.04 input / $0.04 output per 1M tokens
- QVQ Max — 131K context; $1.15 input / $4.59 output per 1M tokens
- Qwen 2.5 32B Abliterated by Huihui AI — 33K context; $0.7 input / $0.7 output per 1M tokens
- Qwen 2.5 7B Instruct Turbo by Alibaba — 33K context; $0.3 input / $0.3 output per 1M tokens
- Qwen 2.5 7B Vision Instruct by Alibaba — 125K context; $0.2 input / $0.2 output per 1M tokens
- Qwen 2.5 Max by Alibaba — 32K context; $1.6 input / $6.39 output per 1M tokens
- Qwen 2.5 VL 32B Instruct by Alibaba — 32K context; $0.05 input / $0.22 output per 1M tokens
- Qwen 3 235b A22B 2507 by Alibaba — 262K context; $0.072 input / $0.464 output per 1M tokens
- Qwen 3 235B Thinking by Alibaba — 262K context; $0.11 input / $0.6 output per 1M tokens
- Qwen 3 Coder 480B by Alibaba — 262K context; $0.13 input / $0.5 output per 1M tokens
- Qwen 3 Coder 480B Turbo — 256K context; $0.35 input / $1.5 output per 1M tokens
- Qwen 3 Embedding 4B by Alibaba — 32K context; $0.01 input / free output per 1M tokens
- Qwen 3 Max Thinking by Alibaba — 256K context; $0.78 input / $3.9 output per 1M tokens
- Qwen 3 Next 80B by Alibaba — 262K context; $0.1 input / $0.8 output per 1M tokens
- Qwen 3.5 397B — 262K context; $0.75 input / $4.5 output per 1M tokens
- Qwen 3.5 9B (MLX 4-bit) — 33K context; free input / free output per 1M tokens
- Qwen 3.5 9B (Q4_K_M) — 33K context; free input / free output per 1M tokens
- Qwen 3.6 35B A3B Uncensored by Alibaba — 262K context; $0.15 input / $0.95 output per 1M tokens
- Qwen 3.6 35B A3B Uncensored Thinking by Alibaba — 262K context; $0.15 input / $0.95 output per 1M tokens
- Qwen 3.6 Plus Uncensored — 1M context; $0.625 input / $3.75 output per 1M tokens
- Qwen 3.8 2.4T — 262K context; $2.5 input / $7.5 output per 1M tokens
- Qwen 3.8 27B Fable by Alibaba — 262K context; $0.25 input / $1.5 output per 1M tokens
- Qwen 3.8 27B Obliterated by Alibaba — 262K context; $0.25 input / $1.5 output per 1M tokens
- Qwen 3.8 27B Obliterated Thinking by Alibaba — 262K context; $0.25 input / $1.5 output per 1M tokens
- Qwen 3.8 27B Uncensored by Alibaba — 262K context; $0.25 input / $1.5 output per 1M tokens
- Qwen 3.8 27B Uncensored Thinking by Alibaba — 262K context; $0.25 input / $1.5 output per 1M tokens
- Qwen Coder Plus by Alibaba — 131K context; $0.502 input / $1 output per 1M tokens
- Qwen Deep Research — 1M context; $7.74 input / $23.37 output per 1M tokens
- Qwen Doc Turbo — 131K context; $0.087 input / $0.144 output per 1M tokens
- Qwen Flash by Alibaba — 1M context; $0.022 input / $0.216 output per 1M tokens
- Qwen Long — 10M context; $0.072 input / $0.287 output per 1M tokens
- Qwen Long 10M — 10M context; $0.1 input / $0.408 output per 1M tokens
- Qwen Math Plus — 4K context; $0.574 input / $1.72 output per 1M tokens
- Qwen Math Turbo — 4K context; $0.287 input / $0.861 output per 1M tokens
- Qwen Max by Alibaba — 33K context; $0.345 input / $1.38 output per 1M tokens
- Qwen Plus by Alibaba — 1M context; $0.115 input / $0.287 output per 1M tokens
- Qwen Plus 0728 by Alibaba — 1M context; $0.26 input / $0.78 output per 1M tokens
- Qwen Plus Character — 33K context; $0.115 input / $0.287 output per 1M tokens
- Qwen Plus Latest — 1M context; $0.4 input / $1.2 output per 1M tokens
- Qwen Plus Latest (Alibaba Cloud) by Alibaba — 1M context; $0.4 input / $1.2 output per 1M tokens
- Qwen Turbo by Alibaba — 1M context; $0.044 input / $0.087 output per 1M tokens
- Qwen VL-MAX-2025-01-25 — 128K context; not published input / not published output per 1M tokens
- Qwen-Max-Latest — 131K context; $0.343 input / $1.37 output per 1M tokens
- Qwen-MT Plus by Alibaba — 16K context; $0.25 input / $0.75 output per 1M tokens
- Qwen-MT Turbo — 16K context; $0.101 input / $0.28 output per 1M tokens
- Qwen-Omni Turbo by Alibaba — 33K context; $0.058 input / $0.23 output per 1M tokens
- Qwen-Omni Turbo Realtime — 33K context; $0.23 input / $0.918 output per 1M tokens
- Qwen-VL Max by Alibaba — 131K context; $0.23 input / $0.574 output per 1M tokens
- Qwen-VL OCR — 34K context; $0.717 input / $0.717 output per 1M tokens
- Qwen-VL Plus by Alibaba — 131K context; $0.115 input / $0.287 output per 1M tokens
- qwen-vl-max-2025-08-13 — not published context; not published input / not published output per 1M tokens
- Qwen/Qwen2.5-72B-Instruct by Alibaba — 33K context; $0.59 input / $0.59 output per 1M tokens
- Qwen/Qwen2.5-7B-Instruct by Alibaba — 33K context; $0.05 input / $0.05 output per 1M tokens
- Qwen/Qwen3-14B by Alibaba — 131K context; $0.07 input / $0.28 output per 1M tokens
- Qwen/Qwen3-235B-A22B-Thinking-2507 by Alibaba — 262K context; $0.13 input / $0.6 output per 1M tokens
- Qwen/Qwen3-30B-A3B-Instruct-2507 by Alibaba — 262K context; $0.09 input / $0.3 output per 1M tokens
- Qwen/Qwen3-32B by Alibaba — 131K context; $0.14 input / $0.57 output per 1M tokens
- Qwen/Qwen3-8B by Alibaba — 131K context; $0.06 input / $0.06 output per 1M tokens
- Qwen/Qwen3-Coder-30B-A3B-Instruct by Alibaba — 262K context; $0.07 input / $0.28 output per 1M tokens
- Qwen/Qwen3-Coder-480B-A35B by Alibaba — 262K context; $0.25 input / $1 output per 1M tokens