LLM List: AI models compared
Browse all LLMs in one comprehensive list: 1,781 large language models and AI models across 214 API providers. Compare token prices, context windows, capabilities, benchmark scores, open-weight availability, and local hardware requirements. Updated every six hours.
Explore the LLM benchmark leaderboard · How prices, benchmarks and model identities are calculated
- DeepSeek-V4-Flash-EL — 1M context; $0.14 input / $0.28 output per 1M tokens
- deepseek-v4-flash-high-20260731 — not published context; not published input / not published output per 1M tokens
- deepseek-v4-flash-low-20260731 — not published context; not published input / not published output per 1M tokens
- deepseek-v4-flash-max-20260731 — not published context; not published input / not published output per 1M tokens
- deepseek-v4-flash-vision-exp-high — not published context; not published input / not published output per 1M tokens
- deepseek-v4-flash-vision-exp-low — not published context; not published input / not published output per 1M tokens
- deepseek-v4-flash-vision-exp-max — not published context; not published input / not published output per 1M tokens
- DeepSeek-V4-Pro-EL — 1M context; $1.67 input / $3.33 output per 1M tokens
- deepseek-v4-pro-high — not published context; not published input / not published output per 1M tokens
- deepseek-v4-pro-low — not published context; not published input / not published output per 1M tokens
- deepseek-v4-pro-max — not published context; not published input / not published output per 1M tokens
- Devstral 2 by Mistral — 262K context; $0.4 input / $2 output per 1M tokens
- Devstral 2 123B by Mistral — 262K context; $0.4 input / $1.4 output per 1M tokens
- Devstral Medium by Mistral — 128K context; $0.4 input / $2 output per 1M tokens
- Devstral Small by Mistral — 128K context; $0.1 input / $0.3 output per 1M tokens
- Devstral Small 2 by Mistral — 256K context; $0.1 input / $0.3 output per 1M tokens
- Devstral Small 2505 by Mistral — 128K context; $0.06 input / $0.06 output per 1M tokens
- Devstral-2-123B-Instruct-2512-int4-AutoRound — 128K context; free input / free output per 1M tokens
- devstral-latest — 256K context; $0.44 input / $2.2 output per 1M tokens
- devstral-latest@eu — 256K context; $0.44 input / $2.2 output per 1M tokens
- devstral-medium-2507 — not published context; not published input / not published output per 1M tokens
- DiffusionGemma 26B-A4B IT by Google — 262K context; $0.5 input / $0.5 output per 1M tokens
- Dola Seed 2.0 Code (preview) by ByteDance — 131K context; $0.5 input / $3 output per 1M tokens
- dola-seed-2.0-preview-text — not published context; not published input / not published output per 1M tokens
- Dots Studio: Dots3-Note Preview (free) — 512K context; free input / free output per 1M tokens
- Dots3-Note Preview (free) — 512K context; free input / free output per 1M tokens
- Doubao 1.5 Pro 256k — 256K context; $0.799 input / $1.45 output per 1M tokens
- Doubao 1.5 Pro 32k — 128K context; $0.134 input / $0.335 output per 1M tokens
- Doubao 1.5 Thinking Pro — 128K context; not published input / not published output per 1M tokens
- Doubao 1.5 Vision Pro — 128K context; not published input / not published output per 1M tokens
- Doubao 1.5 Vision Pro 32k — 32K context; $0.459 input / $1.38 output per 1M tokens
- Doubao Seed 1.6 — 256K context; $0.204 input / $0.51 output per 1M tokens
- Doubao Seed 1.6 Flash — 256K context; $0.037 input / $0.374 output per 1M tokens
- Doubao Seed 2.0 Code — 256K context; $0.9 input / $4.48 output per 1M tokens
- Doubao Seed 2.0 Code Preview — 256K context; $0.48 input / $2.41 output per 1M tokens
- Doubao Seed 2.0 Lite — 256K context; $0.09 input / $0.51 output per 1M tokens
- Doubao Seed 2.0 Lite 260428 — 256K context; $0.08 input / $0.51 output per 1M tokens
- Doubao Seed 2.0 Mini — 256K context; $0.03 input / $0.28 output per 1M tokens
- Doubao Seed 2.0 Mini 260428 — 256K context; $0.03 input / $0.28 output per 1M tokens
- Doubao Seed 2.0 Pro — 256K context; $0.45 input / $2.24 output per 1M tokens
- Doubao Seed 2.1 Pro by ByteDance — 256K context; $1 input / $5 output per 1M tokens
- Doubao Seed 2.1 Turbo by ByteDance — 256K context; $0.5 input / $2.5 output per 1M tokens
- Doubao Seed Character by ByteDance — 128K context; $0.118 input / $0.295 output per 1M tokens
- Doubao-Seed 1.6 Thinking — 256K context; not published input / not published output per 1M tokens
- doubao-seed-1-6-thinking-250715 — 256K context; $0.121 input / $1.21 output per 1M tokens
- doubao-seed-1-6-vision-250815 — 256K context; $0.114 input / $1.14 output per 1M tokens
- doubao-seed-1-8-251215 — 224K context; $0.114 input / $0.286 output per 1M tokens
- Doubao-Seed-1.8 — 256K context; $0.11 input / $0.28 output per 1M tokens
- Doubao-Seed-Code by ByteDance Seed — 256K context; $0.17 input / $1.12 output per 1M tokens
- doubao-seed-code-preview-251028 — 256K context; $0.17 input / $1.14 output per 1M tokens
- dracarys-llama-3.1-70b-instruct — 128K context; free input / free output per 1M tokens
- E5 Large v2 — 512 context; $0.02 input / free output per 1M tokens
- E5 Mistral 7B — 4K context; $0.02 input / $0.02 output per 1M tokens
- E5 Multi-Lingual Large Embeddings 0.6B — 512 context; $0.114 input / $0.114 output per 1M tokens
- Echo — 262K context; $10 input / $50 output per 1M tokens
- Embed v1 0.6b by Perplexity — 32K context; not published input / not published output per 1M tokens
- Embed v1 4b by Perplexity — 32K context; not published input / not published output per 1M tokens
- Embed v3 English — 512 context; $0.1 input / free output per 1M tokens
- Embed v3 Multilingual — 512 context; $0.1 input / free output per 1M tokens
- Embed v4 — 128K context; $0.12 input / free output per 1M tokens
- Embed v4.0 by Cohere — 128K context; not published input / not published output per 1M tokens
- End-to-End Encrypted — 1M context; not published input / not published output per 1M tokens
- ERNIE 4.5 21B A3B by Baidu — 12K context; $0.07 input / $0.28 output per 1M tokens
- Ernie 4.5 21B A3B Thinking by Baidu — 131K context; $0.07 input / $0.28 output per 1M tokens
- ERNIE 4.5 300B A47B by Baidu — 131K context; $0.28 input / $1.1 output per 1M tokens
- ERNIE 4.5 VL 28B A3B by Baidu — 3K context; $0.14 input / $0.56 output per 1M tokens
- ERNIE 4.5 VL 424B A47B by Baidu — 123K context; $0.42 input / $1.25 output per 1M tokens
- ERNIE 5.0 by Baidu — 128K context; $0.84 input / $3.37 output per 1M tokens
- Ernie 5.0 Thinking Preview by Baidu — 128K context; $1 input / $3.5 output per 1M tokens
- ERNIE 5.1 — 119K context; $0.75 input / $3 output per 1M tokens
- ERNIE 5.1 Thinking — 119K context; $0.75 input / $3 output per 1M tokens
- ERNIE X1.1 — 64K context; $0.15 input / $0.6 output per 1M tokens
- ERNIE-4.5-VL-28B-A3B-Thinking by Baidu — 131K context; $0.39 input / $0.39 output per 1M tokens
- esm2-650m by Meta — 128K context; free input / free output per 1M tokens
- esmfold by Meta — 128K context; free input / free output per 1M tokens
- EVA Llama 3.33 70B — 33K context; $2.01 input / $2.01 output per 1M tokens
- EVA-LLaMA-3.33-70B-v0.1 — 33K context; $2.01 input / $2.01 output per 1M tokens
- EVA-Qwen2.5-32B-v0.2 — 16K context; $0.799 input / $0.799 output per 1M tokens
- EVA-Qwen2.5-72B-v0.2 — 16K context; $0.799 input / $0.799 output per 1M tokens
- Evayale 70b — 16K context; $0.493 input / $0.493 output per 1M tokens
- Exa (Answer) — 4K context; $2.5 input / $2.5 output per 1M tokens
- Fabled — 262K context; $0.1 input / $0.45 output per 1M tokens
- Fast — 1M context; not published input / not published output per 1M tokens
- Faster Whisper Large v3 — 448 context; free input / free output per 1M tokens
- Free Models Router — 2K context; free input / free output per 1M tokens
- Fugu — 1M context; not published input / not published output per 1M tokens
- Fugu Ultra by Sakana AI — 1M context; $5 input / $30 output per 1M tokens
- Fugu Ultra v1.0 — 1M context; $7.5 input / $45 output per 1M tokens
- Fugu Ultra v1.1 by Sakana AI — 1M context; $5 input / $30 output per 1M tokens
- Fusion — 1M context; not published input / not published output per 1M tokens
- Garnet — 262K context; $0.1 input / $0.45 output per 1M tokens
- Gembrain — 262K context; $0.1 input / $0.45 output per 1M tokens
- Gemini 2.0 Flash by Google — 1.05M context; $0.1 input / $0.42 output per 1M tokens
- Gemini 2.0 Flash Lite by Google — 2M context; $0.052 input / $0.21 output per 1M tokens
- Gemini 2.0 Pro 0205 — 2.1M context; $1.99 input / $7.96 output per 1M tokens
- Gemini 2.0 Pro 1206 — 2.1M context; $1.26 input / $5 output per 1M tokens
- Gemini 2.0 Pro Reasoner — 128K context; $1.29 input / $5 output per 1M tokens
- Gemini 2.5 Computer Use Preview (10-2025) by Google — 131K context; $1.25 input / $10 output per 1M tokens
- Gemini 2.5 Flash by Google — 1.05M context; $0.09 input / $0.71 output per 1M tokens
- Gemini 2.5 Flash 0520 — 1.05M context; $0.15 input / $0.6 output per 1M tokens