llm-list.com

LLM List: AI models compared

Browse all LLMs in one comprehensive list: 1,781 large language models and AI models across 214 API providers. Compare token prices, context windows, capabilities, benchmark scores, open-weight availability, and local hardware requirements. Updated every six hours.

Explore the LLM benchmark leaderboard · How prices, benchmarks and model identities are calculated

  1. DeepSeek-V4-Flash-EL — 1M context; $0.14 input / $0.28 output per 1M tokens
  2. deepseek-v4-flash-high-20260731 — not published context; not published input / not published output per 1M tokens
  3. deepseek-v4-flash-low-20260731 — not published context; not published input / not published output per 1M tokens
  4. deepseek-v4-flash-max-20260731 — not published context; not published input / not published output per 1M tokens
  5. deepseek-v4-flash-vision-exp-high — not published context; not published input / not published output per 1M tokens
  6. deepseek-v4-flash-vision-exp-low — not published context; not published input / not published output per 1M tokens
  7. deepseek-v4-flash-vision-exp-max — not published context; not published input / not published output per 1M tokens
  8. DeepSeek-V4-Pro-EL — 1M context; $1.67 input / $3.33 output per 1M tokens
  9. deepseek-v4-pro-high — not published context; not published input / not published output per 1M tokens
  10. deepseek-v4-pro-low — not published context; not published input / not published output per 1M tokens
  11. deepseek-v4-pro-max — not published context; not published input / not published output per 1M tokens
  12. Devstral 2 by Mistral — 262K context; $0.4 input / $2 output per 1M tokens
  13. Devstral 2 123B by Mistral — 262K context; $0.4 input / $1.4 output per 1M tokens
  14. Devstral Medium by Mistral — 128K context; $0.4 input / $2 output per 1M tokens
  15. Devstral Small by Mistral — 128K context; $0.1 input / $0.3 output per 1M tokens
  16. Devstral Small 2 by Mistral — 256K context; $0.1 input / $0.3 output per 1M tokens
  17. Devstral Small 2505 by Mistral — 128K context; $0.06 input / $0.06 output per 1M tokens
  18. Devstral-2-123B-Instruct-2512-int4-AutoRound — 128K context; free input / free output per 1M tokens
  19. devstral-latest — 256K context; $0.44 input / $2.2 output per 1M tokens
  20. devstral-latest@eu — 256K context; $0.44 input / $2.2 output per 1M tokens
  21. devstral-medium-2507 — not published context; not published input / not published output per 1M tokens
  22. DiffusionGemma 26B-A4B IT by Google — 262K context; $0.5 input / $0.5 output per 1M tokens
  23. Dola Seed 2.0 Code (preview) by ByteDance — 131K context; $0.5 input / $3 output per 1M tokens
  24. dola-seed-2.0-preview-text — not published context; not published input / not published output per 1M tokens
  25. Dots Studio: Dots3-Note Preview (free) — 512K context; free input / free output per 1M tokens
  26. Dots3-Note Preview (free) — 512K context; free input / free output per 1M tokens
  27. Doubao 1.5 Pro 256k — 256K context; $0.799 input / $1.45 output per 1M tokens
  28. Doubao 1.5 Pro 32k — 128K context; $0.134 input / $0.335 output per 1M tokens
  29. Doubao 1.5 Thinking Pro — 128K context; not published input / not published output per 1M tokens
  30. Doubao 1.5 Vision Pro — 128K context; not published input / not published output per 1M tokens
  31. Doubao 1.5 Vision Pro 32k — 32K context; $0.459 input / $1.38 output per 1M tokens
  32. Doubao Seed 1.6 — 256K context; $0.204 input / $0.51 output per 1M tokens
  33. Doubao Seed 1.6 Flash — 256K context; $0.037 input / $0.374 output per 1M tokens
  34. Doubao Seed 2.0 Code — 256K context; $0.9 input / $4.48 output per 1M tokens
  35. Doubao Seed 2.0 Code Preview — 256K context; $0.48 input / $2.41 output per 1M tokens
  36. Doubao Seed 2.0 Lite — 256K context; $0.09 input / $0.51 output per 1M tokens
  37. Doubao Seed 2.0 Lite 260428 — 256K context; $0.08 input / $0.51 output per 1M tokens
  38. Doubao Seed 2.0 Mini — 256K context; $0.03 input / $0.28 output per 1M tokens
  39. Doubao Seed 2.0 Mini 260428 — 256K context; $0.03 input / $0.28 output per 1M tokens
  40. Doubao Seed 2.0 Pro — 256K context; $0.45 input / $2.24 output per 1M tokens
  41. Doubao Seed 2.1 Pro by ByteDance — 256K context; $1 input / $5 output per 1M tokens
  42. Doubao Seed 2.1 Turbo by ByteDance — 256K context; $0.5 input / $2.5 output per 1M tokens
  43. Doubao Seed Character by ByteDance — 128K context; $0.118 input / $0.295 output per 1M tokens
  44. Doubao-Seed 1.6 Thinking — 256K context; not published input / not published output per 1M tokens
  45. doubao-seed-1-6-thinking-250715 — 256K context; $0.121 input / $1.21 output per 1M tokens
  46. doubao-seed-1-6-vision-250815 — 256K context; $0.114 input / $1.14 output per 1M tokens
  47. doubao-seed-1-8-251215 — 224K context; $0.114 input / $0.286 output per 1M tokens
  48. Doubao-Seed-1.8 — 256K context; $0.11 input / $0.28 output per 1M tokens
  49. Doubao-Seed-Code by ByteDance Seed — 256K context; $0.17 input / $1.12 output per 1M tokens
  50. doubao-seed-code-preview-251028 — 256K context; $0.17 input / $1.14 output per 1M tokens
  51. dracarys-llama-3.1-70b-instruct — 128K context; free input / free output per 1M tokens
  52. E5 Large v2 — 512 context; $0.02 input / free output per 1M tokens
  53. E5 Mistral 7B — 4K context; $0.02 input / $0.02 output per 1M tokens
  54. E5 Multi-Lingual Large Embeddings 0.6B — 512 context; $0.114 input / $0.114 output per 1M tokens
  55. Echo — 262K context; $10 input / $50 output per 1M tokens
  56. Embed v1 0.6b by Perplexity — 32K context; not published input / not published output per 1M tokens
  57. Embed v1 4b by Perplexity — 32K context; not published input / not published output per 1M tokens
  58. Embed v3 English — 512 context; $0.1 input / free output per 1M tokens
  59. Embed v3 Multilingual — 512 context; $0.1 input / free output per 1M tokens
  60. Embed v4 — 128K context; $0.12 input / free output per 1M tokens
  61. Embed v4.0 by Cohere — 128K context; not published input / not published output per 1M tokens
  62. End-to-End Encrypted — 1M context; not published input / not published output per 1M tokens
  63. ERNIE 4.5 21B A3B by Baidu — 12K context; $0.07 input / $0.28 output per 1M tokens
  64. Ernie 4.5 21B A3B Thinking by Baidu — 131K context; $0.07 input / $0.28 output per 1M tokens
  65. ERNIE 4.5 300B A47B by Baidu — 131K context; $0.28 input / $1.1 output per 1M tokens
  66. ERNIE 4.5 VL 28B A3B by Baidu — 3K context; $0.14 input / $0.56 output per 1M tokens
  67. ERNIE 4.5 VL 424B A47B by Baidu — 123K context; $0.42 input / $1.25 output per 1M tokens
  68. ERNIE 5.0 by Baidu — 128K context; $0.84 input / $3.37 output per 1M tokens
  69. Ernie 5.0 Thinking Preview by Baidu — 128K context; $1 input / $3.5 output per 1M tokens
  70. ERNIE 5.1 — 119K context; $0.75 input / $3 output per 1M tokens
  71. ERNIE 5.1 Thinking — 119K context; $0.75 input / $3 output per 1M tokens
  72. ERNIE X1.1 — 64K context; $0.15 input / $0.6 output per 1M tokens
  73. ERNIE-4.5-VL-28B-A3B-Thinking by Baidu — 131K context; $0.39 input / $0.39 output per 1M tokens
  74. esm2-650m by Meta — 128K context; free input / free output per 1M tokens
  75. esmfold by Meta — 128K context; free input / free output per 1M tokens
  76. EVA Llama 3.33 70B — 33K context; $2.01 input / $2.01 output per 1M tokens
  77. EVA-LLaMA-3.33-70B-v0.1 — 33K context; $2.01 input / $2.01 output per 1M tokens
  78. EVA-Qwen2.5-32B-v0.2 — 16K context; $0.799 input / $0.799 output per 1M tokens
  79. EVA-Qwen2.5-72B-v0.2 — 16K context; $0.799 input / $0.799 output per 1M tokens
  80. Evayale 70b — 16K context; $0.493 input / $0.493 output per 1M tokens
  81. Exa (Answer) — 4K context; $2.5 input / $2.5 output per 1M tokens
  82. Fabled — 262K context; $0.1 input / $0.45 output per 1M tokens
  83. Fast — 1M context; not published input / not published output per 1M tokens
  84. Faster Whisper Large v3 — 448 context; free input / free output per 1M tokens
  85. Free Models Router — 2K context; free input / free output per 1M tokens
  86. Fugu — 1M context; not published input / not published output per 1M tokens
  87. Fugu Ultra by Sakana AI — 1M context; $5 input / $30 output per 1M tokens
  88. Fugu Ultra v1.0 — 1M context; $7.5 input / $45 output per 1M tokens
  89. Fugu Ultra v1.1 by Sakana AI — 1M context; $5 input / $30 output per 1M tokens
  90. Fusion — 1M context; not published input / not published output per 1M tokens
  91. Garnet — 262K context; $0.1 input / $0.45 output per 1M tokens
  92. Gembrain — 262K context; $0.1 input / $0.45 output per 1M tokens
  93. Gemini 2.0 Flash by Google — 1.05M context; $0.1 input / $0.42 output per 1M tokens
  94. Gemini 2.0 Flash Lite by Google — 2M context; $0.052 input / $0.21 output per 1M tokens
  95. Gemini 2.0 Pro 0205 — 2.1M context; $1.99 input / $7.96 output per 1M tokens
  96. Gemini 2.0 Pro 1206 — 2.1M context; $1.26 input / $5 output per 1M tokens
  97. Gemini 2.0 Pro Reasoner — 128K context; $1.29 input / $5 output per 1M tokens
  98. Gemini 2.5 Computer Use Preview (10-2025) by Google — 131K context; $1.25 input / $10 output per 1M tokens
  99. Gemini 2.5 Flash by Google — 1.05M context; $0.09 input / $0.71 output per 1M tokens
  100. Gemini 2.5 Flash 0520 — 1.05M context; $0.15 input / $0.6 output per 1M tokens