llm-list.com

LLM List: AI models compared

Browse all LLMs in one comprehensive list: 1,781 large language models and AI models across 214 API providers. Compare token prices, context windows, capabilities, benchmark scores, open-weight availability, and local hardware requirements. Updated every six hours.

Explore the LLM benchmark leaderboard · How prices, benchmarks and model identities are calculated

  1. Gemini 2.5 Flash 0520 Thinking — 1.05M context; $0.15 input / $3.5 output per 1M tokens
  2. Gemini 2.5 Flash Image by Google — 33K context; $0.3 input / $2.5 output per 1M tokens
  3. Gemini 2.5 Flash Lite Preview by Google — 1.05M context; $0.15 input / $0.6 output per 1M tokens
  4. Gemini 2.5 Flash Lite Preview (09/2025) – Thinking — 1.05M context; $0.1 input / $0.4 output per 1M tokens
  5. Gemini 2.5 Flash Preview by Google — 1.05M context; $0.15 input / $0.6 output per 1M tokens
  6. Gemini 2.5 Flash Preview (09/2025) — 1.05M context; $0.3 input / $2.5 output per 1M tokens
  7. Gemini 2.5 Flash Preview (09/2025) – Thinking — 1.05M context; $0.3 input / $2.5 output per 1M tokens
  8. Gemini 2.5 Flash Preview Thinking — 1.05M context; $0.15 input / $3.5 output per 1M tokens
  9. Gemini 2.5 Flash-Lite by Google — 1.05M context; $0.1 input / $0.1 output per 1M tokens
  10. Gemini 2.5 Pro by Google — 1.05M context; $0.625 input / $5 output per 1M tokens
  11. Gemini 2.5 Pro Experimental 0325 — 1.05M context; $2.5 input / $10 output per 1M tokens
  12. Gemini 2.5 Pro Preview 0325 — 1.05M context; $2.5 input / $10 output per 1M tokens
  13. Gemini 2.5 Pro Preview 05-06 by Google — 1.05M context; $1.25 input / $10 output per 1M tokens
  14. Gemini 2.5 Pro Preview 06-05 by Google — 1.05M context; $1.12 input / $9 output per 1M tokens
  15. Gemini 3 Flash by Google — 1.05M context; $0.4 input / $2.4 output per 1M tokens
  16. Gemini 3 Flash (Preview) (Google AI Studio) — 1.05M context; $0.5 input / $3 output per 1M tokens
  17. Gemini 3 Flash (Preview) (Google Vertex AI) — 1.05M context; $0.5 input / $3 output per 1M tokens
  18. Gemini 3 Flash Preview by Google — 1.05M context; $0.07 input / $0.43 output per 1M tokens
  19. Gemini 3 Flash Thinking by Google — 1.05M context; $0.5 input / $3 output per 1M tokens
  20. Gemini 3 Pro by Google — 1.05M context; $1.6 input / $9.6 output per 1M tokens
  21. Gemini 3 Pro Image by Google — 1.05M context; $2 input / $12 output per 1M tokens
  22. Gemini 3 Pro Preview by Google — 1.05M context; $0.57 input / $3.43 output per 1M tokens
  23. Gemini 3.0 Flash Preview — 1M context; not published input / not published output per 1M tokens
  24. Gemini 3.0 Pro Image Preview — 33K context; not published input / not published output per 1M tokens
  25. Gemini 3.0 Pro Preview — 1M context; not published input / not published output per 1M tokens
  26. Gemini 3.1 Flash Image by Google — 33K context; $0.5 input / $3 output per 1M tokens
  27. Gemini 3.1 Flash Image Preview (Nano Banana 2) by Google — 131K context; $0.5 input / $3 output per 1M tokens
  28. Gemini 3.1 Flash Lite by Google — 1.05M context; $0.125 input / $0.75 output per 1M tokens
  29. Gemini 3.1 Flash Lite Image (Nano Banana 2 Lite) by Google — 66K context; $0.25 input / $1.5 output per 1M tokens
  30. Gemini 3.1 Flash Lite Preview by Google — 1.05M context; $0.125 input / $0.75 output per 1M tokens
  31. Gemini 3.1 Flash Live Preview — 131K context; $0.75 input / $4.5 output per 1M tokens
  32. Gemini 3.1 Pro (Preview High) by Google — 1.05M context; $2 input / $12 output per 1M tokens
  33. Gemini 3.1 Pro (Preview Low) by Google — 1.05M context; $2 input / $12 output per 1M tokens
  34. Gemini 3.1 Pro (Preview) (Google AI Studio) — 1.05M context; $2 input / $12 output per 1M tokens
  35. Gemini 3.1 Pro (Preview) (Google Vertex AI) — 1.05M context; $2 input / $12 output per 1M tokens
  36. Gemini 3.1 Pro (Preview) (Quartz) — 1.05M context; $2 input / $12 output per 1M tokens
  37. Gemini 3.1 Pro Preview by Google — 1.05M context; $1 input / $6 output per 1M tokens
  38. Gemini 3.1 Pro Preview Custom Tools by Google — 1.05M context; $2 input / $12 output per 1M tokens
  39. Gemini 3.5 Flash by Google — 1.05M context; $0.186 input / $1.11 output per 1M tokens
  40. Gemini 3.5 Flash Lite by Google — 1.05M context; $0.15 input / $1.25 output per 1M tokens
  41. Gemini 3.5 Flash Thinking by Google — 1.05M context; $1.5 input / $9 output per 1M tokens
  42. Gemini 3.5 Live Translate Preview — 16K context; $3.5 input / $21 output per 1M tokens
  43. Gemini 3.5 Transcribe by Google — not published context; $2 input / $12 output per 1M tokens
  44. Gemini 3.5 Transcribe Live by Google — not published context; not published input / not published output per 1M tokens
  45. Gemini 3.6 Flash by Google — 1.05M context; $0.375 input / $1.88 output per 1M tokens
  46. Gemini 3.7 Flash by Google — 1.05M context; $0.375 input / $1.88 output per 1M tokens
  47. Gemini 3.8 Flash by Google — 1.05M context; $0.375 input / $1.88 output per 1M tokens
  48. Gemini Embedding 001 by Google — 2K context; $0.15 input / free output per 1M tokens
  49. Gemini Embedding 2 by Google — 8K context; $0.2 input / free output per 1M tokens
  50. Gemini Flash Latest by Google — 1.05M context; $0.5 input / $3 output per 1M tokens
  51. Gemini Flash-Lite Latest by Google — 1.05M context; $0.25 input / $1.5 output per 1M tokens
  52. Gemini Omni Flash Preview by Google — 1M context; $1.5 input / $9 output per 1M tokens
  53. Gemini Pro Latest by Google — 1.05M context; $2 input / $12 output per 1M tokens
  54. Gemini Robotics-ER 1.6 Preview by Google — 131K context; $1 input / $5 output per 1M tokens
  55. gemini-2.0-flash-001 — not published context; not published input / not published output per 1M tokens
  56. gemini-2.5-flash-lite-preview-06-17 — 1.05M context; $0.09 input / $0.36 output per 1M tokens
  57. gemini-2.5-flash-lite-preview-09-2025 — 1.05M context; $0.09 input / $0.36 output per 1M tokens
  58. gemini-2.5-flash-nothink — 1M context; $0.3 input / $2.5 output per 1M tokens
  59. gemini-2.5-flash-preview-05-20 — 1.05M context; $0.135 input / $3.15 output per 1M tokens
  60. gemini-2.5-pro-grounding — not published context; not published input / not published output per 1M tokens
  61. gemini-3-flash (thinking-minimal) — not published context; not published input / not published output per 1M tokens
  62. gemini-3-flash-grounding — not published context; not published input / not published output per 1M tokens
  63. gemini-3-pro-image-preview — 33K context; $2 input / $120 output per 1M tokens
  64. gemini-3.1-flash-image-preview — 131K context; $0.5 input / $60 output per 1M tokens
  65. Gemini-3.1-Pro by Google — 1.05M context; $2 input / $12 output per 1M tokens
  66. gemini-3.1-pro-grounding — not published context; not published input / not published output per 1M tokens
  67. gemini-3.5-flash-high — not published context; not published input / not published output per 1M tokens
  68. gemini-3.8-flash-high — not published context; not published input / not published output per 1M tokens
  69. gemini-3.8-flash-low — not published context; not published input / not published output per 1M tokens
  70. gemini-3.8-flash-medium — not published context; not published input / not published output per 1M tokens
  71. gemini-deep-research by Google — 1.05M context; $1.6 input / $9.6 output per 1M tokens
  72. Gemma 2 — 8K context; $0.01 input / $0.03 output per 1M tokens
  73. Gemma 2 27B by Google — 8K context; $0.65 input / $0.65 output per 1M tokens
  74. Gemma 2 2b It by Google — 128K context; free input / free output per 1M tokens
  75. Gemma 3 by Google — 125K context; $0.15 input / $0.3 output per 1M tokens
  76. Gemma 3 12B by Google — 131K context; $0.05 input / $0.1 output per 1M tokens
  77. Gemma 3 27B by Google — 131K context; $0.08 input / $0.16 output per 1M tokens
  78. Gemma 3 4B by Google — 131K context; $0.04 input / $0.08 output per 1M tokens
  79. Gemma 3 4B (Pretrained) by Google — 33K context; $0.15 input / $0.15 output per 1M tokens
  80. Gemma 3n E2b It by Google — 128K context; free input / free output per 1M tokens
  81. Gemma 3N E4B Instruct by Google — 33K context; $0.06 input / $0.12 output per 1M tokens
  82. Gemma 3n E4b It by Google — 128K context; free input / free output per 1M tokens
  83. Gemma 4 — 262K context; $0.18 input / $0.5 output per 1M tokens
  84. Gemma 4 12B Instruct by Google — 262K context; $0.05 input / $0.25 output per 1M tokens
  85. Gemma 4 12B IT by Google — 33K context; $0.25 input / $0.25 output per 1M tokens
  86. Gemma 4 26B A4B by Google — 262K context; $0.042 input / $0.22 output per 1M tokens
  87. Gemma 4 26B A4B IT — 262K context; $0.07 input / $0.34 output per 1M tokens
  88. Gemma 4 26B A4B MeroMero — 262K context; $0.12 input / $0.38 output per 1M tokens
  89. Gemma 4 26B A4B MeroMero Thinking — 262K context; $0.12 input / $0.38 output per 1M tokens
  90. Gemma 4 26B A4B Thinking by Google — 262K context; $0.13 input / $0.4 output per 1M tokens
  91. Gemma 4 26B A4B Uncensored — 262K context; $0.12 input / $0.38 output per 1M tokens
  92. Gemma 4 26B A4B Uncensored TEE — 66K context; $0.15 input / $0.7 output per 1M tokens
  93. Gemma 4 26B A4B Uncensored Thinking — 262K context; $0.12 input / $0.38 output per 1M tokens
  94. Gemma 4 31B by Google — 262K context; $0.102 input / $0.297 output per 1M tokens
  95. Gemma 4 31B Claude 4.6 Opus Reasoning Distilled — 262K context; $0.306 input / $0.306 output per 1M tokens
  96. Gemma 4 31B Cognitive Unshackled — 262K context; $0.306 input / $0.306 output per 1M tokens
  97. Gemma 4 31B DarkIdol — 262K context; $0.306 input / $0.306 output per 1M tokens
  98. Gemma 4 31B Garnet V2 — 262K context; $0.306 input / $0.306 output per 1M tokens
  99. Gemma 4 31B IT — 262K context; $0.102 input / $0.297 output per 1M tokens
  100. Gemma 4 31B IT FP8 — 262K context; free input / free output per 1M tokens