llm-list.com

LLM List: AI models compared

Browse all LLMs in one comprehensive list: 1,781 large language models and AI models across 214 API providers. Compare token prices, context windows, capabilities, benchmark scores, open-weight availability, and local hardware requirements. Updated every six hours.

Explore the LLM benchmark leaderboard · How prices, benchmarks and model identities are calculated

  1. Gemma 4 31B IT TEE — 262K context; $0.15 input / $0.46 output per 1M tokens
  2. Gemma 4 31B MeroMero v2 — 262K context; $0.1 input / $0.45 output per 1M tokens
  3. Gemma 4 31B MeroMero v2 Thinking — 262K context; $0.1 input / $0.45 output per 1M tokens
  4. Gemma 4 31B Queen — 262K context; $0.306 input / $0.306 output per 1M tokens
  5. Gemma 4 31B Thinking by Google — 262K context; $0.1 input / $0.35 output per 1M tokens
  6. Gemma 4 31B Thinking TEE — 262K context; $0.4 input / $1 output per 1M tokens
  7. gemma 4 31B turbo TEE by Google — 131K context; $0.12 input / $0.37 output per 1M tokens
  8. Gemma 4 E2B Instruct by Google — 131K context; $0.02 input / $0.1 output per 1M tokens
  9. Gemma 4 E2B IT by Google — 33K context; $0.1 input / $0.1 output per 1M tokens
  10. Gemma 4 E4B Instruct by Google — 33K context; $0.04 input / $0.2 output per 1M tokens
  11. Gemma 4 E4B IT by Google — 131K context; $0.02 input / $0.1 output per 1M tokens
  12. Gemma 4 Uncensored — 256K context; $0.163 input / $0.5 output per 1M tokens
  13. Gemma Sea Lion V4 27B It by AI Singapore — 128K context; $0.351 input / $0.555 output per 1M tokens
  14. Gemsicle — 262K context; $0.1 input / $0.45 output per 1M tokens
  15. GLiGuard LLM Guardrails 300M by Fastino — 8K context; $0.15 input / $0.15 output per 1M tokens
  16. gliner-pii — 128K context; free input / free output per 1M tokens
  17. GLiNER2 Base by Fastino — 8K context; $0.15 input / $0.15 output per 1M tokens
  18. GLiNER2 Large by Fastino — 8K context; $0.15 input / $0.15 output per 1M tokens
  19. GLiNER2 Multi by Fastino — 8K context; $0.15 input / $0.15 output per 1M tokens
  20. GLiNER2 Multi Large by Fastino — 8K context; $0.15 input / $0.15 output per 1M tokens
  21. GLiNER2 Privacy Filter PII (Multi) by Fastino — 8K context; $0.15 input / $0.15 output per 1M tokens
  22. GLM 4 32B 0414 by THUDM — 128K context; $0.2 input / $0.2 output per 1M tokens
  23. GLM 4 9B 0414 by THUDM — 32K context; $0.2 input / $0.2 output per 1M tokens
  24. GLM 4 Air 0111 — 128K context; $0.139 input / $0.139 output per 1M tokens
  25. GLM 4 Plus 0111 — 128K context; $10 input / $10 output per 1M tokens
  26. GLM 4.1V Thinking Flash — 64K context; $0.3 input / $0.3 output per 1M tokens
  27. GLM 4.1V Thinking FlashX — 64K context; $0.3 input / $0.3 output per 1M tokens
  28. GLM 4.5 (Thinking) by Z.AI — 128K context; $0.3 input / $1.3 output per 1M tokens
  29. GLM 4.5 Air (Thinking) by Z.AI — 128K context; $0.12 input / $0.8 output per 1M tokens
  30. GLM 4.5 FP8 by Z.AI — 131K context; $0.2 input / $0.8 output per 1M tokens
  31. GLM 4.5V Thinking by Z.AI — 66K context; $0.6 input / $1.8 output per 1M tokens
  32. GLM 4.6 Derestricted v5 — 131K context; $0.4 input / $1.5 output per 1M tokens
  33. GLM 4.6 Original by Z.AI — 256K context; $0.35 input / $1.4 output per 1M tokens
  34. GLM 4.6 Thinking by Z.AI — 2K context; $0.35 input / $1.4 output per 1M tokens
  35. GLM 4.6 Turbo by Z.AI — 205K context; $1 input / $3 output per 1M tokens
  36. GLM 4.6 Turbo (Thinking) by Z.AI — 205K context; $1 input / $3 output per 1M tokens
  37. GLM 4.6V Flash by Z.AI — 2K context; $0.3 input / $0.9 output per 1M tokens
  38. GLM 4.6V Original by Z.AI — 128K context; $0.6 input / $0.9 output per 1M tokens
  39. GLM 4.7 Flash Heretic — 2K context; $0.07 input / $0.4 output per 1M tokens
  40. GLM 4.7 Flash Original by Z.AI — 2K context; $0.07 input / $0.4 output per 1M tokens
  41. GLM 4.7 Flash Original Thinking by Z.AI — 2K context; $0.07 input / $0.4 output per 1M tokens
  42. GLM 4.7 Flash Thinking by Z.AI — 2K context; $0.07 input / $0.4 output per 1M tokens
  43. GLM 4.7 Original by Z.AI — 2K context; $0.6 input / $2.2 output per 1M tokens
  44. GLM 4.7 Original Thinking by Z.AI — 2K context; $0.6 input / $2.2 output per 1M tokens
  45. GLM 4.7 Thinking by Z.AI — 2K context; $0.2 input / $0.8 output per 1M tokens
  46. GLM 5 Original by Z.AI — 2K context; $1 input / $3.2 output per 1M tokens
  47. GLM 5 Original Thinking by Z.AI — 2K context; $1 input / $3.2 output per 1M tokens
  48. GLM 5 Thinking by Z.AI — 2K context; $0.5 input / $2.55 output per 1M tokens
  49. GLM 5 Vision Turbo — 2K context; $0.704 input / $3.1 output per 1M tokens
  50. GLM 5.1 TEE by Z.AI — 203K context; $0.98 input / $3.08 output per 1M tokens
  51. GLM 5.1 Thinking by Z.AI — 2K context; $0.75 input / $2.6 output per 1M tokens
  52. GLM 5.1 Thinking TEE — 203K context; $1.5 input / $5.25 output per 1M tokens
  53. GLM 5.2 Fast by Z.AI — 1M context; $1.45 input / $4.5 output per 1M tokens
  54. GLM 5.2 Flex — 1.05M context; $0.943 input / $2.92 output per 1M tokens
  55. GLM 5.2 Nitro — 2K context; $0.8 input / $2.4 output per 1M tokens
  56. GLM 5.2 Short — 2K context; $1.45 input / $4.5 output per 1M tokens
  57. GLM 5.2 Short Fast — 2K context; $1.45 input / $4.5 output per 1M tokens
  58. GLM 5.2 Short Fast Flex — 2K context; $0.943 input / $2.92 output per 1M tokens
  59. GLM 5.2 Short Flex — 2K context; $0.943 input / $2.92 output per 1M tokens
  60. GLM 5.2 TEE by Z.AI — 1.05M context; $1.25 input / $3.95 output per 1M tokens
  61. GLM 5.2 Thinking by Z.AI — 1.05M context; $0.42 input / $1.32 output per 1M tokens
  62. GLM 5.2 Thinking TEE — 1.05M context; $1.4 input / $4.6 output per 1M tokens
  63. GLM 5.3 Fast by Z.AI — 1.05M context; $2.1 input / $6.6 output per 1M tokens
  64. GLM 5.3 Flash TEE — 1.05M context; $0.15 input / $0.5 output per 1M tokens
  65. GLM 5.3 Flash Uncensored by Z.AI — 1.05M context; $0.35 input / $1.4 output per 1M tokens
  66. GLM 5.3 FP4 — 1.05M context; $1.12 input / $4.4 output per 1M tokens
  67. GLM 5.3 TEE — 1.05M context; $1.4 input / $4.4 output per 1M tokens
  68. GLM 5.3 Thinking by Z.AI — 1.05M context; $1 input / $3.2 output per 1M tokens
  69. GLM 5V Turbo Thinking by Z.AI — 203K context; $1.2 input / $4 output per 1M tokens
  70. GLM Flash Latest by Z.AI — 1.31M context; $0.075 input / $0.25 output per 1M tokens
  71. GLM Latest by Z.AI — 1.31M context; $1 input / $3.2 output per 1M tokens
  72. GLM Z1 9B 0414 by THUDM — 32K context; $0.2 input / $0.2 output per 1M tokens
  73. GLM Z1 AirX — 32K context; $0.7 input / $0.7 output per 1M tokens
  74. GLM-4 32B (0414-128k) — 128K context; $0.1 input / $0.1 output per 1M tokens
  75. GLM-4 32B (0414-128k) (Z AI) by Z.AI — 128K context; $0.1 input / $0.1 output per 1M tokens
  76. GLM-4 Long — 1M context; $0.201 input / $0.201 output per 1M tokens
  77. GLM-4.5 by Z.AI — 131K context; $0.286 input / $1.14 output per 1M tokens
  78. GLM-4.5 AirX by Z.AI — 128K context; $0.572 input / $1.71 output per 1M tokens
  79. GLM-4.5 X by Z.AI — 128K context; $1.14 input / $2.29 output per 1M tokens
  80. GLM-4.5-Air by Z.AI — 131K context; $0.114 input / $0.286 output per 1M tokens
  81. GLM-4.5-Flash — 131K context; free input / free output per 1M tokens
  82. GLM-4.5V by Z.AI — 66K context; $0.29 input / $0.86 output per 1M tokens
  83. GLM-4.6 by Z.AI — 205K context; $0.286 input / $1.14 output per 1M tokens
  84. GLM-4.6V by Z.AI — 131K context; $0.14 input / $0.42 output per 1M tokens
  85. GLM-4.6V FlashX by Z.AI — 128K context; $0.02 input / $0.21 output per 1M tokens
  86. GLM-4.7 by Z.AI — 205K context; $0.2 input / $0.8 output per 1M tokens
  87. GLM-4.7 Free — 205K context; free input / free output per 1M tokens
  88. GLM-4.7-Flash by Z.AI — 2K context; $0.06 input / $0.4 output per 1M tokens
  89. GLM-4.7-FlashX by Z.AI — 2K context; $0.06 input / $0.4 output per 1M tokens
  90. glm-4.7-n — 205K context; not published input / not published output per 1M tokens
  91. GLM-5 by Z.AI — 203K context; $0.6 input / $1.92 output per 1M tokens
  92. GLM-5 Free — 205K context; free input / free output per 1M tokens
  93. GLM-5-Turbo by Z.AI — 2K context; $0.72 input / $3.2 output per 1M tokens
  94. GLM-5.1 by Z.AI — 2K context; $0.45 input / $2.15 output per 1M tokens
  95. GLM-5.1 FP8 by Z.AI — 203K context; $0.85 input / $3.3 output per 1M tokens
  96. GLM-5.2 by Z.AI — 1M context; $0.3 input / $1.05 output per 1M tokens
  97. GLM-5.2 Caveman — 1M context; $1.25 input / $5.02 output per 1M tokens
  98. GLM-5.2 Caveman Lite — 1M context; $1.25 input / $5.02 output per 1M tokens
  99. GLM-5.2 Caveman Ultra — 1M context; $1.25 input / $5.02 output per 1M tokens
  100. GLM-5.2 Highspeed — 1M context; free input / free output per 1M tokens