llm-list.com

LLM List: AI models compared

Browse all LLMs in one comprehensive list: 1,781 large language models and AI models across 214 API providers. Compare token prices, context windows, capabilities, benchmark scores, open-weight availability, and local hardware requirements. Updated every six hours.

Explore the LLM benchmark leaderboard · How prices, benchmarks and model identities are calculated

  1. Qwen3.5 122B A10B Thinking — 131K context; $0.437 input / $3.5 output per 1M tokens
  2. Qwen3.5 122B-A10B by Alibaba — 262K context; $0.115 input / $0.917 output per 1M tokens
  3. Qwen3.5 122B-A10B FP8 by Alibaba — 2K context; $0.5 input / $3.97 output per 1M tokens
  4. Qwen3.5 27B by Alibaba — 262K context; $0.086 input / $0.688 output per 1M tokens
  5. Qwen3.5 27B BlueStar v3 Derestricted — 262K context; $0.306 input / $0.306 output per 1M tokens
  6. Qwen3.5 27B Queen Derestricted — 262K context; $0.306 input / $0.306 output per 1M tokens
  7. Qwen3.5 27B TEE — 262K context; $0.3 input / $2.4 output per 1M tokens
  8. Qwen3.5 27B Thinking — 26K context; $0.27 input / $2.16 output per 1M tokens
  9. Qwen3.5 2B by Alibaba — 262K context; $0.02 input / $0.1 output per 1M tokens
  10. Qwen3.5 35B A3B Thinking — 26K context; $0.225 input / $1.8 output per 1M tokens
  11. Qwen3.5 35B-A3B by Alibaba — 262K context; $0.057 input / $0.459 output per 1M tokens
  12. Qwen3.5 397B A17B (Alibaba Cloud) by Alibaba — 262K context; $0.6 input / $3.6 output per 1M tokens
  13. Qwen3.5 397B A17B TEE by Alibaba — 262K context; $0.45 input / $3 output per 1M tokens
  14. Qwen3.5 397B A17B Thinking by Alibaba — 258K context; $0.6 input / $3.6 output per 1M tokens
  15. Qwen3.5 397B-A17B by Alibaba — 262K context; $0.172 input / $1.03 output per 1M tokens
  16. Qwen3.5 397B-A17B FP8 by Alibaba — 2K context; $0.99 input / $4.46 output per 1M tokens
  17. Qwen3.5 4B by Alibaba — 262K context; $0.04 input / $0.07 output per 1M tokens
  18. Qwen3.5 9B by Alibaba — 262K context; $0.04 input / $0.15 output per 1M tokens
  19. Qwen3.5 Flash by Alibaba — 1M context; $0.029 input / $0.287 output per 1M tokens
  20. Qwen3.5 Flash Thinking — 992K context; $0.1 input / $0.4 output per 1M tokens
  21. Qwen3.5 Omni Flash by Alibaba — 49K context; free input / free output per 1M tokens
  22. Qwen3.5 Omni Plus by Alibaba — 984K context; free input / free output per 1M tokens
  23. Qwen3.5 Plus by Alibaba — 1M context; $0.115 input / $0.688 output per 1M tokens
  24. Qwen3.5 Plus 2026-02-15 by Alibaba — 1M context; $0.26 input / $1.56 output per 1M tokens
  25. Qwen3.5 Plus 2026-04-20 by Alibaba — 1M context; $0.3 input / $1.8 output per 1M tokens
  26. Qwen3.5 Plus Thinking by Alibaba — 984K context; $0.4 input / $2.4 output per 1M tokens
  27. Qwen3.5-122B — 262K context; $0.9 input / $3.6 output per 1M tokens
  28. qwen3.5-122b-a10b-code — not published context; not published input / not published output per 1M tokens
  29. qwen3.5-27b-code — not published context; not published input / not published output per 1M tokens
  30. qwen3.5-35b-a3b-code — not published context; not published input / not published output per 1M tokens
  31. qwen3.5-max-preview — not published context; not published input / not published output per 1M tokens
  32. Qwen3.6 27B by Alibaba — 262K context; $0.6 input / $0.6 output per 1M tokens
  33. Qwen3.6 27B FP8 — 262K context; free input / free output per 1M tokens
  34. Qwen3.6 27B TEE by Alibaba — 262K context; $0.3 input / $2 output per 1M tokens
  35. Qwen3.6 27B Thinking by Alibaba — 26K context; $0.203 input / $2.24 output per 1M tokens
  36. Qwen3.6 35B — 131K context; $0.29 input / $1.15 output per 1M tokens
  37. Qwen3.6 35B A3B FP8 by Alibaba — 262K context; $0.17 input / $1.1 output per 1M tokens
  38. Qwen3.6 35B A3B TEE — 262K context; $0.2 input / $1.27 output per 1M tokens
  39. Qwen3.6 35B A3B Thinking by Alibaba — 262K context; $0.112 input / $0.8 output per 1M tokens
  40. Qwen3.6 35B A3B Uncensored TEE — 131K context; $0.3 input / $1.5 output per 1M tokens
  41. Qwen3.6 35B Fast — 131K context; $0.29 input / $1.15 output per 1M tokens
  42. Qwen3.6 35B-A3B by Alibaba — 262K context; $0.07 input / $0.42 output per 1M tokens
  43. Qwen3.6 Flash by Alibaba — 1M context; $0.165 input / $0.99 output per 1M tokens
  44. Qwen3.6 Max Preview by Alibaba — 262K context; $1.03 input / $6.16 output per 1M tokens
  45. Qwen3.6 Plus by Alibaba — 1M context; $0.276 input / $1.65 output per 1M tokens
  46. Qwen3.6 Plus Free — 262K context; free input / free output per 1M tokens
  47. Qwen3.6-35B-A3B-fp8-no-thinking — 262K context; free input / free output per 1M tokens
  48. qwen3.6-plus-preview — not published context; not published input / not published output per 1M tokens
  49. Qwen3.7 Flash by Alibaba — 1M context; $0.028 input / $0.113 output per 1M tokens
  50. Qwen3.7 Flash Thinking — 984K context; $0.03 input / $0.13 output per 1M tokens
  51. Qwen3.7 Max by Alibaba — 1M context; $0.825 input / $2.48 output per 1M tokens
  52. Qwen3.7 Max Thinking — 1M context; $2.5 input / $7.5 output per 1M tokens
  53. Qwen3.7 Plus by Alibaba — 1M context; $0.282 input / $1.13 output per 1M tokens
  54. Qwen3.7 Plus Thinking — 984K context; $0.4 input / $1.6 output per 1M tokens
  55. qwen3.7-max-20260517 — not published context; not published input / not published output per 1M tokens
  56. qwen3.7-max-preview — not published context; not published input / not published output per 1M tokens
  57. qwen3.7-plus-preview — not published context; not published input / not published output per 1M tokens
  58. Qwen3.8 2.4T A95B by Alibaba — 262K context; $2 input / $6 output per 1M tokens
  59. Qwen3.8 27B by Alibaba — 262K context; $0.1 input / $0.4 output per 1M tokens
  60. Qwen3.8 27B TEE by Alibaba — 262K context; $0.32 input / $2.5 output per 1M tokens
  61. Qwen3.8 27B Thinking — 262K context; $0.15 input / $0.7 output per 1M tokens
  62. Qwen3.8 Flash by Alibaba — 1M context; $0.12 input / $0.38 output per 1M tokens
  63. Qwen3.8 Flash Next by Alibaba — 262K context; $0.12 input / $0.4 output per 1M tokens
  64. Qwen3.8 Max by Alibaba — 1M context; $1.6 input / $4.8 output per 1M tokens
  65. Qwen3.8 Max 0902 by Alibaba — 1M context; $2 input / $6 output per 1M tokens
  66. Qwen3.8 Max Preview by Alibaba — 1M context; $1.81 input / $5.45 output per 1M tokens
  67. Qwen3.8 Max Thinking — 991K context; $2 input / $6 output per 1M tokens
  68. Qwen3Guard-Gen-0.6B — 33K context; free input / free output per 1M tokens
  69. Qwen3Guard-Gen-8B — 33K context; free input / free output per 1M tokens
  70. Qwen: QvQ Max — 128K context; $1.2 input / $4.8 output per 1M tokens
  71. Qwen: Qwen Plus 0728 by Alibaba — 1M context; $0.26 input / $0.78 output per 1M tokens
  72. Qwen: Qwen2.5 VL 72B Instruct by Alibaba — 128K context; $0.8 input / $1 output per 1M tokens
  73. Qwen: Qwen3 235B A22B Instruct 2507 by Alibaba — 262K context; $0.149 input / $0.598 output per 1M tokens
  74. Qwen: Qwen3 30B A3B Thinking 2507 by Alibaba — 82K context; $0.2 input / $2.4 output per 1M tokens
  75. Qwen: Qwen3 Max Thinking by Alibaba — 262K context; $0.78 input / $3.9 output per 1M tokens
  76. Qwen: Qwen3 VL 8B Thinking by Alibaba — 131K context; $0.18 input / $2.1 output per 1M tokens
  77. Qwen: Qwen3.5 Plus 2026-02-15 by Alibaba — 1M context; $0.26 input / $1.56 output per 1M tokens
  78. Qwen: Qwen3.5 Plus 2026-04-20 by Alibaba — 1M context; $0.3 input / $1.8 output per 1M tokens
  79. Qwen: Qwen3.5-Flash by Alibaba — 1M context; $0.065 input / $0.26 output per 1M tokens
  80. Qwerky 72B — 32K context; $0.5 input / $0.5 output per 1M tokens
  81. QwQ 32B by Alibaba — 131K context; $0.4 input / $0.4 output per 1M tokens
  82. QwQ Plus by Alibaba — 131K context; $0.23 input / $0.574 output per 1M tokens
  83. R1 0528 by DeepSeek — 164K context; $0.5 input / $2.15 output per 1M tokens
  84. R1 Distill Llama 70B by DeepSeek — 8K context; $0.8 input / $0.8 output per 1M tokens
  85. Reka Edge by Reka — 16K context; $0.1 input / $0.1 output per 1M tokens
  86. Reka Flash 3 by Reka — 66K context; $0.1 input / $0.2 output per 1M tokens
  87. Relace Apply 3 — 256K context; $0.85 input / $1.25 output per 1M tokens
  88. Relace Search — 256K context; $1 input / $3 output per 1M tokens
  89. Relace: Relace Apply 3 — 256K context; $0.85 input / $1.25 output per 1M tokens
  90. Relace: Relace Search — 256K context; $1 input / $3 output per 1M tokens
  91. ReMM SLERP 13B — 6K context; $0.35 input / $0.65 output per 1M tokens
  92. Rerank 4 Fast by Cohere — 32K context; not published input / not published output per 1M tokens
  93. Rerank 4 Pro by Cohere — 32K context; not published input / not published output per 1M tokens
  94. rerank-qa-mistral-4b — 128K context; free input / free output per 1M tokens
  95. Ring 2.6 1T Free — 262K context; free input / free output per 1M tokens
  96. Ring-1T by InclusionAI — 128K context; $0.56 input / $2.24 output per 1M tokens
  97. ring-2.5-1t — not published context; not published input / not published output per 1M tokens
  98. Ring-2.6-1T by InclusionAI — 262K context; $0.3 input / $2.5 output per 1M tokens
  99. ring-flash-2.0 by InclusionAI — not published context; not published input / not published output per 1M tokens
  100. riva-translate-4b-instruct-v1_1 — 128K context; free input / free output per 1M tokens