llm-list.com

LLM List: AI models compared

Browse all LLMs in one comprehensive list: 1,781 large language models and AI models across 214 API providers. Compare token prices, context windows, capabilities, benchmark scores, open-weight availability, and local hardware requirements. Updated every six hours.

Explore the LLM benchmark leaderboard · How prices, benchmarks and model identities are calculated

  1. Laguna XS.2 — 33K context; free input / free output per 1M tokens
  2. leanstral-1-5 — 262K context; free input / free output per 1M tokens
  3. leanstral-1-5@eu — 262K context; free input / free output per 1M tokens
  4. LFM2 24B A2B by Liquid AI — 33K context; $0.03 input / $0.12 output per 1M tokens
  5. LFM2.5 2.6B by Liquid AI — 128K context; $0.1 input / $0.2 output per 1M tokens
  6. Ling 2.6 Flash Free — 262K context; free input / free output per 1M tokens
  7. Ling 3.0 Flash by InclusionAI — 262K context; $0.021 input / $0.063 output per 1M tokens
  8. Ling 3.0 Flash Fin by InclusionAI — 262K context; $0.06 input / $0.18 output per 1M tokens
  9. Ling 3.0 Flash Fin (Free) — 262K context; free input / free output per 1M tokens
  10. Ling 3.0 Flash Thinking by InclusionAI — 262K context; $0.075 input / $0.22 output per 1M tokens
  11. Ling-1T by InclusionAI — 128K context; $0.56 input / $2.24 output per 1M tokens
  12. ling-2.5-1t — not published context; not published input / not published output per 1M tokens
  13. Ling-2.6-1T by InclusionAI — 262K context; $0.3 input / $2.5 output per 1M tokens
  14. Ling-2.6-flash by InclusionAI — 262K context; $0.1 input / $0.3 output per 1M tokens
  15. Ling-3.0-flash Free — 262K context; free input / free output per 1M tokens
  16. ling-3.0-tiny by InclusionAI — 262K context; free input / free output per 1M tokens
  17. Ling-3.0-tiny Free — 262K context; free input / free output per 1M tokens
  18. Ling-flash-2.0 by InclusionAI — 131K context; $0.14 input / $0.57 output per 1M tokens
  19. LiquidAI: LFM2.5-2.6B (free) by Liquid AI — 66K context; free input / free output per 1M tokens
  20. Llama 3 70B abliterated — 8K context; $0.7 input / $0.7 output per 1M tokens
  21. Llama 3 70B Instruct by Meta — 8K context; $0.51 input / $0.74 output per 1M tokens
  22. Llama 3 8B Instruct by Meta — 8K context; $0.04 input / $0.04 output per 1M tokens
  23. Llama 3 8B Lunaris by Sao10K — 8K context; $0.04 input / $0.05 output per 1M tokens
  24. Llama 3.05 Storybreaker Ministral 70b — 16K context; $0.493 input / $0.493 output per 1M tokens
  25. Llama 3.1 405B Instruct by Meta — 128K context; $1.95 input / $1.95 output per 1M tokens
  26. Llama 3.1 405B Instruct Turbo by Meta — 128K context; $3.5 input / $3.5 output per 1M tokens
  27. Llama 3.1 70B by Meta — 128K context; $0.4 input / $0.4 output per 1M tokens
  28. Llama 3.1 70B Celeste v0.1 — 33K context; $0.493 input / $0.493 output per 1M tokens
  29. Llama 3.1 70B Dracarys 2 — 33K context; $0.493 input / $0.493 output per 1M tokens
  30. Llama 3.1 70B Euryale by Sao10K — 2K context; $0.306 input / $0.357 output per 1M tokens
  31. Llama 3.1 70B Hanami by Sao10K — 33K context; $0.493 input / $0.493 output per 1M tokens
  32. Llama 3.1 70B Instruct — 128K context; $0.72 input / $0.72 output per 1M tokens
  33. Llama 3.1 8B by Meta — 131K context; $0.025 input / $0.025 output per 1M tokens
  34. Llama 3.1 8B (decentralized) — 128K context; $0.02 input / $0.03 output per 1M tokens
  35. Llama 3.1 8b (uncensored) by Aion Labs — 33K context; $0.8 input / $1.6 output per 1M tokens
  36. Llama 3.1 8B Instruct fp8 by Meta — 32K context; $0.152 input / $0.287 output per 1M tokens
  37. Llama 3.1 Euryale 70B v2.2 by Sao10K — 131K context; $0.85 input / $0.85 output per 1M tokens
  38. Llama 3.1 Nemotron 70B Instruct by NVIDIA — 131K context; $0.6 input / $0.6 output per 1M tokens
  39. Llama 3.1 Nemotron Nano 8B v1 — 131K context; free input / free output per 1M tokens
  40. Llama 3.1 Nemotron Nano VL 8B v1 — 33K context; free input / free output per 1M tokens
  41. Llama 3.1 Nemotron Ultra 253B — 128K context; free input / free output per 1M tokens
  42. Llama 3.2 11B Instruct by Meta — 128K context; $0.07 input / $0.33 output per 1M tokens
  43. Llama 3.2 11B Vision Instruct by Meta — 128K context; $0.055 input / $0.055 output per 1M tokens
  44. Llama 3.2 1B by Meta — 6K context; $0.01 input / $0.01 output per 1M tokens
  45. Llama 3.2 3B by Meta — 131K context; $0.02 input / $0.02 output per 1M tokens
  46. Llama 3.2 3B Instruct — 33K context; $0.03 input / $0.05 output per 1M tokens
  47. Llama 3.2 90B Vision Instruct by Meta — 128K context; $0.35 input / $0.4 output per 1M tokens
  48. Llama 3.3 70B by Meta — 128K context; $0.05 input / $0.23 output per 1M tokens
  49. Llama 3.3 70B Cu Mai — 33K context; $0.493 input / $0.493 output per 1M tokens
  50. Llama 3.3 70B Euryale by Sao10K — 131K context; $0.493 input / $0.493 output per 1M tokens
  51. Llama 3.3 70B Instruct — 131K context; $0.135 input / $0.4 output per 1M tokens
  52. Llama 3.3 70B Instruct abliterated by Huihui AI — 33K context; $0.7 input / $0.7 output per 1M tokens
  53. Llama 3.3 70B Instruct fp8 Fast by Meta — 24K context; $0.293 input / $2.25 output per 1M tokens
  54. Llama 3.3 70B Turbo by Meta — 131K context; $0.1 input / $0.32 output per 1M tokens
  55. Llama 3.3 70B Versatile — 131K context; $0.59 input / $0.79 output per 1M tokens
  56. Llama 3.3 70B Wayfarer — 33K context; $0.7 input / $0.7 output per 1M tokens
  57. Llama 3.3 Nemotron Super 49B v1 by NVIDIA — 131K context; free input / free output per 1M tokens
  58. Llama 3.3 Nemotron Super 49B v1.5 by NVIDIA — 131K context; $0.4 input / $0.4 output per 1M tokens
  59. Llama 4 Maverick by Meta — 1.05M context; $0.15 input / $0.6 output per 1M tokens
  60. Llama 4 Maverick 17B 128E Instruct by Meta — 524K context; $0.15 input / $0.6 output per 1M tokens
  61. Llama 4 Maverick 17B 128E Instruct FP8 by Meta — 1M context; $0.25 input / $1 output per 1M tokens
  62. Llama 4 Maverick 17B FP8 by Meta — 1.05M context; $0.2 input / $0.8 output per 1M tokens
  63. Llama 4 Maverick 17B Instruct by Meta — 1.05M context; $0.14 input / $0.59 output per 1M tokens
  64. Llama 4 Scout by Meta — 1.31M context; $0.1 input / $0.3 output per 1M tokens
  65. Llama 4 Scout 17B by Meta — 3.5M context; $0.1 input / $0.3 output per 1M tokens
  66. Llama 4 Scout 17B 16E Instruct by Meta — 128K context; $0.2 input / $0.78 output per 1M tokens
  67. Llama 4 Scout 17B Instruct — 3.5M context; $0.18 input / $0.59 output per 1M tokens
  68. Llama Guard 4 12B by Meta — 164K context; $0.18 input / $0.18 output per 1M tokens
  69. Llama Prompt Guard 2 22M by Meta — 512 context; $0.01 input / $0.01 output per 1M tokens
  70. Llama-3.1-8B-CS — 128K context; $0.1 input / $0.1 output per 1M tokens
  71. llama-3.1-nemotron-safety-guard-8b-v3 — 128K context; free input / free output per 1M tokens
  72. llama-3.1-nemotron-ultra-253b-v1 by NVIDIA — 128K context; $0.598 input / $1.79 output per 1M tokens
  73. llama-3.3-70b-cs — not published context; not published input / not published output per 1M tokens
  74. Llama-3.3-8B-Instruct — 128K context; free input / free output per 1M tokens
  75. llama-3_2-nemoretriever-300m-embed-v1 — 33K context; free input / free output per 1M tokens
  76. Llama-4-Scout-17B-16E-Instruct-FP8 by Meta — 128K context; free input / free output per 1M tokens
  77. Llama-Guard-3-8B by Meta — 131K context; $0.055 input / $0.055 output per 1M tokens
  78. llama-nemotron-embed-vl-1b-v2 — 33K context; free input / free output per 1M tokens
  79. llama-nemotron-rerank-vl-1b-v2 — 128K context; free input / free output per 1M tokens
  80. Llama-xLAM-2 70B fc-r — 128K context; $2.5 input / $2.5 output per 1M tokens
  81. LongCat 2.0 Thinking — 1.05M context; $0.75 input / $3 output per 1M tokens
  82. LongCat-2.0 by Meituan — 1.05M context; $0.3 input / $1.2 output per 1M tokens
  83. LongCat-2.0 Free — 1M context; free input / free output per 1M tokens
  84. Longcat-Flash-Chat by Meituan — 131K context; not published input / not published output per 1M tokens
  85. LucidNova RF1 100B — 12K context; $2 input / $5 output per 1M tokens
  86. LucidQuery Nexus Coder — 25K context; $2 input / $5 output per 1M tokens
  87. Lumimaid v0.2 — 16K context; $1 input / $1.5 output per 1M tokens
  88. Luminous Mirror — 262K context; $0.12 input / $0.38 output per 1M tokens
  89. Lynkr Auto (complexity routing) — 128K context; free input / free output per 1M tokens
  90. Lyria 3 Clip Preview by Google — 1.05M context; free input / free output per 1M tokens
  91. Lyria 3 Pro Preview by Google — 1.05M context; free input / free output per 1M tokens
  92. Mag Mell R1 — 16K context; $0.493 input / $0.493 output per 1M tokens
  93. Magibu 11B v8 — 8K context; $0.1 input / $0.5 output per 1M tokens
  94. Magistral Medium (latest) by Mistral — 128K context; $2 input / $5 output per 1M tokens
  95. Magistral Small by Mistral — 128K context; $0.5 input / $1.5 output per 1M tokens
  96. Magistral Small 1.2 by Mistral — 128K context; $0.5 input / $1.5 output per 1M tokens
  97. Magistral Small 2506 by Mistral — 128K context; $0.5 input / $1.5 output per 1M tokens
  98. Magnum V2 72B by Anthracite — 16K context; $2.01 input / $2.99 output per 1M tokens
  99. Magnum v4 72B by Anthracite — 33K context; $2.01 input / $2.99 output per 1M tokens
  100. MAI-Code-1-Flash — 256K context; $0.75 input / $4.5 output per 1M tokens